← all papers · overview

Learning From Students: Applying T-distributions To Explore Accurate And Efficient Formats For Llms

Abstract

The increasing size of large language models (LLMs) traditionally requires low-precision integer formats to meet strict latency and power demands. Yet recently, alternative formats such as Normal Float (NF4) have increased model accuracy at the cost of increased chip area. In this work, we first conduct a large-scale analysis of LLM weights and activations across 30 networks and conclude that most

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).