← all papers · overview

On The Importance Of A Multi-scale Calibration For Quantization

Abstract

Post-training quantization (PTQ) is a cornerstone for efficiently deploying large language models (LLMs), where a small calibration set critically affects quantization performance. However, conventional practices rely on random sequences of fixed length, overlooking the variable-length nature of LLM inputs. Input length directly influences the activation distribution and, consequently, the weight

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).