← all papers · overview

Ternaryllm: Ternarized Large Language Model

Abstract

Large language models (LLMs) have achieved remarkable performance on Natural Language Processing (NLP) tasks, but they are hindered by high computational costs and memory requirements. Ternarization, an extreme form of quantization, offers a solution by reducing memory usage and enabling energy-efficient floating-point additions. However, applying ternarization to LLMs faces challenges stemming fr

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).