← all papers · overview

Unlocking Full Efficiency Of Token Filtering In Large Language Model Training

Abstract

Token filtering has been proposed to enhance the utility of large language models (LLMs) by eliminating inconsequential tokens during training. While usingfewer tokens is expected to reduce computational workloads, existing methods have not yet achieved a real-world efficiency boost. This is primarily due to two factors: (1) existing work has inadequate sparsity for speedup, and (2) token filterin

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).