← all papers · overview

Beyond Single-task: Robust Multi-task Length Generalization For Llms

Abstract

Length generalization, the ability to solve problems longer than those seen during training, remains a critical challenge for large language models (LLMs). Previous work modifies positional encodings (PEs) and data formats to improve length generalization on specific symbolic tasks such as addition and sorting. However, these approaches are fundamentally limited to special tasks, often degrading g

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).