← all papers · overview

Merging Beyond: Streaming LLM Updates Via Activation-guided Rotations

Abstract

The escalating scale of Large Language Models (LLMs) necessitates efficient adaptation techniques. Model merging has gained prominence for its efficiency and controllability. However, existing merging techniques typically serve as post-hoc refinements or focus on mitigating task interference, often failing to capture the dynamic optimization benefits of supervised fine-tuning (SFT). In this work,

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).