← all papers · overview

Not Everything Is All You Need: Toward Low-redundant Optimization For Large Language Model Alignment

Abstract

Large language models (LLMs) are still struggling in aligning with human preference in complex tasks and scenarios. They are prone to overfit into the unexpected patterns or superficial styles in the training data. We conduct an empirical study that only selects the top-10% most updated parameters in LLMs for alignment training, and see improvements in the convergence process and final performance

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).