← all papers · overview

Align-pro: A Principled Approach To Prompt Optimization For LLM Alignment

Abstract

The alignment of large language models (LLMs) with human values is critical as these models become increasingly integrated into various societal and decision-making processes. Traditional methods, such as reinforcement learning from human feedback (RLHF), achieve alignment by fine-tuning model parameters, but these approaches are often computationally expensive and impractical when models are froz

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).