← all papers · overview

The Real, The Better: Aligning Large Language Models With Online Human Behaviors

Abstract

Large language model alignment is widely used and studied to avoid LLM producing unhelpful and harmful responses. However, the lengthy training process and predefined preference bias hinder adaptation to online diverse human preferences. To this end, this paper proposes an alignment framework, called Reinforcement Learning with Human Behavior (RLHB), to align LLMs by directly leveraging real onlin

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).