← all papers · overview

Supervised Fine-tuning As Inverse Reinforcement Learning

Hao Sun·2024

Abstract

The prevailing approach to aligning Large Language Models (LLMs) typically relies on human or AI feedback and assumes access to specific types of preference datasets. In our work, we question the efficacy of such datasets and explore various scenarios where alignment with expert demonstrations proves more realistic. We build a sequential decision-making framework to formulate the problem of aligni

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).