← all papers · overview

Aligning Neural Machine Translation Models: Human Feedback In Training And Inference

Abstract

Reinforcement learning from human feedback (RLHF) is a recent technique to improve the quality of the text generated by a language model, making it closer to what humans would generate. A core ingredient in RLHF's success in aligning and improving large language models (LLMs) is its reward model, trained using human feedback on model outputs. In machine translation (MT), where metrics trained from

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).