← all papers · overview

Dmoerm: Recipes Of Mixture-of-experts For Effective Reward Modeling

Abstract

The performance of the reward model (RM) is a critical factor in improving the effectiveness of the large language model (LLM) during alignment fine-tuning. There remain two challenges in RM training: 1) training the same RM using various categories of data may cause its generalization performance to suffer from multi-task disturbance, and 2) the human annotation consistency rate is generally only

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).