Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
Human
loadingβ¦
π€
Ask AI
Awesome Human β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
Human
12 papers tagged Human β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
12 papers Β· trending (default)
numbers = π₯ heat
RLHS: Mitigating Misalignment in RLHF with Hindsight Simulation
(2025)
Kaiqu Liang et al.
1.28
PILAF: Optimal Human Preference Sampling for Reward Modeling
(2025)
Yunzhen Feng et al.
1.28
Iterative Value Function Optimization for Guided Decoding
(2025)
Zhenhua Liu et al.
1.28
Accelerating Nash Learning from Human Feedback via Mirror Prox
(2025)
Daniil Tiapkin et al.
1.28
RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards
(2025)
Zhilin Wang et al.
1.28
The Era of Real-World Human Interaction: RL from User Conversations
(2025)
Chuanyang Jin et al.
1.28
Efficient RLHF: Reducing the Memory Usage of PPO
(2023)
Michael Santacroce et al.
β
ICE-GRT: Instruction Context Enhancement by Generative Reinforcement based Transformers
(2024)
Chen Zheng et al.
β
MusicRL: Aligning Music Generation to Human Preferences
(2024)
Geoffrey Cideron et al.
β
NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment
(2024)
Gerald Shen et al.
β
Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models
(2024)
Guanting Dong et al.
β
Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
(2024)
Huanqian Wang et al.
β