Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
DPO
loadingβ¦
π€
Ask AI
Awesome DPO β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
DPO
15 papers tagged DPO β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
15 papers Β· trending (default)
numbers = π₯ heat
Building a Precise Video Language with Human-AI Oversight
(2026)
Zhiqiu Lin et al.
1.94
ViPO: Visual Preference Optimization at Scale
(2026)
Ming Li et al.
1.83
DPO-Shift: Shifting the Distribution of Direct Preference Optimization
(2025)
Xiliang Yang et al.
1.28
IterPref: Focal Preference Learning for Code Generation via Iterative Debugging
(2025)
Jie Wu et al.
1.28
DRAGON: Distributional Rewards Optimize Diffusion Generative Models
(2025)
Yatong Bai et al.
1.28
The Aloe Family Recipe for Open and Specialized Healthcare LLMs
(2025)
Dario Garcia-Gasulla et al.
1.28
Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning
(2025)
Shuyao Xu et al.
1.28
CAMS: A CityGPT-Powered Agentic Framework for Urban Human Mobility Simulation
(2025)
Yuwei Du et al.
1.28
Ovis2.5 Technical Report
(2025)
Shiyin Lu et al.
1.28
DRIFT: Learning from Abundant User Dissatisfaction in Real-World Preference Learning
(2025)
Yifan Wang et al.
1.28
Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models
(2024)
Guanting Dong et al.
β
Qwen2-Audio Technical Report
(2024)
Yunfei Chu et al.
β
xGen-MM (BLIP-3): A Family of Open Large Multimodal Models
(2024)
Le Xue et al.
β
Building Math Agents with Multi-Turn Iterative Preference Learning
(2024)
Wei Xiong et al.
β
MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models
(2024)
Ziyu Liu et al.
β