Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
Direct
loadingβ¦
π€
Ask AI
Awesome Direct β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
Direct
22 papers tagged Direct β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
22 papers Β· trending (default)
numbers = π₯ heat
DocAtlas: Multilingual Document Understanding Across 80+ Languages
(2026)
Ahmed Heakl et al.
1.89
RLHS: Mitigating Misalignment in RLHF with Hindsight Simulation
(2025)
Kaiqu Liang et al.
1.28
Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step
(2025)
Ziyu Guo et al.
1.28
YINYANG-ALIGN: Benchmarking Contradictory Objectives and Proposing Multi-Objective Optimization based DPO for Text-to-Image Alignment
(2025)
Amitava Das et al.
1.28
DPO-Shift: Shifting the Distribution of Direct Preference Optimization
(2025)
Xiliang Yang et al.
1.28
LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models
(2025)
Shangqing Tu et al.
1.28
OmniAlign-V: Towards Enhanced Alignment of MLLMs with Human Preference
(2025)
Xiangyu Zhao et al.
1.28
DiaTool-DPO: Multi-Turn Direct Preference Optimization for Tool-Augmented Large Language Models
(2025)
Sunghee Jung et al.
1.28
MM-IFEngine: Towards Multimodal Instruction Following
(2025)
Shengyuan Ding et al.
1.28
Self-alignment of Large Video Language Models with Refined Regularized Preference Optimization
(2025)
Pritam Sarkar et al.
1.28
WebThinker: Empowering Large Reasoning Models with Deep Research Capability
(2025)
Xiaoxi Li et al.
1.28
The Aloe Family Recipe for Open and Specialized Healthcare LLMs
(2025)
Dario Garcia-Gasulla et al.
1.28
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization
(2025)
Yunxin Li et al.
1.28
ARM: Adaptive Reasoning Model
(2025)
Siye Wu et al.
1.28
Afterburner: Reinforcement Learning Facilitates Self-Improving Code Efficiency Optimization
(2025)
Mingzhe Du et al.
1.28
Technical Report of TeleChat2, TeleChat2.5 and T1
(2025)
Zihan Wang et al.
1.28
Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning
(2025)
Yifei Chen et al.
1.28
RewardBench: Evaluating Reward Models for Language Modeling
(2024)
Nathan Lambert et al.
β
NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment
(2024)
Gerald Shen et al.
β
Aligning Diffusion Models with Noise-Conditioned Perception
(2024)
Alexander Gambashidze et al.
β
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System
(2024)
Weize Chen et al.
β
MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models
(2024)
Ziyu Liu et al.
β