Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Dian Yu — most-cited papers & profile · Large Language Models
← authors
·
overview
Dian Yu
41
papers ·
7953
citations ·
13
h-index
Oklahoma State University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
React: Synergizing Reasoning And Acting In Language Models
2022 · 7191 citations
Tree of Thoughts: Deliberate Problem Solving with Large Language Models
2023 · 104 citations
Iterative Nash Policy Optimization: Aligning Llms With General Preferences Via No-regret Learning
2024 · 40 citations
One Token to Fool LLM-as-a-Judge
2025 · 1 citations
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
2026
DeltaRubric: Generative Multimodal Reward Modeling via Joint Planning and Verification
2026
Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs
2025
Expanding RL with Verifiable Rewards Across Diverse Domains
2025
DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
2025
CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models
2025
Evolving Language Models without Labels: Majority Drives Selection, Novelty Promotes Variation
2025
VOGUE: Guiding Exploration with Visual Uncertainty Improves Multimodal Reasoning
2025
CLUE: Non-parametric Verification from Experience via Hidden-State Clustering
2025
Every Question Has Its Own Value: Reinforcement Learning with Explicit Human Values
2025
Improving LLM General Preference Alignment via Optimistic Online Mirror Descent
2025
Top co-authors
Haitao Mi
· 17
Dong Yu
· 16
Linfeng Song
· 14
Zhenwen Liang
· 7
Zhaopeng Tu
· 6
Baolin Peng
· 5
Jiahao Xu
· 4
Tian Liang
· 4
Ye Tian
· 4
Jeffrey Zhao
· 3
Kishan Panaganti
· 3
Qiuzhi Liu
· 3
Topics
Training Techniques
Reinforcement Learning
Evaluation
In-Context Learning
Efficiency
Fine-Tuning
Safety & Alignment
Prompting
Vision-Language
Model Architecture