Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Haitao Mi — most-cited papers & profile · AI for Code
← authors
·
overview
Haitao Mi
56
papers ·
160
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
R-Zero: Self-Evolving Reasoning LLM from Zero Data
2025 · 115 citations
CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models
2025 · 26 citations
DOTS: Learning To Reason Dynamically In Llms Via Optimal Reasoning Trajectories Search
2024 · 18 citations
The Trickle-down Impact of Reward (In-)consistency on RLHF
2023 · 1 citations
MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment
2025
FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention
2026
Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning
2026
Locas: Your Models are Principled Initializers of Locally-Supported Parametric Memories
2026
Free(): Learning to Forget in Malloc-Only Reasoning Models
2026
The Pensieve Paradigm: Stateful Language Models Mastering Their Own Context
2026
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
2026
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
2026
DeltaRubric: Generative Multimodal Reward Modeling via Joint Planning and Verification
2026
Save the Good Prefix: Precise Error Penalization via Process-Supervised RL to Enhance LLM Reasoning
2026
Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning
2026
Top co-authors
Junyao Yang
· 2
Kishan Panaganti
· 2
Ruhan Wang
· 2
Yucheng Shi
· 2
Zhongzhi Li
· 2
Zongxia Li
· 2
Dongruo Zhou
· 1
Fan Zhang
· 1
Leoweiliang
· 1
Leowei Liang
· 1
Shijue Huang
· 1
Sixiang Chen
· 1
Topics
Training Techniques
Efficiency
Reinforcement Learning
Evaluation
Fine-Tuning
In-Context Learning
Model Architecture
cs.LG
cs.CL
cs.AI