Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Lifan Yuan — most-cited papers & profile · AI for Code
← authors
·
overview
Lifan Yuan
23
papers ·
840
citations ·
3
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Executable Code Actions Elicit Better LLM Agents
2024 · 421 citations
Ultrafeedback: Boosting Language Models With Scaled AI Feedback
2023 · 277 citations
RLPR: Extrapolating RLVR to General Domains without Verifiers
2025 · 69 citations
NFT: Bridging Supervised Learning and Reinforcement Learning in Math Reasoning
2025 · 28 citations
A Unified Evaluation Of Textual Backdoor Learning: Frameworks And Benchmarks
2022 · 24 citations
Process Reinforcement through Implicit Rewards
2025 · 1 citations
The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning
2025 · 1 citations
The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning
2025 · 1 citations
Free Process Rewards without Process Labels
2024 · 1 citations
RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments
2025
Do We Need Adam? Surprisingly Strong and Sparse Reinforcement Learning with SGD in LLMs
2026
Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
2025
RLPR: Extrapolating RLVR to General Domains without Verifiers
2025
From f(x) and g(x) to f(g(x)): LLMs Learn New Skills in RL by Composing Old Ones
2025
Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark
2025
Topics
Training Techniques
RLHF & Alignment
Policy Gradient
Fine-Tuning
Model-Based RL
Reinforcement Learning
Code
Evaluation
Safety & Alignment
Efficiency