Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiaodong Liu — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Xiaodong Liu
36
papers ·
6007
citations ·
0
h-index
Edinburgh Napier University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
2025 · 5456 citations
DeepSeek-V3 Technical Report
2024 · 248 citations
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
2024 · 158 citations
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
2024 · 105 citations
Reeval: Automatic Hallucination Evaluation For Retrieval-augmented Large Language Models Via Transferable Adversarial Attacks
2023 · 33 citations
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
2025 · 5 citations
LSAQ: Layer-specific Adaptive Quantization For Large Language Model Deployment
2024 · 2 citations
DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
2026
Statistical Estimation Of Adversarial Risk In Large Language Models Under Best-of-n Sampling
2026
MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety
2026
Does a Global Perspective Help Prune Sparse MoEs Elegantly?
2026
Statistical Estimation of Adversarial Risk in Large Language Models under Best-of-N Sampling
2026
FlowRL: Matching Reward Distributions for LLM Reasoning
2025
EMSEdit: Efficient Multi-Step Meta-Learning-based Model Editing
2025
Stand on The Shoulders of Giants: Building JailExpert from Previous Attack Experience
2025
Topics
Efficiency
Training Techniques
cs.CL
Evaluation
Model Architecture
Prompting
In-Context Learning
Reinforcement Learning
RAG
cs.LG