Awesome AI Agents
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Qian Liu — most-cited papers & profile · AI Agents
← authors
·
overview
Qian Liu
41
papers ·
635
citations ·
17
h-index
Chongqing University of Posts and Telecommunications · Southwest Jiaotong University · Beihang University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
2025 · 471 citations
Hyperentanglement purification for two-photon six-qubit quantum systems
2016 · 93 citations
StarCoder 2 and The Stack v2: The Next Generation
2024 · 59 citations
Mercury: A Code Efficiency Benchmark for Code Large Language Models
2024 · 7 citations
CodeArena: A Collective Evaluation Platform for LLM Code Generation
2025 · 2 citations
Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?
2024 · 2 citations
On Grounded Planning for Embodied Tasks with Language Models
2022 · 1 citations
The Optimal Token Baseline: Variance Reduction for Long-Horizon LLM-RL
2026
Trust Region Masking for Long-Horizon LLM Reinforcement Learning
2025
Trust Region Masking for Long-Horizon LLM Reinforcement Learning
2025
Proxy Compression for Language Modeling
2026
The Optimal Token Baseline: Variance Reduction for Long-Horizon LLM-RL
2026
Dynamic Vocabulary Pruning: Stable LLM-RL by Taming the Tail
2025
TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling
2025
History Rhymes: Accelerating LLM Reinforcement Learning with RhymeRL
2025
Top co-authors
Dong Huang
· 1
Luu Anh Tuan
· 1
Mingzhe Du
· 1
See-Kiong Ng
· 1
Xinyi He
· 1
Yue Liu
· 1
Yuhao Qing
· 1
Zejun Ma
· 1
Topics
Code
Training Techniques
Code Models
Efficiency
Fine-Tuning
Code Generation
Software Engineering
Evaluation
RLHF & Alignment
Model Architecture