Awesome Multimodal
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Qi Gu β most-cited papers & profile Β· Multimodal
β authors
Β·
overview
Qi Gu
23
papers Β·
0
citations Β·
0
h-index
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
When to Stop Reusing: Dynamic Gradient Gating for Sample-Efficient RLVR
2026
CAST: Game Solvers as Turn-Level Teachers for LLM Agents
2026
CAST: Game Solvers as Turn-Level Teachers for LLM Agents
2026
TopoCurate:Modeling Interaction Topology for Tool-Use Agent Training
2026
SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization
2026
CoBA-RL: Capability-Oriented Budget Allocation for Reinforcement Learning in LLMs
2026
Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning
2026
Self-Distilled Agentic Reinforcement Learning
2026
Look Before You Leap: Autonomous Exploration for LLM Agents
2026
Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning
2026
SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization
2026
AJ-Bench: Benchmarking Agent-as-a-Judge for Environment-Aware Evaluation
2026
DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training
2026
TopoCurate:Modeling Interaction Topology for Tool-Use Agent Training
2026
$V_0.5$: Generalist Value Model as a Prior for Sparse RL Rollouts
2026
Top co-authors
Fuli Feng
Β· 1
Han-Jia Ye
Β· 1
Lan-Zhe Guo
Β· 1
Wentao Shi
Β· 1
Xunliang Cai
Β· 1
Yi-Kai Zhang
Β· 1
Yuchun Miao
Β· 1
Yueqing Sun
Β· 1
Yu Wang
Β· 1
Ziang Ye
Β· 1
Topics
cs.AI
cs.LG
cs.CL
Reinforcement Learning
In-Context Learning
Multi-Agent
Efficiency
Training Techniques
Agentic
Code Agents