Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guojun Yin — most-cited papers & profile · AI for Code
← authors
·
overview
Guojun Yin
25
papers ·
4
citations ·
16
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Beyond Static Testbeds: An Interaction-Centric Agent Simulation Platform for Dynamic Recommender Systems
2025 · 4 citations
Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration
2026
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
2026
SRFT: A Single-Stage Method with Supervised and Reinforcement Fine-Tuning for Reasoning
2025
When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards
2026
CDRRM: Contrast-Driven Rubric Generation for Reliable and Interpretable Reward Modeling
2026
UniMem: Complementary Episodic-to-Parametric Memory for Boundary-Agnostic Task Streams
2026
ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning
2026
ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning
2026
ZipRL: Adaptive Multi-Turn Context Compression with Hindsight Response Replay
2026
Policy Improvement Reinforcement Learning
2026
AutoSearch: Adaptive Search Depth for Efficient Agentic RAG via Reinforcement Learning
2026
LocalSearchBench: Benchmarking Agentic Search in Real-World Local Life Services
2025
Training Multi-Image Vision Agents via End2End Reinforcement Learning
2025
From Experience to Strategy: Empowering LLM Agents with Trainable Graph Memory
2025
Top co-authors
Chenheng Zhang
· 1
Haifeng Zhang
· 1
Haoxuan Li
· 1
Jiajun Chai
· 1
Jun Wang
· 1
Siyu Xia
· 1
Wei Lin
· 1
Xiaohan Wang
· 1
Yanting Wu
· 1
Zhouchen Lin
· 1
Topics
cs.AI
cs.LG
Evaluation
Multi-Agent
Policy Gradient
cs.CL
Reinforcement Learning
Training Techniques
Memory
Code Agents