Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jiajun Chai — most-cited papers & profile · AI for Code
← authors
·
overview
Jiajun Chai
25
papers ·
2
citations ·
6
h-index
Chinese Academy of Sciences · Shandong Institute of Automation · Beijing Academy of Artificial Intelligence · Institute of Automation · University of Chinese Academy of Sciences
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
NVIF: Neighboring Variational Information Flow for Large-Scale Cooperative Multi-Agent Scenarios
2022 · 1 citations
A Hierarchical Deep Reinforcement Learning Framework for 6-DOF UCAV Air-to-Air Combat
2022 · 1 citations
Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration
2026
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
2026
SRFT: A Single-Stage Method with Supervised and Reinforcement Fine-Tuning for Reasoning
2025
When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards
2026
CDRRM: Contrast-Driven Rubric Generation for Reliable and Interpretable Reward Modeling
2026
UniMem: Complementary Episodic-to-Parametric Memory for Boundary-Agnostic Task Streams
2026
ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning
2026
ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning
2026
ZipRL: Adaptive Multi-Turn Context Compression with Hindsight Response Replay
2026
Policy Improvement Reinforcement Learning
2026
AutoSearch: Adaptive Search Depth for Efficient Agentic RAG via Reinforcement Learning
2026
LocalSearchBench: Benchmarking Agentic Search in Real-World Local Life Services
2025
Training Multi-Image Vision Agents via End2End Reinforcement Learning
2025
Top co-authors
Chenheng Zhang
· 1
Guojun Yin
· 1
Haifeng Zhang
· 1
Haoxuan Li
· 1
Jun Wang
· 1
Siyu Xia
· 1
Wei Lin
· 1
Xiaohan Wang
· 1
Yanting Wu
· 1
Zhouchen Lin
· 1
Topics
cs.AI
Policy Gradient
Multi-Agent
cs.LG
Value-Based
Evaluation
Fine-Tuning
RLHF & Alignment
cs.CL
Reinforcement Learning