Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Junwu Xiong — most-cited papers & profile · AI for Code
← authors
·
overview
Junwu Xiong
15
papers ·
38
citations ·
6
h-index
South China Agricultural University · Zhejiang Chinese Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Value Propagation for Decentralized Networked Deep Multi-agent Reinforcement Learning
2019 · 27 citations
Reinforcement Learning for Uplift Modeling
2018 · 6 citations
AI Agents Under Threat: A Survey of Key Security Challenges and Future Pathways
2024 · 4 citations
Variational Policy Propagation for Multi-agent Reinforcement Learning
2020 · 1 citations
Building a Scalable, Reproducible, Evaluatable, and Closed-Loop Simulation Environment Foundation for Embodied Intelligence
2026
Ring-lite: Scalable Reasoning via C3PO-Stabilized Reinforcement Learning for LLMs
2025
NoiseGate: Learning Per-Latent Timestep Schedules as Information Gating in World Action Models
2026
D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models
2026
JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy
2026
Thousand-GPU Large-Scale Training and Optimization Recipe for AI-Native Cloud Embodied Intelligence Infrastructure
2026
RL-VLA$^3$: A Flexible and Asynchronous Reinforcement Learning Framework for VLA Training
2026
RL-VLA$^3$: A Flexible and Asynchronous Reinforcement Learning Framework for VLA Training
2026
Ring-lite: Scalable Reasoning via C3PO-Stabilized Reinforcement Learning for LLMs
2025
Model Embedding Model-Based Reinforcement Learning
2020
Top co-authors
Chen Zhao
· 1
Chen Zhou
· 1
Haoran Sun
· 1
Hedan Yang
· 1
Hui Zhang
· 1
Jing Long
· 1
Junyang Hua
· 1
Mingxi Luo
· 1
Qiming Yang
· 1
Shuai Di
· 1
Song Wang
· 1
Wanting Xu
· 1
Topics
Control
Model-Based RL
Policy Gradient
Manipulation
Sim-to-Real
Multi-Robot
cs.AI
Multi-Agent
Value-Based
RLHF & Alignment