Awesome Cybersecurity
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xipeng Qiu — most-cited papers & profile · Cybersecurity
← authors
·
overview
Xipeng Qiu
44
papers ·
43
citations ·
58
h-index
Fudan University · Open Society · Shanghai Artificial Intelligence Laboratory
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Secrets of RLHF in Large Language Models Part I: PPO
2023 · 19 citations
DiffusionBERT: Improving Generative Masked Language Models with Diffusion Models
2022 · 13 citations
Secrets of RLHF in Large Language Models Part II: Reward Modeling
2024 · 7 citations
Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
2024 · 4 citations
In-Context World Modeling for Robotic Control
2026
Advancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical Autonomy
2026
RLoop: An Self-Improving Framework for Reinforcement Learning with Iterative Policy Initialization
2025
Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models
2026
Two Bridges, One Pathway: From VLMs to Generalizable VLAs with Embodied Trajectory-Coupled Data
2026
Investigating effective LLM-based in-context tool use: what matters and how to improve
2026
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
2026
Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO
2026
World Action Models: The Next Frontier in Embodied AI
2026
BandPO: Bridging Trust Regions and Ratio Clipping via Probability-Aware Bounds for LLM Reinforcement Learning
2026
DFPO: Scaling Value Modeling via Distributional Flow towards Robust and Generalizable LLM Post-Training
2026
Topics
Manipulation
Control
Human-Robot Interaction
Model-Based RL
Training Techniques
Policy Gradient
Perception
Model Architecture
Navigation
Exploration