Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Muhao Chen — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Muhao Chen
14
papers ·
788
citations ·
31
h-index
Microsoft (United States) · California Southern University · University of California, Davis
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Autodan: Generating Stealthy Jailbreak Prompts On Aligned Large Language Models
2023 · 698 citations
Deepedit: Knowledge Editing As Decoding With Constraints
2024 · 42 citations
Rlhfpoison: Reward Poisoning Attack For Reinforcement Learning With Human Feedback In Large Language Models
2023 · 35 citations
AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
2024 · 9 citations
mDPO: Conditional Preference Optimization for Multimodal Large Language Models
2024 · 2 citations
SoftSnap: Rapid Prototyping of Untethered Soft Robots Using Snap-Together Modules
2024 · 2 citations
FutureWeaver: Planning Test-Time Compute for Multi-Agent Systems with Modularized Collaboration
2025
Be My Eyes: Extending Large Language Models to New Modalities Through Multi-Agent Collaboration
2025
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
2025
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
2025
QLIP: A Dynamic Quadtree Vision Prior Enhances MLLM Performance Without Retraining
2025
Semantic-Clipping: Efficient Vision-Language Modeling with Semantic-Guidedd Visual Selection
2025
Bio-JOIE: Joint Representation Learning of Biological Knowledge Bases
2021
An Untethered Bioinspired Robotic Tensegrity Dolphin with Multi-Flexibility Design for Aquatic Locomotion
2024
Topics
Safety & Alignment
Vision-Language Models
Reinforcement Learning
Training Techniques
Multi-Agent
Evaluation
Benchmarks
Video-Language
Visual QA & Reasoning
Locomotion