Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Bin Yu — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Bin Yu
16
papers ·
0
citations ·
37
h-index
University of Science and Technology Liaoning · Liaoning University · Liaoning Technical University · Liaoning University of Technology · Chan Zuckerberg Initiative (United States) · University of California, Berkeley
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution
2026
Does a Global Perspective Help Prune Sparse MoEs Elegantly?
2026
PhysBrain: Human Egocentric Data as a Bridge from Vision Language Models to Physical Intelligence
2025
LLMBoost: Make Large Language Models Stronger with Boosting
2025
The Future of Artificial Intelligence and the Mathematical and Physical Sciences (AI+MPS)
2025
TrajSelector: Harnessing Latent Representations for Efficient and Effective Best-of-N in Large Reasoning Model
2025
MR-Align: Meta-Reasoning Informed Factuality Alignment for Large Reasoning Models
2025
ProxySPEX: Inference-Efficient Interpretability via Sparse Feature Interactions in LLMs
2025
Mobile Robot Localisation and Navigation Using LEGO NXT and Ultrasonic Sensor
2018
A Survey on Dynamic Network Embedding
2020
SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
2023
Improving Prototypical Visual Explanations with Reward Reweighing, Reselection, and Retraining
2023
Tell Your Model Where to Attend: Post-hoc Attention Steering for LLMs
2023
KnowGraph: Knowledge-Enabled Anomaly Detection via Logical Reasoning on Graph Data
2024
Explaining black box text modules in natural language with language models
2023
Topics
RAG
Efficiency
Evaluation
Model Architecture
Perception
Control
Fine-Tuning
Reinforcement Learning
Training Techniques
Code