Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Elias Stengel-Eskin — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Elias Stengel-Eskin
16
papers ·
75
citations ·
8
h-index
University of North Carolina at Chapel Hill · University of North Carolina Health Care · The University of Texas at Austin
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Magdi: Structured Distillation Of Multi-agent Interaction Graphs Improves Reasoning In Smaller Language Models
2024 · 30 citations
Soft Self-consistency Improves Language Model Agents
2024 · 24 citations
LACIE: Listener-aware Finetuning For Confidence Calibration In Large Language Models
2024 · 15 citations
Laser: Learning To Adaptively Select Reward Models With Multi-armed Bandits
2024 · 6 citations
PRInTS: Reward Modeling for Long-Horizon Information Seeking
2025
Gistify! Codebase-Level Understanding via Runtime Execution
2025
GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs
2025
Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact
2025
Posh: Using Scene Graphs To Guide Llms-as-a-judge For Detailed Image Descriptions
2025
Symbolic Mixture-of-Experts: Adaptive Skill-based Routing for Heterogeneous Reasoning
2025
The Sum Leaks More Than Its Parts: Compositional Privacy Risks and Mitigations in Multi-Agent Collaboration
2025
Generalized Correctness Models: Learning Calibrated and Model-Agnostic Correctness Predictors from Historical Patterns
2025
Rotbench: Evaluating Multimodal Large Language Models On Identifying Image Rotation
2025
Clamr: Contextualized Late-interaction For Multimodal Content Retrieval
2025
Teaching Models to Balance Resisting and Accepting Persuasion
2024
Topics
Evaluation
Training Techniques
Safety & Alignment
Benchmarks
Efficiency
Vision-Language Models
In-Context Learning
Model Architecture
Fine-Tuning
Reinforcement Learning