Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shalabh Bhatnagar — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Shalabh Bhatnagar
32
papers ·
51
citations ·
30
h-index
Manipal Academy of Higher Education · Indian Institute of Science Bangalore · Kasturba Medical College, Manipal
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Model-based Safe Deep Reinforcement Learning via a Constrained Proximal Policy Optimization Algorithm
2022 · 19 citations
Actor-Critic Algorithms for Constrained Multi-agent Reinforcement Learning
2019 · 5 citations
Attention Actor-Critic algorithm for Multi-Agent Constrained Co-operative Reinforcement Learning
2021 · 5 citations
Learning Active Spine Behaviors for Dynamic and Efficient Locomotion in Quadruped Robots
2019 · 4 citations
An Online Prediction Algorithm for Reinforcement Learning with Linear Function Approximation using Cross Entropy Method
2018 · 2 citations
Memory-based Deep Reinforcement Learning for Obstacle Avoidance in UAV with Limited Environment Knowledge
2018 · 2 citations
A Convergent Off-Policy Temporal Difference Algorithm
2019 · 2 citations
An Actor-Critic Algorithm with Function Approximation for Risk Sensitive Cost Markov Decision Processes
2025 · 1 citations
A policy gradient approach for Finite Horizon Constrained Markov Decision Processes
2022 · 1 citations
Hindsight Experience Replay with Kronecker Product Approximate Curvature
2020 · 1 citations
Actor-Critic or Critic-Actor? A Tale of Two Time Scales
2022 · 1 citations
A Framework for Provably Stable and Consistent Training of Deep Feedforward Networks
2023 · 1 citations
Enabling Off-Policy Imitation Learning with Deep Actor Critic Stabilization
2025
Convergent Reinforcement Learning Algorithms for Stochastic Shortest Path Problem
2025
Finite-Time Analysis of Three-Timescale Constrained Actor-Critic and Constrained Natural Actor-Critic Algorithms
2023
Top co-authors
Raghuram Bharadwaj Diddigi
· 5
Abhik Singla
· 4
Soumyajit Guin
· 4
Shishir Kolathaya
· 3
Ajin George Joseph
· 2
Ashitava Ghosal
· 2
Bharadwaj Amrutur
· 2
Dhaivat Dholakiya
· 2
Meet Gandhi
· 2
Naman Saxena
· 2
Prabuchandran K.J
· 2
Prashansa Panda
· 2
Topics
Policy Gradient
Value-Based
Safe RL
Model-Based RL
Offline RL
Exploration
Multi-Agent