Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
A. Rupam Mahmood — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
A. Rupam Mahmood
18
papers ·
56
citations ·
7
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Benchmarking Reinforcement Learning Algorithms on Real-World Robots
2018 · 46 citations
Memory-efficient Reinforcement Learning with Value-based Knowledge Consolidation
2022 · 4 citations
Learning to Optimize for Reinforcement Learning
2023 · 2 citations
Autoregressive Policies for Continuous Control Deep Reinforcement Learning
2019 · 1 citations
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
2024 · 1 citations
On Generalized Bellman Equations and Temporal-Difference Learning
2017
An Alternate Policy Gradient Estimator for Softmax Policies
2021
A Temporal-Difference Approach to Policy Gradient Estimation
2022
Asynchronous Reinforcement Learning for Real-Time Control of Physical Robots
2022
Real-Time Reinforcement Learning for Vision-Based Robotics Utilizing Local and Remote Computers
2022
Dynamic Decision Frequency with Continuous Options
2022
Reducing the Cost of Cycle-Time Tuning for Real-World Policy Optimization
2023
MaDi: Learning to Mask Distractions for Generalization in Visual Deep Reinforcement Learning
2023
Target Networks and Over-parameterization Stabilize Off-policy Bootstrapping with Function Approximation
2024
Deep Policy Gradient Methods Without Batch Updates, Target Networks, or Replay Buffers
2024
Top co-authors
Gautham Vasan
· 6
James Bergstra
· 3
Martha White
· 3
Samuele Tosatto
· 3
Dmytro Korenkevych
· 2
Jun Luo
· 2
Martin Jägersand
· 2
Qingfeng Lan
· 2
Yangchen Pan
· 2
Yan Wang
· 2
Alireza Azimi
· 1
Amirmohammad Karimi
· 1
Topics
Policy Gradient
Value-Based
Exploration
Offline RL
Model-Based RL
Meta-RL
Safe RL