Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Akifumi Wachi — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Akifumi Wachi
13
papers ·
157
citations ·
7
h-index
Line Corporation (Japan)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Failure-Scenario Maker for Rule-Based Agent using Multi-agent Adversarial Reinforcement Learning and its Application to Autonomous Driving
2019 · 60 citations
Safe Reinforcement Learning in Constrained Markov Decision Processes
2020 · 54 citations
Neuro-Symbolic Reinforcement Learning with First-Order Logic
2021 · 21 citations
Reinforcement Learning with External Knowledge by using Logical Neural Networks
2021 · 8 citations
Safe Exploration in Markov Decision Processes with Time-Variant Safety using Spatio-Temporal Gaussian Process
2018 · 4 citations
Safe Exploration in Reinforcement Learning: A Generalized Formulation and Algorithms
2023 · 4 citations
Long-term Safe Reinforcement Learning with Binary Feedback
2024 · 3 citations
Safe Policy Optimization with Local Generalized Linear Function Approximations
2021 · 2 citations
A Survey of Constraint Formulations in Safe Reinforcement Learning
2024 · 1 citations
Offline Guarded Safe Reinforcement Learning for Medical Treatment Optimization Strategies
2025
A Provable Approach for End-to-End Safe Reinforcement Learning
2025
LOA: Logical Optimal Actions for Text-based Interaction Games
2021
Flipping-based Policy for Chance-Constrained Markov Decision Processes
2024
Top co-authors
Asim Munawar
· 4
Xun Shen
· 4
Alexander Gray
· 3
Daiki Kimura
· 3
Michiaki Tatsubori
· 3
Ryosuke Kohita
· 3
Subhajit Chaudhury
· 3
Yanan Sui
· 3
Don Joven Agravante
· 2
Kazumune Hashimoto
· 2
Masaki Ono
· 2
Wataru Hashimoto
· 2
Topics
Safe RL
Model-Based RL
Exploration
Offline RL
Value-Based
Game AI
Multi-Agent
Meta-RL
Policy Gradient