Awesome Reinforcement Learning
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Reading Packs
News
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Reading Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Vasilis Syrgkanis — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Vasilis Syrgkanis
30
papers ·
0
citations ·
25
h-index
Stanford University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Policy Learning with Abstention
2025
Inference on Optimal Policy Values and Other Irregular Functionals via Softmax Smoothing
2025
Estimation of Treatment Effects in Extreme and Unobserved Data
2025
Preference Learning with Response Time: Robust Losses and Guarantees
2025
Efficient Algorithms for Adversarial Contextual Learning
2016
Bayesian Exploration: Incentivizing Exploration in Bayesian Games
2016
Improved Regret Bounds for Oracle-Based Adversarial Contextual Bandits
2016
Optimal and Myopic Information Acquisition
2017
Optimal Data Acquisition for Statistical Estimation
2017
Accurate Inference for Adaptive Linear Models
2017
Low-Rank Bandit Methods for High-Dimensional Dynamic Pricing
2018
Semiparametric Contextual Bandits
2018
Machine Learning Estimation of Heterogeneous Treatment Effects with Instruments
2019
Dynamically Aggregating Diverse Information
2019
Estimating the Long-Term Effects of Novel Treatments
2021
Top co-authors
Zhiwei Steven Wu
· 4
Akshay Krishnamurthy
· 3
Anish Agarwal
· 2
Annie Liang
· 2
Ayush Sawarni
· 2
Daniel Ngo
· 2
Greg Lewis
· 2
Justin Whitehouse
· 2
Keertana Chidambaram
· 2
Keith Battocchi
· 2
Maggie Hei
· 2
Matt Taddy
· 2
Topics
cs.LG
econ.EM
stat.ME
stat.ML
Uncategorized
math.ST
stat.TH
Exploration
Value-Based
Game AI