Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Tal Lancewicki β most-cited papers & profile Β· Large Language Models
β authors
Β·
overview
Tal Lancewicki
6
papers Β·
1
citations Β·
2
h-index
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
Near-Optimal Regret for Adversarial MDP with Delayed Bandit Feedback
2022 Β· 1 citations
Online Learning In Mdps With Partially Adversarial Transitions And Losses
2026
Near-optimal Regret Using Policy Optimization in Online MDPs with Aggregate Bandit Feedback
2025
Learning Adversarial Markov Decision Processes with Delayed Feedback
2020
Cooperative Online Learning in Stochastic and Adversarial MDPs
2022
Delay-Adapted Policy Optimization and Improved Regret for Adversarial MDP with Delayed Bandit Feedback
2023
Topics
Model-Based RL
Exploration
Policy Gradient
Value-Based
Safe RL
Multi-Agent