Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Scott Sanner — most-cited papers & profile · Large Language Models
← authors
·
overview
Scott Sanner
21
papers ·
39
citations ·
40
h-index
University of Toronto · Vector Institute
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
ε-BMC: A Bayesian Ensemble Approach to Epsilon-Greedy Exploration in Model-Free Reinforcement Learning
2020 · 11 citations
Bayesian Knowledge-driven Critiquing with Indirect Evidence
2023 · 10 citations
Risk-Aware Transfer in Reinforcement Learning using Successor Features
2021 · 9 citations
Multimodal Item Scoring for Natural Language Recommendation via Gaussian Process Regression with LLM Relevance Judgments
2025 · 4 citations
Learning to Follow Instructions in Text-Based Games
2022 · 2 citations
Conditional Inference in Pre-trained Variational Autoencoders via Cross-coding
2018 · 1 citations
Contextual Policy Transfer in Reinforcement Learning Domains via Deep Mixtures-of-Experts
2020 · 1 citations
Diffusion on the Probability Simplex
2023 · 1 citations
ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning
2026
Reflect-then-Plan: Offline Model-Based Planning through a Doubly Bayesian Lens
2025
Reward Potentials for Planning with Learned Neural Network Transition Models
2019
RAPTOR: End-to-end Risk-Aware MDP Planning and Policy Learning by Backpropagation
2021
TransCAM: Transformer Attention-based CAM Refinement for Weakly Supervised Semantic Segmentation
2022
Conservative Bayesian Model-Based Value Expansion for Offline Policy Optimization
2022
Perimeter Control Using Deep Reinforcement Learning: A Model-free Approach towards Homogeneous Flow Rate Optimization
2023
Top co-authors
Anton Korikov
· 1
Armin Toroghi
· 1
Jiazhou Liang
· 1
Junyoung Kim
· 1
Justin Cui
· 1
Mark Zhao
· 1
Qianfeng Wen
· 1
Yifan Liu
· 1
Topics
Model-Based RL
Safe RL
Value-Based
Policy Gradient
Offline RL
Exploration
Ranking & CTR
Conditioning & Control
Training & Sampling
Meta-RL