Awesome Large Language Models
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Xuezhou Zhang โ most-cited papers & profile ยท Large Language Models
โ authors
ยท
overview
Xuezhou Zhang
31
papers ยท
158
citations ยท
13
h-index
Boston University ยท Beijing Anzhen Hospital
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Policy Poisoning in Batch Reinforcement Learning and Control
2019 ยท 43 citations
Adaptive Reward-Poisoning Attacks against Reinforcement Learning
2020 ยท 34 citations
Reward Poisoning in Reinforcement Learning: Attacks Against Unknown Learners in Unknown Environments
2021 ยท 16 citations
Task-agnostic Exploration in Reinforcement Learning
2020 ยท 13 citations
Robust Policy Gradient against Strong Data Corruption
2021 ยท 10 citations
Corruption-Robust Offline Reinforcement Learning
2021 ยท 7 citations
Provable Defense against Backdoor Policies in Reinforcement Learning
2022 ยท 7 citations
Using Machine Teaching to Investigate Human Assumptions when Teaching Reinforcement Learners
2020 ยท 6 citations
Representation Learning for Online and Offline RL in Low-rank MDPs
2021 ยท 5 citations
Bandit Theory and Thompson Sampling-Guided Directed Evolution for Sequence Optimization
2022 ยท 4 citations
Decentralized Gossip-Based Stochastic Bilevel Optimization over Communication Networks
2022 ยท 4 citations
Decentralized Gossip-Based Stochastic Bilevel Optimization over Communication Networks
2022 ยท 4 citations
Provable Benefits of Representational Transfer in Reinforcement Learning
2022 ยท 2 citations
Byzantine-Robust Online and Offline Distributed Reinforcement Learning
2022 ยท 2 citations
The Sample Complexity of Teaching-by-Reinforcement on Q-Learning
2020 ยท 1 citations
Top co-authors
Mingyu Chen
ยท 2
Aldo Pacchiano
ยท 1
Hejian Sang
ยท 1
Ioannis Paschalidis
ยท 1
Jason D. Lee
ยท 1
Kiant\'e Brantley
ยท 1
Sirou Zhu
ยท 1
Wenhao Zhan
ยท 1
Wen Sun
ยท 1
Xiaofeng Lin
ยท 1
Yilei Chen
ยท 1
Zhaolin Gao
ยท 1
Topics
Exploration
Safe RL
Model-Based RL
Value-Based
Offline RL
Multi-Agent
Meta-RL
Policy Gradient
Evaluation
Benchmarks