Awesome Reinforcement Learning
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Reading Packs
News
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Reading Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Tat-Seng Chua — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Tat-Seng Chua
289
papers ·
17295
citations ·
2
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
RLPR: Extrapolating RLVR to General Domains without Verifiers
2025 · 69 citations
Positive-Unlabeled Reinforcement Learning Distillation for On-Premise Small Models
2026 · 1 citations
Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments
2026
Towards Knowledgeable Deep Research: Framework and Benchmark
2026
Less Approximates More: Harmonizing Performance and Confidence Faithfulness via Hybrid Post-Training for High-Stakes Tasks
2026
NextMem: Towards Latent Factual Memory for LLM-based Agents
2026
Self-Guard: Defending Large Reasoning Models via enhanced self-reflection
2026
The Missing Half: Unveiling Training-time Implicit Safety Risks Beyond Deployment
2026
Learning to Self-Verify Makes Language Models Better Reasoners
2026
AgentNoiseBench: Benchmarking Robustness of Tool-Using LLM Agents Under Noisy Condition
2026
FinDeepForecast: A Live Multi-Agent System for Benchmarking Deep Research Agents in Financial Forecasting
2026
Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement
2025
SafeMLRM: Demystifying Safety in Multi-modal Large Reasoning Models
2025
Rethinking Dialogue State Tracking with Reasoning
2020
Top co-authors
An Zhang
· 5
Xiang Wang
· 5
Junfeng Fang
· 4
Qi Gu
· 3
Xunliang Cai
· 3
Yuxin Chen
· 3
Chao Wang
· 2
Fengbin Zhu
· 2
Fuli Feng
· 2
Hui Su
· 2
Ruipeng Wang
· 2
Wenjie Wang
· 2
Topics
cs.AI
cs.LG
cs.CL
cs.IR
cs.MA
cs.CR
Uncategorized