Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Bill Yuchen Lin — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Bill Yuchen Lin
22
papers ·
168
citations ·
24
h-index
University of Cambridge
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion
2023 · 99 citations
Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models
2024 · 40 citations
LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition
2023 · 7 citations
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
2024 · 7 citations
OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
2024 · 4 citations
Faith and Fate: Limits of Transformers on Compositionality
2023 · 4 citations
FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks
2021 · 3 citations
SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
2024 · 3 citations
On Grounded Planning for Embodied Tasks with Language Models
2022 · 1 citations
ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
2025
Small Models Struggle to Learn from Strong Reasoners
2025
CrossWordBench: Evaluating the Reasoning Capabilities of LLMs and LVLMs with Controllable Puzzle Generation
2025
TinyV: Reducing False Negatives in Verification Improves RL for LLM Reasoning
2025
Temporal Sampling for Forgotten Reasoning in LLMs
2025
Latent Action Pretraining from Videos
2024
Topics
Evaluation
Training Techniques
Fine-Tuning
Model Architecture
cs.AI
cs.CL
Code Models
Code
Safety & Alignment
Large