Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Graham Neubig — most-cited papers & profile · Large Language Models
← authors
·
overview
Graham Neubig
46
papers ·
232
citations ·
65
h-index
Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
The Devil is in the Errors: Leveraging Large Language Models for Fine-grained Machine Translation Evaluation
2023 · 31 citations
NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples
2024 · 1 citations
Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs
2026
What do Language Models Learn and When? The Implicit Curriculum Hypothesis
2026
VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge
2025
Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning
2025
RefineBench: Evaluating Refinement Capability of Language Models via Checklists
2025
On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models
2025
Scaling Evaluation-time Compute with Reasoning Models as Evaluators
2025
WebArena: A Realistic Web Environment for Building Autonomous Agents
2023
Alignment for Honesty
2023
Instruction-tuned Language Models are Better Knowledge Learners
2024
Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models
2024
OpenDevin: An Open Platform for AI Software Developers as Generalist Agents
2024
Harnessing Webpage UIs for Text-Rich Visual Understanding
2024
Top co-authors
Xiang Yue
· 8
Seungone Kim
· 6
Sean Welleck
· 5
Frank F. Xu
· 4
Akari Asai
· 2
Boxuan Li
· 2
Carolin Lawrence
· 2
Juyoung Suk
· 2
Kiril Gashteovski
· 2
Seongyun Lee
· 2
Shuyan Zhou
· 2
Tianyue Ou
· 2
Topics
Evaluation
Fine-Tuning
Vision-Language
Training Techniques
Code
In-Context Learning
RAG
Efficiency
Agentic
Model Architecture