Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
error
loadingβ¦
π€
Ask AI
Awesome error β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
error
20 papers tagged error β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
20 papers Β· trending (default)
numbers = π₯ heat
Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing
(2026)
Tommaso Cerruti et al.
2.00
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
(2026)
Dingjie Song et al.
1.94
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
(2026)
Yawen Luo et al.
1.94
Improving Vision-language Models with Perception-centric Process Reward Models
(2026)
Yingqian Min et al.
1.94
Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems
(2026)
Shihao Qi et al.
1.94
ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attention
(2026)
Joe Sharratt
1.94
EmpiriGraph-Psy: A Dataset and LLM Pipeline for Extracting Empirical Relation Graphs from Psychology Abstracts
(2026)
Danqin Zhao et al.
1.94
Learning Conformal Abstention Policies for Adaptive Risk Management in Large Language and Vision-Language Models
(2025)
Sina Tayebati et al.
1.28
IterPref: Focal Preference Learning for Code Generation via Iterative Debugging
(2025)
Jie Wu et al.
1.28
LLM Context Conditioning and PWP Prompting for Multimodal Validation of Chemical Formulas
(2025)
Evgeny Markhasin
1.28
ChartMuseum: Testing Visual Reasoning Capabilities of Large Vision-Language Models
(2025)
Liyan Tang et al.
1.28
MMMR: Benchmarking Massive Multi-Modal Reasoning Tasks
(2025)
Guiyao Tie et al.
1.28
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos
(2025)
Tingyu Song et al.
1.28
Mitigating Attention Sinks and Massive Activations in Audio-Visual Speech Recognition with LLMS
(2025)
Anand et al.
1.28
Thinking with Programming Vision: Towards a Unified View for Thinking with Images
(2025)
Zirun Guo et al.
1.28
Towards a Science of Scaling Agent Systems
(2025)
Yubin Kim et al.
1.28
EVOLVE-VLA: Test-Time Training from Environment Feedback for Vision-Language-Action Models
(2025)
Zechen Bai et al.
1.28
PPTC Benchmark: Evaluating Large Language Models for PowerPoint Task Completion
(2023)
Yiduo Guo et al.
β
Denoising LM: Pushing the Limits of Error Correction Models for Speech Recognition
(2024)
Zijin Gu et al.
β
Are We Done with MMLU?
(2024)
Aryo Pradipta Gema et al.
β