Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
input
loadingβ¦
π€
Ask AI
Awesome input β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
input
13 papers tagged input β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
13 papers Β· trending (default)
numbers = π₯ heat
VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents
(2026)
Udi Barzelay et al.
1.94
RAGEN-2: Reasoning Collapse in Agentic RL
(2026)
Zihan Wang et al.
1.94
Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training
(2026)
Michal Chudoba et al.
1.94
GIMMICK -- Globally Inclusive Multimodal Multitask Cultural Knowledge Benchmarking
(2025)
Florian Schneider et al.
1.28
Beyond One-Size-Fits-All: Inversion Learning for Highly Effective NLG Evaluation Prompts
(2025)
Hanhua Hong et al.
1.28
An Agentic System for Rare Disease Diagnosis with Traceable Reasoning
(2025)
Weike Zhao et al.
1.28
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
(2025)
Ke Wang et al.
1.28
DialectGen: Benchmarking and Improving Dialect Robustness in Multimodal Generation
(2025)
Yu Zhou et al.
1.28
A Suite of Generative Tasks for Multi-Level Multimodal Webpage Understanding
(2023)
Andrea Burns et al.
β
Improved Distribution Matching Distillation for Fast Image Synthesis
(2024)
Tianwei Yin et al.
β
Summary of a Haystack: A Challenge to Long-Context LLMs and RAG Systems
(2024)
Philippe Laban et al.
β
MovieSum: An Abstractive Summarization Dataset for Movie Screenplays
(2024)
Rohit Saxena et al.
β
CAD-MLLM: Unifying Multimodality-Conditioned CAD Generation With MLLM
(2024)
Jingwei Xu et al.
β