Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
large-scale
loadingβ¦
π€
Ask AI
Awesome large-scale β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
large-scale
23 papers tagged large-scale β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
23 papers Β· trending (default)
numbers = π₯ heat
Qwen3-ASR Technical Report
(2026)
Xian Shi et al.
1.94
The PokeAgent Challenge: Competitive and Long-Context Learning at Scale
(2026)
Seth Karten et al.
1.94
Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music
(2026)
Sreyan Ghosh et al.
1.94
LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence
(2026)
Xiang An et al.
1.94
DeNovoSWE: Scaling Long-Horizon Environments for Generating Entire Repositories from Scratch
(2026)
Jiale Zhao et al.
1.94
SWE-INTERACT: Reimagining SWE Benchmarks as User-Driven Long-Horizon Coding Sessions
(2026)
Mohit Raghavendra et al.
1.94
MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale
(2026)
Bin Wang et al.
1.83
Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing
(2026)
Zeyue Tian et al.
1.83
ACECODER: Acing Coder RL via Automated Test-Case Synthesis
(2025)
Huaye Zeng et al.
1.28
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
(2025)
Qiying Yu et al.
1.28
Wan: Open and Advanced Large-Scale Video Generative Models
(2025)
WanTeam et al.
1.28
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
(2025)
Anjiang Wei et al.
1.28
UniBiomed: A Universal Foundation Model for Grounded Biomedical Image Interpretation
(2025)
Linshan Wu et al.
1.28
Decoding Open-Ended Information Seeking Goals from Eye Movements in Reading
(2025)
Cfir Avraham Hadar et al.
1.28
Unified Multimodal Chain-of-Thought Reward Model through Reinforcement Fine-Tuning
(2025)
Yibin Wang et al.
1.28
Are We on the Right Way for Assessing Document Retrieval-Augmented Generation?
(2025)
Wenxuan Shen et al.
1.28
SAIL-VL2 Technical Report
(2025)
Weijie Yin et al.
1.28
DialectGen: Benchmarking and Improving Dialect Robustness in Multimodal Generation
(2025)
Yu Zhou et al.
1.28
DataFlow: An LLM-Driven Framework for Unified Data Preparation and Workflow Automation in the Era of Data-Centric AI
(2025)
Hao Liang et al.
1.28
VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
(2023)
Wenlong Huang et al.
β
ORES: Open-vocabulary Responsible Visual Synthesis
(2023)
Minheng Ni et al.
β
MyVLM: Personalizing VLMs for User-Specific Queries
(2024)
Yuval Alaluf et al.
β
Compositional 3D-aware Video Generation with LLM Director
(2024)
Hanxin Zhu et al.
β