Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
two-stage
loadingβ¦
π€
Ask AI
Awesome two-stage β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
two-stage
12 papers tagged two-stage β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
12 papers Β· trending (default)
numbers = π₯ heat
Rethinking UMM Visual Generation: Masked Modeling for Efficient Image-Only Pre-training
(2026)
Peng Sun et al.
1.94
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
(2026)
Yawen Luo et al.
1.94
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer
(2026)
Haoyi Zhu et al.
1.94
LongAttnComp: Cross-Family Context Compression for Long-Context Reasoning
(2026)
Mengmeng Ji et al.
1.94
AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting
(2025)
Shijue Huang et al.
1.28
Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue
(2025)
Xingyao Lin et al.
1.28
CapRL: Stimulating Dense Image Caption Capabilities via Reinforcement Learning
(2025)
Long Xing et al.
1.28
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
(2025)
Chenghao Zhang et al.
1.28
V-Thinker: Interactive Thinking with Images
(2025)
Runqi Qiao et al.
1.28
MIRA: Multimodal Iterative Reasoning Agent for Image Editing
(2025)
Ziyun Zeng et al.
1.28
OpenREAD: Reinforced Open-Ended Reasoing for End-to-End Autonomous Driving with LLM-as-Critic
(2025)
Songyan Zhang et al.
1.28
VIMI: Grounding Video Generation through Multi-modal Instruction
(2024)
Yuwei Fang et al.
β