Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
complexity
loadingβ¦
π€
Ask AI
Awesome complexity β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
complexity
17 papers tagged complexity β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
17 papers Β· trending (default)
numbers = π₯ heat
Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing
(2026)
Tommaso Cerruti et al.
2.00
Scaling Small Agents Through Strategy Auctions
(2026)
Lisa Alazraki et al.
1.94
Memanto: Typed Semantic Memory with Information-Theoretic Retrieval for Long-Horizon Agents
(2026)
Seyed Moein Abtahi et al.
1.94
Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators
(2026)
Zachary Novack et al.
1.94
ARWKV: Pretrain is not what we need, an RNN-Attention-Based Language Model Born from Transformer
(2025)
Lin Yueyu et al.
1.28
Pixel-SAIL: Single Transformer For Pixel-Grounded Understanding
(2025)
Tao Zhang et al.
1.28
UFT: Unifying Supervised and Reinforcement Fine-Tuning
(2025)
Mingyang Liu et al.
1.28
SageAttention2++: A More Efficient Implementation of SageAttention2
(2025)
Jintao Zhang et al.
1.28
Diagonal Batching Unlocks Parallelism in Recurrent Memory Transformers for Long Contexts
(2025)
Danil Sivtsov et al.
1.28
MovieCORE: COgnitive REasoning in Movies
(2025)
Gueter Josmy Faure et al.
1.28
MemMamba: Rethinking Memory Patterns in State Space Model
(2025)
Youjin Wang et al.
1.28
Aligned but Stereotypical? The Hidden Influence of System Prompts on Social Bias in LVLM-Based Text-to-Image Models
(2025)
NaHyeon Park et al.
1.28
On the Reliability of Watermarks for Large Language Models
(2023)
John Kirchenbauer et al.
β
Weaver: Foundation Models for Creative Writing
(2024)
Tiannan Wang et al.
β
LoRA Land: 310 Fine-tuned LLMs that Rival GPT-4, A Technical Report
(2024)
Justin Zhao et al.
β
Cottention: Linear Transformers With Cosine Attention
(2024)
Gabriel Mongaras et al.
β
Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models
(2024)
Alex Havrilla et al.
β