Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
decoder
loadingβ¦
π€
Ask AI
Awesome decoder β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
decoder
15 papers tagged decoder β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
15 papers Β· trending (default)
numbers = π₯ heat
NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priors
(2026)
Lingfeng Ren et al.
1.94
Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models
(2026)
Cheng-Yu Yang et al.
1.94
One Forward Beats Two: InnerZoom for Accurate and Efficient GUI Grounding
(2026)
Chen Liu et al.
1.94
QuoTA: Query-oriented Token Assignment via CoT Query Decouple for Long Video Comprehension
(2025)
Yongdong Luo et al.
1.28
Prot2Token: A Unified Framework for Protein Modeling via Next-Token Prediction
(2025)
Mahdi Pourmirzaei et al.
1.28
X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again
(2025)
Zigang Geng et al.
1.28
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
(2025)
Peiyu Wang et al.
1.28
Can Understanding and Generation Truly Benefit Together -- or Just Coexist?
(2025)
Zhiyuan Yan et al.
1.28
ECLIPSE: A Resource-Efficient Text-to-Image Prior for Image Generations
(2023)
Maitreya Patel et al.
β
BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
(2024)
Mateusz Εajszczak et al.
β
LoGAH: Predicting 774-Million-Parameter Transformers using Graph HyperNetworks with 1/100 Parameters
(2024)
Xinyu Zhou et al.
β
OmniJARVIS: Unified Vision-Language-Action Tokenization Enables Open-World Instruction Following Agents
(2024)
Zihao Wang et al.
β
VideoGLaMM: A Large Multimodal Model for Pixel-Level Visual Grounding in Videos
(2024)
Shehan Munasinghe et al.
β
GeoX: Geometric Problem Solving Through Unified Formalized Vision-Language Pre-training
(2024)
Renqiu Xia et al.
β
Compressed Chain of Thought: Efficient Reasoning Through Dense Representations
(2024)
Jeffrey Cheng et al.
β