Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
window
loadingβ¦
π€
Ask AI
Awesome window β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
window
13 papers tagged window β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
13 papers Β· trending (default)
numbers = π₯ heat
Adapting Multilingual Embedding Models to Turkish via Cross-Lingual Tokenizer Surgery and Offline Distillation
(2026)
M. Ali Bayram et al.
1.89
Model Context Protocol (MCP) Tool Descriptions Are Smelly! Towards Improving AI Agent Efficiency with Augmented MCP Tool Descriptions
(2026)
Mohammed Mehedi Hasan et al.
1.72
Revisiting Long-context Modeling from Context Denoising Perspective
(2025)
Zecheng Tang et al.
1.28
Native Hybrid Attention for Efficient Sequence Modeling
(2025)
Jusen Du et al.
1.28
Artificial Hippocampus Networks for Efficient Long-Context Modeling
(2025)
Yunhao Fang et al.
1.28
VR-Thinker: Boosting Video Reward Models through Thinking-with-Image Reasoning
(2025)
Qunzhong Wang et al.
1.28
On Pretraining for Project-Level Code Completion
(2025)
Maksim Sapronov et al.
1.28
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
(2025)
Yunhong Lu et al.
1.28
InfiniteVL: Synergizing Linear and Sparse Attention for Highly-Efficient, Unlimited-Input Vision-Language Models
(2025)
Hongyuan Tao et al.
1.28
Fast Chain-of-Thought: A Glance of Future from Parallel Decoding Leads to Answers Faster
(2023)
Hongxuan Zhang et al.
β
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
(2024)
Peng Xu et al.
β
Your Context Is Not an Array: Unveiling Random Access Limitations in Transformers
(2024)
MohammadReza Ebrahimi et al.
β
Gated Delta Networks: Improving Mamba2 with Delta Rule
(2024)
Songlin Yang et al.
β