Awesome Large Language Models
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Reading Packs
News
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Reading Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Piotr Nawrot — most-cited papers & profile · Large Language Models
← authors
·
overview
Piotr Nawrot
6
papers ·
1
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
The Sparse Frontier: Sparse Attention Trade-offs in Transformer LLMs
2025 · 1 citations
Fast and Expressive Multi-Byte Prediction with Probabilistic Circuits
2025
Inference-Time Hyper-Scaling with KV Cache Compression
2025
Hierarchical Transformers Are More Efficient Language Models
2021
No Train No Gain: Revisiting Efficient Training Algorithms For Transformer-based Language Models
2023
nanoT5: A PyTorch Framework for Pre-training and Fine-tuning T5-style Models with Limited Resources
2023
Top co-authors
Edoardo M. Ponti
· 2
Pasquale Minervini
· 2
Adrian {\L}a\'ncucki
· 1
Andreas Grivas
· 1
Antonio Vergari
· 1
Christian Szegedy
· 1
Edoardo Ponti
· 1
Emile van Krieken
· 1
Euan Wielewski
· 1
Henryk Michalewski
· 1
Jean Kaddour
· 1
Kelly Marchisio
· 1
Topics
Model Architecture
Efficiency
cs.LG
Evaluation
cs.CL
RAG
Vision-Language
Fine-Tuning
Code