Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
neural
loadingβ¦
π€
Ask AI
Awesome neural β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
neural
21 papers tagged neural β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
21 papers Β· trending (default)
numbers = π₯ heat
Hypencoder: Hypernetworks for Information Retrieval
(2025)
Julian Killingback et al.
4.47
Hybrid Linear Attention Done Right: Efficient Distillation and Effective Architectures for Extremely Long Contexts
(2026)
Yingfa Chen et al.
1.94
MemoryLLM: Plug-n-Play Interpretable Feed-Forward Memory for Transformers
(2026)
Ajay Jaiswal et al.
1.94
ReMix: Reinforcement routing for mixtures of LoRAs in LLM finetuning
(2026)
Ruizhong Qiu et al.
1.94
Graph-Aware Isomorphic Attention for Adaptive Dynamics in Transformers
(2025)
Markus J. Buehler
1.28
Instruction-Guided Autoregressive Neural Network Parameter Generation
(2025)
Soro Bedionita et al.
1.28
Advancing Arabic Reverse Dictionary Systems: A Transformer-Based Approach with Dataset Construction Guidelines
(2025)
Serry Sibaee et al.
1.28
SingLoRA: Low Rank Adaptation Using a Single Matrix
(2025)
David BensaΓ―d et al.
1.28
MemMamba: Rethinking Memory Patterns in State Space Model
(2025)
Youjin Wang et al.
1.28
Visionary: The World Model Carrier Built on WebGPU-Powered Gaussian Splatting Platform
(2025)
Yuning Gong et al.
1.28
Interpretability at Scale: Identifying Causal Mechanisms in Alpaca
(2023)
Zhengxuan Wu et al.
β
Brainformers: Trading Simplicity for Efficiency
(2023)
Yanqi Zhou et al.
β
Mixture-of-Supernets: Improving Weight-Sharing Supernet Training with Architecture-Routed Mixture-of-Experts
(2023)
Ganesh Jawahar et al.
β
DrugChat: Towards Enabling ChatGPT-Like Capabilities on Drug Molecule Graphs
(2023)
Youwei Liang et al.
β
UniAudio: An Audio Foundation Model Toward Universal Audio Generation
(2023)
Dongchao Yang et al.
β
LoGAH: Predicting 774-Million-Parameter Transformers using Graph HyperNetworks with 1/100 Parameters
(2024)
Xinyu Zhou et al.
β
LETS-C: Leveraging Language Embedding for Time Series Classification
(2024)
Rachneet Kaur et al.
β
A Web-Based Solution for Federated Learning with LLM-Based Automation
(2024)
Chamith Mawela et al.
β
Cottention: Linear Transformers With Cosine Attention
(2024)
Gabriel Mongaras et al.
β
Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis
(2024)
Hippolyte Gisserot-Boukhlef et al.
β
Whisper-GPT: A Hybrid Representation Audio Large Language Model
(2024)
Prateek Verma
β