Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
interactions
loadingβ¦
π€
Ask AI
Awesome interactions β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
interactions
23 papers tagged interactions β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
23 papers Β· trending (default)
numbers = π₯ heat
DSAEval: Evaluating Data Science Agents on a Wide Range of Real-World Data Science Problems
(2026)
Maojun Sun et al.
1.94
Enhancing Multi-Image Understanding through Delimiter Token Scaling
(2026)
Minyoung Lee et al.
1.94
MINTEval: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems
(2026)
Hyunji Lee et al.
1.94
ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attention
(2026)
Joe Sharratt
1.94
SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking
(2026)
Guohong Liu et al.
1.94
SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents
(2026)
Wenxuan Wang et al.
1.94
APPO: Agentic Procedural Policy Optimization
(2026)
Xucong Wang et al.
1.94
Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale
(2026)
Ang Li et al.
1.94
Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention
(2026)
Vishesh Tripathi et al.
1.94
TRIP-Bench: A Benchmark for Long-Horizon Interactive Agents in Real-World Scenarios
(2026)
Yuanzhe Shen et al.
1.72
LLMVoX: Autoregressive Streaming Text-to-Speech Model for Any LLM
(2025)
Sambal Shikhar et al.
1.28
Lumina-Image 2.0: A Unified and Efficient Image Generative Framework
(2025)
Qi Qin et al.
1.28
X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
(2025)
Salman Rahman et al.
1.28
Build the web for agents, not agents for the web
(2025)
Xing Han LΓΉ et al.
1.28
VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
(2023)
Wenlong Huang et al.
β
Hiformer: Heterogeneous Feature Interactions Learning with Transformers for Recommender Systems
(2023)
Huan Gui et al.
β
Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
(2023)
Bin Lin et al.
β
Very Large-Scale Multi-Agent Simulation in AgentScope
(2024)
Xuchen Pan et al.
β
Text2SQL is Not Enough: Unifying AI and Databases with TAG
(2024)
Asim Biswal et al.
β
MentalArena: Self-play Training of Language Models for Diagnosis and Treatment of Mental Health Disorders
(2024)
Cheng Li et al.
β
MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
(2024)
Hang Hua et al.
β
Mitigating Object Hallucination via Concentric Causal Attention
(2024)
Yun Xing et al.
β
Taipan: Efficient and Expressive State Space Language Models with Selective Attention
(2024)
Chien Van Nguyen et al.
β