Awesome AI for Code
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
LLM
loadingβ¦
π€
Ask AI
Awesome LLM β curated papers, datasets & benchmarks Β· Awesome AI for Code
β all topics
overview
LLM
22 papers tagged LLM β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
22 papers Β· trending (default)
numbers = π₯ heat
Human-in-the-Loop Large Language Model Framework for Identification of Cutaneous Immune-Related Adverse Events
(2026)
Charles Lu et al.
5.01
Frontier Financial Judgement: Can agents tell what might move a stock?
(2026)
Joshua Harris
5.01
AppWorld-UL: Benchmarking Diverse Agent-User Interactions for Tool-Use
(2026)
Junzhi Chen et al.
4.39
Agentic systems for breast cancer treatment recommendations
(2026)
Vinicius Anjos de Almeida et al.
3.51
Fin-Analyst at FinMMEval 2026 Task 3: A Live Hybrid Trading Agent with LLM Specialists and Rule-Based Signals
(2026)
Mohotarema Rashid et al.
3.51
LLMAID: Identifying AI Capabilities in Android Apps with LLMs
(2025)
Pei Liu et al.
3.15
SkillSelect-Serve: QoS-Aware Budgeted Skill Service Recommendation for LLM Agents
(2026)
Jingyuan Zheng et al.
2.00
WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory
(2026)
Hanlin Wang et al.
2.00
Revisiting Observation Reduction for Web Agents: Comprehensive Evaluation with a Lightweight Framework
(2026)
Masafumi Enomoto et al.
1.89
SEVerA: Verified Synthesis of Self-Evolving Agents
(2026)
Debangshu Banerjee et al.
1.78
Favia: Forensic Agent for Vulnerability-fix Identification and Analysis
(2026)
AndrΓ© Storhaug et al.
1.72
API Agents vs. GUI Agents: Divergence and Convergence
(2025)
Chaoyun Zhang et al.
1.28
PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines
(2025)
Reya Vir et al.
1.28
Which Agent Causes Task Failures and When? On Automated Failure Attribution of LLM Multi-Agent Systems
(2025)
Shaokun Zhang et al.
1.28
Skywork-SWE: Unveiling Data Scaling Laws for Software Engineering in LLMs
(2025)
Liang Zeng et al.
1.28
ReLook: Vision-Grounded RL with a Multimodal LLM Critic for Agentic Web Coding
(2025)
Yuhang Li et al.
1.28
Tailoring with Targeted Precision: Edit-Based Agents for Open-Domain Procedure Customization
(2023)
Yash Kumar Lal and Li Zhang and Faeze Brahman and Bodhisattwa Prasad Majumder and Peter Clark and Niket Tandon
β
LLM as OS, Agents as Apps: Envisioning AIOS, Agents and the AIOS-Agent Ecosystem
(2023)
Yingqiang Ge et al.
β
Intelligent Virtual Assistants with LLM-based Process Automation
(2023)
Yanchu Guan et al.
β
CogAgent: A Visual Language Model for GUI Agents
(2023)
Wenyi Hong et al.
β
Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction
(2024)
Zhenmei Shi et al.
β
Hardware and Software Platform Inference
(2024)
Cheng Zhang et al.
β