Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
executable
loadingβ¦
π€
Ask AI
Awesome executable β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
executable
9 papers tagged executable β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
9 papers Β· trending (default)
numbers = π₯ heat
Unified Thinker: A General Reasoning Modular Core for Image Generation
(2026)
Sashuai Zhou et al.
1.94
ABC-Bench: Benchmarking Agentic Backend Coding in Real-World Development
(2026)
Jie Yang et al.
1.94
FinToolBench: Evaluating LLM Agents for Real-World Financial Tool Use
(2026)
Jiaxuan Lu et al.
1.94
CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents
(2026)
Bowen Wang et al.
1.94
Beyond NL2Code: A Structured Survey of Multimodal Code Intelligence
(2026)
Xuanle Zhao et al.
1.94
CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation
(2026)
Haodong Li et al.
1.78
GitChameleon: Evaluating AI Code Generation Against Python Library Version Incompatibilities
(2025)
Diganta Misra et al.
1.28
OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
(2024)
Raghav Kapoor et al.
β
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
(2024)
Zuxin Liu et al.
β