Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
multi-step
loadingβ¦
π€
Ask AI
Awesome multi-step β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
multi-step
19 papers tagged multi-step β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
19 papers Β· trending (default)
numbers = π₯ heat
Deep Search with Hierarchical Meta-Cognitive Monitoring Inspired by Cognitive Neuroscience
(2026)
Zhongxiang Sun et al.
1.94
Claw-Eval: Toward Trustworthy Evaluation of Autonomous Agents
(2026)
Bowen Ye et al.
1.94
From Prompt Injection to Persistent Control: Defending Agentic Harness Against Trojan Backdoors
(2026)
Jiejun Tan et al.
1.94
Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation
(2026)
Amin Karimi Monsefi et al.
1.89
Signals: Trajectory Sampling and Triage for Agentic Interactions
(2026)
Shuguang Chen et al.
1.83
ClawGym: A Scalable Framework for Building Effective Claw Agents
(2026)
Fei Bai et al.
1.83
Agentic Search in the Wild: Intents and Trajectory Dynamics from 14M+ Real Search Requests
(2026)
Jingjie Ning et al.
1.67
ASTRA: Automated Synthesis of agentic Trajectories and Reinforcement Arenas
(2026)
Xiaoyu Tian et al.
1.67
SegAgent: Exploring Pixel Understanding Capabilities in MLLMs by Imitating Human Annotator Trajectories
(2025)
Muzhi Zhu et al.
1.28
FREESON: Retriever-Free Retrieval-Augmented Reasoning via Corpus-Traversing MCTS
(2025)
Chaeeun Kim et al.
1.28
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos
(2025)
Jiashuo Yu et al.
1.28
Specification Self-Correction: Mitigating In-Context Reward Hacking Through Test-Time Refinement
(2025)
VΓctor Gallego
1.28
Scaling Agents via Continual Pre-training
(2025)
Liangcai Su et al.
1.28
Multimodal Web Navigation with Instruction-Finetuned Foundation Models
(2023)
Hiroki Furuta et al.
β
GPT Can Solve Mathematical Problems Without a Calculator
(2023)
Zhen Yang et al.
β
RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval
(2024)
Parth Sarthi et al.
β
Improved Distribution Matching Distillation for Fast Image Synthesis
(2024)
Tianwei Yin et al.
β
HEMM: Holistic Evaluation of Multimodal Foundation Models
(2024)
Paul Pu Liang et al.
β
Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation
(2024)
Satyapriya Krishna et al.
β