Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
evidence
loadingβ¦
π€
Ask AI
Awesome evidence β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
evidence
21 papers tagged evidence β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
21 papers Β· trending (default)
numbers = π₯ heat
Unified Personalized Reward Model for Vision Generation
(2026)
Yibin Wang et al.
1.94
GraphAgents: Knowledge Graph-Guided Agentic AI for Cross-Domain Materials Design
(2026)
Isabella A. Stewart et al.
1.94
REDSearcher: A Scalable and Cost-Efficient Framework for Long-Horizon Search Agents
(2026)
Zheng Chu et al.
1.94
BubbleRAG: Evidence-Driven Retrieval-Augmented Generation for Black-Box Knowledge Graphs
(2026)
Duyi Pan et al.
1.94
VideoZeroBench: Probing the Limits of Video MLLMs with Spatio-Temporal Evidence Verification
(2026)
Jiahao Meng et al.
1.94
Learning to Retrieve from Agent Trajectories
(2026)
Yuqi Zhou et al.
1.94
MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments
(2026)
Han Wang et al.
1.94
ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning
(2026)
Juncheng Wu et al.
1.94
Advancing Creative Physical Intelligence in Large Multimodal Models
(2026)
Cheng Qian et al.
1.94
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
(2026)
Chengzhi Liu et al.
1.94
Struct-Searcher: Agentic Structural Thinking Advances Multimodal Deep Information Seeking
(2026)
Fan Zhang et al.
1.94
Beyond Monolingual Deep Research: Evaluating Agents and Retrievers with Cross-Lingual BrowseComp-Plus
(2026)
Yuheng Lu et al.
1.94
One Forward Beats Two: InnerZoom for Accurate and Efficient GUI Grounding
(2026)
Chen Liu et al.
1.94
RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards
(2026)
Gaotang Li et al.
1.89
Diversed Model Discovery via Structured Table Discovery
(2026)
Zhengyuan Dong et al.
1.89
The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure
(2026)
Yubo Li et al.
1.89
Agentic Search in the Wild: Intents and Trajectory Dynamics from 14M+ Real Search Requests
(2026)
Jingjie Ning et al.
1.67
On-Policy RL Meets Off-Policy Experts: Harmonizing Supervised Fine-Tuning and Reinforcement Learning via Dynamic Weighting
(2025)
Wenhao Zhang et al.
1.28
VA-Ο: Variational Policy Alignment for Pixel-Aware Autoregressive Generation
(2025)
Xinyao Liao et al.
1.28
LoRA-Contextualizing Adaptation of Large Multimodal Models for Long Document Understanding
(2024)
Jian Chen et al.
β
M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
(2024)
Jaemin Cho et al.
β