Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
extraction
loadingβ¦
π€
Ask AI
Awesome extraction β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
extraction
16 papers tagged extraction β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
16 papers Β· trending (default)
numbers = π₯ heat
VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents
(2026)
Udi Barzelay et al.
1.94
What Really Controls Temporal Reasoning in Large Language Models: Tokenisation or Representation of Time?
(2026)
Gagan Bhatia et al.
1.94
FactReview: Evidence-Grounded Reviews with Literature Positioning and Execution-Based Claim Verification
(2026)
Hang Xu et al.
1.94
Memanto: Typed Semantic Memory with Information-Theoretic Retrieval for Long-Horizon Agents
(2026)
Seyed Moein Abtahi et al.
1.94
From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills
(2026)
Zisu Huang et al.
1.94
EmpiriGraph-Psy: A Dataset and LLM Pipeline for Extracting Empirical Relation Graphs from Psychology Abstracts
(2026)
Danqin Zhao et al.
1.94
Bridging VideoQA and Video-Guided Agentic Tasks via Generalized Keyframe Extraction
(2026)
Sunqi Fan et al.
1.94
AtomiMed: Hierarchical Atomic Fact-Checking for Universal Clinical-Aware Medical Report Evaluation
(2026)
Yuan Wang et al.
1.94
Qianfan-OCR: A Unified End-to-End Model for Document Intelligence
(2026)
Daxiang Dong et al.
1.78
PlainQAFact: Automatic Factuality Evaluation Metric for Biomedical Plain Language Summaries Generation
(2025)
Zhiwen You et al.
1.28
MegaScience: Pushing the Frontiers of Post-Training Datasets for Science Reasoning
(2025)
Run-Ze Fan et al.
1.28
Seeing, Listening, Remembering, and Reasoning: A Multimodal Agent with Long-Term Memory
(2025)
Lin Long et al.
1.28
AutoPR: Let's Automate Your Academic Promotion!
(2025)
Qiguang Chen et al.
1.28
MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
(2025)
Massimo Bini et al.
1.28
DataFlow: An LLM-Driven Framework for Unified Data Preparation and Workflow Automation in the Era of Data-Centric AI
(2025)
Hao Liang et al.
1.28
Genie: Achieving Human Parity in Content-Grounded Datasets Generation
(2024)
Asaf Yehudai et al.
β