Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
scenarios
loadingβ¦
π€
Ask AI
Awesome scenarios β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
scenarios
20 papers tagged scenarios β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
20 papers Β· trending (default)
numbers = π₯ heat
Same Claim, Different Judgment: Benchmarking Scenario-Induced Bias in Multilingual Financial Misinformation Detection
(2026)
Zhiwei Liu et al.
1.94
TRIP-Bench: A Benchmark for Long-Horizon Interactive Agents in Real-World Scenarios
(2026)
Yuanzhe Shen et al.
1.72
Exploring the sustainable scaling of AI dilemma: A projective study of corporations' AI environmental impacts
(2025)
ClΓ©ment Desroches et al.
1.28
X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
(2025)
Salman Rahman et al.
1.28
Thinking with Generated Images
(2025)
Ethan Chern et al.
1.28
Prefix Grouper: Efficient GRPO Training through Shared-Prefix Forward
(2025)
Zikang Liu et al.
1.28
An Agentic System for Rare Disease Diagnosis with Traceable Reasoning
(2025)
Weike Zhao et al.
1.28
ModelCitizens: Representing Community Voices in Online Safety
(2025)
Ashima Suvarna et al.
1.28
Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation
(2025)
Junyan Ye et al.
1.28
GraphTracer: Graph-Guided Failure Tracing in LLM Agents for Robust Multi-Turn Deep Search
(2025)
Heng Zhang et al.
1.28
MMPersuade: A Dataset and Evaluation Framework for Multimodal Persuasion
(2025)
Haoyi Qiu et al.
1.28
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
(2025)
DeepSeek-AI et al.
1.28
Hybrid Attribution Priors for Explainable and Robust Model Training
(2025)
Zhuoran Zhang et al.
1.28
ICE-GRT: Instruction Context Enhancement by Generative Reinforcement based Transformers
(2024)
Chen Zheng et al.
β
Iterative Graph Alignment
(2024)
Fangyuan Yu et al.
β
Agent Workflow Memory
(2024)
Zora Zhiruo Wang et al.
β
VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI
(2024)
Sijie Cheng et al.
β
Cross-Lingual Auto Evaluation for Assessing Multilingual LLMs
(2024)
Sumanth Doddapaneni et al.
β
GATE OpenING: A Comprehensive Benchmark for Judging Open-ended Interleaved Image-Text Generation
(2024)
Pengfei Zhou et al.
β
VideoICL: Confidence-based Iterative In-context Learning for Out-of-Distribution Video Understanding
(2024)
Kangsan Kim et al.
β