Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
complex
loadingβ¦
π€
Ask AI
Awesome complex β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
complex
19 papers tagged complex β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
19 papers Β· trending (default)
numbers = π₯ heat
Beyond Binary Preference: Aligning Diffusion Models to Fine-grained Criteria by Decoupling Attributes
(2026)
Chenye Meng et al.
1.94
Search-o1: Agentic Search-Enhanced Large Reasoning Models
(2025)
Xiaoxi Li et al.
1.28
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
(2025)
Rui Yang et al.
1.28
CoRT: Code-integrated Reasoning within Thinking
(2025)
Chengpeng Li et al.
1.28
Inverse Scaling in Test-Time Compute
(2025)
Aryo Pradipta Gema et al.
1.28
Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation
(2025)
Junyan Ye et al.
1.28
Ovis2.5 Technical Report
(2025)
Shiyin Lu et al.
1.28
Visual Representation Alignment for Multimodal Large Language Models
(2025)
Heeji Yoon et al.
1.28
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
(2025)
Sidharth Surapaneni et al.
1.28
SAIL-VL2 Technical Report
(2025)
Weijie Yin et al.
1.28
Scaling Clinical Trial Matching Using Large Language Models: A Case Study in Oncology
(2023)
Cliff Wong et al.
β
Silkie: Preference Distillation for Large Visual Language Models
(2023)
Lei Li et al.
β
Agent Workflow Memory
(2024)
Zora Zhiruo Wang et al.
β
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System
(2024)
Weize Chen et al.
β
MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
(2024)
Hang Hua et al.
β
CAMEL-Bench: A Comprehensive Arabic LMM Benchmark
(2024)
Sara Ghaboura et al.
β
VLRewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models
(2024)
Lei Li et al.
β
Multi-Dimensional Insights: Benchmarking Real-World Personalization in Large Multimodal Models
(2024)
YiFan Zhang et al.
β
MapQaTor: A System for Efficient Annotation of Map Query Datasets
(2024)
Mahir Labib Dihan et al.
β