Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
modality
loadingβ¦
π€
Ask AI
Awesome modality β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
modality
12 papers tagged modality β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
12 papers Β· trending (default)
numbers = π₯ heat
Enhanced OoD Detection through Cross-Modal Alignment of Multi-Modal Representations
(2025)
Jeonghyeon Kim et al.
4.47
AVERE: Improving Audiovisual Emotion Reasoning with Preference Optimization
(2026)
Ashutosh Chaubey et al.
1.94
LoMo: Local Modality Substitution for Deeper Vision-Language Fusion
(2026)
Feng Han et al.
1.94
Beyond Language Modeling: An Exploration of Multimodal Pretraining
(2026)
Shengbang Tong et al.
1.78
VisualSimpleQA: A Benchmark for Decoupled Evaluation of Large Vision-Language Models in Fact-Seeking Question Answering
(2025)
Yanling Wang et al.
1.28
Unicorn: Text-Only Data Synthesis for Vision Language Model Training
(2025)
Xiaomin Yu et al.
1.28
Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models
(2025)
Weihao Xuan et al.
1.28
Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability
(2025)
Haiqi Yang et al.
1.28
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
(2025)
Sidharth Surapaneni et al.
1.28
OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs
(2025)
Caorui Li et al.
1.28
Omni-Reward: Towards Generalist Omni-Modal Reward Modeling with Free-Form Preferences
(2025)
Zhuoran Jin et al.
1.28
The Curse of Multi-Modalities: Evaluating Hallucinations of Large Multimodal Models across Language, Visual, and Audio
(2024)
Sicong Leng et al.
β