Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
GPT-4V
loadingβ¦
π€
Ask AI
Awesome GPT-4V β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
GPT-4V
12 papers tagged GPT-4V β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
12 papers Β· trending (default)
numbers = π₯ heat
Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models
(2023)
Haoning Wu et al.
β
Silkie: Preference Distillation for Large Visual Language Models
(2023)
Lei Li et al.
β
Gemini vs GPT-4V: A Preliminary Comparison and Combination of Vision-Language Models Through Qualitative Cases
(2023)
Zhangyang Qi et al.
β
GPT-4V(ision) is a Human-Aligned Evaluator for Text-to-3D Generation
(2024)
Tong Wu et al.
β
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
(2024)
Rohan Wadhawan et al.
β
BBA: Bi-Modal Behavioral Alignment for Reasoning with Large Vision-Language Models
(2024)
Xueliang Zhao et al.
β
Towards Open-ended Visual Quality Comparison
(2024)
Haoning Wu et al.
β
InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
(2024)
Xiaoyi Dong et al.
β
HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing
(2024)
Mude Hui et al.
β
LEGENT: Open Platform for Embodied Agents
(2024)
Zhili Cheng et al.
β
AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
(2024)
Wenhao Chai et al.
β
VLsI: Verbalized Layers-to-Interactions from Large to Small Vision Language Models
(2024)
Byung-Kwan Lee et al.
β