Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiaoyi Dong — most-cited papers & profile · Large Language Models
← authors
·
overview
Xiaoyi Dong
35
papers ·
426
citations ·
9
h-index
Beijing Academy of Artificial Intelligence · ShangHai JiAi Genetics & IVF Institute · Shanghai Artificial Intelligence Laboratory · Nanjing University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
2024 · 230 citations
ShareGPT4Video: Improving Video Understanding and Generation with Better Captions
2024 · 33 citations
InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
2024 · 2 citations
MMDU: A Multi-Turn Multi-Image Dialog Understanding Benchmark and Instruction-Tuning Dataset for LVLMs
2024 · 1 citations
Demo-ICL: In-Context Learning for Procedural Video Knowledge Acquisition
2026
BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning
2025
OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?
2025
InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model
2025
MM-IFEngine: Towards Multimodal Instruction Following
2025
Beyond Fixed: Variable-Length Denoising for Diffusion Large Language Models
2025
SIM-CoT: Supervised Implicit Chain-of-Thought
2025
SPARK: Synergistic Policy And Reward Co-Evolving Framework
2025
CapRL: Stimulating Dense Image Caption Capabilities via Reinforcement Learning
2025
InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
2024
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
2024
Top co-authors
Jiaqi Wang
· 20
Yuhang Zang
· 19
Dahua Lin
· 18
Yuhang Cao
· 16
Pan Zhang
· 14
Haodong Duan
· 11
Conghui He
· 8
Yu Qiao
· 7
Bin Wang
· 5
Kai Chen
· 5
Wei Li
· 5
Wenwei Zhang
· 5
Topics
Vision-Language
Training Techniques
Fine-Tuning
Efficiency
Model Architecture
Evaluation
In-Context Learning
Large
Models
Code