Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chong Chen — most-cited papers & profile · Multimodal
← authors
·
overview
Chong Chen
37
papers ·
313
citations ·
24
h-index
Tianjin People's Hospital · South China University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SynthVLM: Towards High-Quality and Efficient Synthesis of Image-Caption Datasets for Vision-Language Models
2024 · 1 citations
PilotRL: Training Language Model Agents via Global Planning-Guided Progressive Reinforcement Learning
2025
HIPPO: Enhancing the Table Understanding Capability of LLMs through Hybrid-Modal Preference Optimization
2025
Top co-authors
Bin Cui
· 1
Bin Cui
· 1
Bozhou Li
· 1
Conghui He
· 1
Fangfang Li
· 1
Ge Yu
· 1
Haolan Wang
· 1
Hao Liang
· 1
Keer Lu
· 1
Qi Shi
· 1
Wentao Xiong
· 1
Wentao Zhang
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Video-Language
Image-Text Retrieval
Benchmarks