Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Manling Li — most-cited papers & profile · Multimodal
← authors
·
overview
Manling Li
10
papers ·
10
citations ·
1
h-index
Guiyang College of Traditional Chinese Medicine · Johns Hopkins University · Guizhou University · Drexel University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
HourVideo: 1-Hour Video-Language Understanding
2024 · 5 citations
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
2025 · 1 citations
Towards Fast Adaptation of Pretrained Contrastive Models for Multi-channel Video-Language Retrieval
2022 · 1 citations
ENACT: Evaluating Embodied Cognition With World Modeling Of Egocentric Interaction
2025
Rethinking Task Sampling for Few-shot Vision-Language Transfer Learning
2022
Top co-authors
Qineng Wang
· 2
Agrim Gupta
· 1
Cheng Qian
· 1
Crist\'obal Eyzaguirre
· 1
Feifei Li
· 1
Hang Yu
· 1
Hanyang Chen
· 1
Han Zhao
· 1
Heng Ji
· 1
Heng Ji
· 1
Heng Ji
· 1
Huan Zhang
· 1
Topics
Vision-Language Models
Video-Language
Benchmarks
Embodied & Agents
Visual QA & Reasoning
Image-Text Retrieval