Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Lidong Bing — most-cited papers & profile · Multimodal
← authors
·
overview
Lidong Bing
18
papers ·
47
citations ·
46
h-index
UF Health Shands Hospital · Shanda Games
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
2023 · 18 citations
Mitigating Object Hallucinations in Large Vision-Language Models through Visual Contrastive Decoding
2023 · 2 citations
Longvt: Incentivizing "thinking With Long Videos" Via Native Tool Calling
2025
ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition Benchmark
2025
The Curse of Multi-Modalities: Evaluating Hallucinations of Large Multimodal Models across Language, Visual, and Audio
2024
Top co-authors
Shijian Lu
· 3
Chunyan Miao
· 2
Sicong Leng
· 2
Boqiang Zhang
· 1
Deli Zhao
· 1
Guanzheng Chen
· 1
Hang Zhang
· 1
Hang Zhang
· 1
Hang Zhang
· 1
K. Zhang
· 1
Liuyi Wang
· 1
Long Li
· 1
Topics
Video-Language
Benchmarks
Vision-Language Models
Visual QA & Reasoning
Audio-Visual
Embodied & Agents
Instruction Tuning