Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhanyu Ma — most-cited papers & profile · Multimodal
← authors
·
overview
Zhanyu Ma
27
papers ·
476
citations ·
46
h-index
Beijing University of Posts and Telecommunications
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Multimodal Conditional Information Bottleneck For Generalizable Ai-generated Image Detection
2025 · 1 citations
Zero-Shot Audio Captioning Using Soft and Hard Prompts
2024 · 1 citations
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions
2026
OVC-Net: Object-Oriented Video Captioning with Temporal Graph and Detail Enhancement
2020
CMF: Cascaded Multi-model Fusion for Referring Image Segmentation
2021
Top co-authors
Bofu Yu
· 1
Dongliang Chang
· 1
Fangyi Zhu
· 1
Guang Chen
· 1
Haohe Liu
· 1
Haolong Yan
· 1
Haotian Qin
· 1
Hongbing Li
· 1
Jenq-Neng Hwang
· 1
Jianhua Yang
· 1
Jun Guo
· 1
Kongming Liang
· 1
Topics
Vision-Language Models
Image-Text Retrieval
Benchmarks
Video-Language
Audio-Visual
eess.AS