Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hongxia Yang — most-cited papers & profile · Multimodal
← authors
·
overview
Hongxia Yang
48
papers ·
1828
citations ·
40
h-index
Southwest University of Science and Technology · Hong Kong Polytechnic University · Shaoxing University · Wuhan University · Shenzhen Goodix Technology (China) · Renmin Hospital of Wuhan University · Sichuan University of Science and Engineering
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework
2022 · 258 citations
InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
2020 · 56 citations
Video-Teller: Enhancing Cross-Modal Generation with Fusion and Decoupling
2023 · 1 citations
InfiMM-Eval: Complex Open-Ended Reasoning Evaluation For Multi-Modal Large Language Models
2023 · 1 citations
Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models
2025
Revisiting Multimodal Representation in Contrastive Learning: From Patch and Token Embeddings to Finite Discrete Tokens
2023
Law of Vision Representation in MLLMs
2024
Top co-authors
An Yang
· 2
Bohan Zhai
· 2
Jianbo Yuan
· 2
Jingren Zhou
· 2
Junyang Lin
· 2
Quanzeng You
· 2
Chang Zhou
· 1
Chenfeng Xu
· 1
Congkai Xie
· 1
Dimitris N. Metaxas
· 1
Ding Zhou
· 1
Fei Wu
· 1
Topics
Vision-Language Models
Benchmarks
Visual QA & Reasoning
Video-Language
Image-Text Retrieval
Instruction Tuning
Audio-Visual