Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Lewei Lu — most-cited papers & profile · Multimodal
← authors
·
overview
Lewei Lu
13
papers ·
2761
citations ·
16
h-index
Group Sense (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
VL-BERT: Pre-training of Generic Visual-Linguistic Representations
2019 · 785 citations
MM-Interleaved: Interleaved Image-Text Generative Modeling via Multi-modal Feature Synchronizer
2024 · 6 citations
Enhancing the Outcome Reward-based RL Training of MLLMs with Self-Consistency Sampling
2025
Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces
2025
Learning 1D Causal Visual Representation with De-focus Attention Networks
2024
HoVLE: Unleashing the Power of Monolithic Vision-Language Models with Holistic Vision-Language Embedding
2024
Top co-authors
Xizhou Zhu
· 4
Jifeng Dai
· 3
Changyao Tian
· 2
Chenxin Tao
· 2
Gao Huang
· 2
Hongsheng Li
· 2
Jie Zhou
· 2
Jifeng Dai
· 2
Shiqian Su
· 2
Wenhai Wang
· 2
Yu Qiao
· 2
Aijun Yang
· 1
Topics
Vision-Language Models
Benchmarks
Visual QA & Reasoning
Image-Text Retrieval
Instruction Tuning
Embodied & Agents
Audio-Visual
Video-Language