Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Volker Tresp — most-cited papers & profile · Multimodal
← authors
·
overview
Volker Tresp
67
papers ·
473
citations ·
58
h-index
Siemens (Germany) · Munich Center for Quantum Science and Technology · Ludwig-Maximilians-Universität München
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
OpenDriveVLA: Towards End-to-end Autonomous Driving with Large Vision Language Action Model
2025 · 1 citations
Enhancing Multimodal Compositional Reasoning of Visual Language Models with Generative Negative Mining
2023 · 1 citations
ProcessThinker: Enhancing Multi-modal Large Language Models Reasoning via Rollout-based Process Reward
2026
True Multimodal In-Context Learning Needs Attention to the Visual Context
2025
PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection
2025
Can Vision-Language Models be a Good Guesser? Exploring VLMs for Times and Location Reasoning
2023
LLaVA Steering: Visual Instruction Tuning with 500x Fewer Parameters through Modality Linear Representation-Steering
2024
Perceive, Query & Reason: Enhancing Video QA with Question-Guided Temporal Queries
2024
Top co-authors
Yunpu Ma
· 3
Daniel Cremers
· 2
Jinhe Bi
· 2
Xun Xiao
· 2
Hang Li
· 1
Haokun Chen
· 1
Jianzhe Liu
· 1
Jindong Gu
· 1
Jingpei Wu
· 1
Mang Ye
· 1
Philip Torr
· 1
Shuo Chen
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Video-Language
Instruction Tuning
Embodied & Agents