Awesome Computer Vision
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Gengyuan Zhang — most-cited papers & profile · Computer Vision
← authors
·
overview
Gengyuan Zhang
13
papers ·
67
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
A Systematic Survey of Prompt Engineering on Vision-Language Foundation Models
2023 · 64 citations
CL-CrossVQA: A Continual Learning Benchmark for Cross-Domain Visual Question Answering
2022 · 2 citations
VideoINSTA: Zero-shot Long Video Understanding via Informative Spatial-Temporal Reasoning with LLMs
2024 · 1 citations
Can Vision-Language Models be a Good Guesser? Exploring VLMs for Times and Location Reasoning
2023
Multi-event Video-Text Retrieval
2023
SPOT! Revisiting Video-Language Models for Event Understanding
2023
Perceive, Query & Reason: Enhancing Video QA with Question-Guided Temporal Queries
2024
Top co-authors
Volker Tresp
· 5
Jindong Gu
· 4
Ruotong Liao
· 2
Ahmad Beirami
· 1
Ahmed Frikha
· 1
Bailan He
· 1
Denis Krompass
· 1
Guangyao Zhai
· 1
Haokun Chen
· 1
Huiyu Wang
· 1
Jinhe Bi
· 1
Jisen Ren
· 1
Topics
Visual Language
3D Vision
Image Generation