Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Daquan Zhou — most-cited papers & profile · AI for Code
← authors
·
overview
Daquan Zhou
17
papers ·
267
citations ·
21
h-index
National University of Singapore · Peking University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
All Tokens Matter: Token Labeling for Training Better Vision Transformers
2021 · 142 citations
MagicVideo: Efficient Video Generation With Latent Diffusion Models
2022 · 63 citations
Understanding The Robustness in Vision Transformers
2022 · 34 citations
Shunted Self-Attention via Multi-Scale Token Aggregation
2021 · 10 citations
StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
2024 · 7 citations
DiffFit: Unlocking Transferability of Large Diffusion Models via Simple Parameter-Efficient Fine-Tuning
2023 · 5 citations
Low-Resolution Self-Attention for Semantic Segmentation
2023 · 2 citations
MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation
2024 · 2 citations
High-Quality Mask Tuning Matters for Open-Vocabulary Segmentation
2024 · 1 citations
LVD-2M: A Long-take Video Dataset with Temporally Dense Captions
2024 · 1 citations
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation
2026
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation
2026
MaskDiffusion: Boosting Text-to-Image Consistency with Conditional Mask
2023
BuboGPT: Enabling Visual Grounding in Multi-Modal LLMs
2023
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
2024
Top co-authors
Enze Xie
· 1
Haozhe Liu
· 1
Jincheng Yu
· 1
Jingyu Xin
· 1
Junsong Chen
· 1
Ping Luo
· 1
Shuchen Xue
· 1
Song Han
· 1
Tian Ye
· 1
Yitong Li
· 1
Yuyang Zhao
· 1
Zhangjie Wu
· 1
Topics
Segmentation
Diffusion Models
Text-to-Video
Audio Generation
Text-to-Image
Model Architecture
Training Techniques
Video Understanding
3D Vision
Object Detection