Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Qi Dai — most-cited papers & profile · Large Language Models
← authors
·
overview
Qi Dai
14
papers ·
121
citations ·
25
h-index
Zhejiang Sci-Tech University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Self-Supervised Learning with Swin Transformers
2021 · 113 citations
Self-supervised Object Motion and Depth Estimation from Video
2019 · 3 citations
VIDiff: Translating Videos via Multi-Modal Instructions with Diffusion Models
2023 · 2 citations
VIDiff: Translating Videos via Multi-Modal Instructions with Diffusion Models
2023 · 2 citations
Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions
2024 · 1 citations
Decoupling Language Guidance from Backbones for Text-Guided Medical Segmentation
2026
Language-Conditioned World Modeling for Visual Navigation
2026
Towards Unified World Models for Visual Navigation via Memory-Augmented Planning and Foresight
2025
PACR: Progressively Ascending Confidence Reward for LLM Reasoning
2025
Viarl: Adaptive Temporal Grounding Via Visual Iterated Amplification Reinforcement Learning
2025
JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion Transformers
2025
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning
2025
Top co-authors
Aoqi Wu
· 1
Chong Luo
· 1
Dongdong Chen
· 1
Liang Hu
· 1
Lili Qiu
· 1
Weiquan Huang
· 1
Xiyang Dai
· 1
Xufang Luo
· 1
Yifan Yang
· 1
Yuqing Yang
· 1
Topics
Segmentation
Navigation
Perception
Control
Image Generation
3D Vision
Object Detection
Video Understanding
Visual Language
Medical Imaging