Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yin Cui — most-cited papers & profile · Multimodal
← authors
·
overview
Yin Cui
11
papers ·
84
citations ·
20
h-index
Guizhou Normal University · Xuzhou Medical College · Nanjing Drum Tower Hospital · Yanbian University Hospital
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models
2022 · 37 citations
NFT: Bridging Supervised Learning and Reinforcement Learning in Math Reasoning
2025 · 28 citations
Towards a Unified Foundation Model: Jointly Pre-Training Transformers on Unpaired Images and Text
2021 · 9 citations
Exploring Temporal Granularity in Self-Supervised Video Representation Learning
2021 · 4 citations
Simple Copy-Paste is a Strong Data Augmentation Method for Instance Segmentation
2020 · 3 citations
Edify 3D: Scalable High-Quality 3D Asset Generation
2024 · 3 citations
Bridging the Gap Between Object Detection and User Intent via Query-Modulation
2021
Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models
2024
Visual Fact Checker: Enabling High-Fidelity Detailed Caption Generation
2024
Topics
Object Detection
Visual Language
Diffusion Models
Audio Generation
3D & NeRF Generation
RLHF & Alignment
Policy Gradient
Value-Based
Segmentation
Video Understanding