Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ying Shan — most-cited papers & profile · Speech Audio
← authors
·
overview
Ying Shan
9
papers ·
20
citations ·
51
h-index
Tencent (China) · Jinhua University of Vocational Technology · Zhejiang University of Finance and Economics
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Recurrent Binary Embedding For Gpu-enabled Exhaustive Retrieval From Billion-scale Semantic Vectors
2018 · 14 citations
GPT4Tools: Teaching Large Language Model to Use Tools via Self-instruction
2023 · 6 citations
DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
2025
GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning
2025
EF-VI: Enhancing End-Frame Injection for Video Inbetweening
2025
EF-VI: Enhancing End-Frame Injection for Video Inbetweening
2025
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
2025
Moto: Latent Motion Token as the Bridging Language for Learning Robot Manipulation from Videos
2024
LLaMA Pro: Progressive LLaMA with Block Expansion
2024
Topics
Video Understanding
Diffusion Models
Conditioning & Control
Model Architecture
Code
Fine-Tuning
Image Retrieval
3D Vision
Model-Based RL
RLHF & Alignment