Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuhang Cao — most-cited papers & profile · Speech Audio
← authors
·
overview
Yuhang Cao
11
papers ·
5
citations ·
12
h-index
Jilin Agricultural University · Tsinghua University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Intern-S1: A Scientific Multimodal Foundation Model
2025 · 5 citations
Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale
2026
CapRL: Stimulating Dense Image Caption Capabilities via Reinforcement Learning
2025
2nd Place Report of MOSEv2 Challenge 2025: Concept Guided Video Object Segmentation via SeC
2025
Intern-S1: A Scientific Multimodal Foundation Model
2025
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning
2025
Advancing Complex Video Object Segmentation via Progressive Concept Construction
2025
BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning
2025
Visual Agentic Reinforcement Fine-Tuning
2025
MM-IFEngine: Towards Multimodal Instruction Following
2025
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
2024
Topics
Model-Based RL
RLHF & Alignment
Chemistry
Materials
Segmentation
Video Understanding
Offline RL
Training Techniques
Drug Discovery
Protein Science