Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jiaolong Yang — most-cited papers & profile · Speech Audio
← authors
·
overview
Jiaolong Yang
15
papers ·
6
citations ·
31
h-index
Politecnico di Torino · Microsoft Research Asia (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
2024 · 5 citations
Fast, Accurate Thin-Structure Obstacle Detection for Autonomous Mobile Robots
2017 · 1 citations
Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale
2026
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation
2026
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation
2026
Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer
2026
Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer
2026
TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning
2026
From Human Videos to Robot Manipulation: A Survey on Scalable Vision-Language-Action Learning with Human-Centric Data
2026
ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
2026
HiMoE-VLA: Hierarchical Mixture-of-Experts for Generalist Vision-Language-Action Policies
2025
Seeing Across Views: Benchmarking Spatial Reasoning of Vision-Language Models in Robotic Scenes
2025
Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis
2025
Every Sample Matters: Leveraging Mixture-of-Experts and High-Quality Data for Efficient and Accurate Code LLM
2025
Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer
2026
Topics
Manipulation
Human-Robot Interaction
Control
Perception
Planning
Multi-Robot
Vision-Language Models
Embodied & Agents
Sim-to-Real
Code Agents