Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ziwei Liu — most-cited papers & profile · Speech Audio
← authors
·
overview
Ziwei Liu
34
papers ·
4
citations ·
82
h-index
Nanyang Technological University · Harbin Institute of Technology · Sichuan University · State Key Laboratory of Robotics and Systems
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Octopus: Embodied Vision-Language Programmer from Environmental Feedback
2023 · 3 citations
Invar-RAG: Invariant LLM-aligned Retrieval for Better Generation
2024 · 1 citations
Prisma-World: Camera-Controllable Multi-Agent Video World Model
2026
Continual GUI Agents
2026
Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond
2026
End-to-End Dexterous Grasp Learning from Single-View Point Clouds via a Multi-Object Scene Dataset
2026
Kinema4D: Kinematic 4D World Modeling for Spatiotemporal Embodied Simulation
2026
MonoArt: Progressive Structural Reasoning for Monocular Articulated 3D Reconstruction
2026
BFA++: Hierarchical Best-Feature-Aware Token Prune for Multi-View Vision Language Action Model
2026
The RoboSense Challenge: Sense Anything, Navigate Anywhere, Adapt Across Platforms
2026
Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future
2025
3EED: Ground Everything Everywhere in 3D
2025
PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image
2025
NEZHA: A Zero-sacrifice and Hyperspeed Decoding Architecture for Generative Recommendations
2025
LLM-EDT: Large Language Model Enhanced Cross-domain Sequential Recommendation with Dual-phase Training
2025
Topics
Perception
Manipulation
Control
Multi-Robot
Benchmarks
Vision-Language Models
Human-Robot Interaction
Navigation
Visual QA & Reasoning
Model Architecture