Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Lu Sheng — most-cited papers & profile · Multimodal
← authors
·
overview
Lu Sheng
17
papers ·
36
citations ·
29
h-index
Beihang University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CAMP: Cross-Modal Adaptive Message Passing for Text-Image Retrieval
2019 · 31 citations
Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic Guidance
2025 · 2 citations
WorldSimBench: Towards Video Generation Models as World Simulators
2024 · 2 citations
MP5: A Multi-modal Open-ended Embodied System in Minecraft via Active Perception
2023 · 1 citations
RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics
2025
TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics
2025
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
2025
Effective LLM Code Refinement via Property-Oriented and Structurally Minimal Feedback
2025
Use Property-Based Testing to Bridge LLM Code Generation and Validation
2025
Use Property-Based Testing to Bridge LLM Code Generation and Validation
2025
Visibility Constrained Generative Model for Depth-based 3D Facial Pose Tracking
2019
Siamese DETR
2023
RH20T-P: A Primitive-Level Robotic Dataset Towards Composable Generalization Agents
2024
Self-Supervised Monocular Depth Estimation in the Dark: Towards Data Distribution Compensation
2024
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
2024
Topics
Perception
Manipulation
Control
Human-Robot Interaction
Multi-Robot
Testing
Evaluation
3D Vision
Image Generation
Image Restoration