Mmf-track: Multi-modal Multi-level Fusion For 3D Single Object Tracking
2023 Β· Zhiheng Li, Yubo Cui, Yu Lin, et al.
Abstract
3D single object tracking plays a crucial role in computer vision. Mainstream methods mainly rely on point clouds to achieve geometry matching between target template and search area. However, textureless and incomplete point clouds make it difficult for single-modal trackers to distinguish objects with similar structures. To overcome the limitations of geometry matching, we propose a Multi-modal Multi-level Fusion Tracker (MMF-Track), which exploits the image texture and geometry characteristic of point clouds to track 3D target. Specifically, we first propose a Space Alignment Module (SAM) to align RGB images with point clouds in 3D space, which is the prerequisite for constructing inter-modal associations. Then, in feature interaction level, we design a Feature Interaction Module (FIM) based on dual-stream structure, which enhances intra-modal features in parallel and constructs inter-modal semantic associations. Meanwhile, in order to refine each modal feature, we introduce a Coars
Authors
(none)
Tags
Stats
Related papers
- MFST: Multi-features Siamese Tracker (2021)0.95
- Locality Aware Appearance Metric For Multi-target Multi-camera Tracking (2019)0.00
- Smiletrack: Similarity Learning For Occlusion-aware Multiple Object Tracking (2022)17.36
- Enhanced Cross-modal 3D Retrieval Via Tri-modal Reconstruction (2025)0.00
- Quasi-dense Similarity Learning For Multiple Object Tracking (2020)19.58
- Latformer: Locality-aware Point-view Fusion Transformer For 3D Shape Recognition (2021)6.34
- Qdtrack: Quasi-dense Similarity Learning For Appearance-only Multiple Object Tracking (2022)15.75
- Sca-pvnet: Self-and-cross Attention Based Aggregation Of Point Cloud And Multi-view For 3D Object Retrieval (2023)10.07