摘要
Detection of small or distant objects in large-scale scenes remains one of the main challenges for high-precision 3D object detection in autonomous driving. Although multiple sensor fusion has become increasingly common for this task, existing fusion methods still struggle with issues like occlusion and weak feature representation of small or distant objects. To this end, we propose MACF-Net, a cross-modal fusion 3D object detection network suitable for small or distant object. Specifically, we propose an image-guided dynamic sampling strategy to enhance point density in distant regions. We further integrate feature information from images, point clouds coupled with voxels by achieving cross-modal alignment through geometric projection. By employing voxel-based and point-based fusion modules, we achieve cross-modal feature fusion at both voxel and point levels, effectively leveraging the texture details from images and the geometric depth from point clouds. Finally, a multi-feature fused module aggregates the multi-level fused features to produce refined bounding box predictions and confidence scores. The experimental results on the KITTI dataset demonstrate that MACF-Net outperforms existing state-of-the-art methods, highlighting its effectiveness in the detection of small or distant objects.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 981-986 |
| 页数 | 6 |
| 期刊 | International Conference on Electronic Measurement and Instruments |
| 期 | 2025 |
| DOI | |
| 出版状态 | 已出版 - 2025 |
| 活动 | 17th IEEE International Conference on Electronic Measurement and Instruments, ICEMI 2025 - Beijing, 中国 期限: 22 8月 2025 → 24 8月 2025 |
学术指纹
探究 'MACF-Net: Multi-Level Aware Cross-Modal Fusion Network for 3D Object Detection' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver