摘要
Deep learning method for 6D object pose estimation based on RGB image and depth (RGB-D) has been successfully applied to robot grasping. The fusion of RGB and depth is one of the most important difficulties. Previous works on the fusion of these two features are mostly concatenated together without considering the different contributions of the two types of features to pose estimation. We propose a selective embedding with gated fusion structure called SEGate, which can adjust the weights of RGB and depth features adaptively. Furthermore, we aggregate the local features of point clouds according to the distance between them. More specifically, the close point clouds contribute a lot to local features, while the distant point clouds contribute a little. Experiments show that our approach achieves the state-of-art performance in both LineMOD and YCB-Video datasets. Meanwhile, our approach is more robust to the pose estimation of occluded objects.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 2417-2436 |
| 页数 | 20 |
| 期刊 | Neural Processing Letters |
| 卷 | 51 |
| 期 | 3 |
| DOI | |
| 出版状态 | 已出版 - 1 6月 2020 |
学术指纹
探究 'Selective Embedding with Gated Fusion for 6D Object Pose Estimation' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver