TY - JOUR
T1 - Self-Supervised deep homography estimation with invertibility constraints
AU - Wang, Chen
AU - Wang, Xiang
AU - Bai, Xiao
AU - Liu, Yun
AU - Zhou, Jun
N1 - Publisher Copyright:
© 2019 Elsevier B.V.
PY - 2019/12/1
Y1 - 2019/12/1
N2 - Remarkable performance of the homography estimation has been achieved by the deep CNN based approaches. These homography estimation methods, more often than not, are supervised methods and rely too much on the ground truth annotations as they aim to learn the mapping between image pairs and homography. On the other hand, the inherent invertibility of homography is helpful to avoid over-fitting and improve the performance, which however is ignored by previous homography estimation methods. In this paper, we propose a novel homography estimation approach, named “Self-Supervised Regression Network(SSR-Net)”, which relaxes the need of ground truth annotations and takes advantage of invertibility constraints. We utilize spatial pyramid pooling modules to improve the quality of extracted features in each image by exploiting context information. To employ the invertibility constraints, we adopt the matrix representation of the homography rather than the commonly used 4-point parameterization in other methods. Our proposed SSR-Net produce homography matrices and synthetic images in a cycled way. The network are trained in a self-supervised way by minimizing the combination of photometric loss and invertibility loss. Experiments on the synthetic dataset generated from MSCOCO dataset show that our proposed method outperforms several state-of-the-art approaches.
AB - Remarkable performance of the homography estimation has been achieved by the deep CNN based approaches. These homography estimation methods, more often than not, are supervised methods and rely too much on the ground truth annotations as they aim to learn the mapping between image pairs and homography. On the other hand, the inherent invertibility of homography is helpful to avoid over-fitting and improve the performance, which however is ignored by previous homography estimation methods. In this paper, we propose a novel homography estimation approach, named “Self-Supervised Regression Network(SSR-Net)”, which relaxes the need of ground truth annotations and takes advantage of invertibility constraints. We utilize spatial pyramid pooling modules to improve the quality of extracted features in each image by exploiting context information. To employ the invertibility constraints, we adopt the matrix representation of the homography rather than the commonly used 4-point parameterization in other methods. Our proposed SSR-Net produce homography matrices and synthetic images in a cycled way. The network are trained in a self-supervised way by minimizing the combination of photometric loss and invertibility loss. Experiments on the synthetic dataset generated from MSCOCO dataset show that our proposed method outperforms several state-of-the-art approaches.
KW - Homography estimation
KW - Invertibility constraint
KW - Self-Supervised deep learning
KW - Spatial pyramid pooling
UR - https://www.scopus.com/pages/publications/85072863609
U2 - 10.1016/j.patrec.2019.09.021
DO - 10.1016/j.patrec.2019.09.021
M3 - 文章
AN - SCOPUS:85072863609
SN - 0167-8655
VL - 128
SP - 355
EP - 360
JO - Pattern Recognition Letters
JF - Pattern Recognition Letters
ER -