TY - JOUR
T1 - A Knowledge Transfer Method for Unsupervised Pose Keypoint Detection Based on Domain Adaptation and CAD Models
AU - Du, Fuzhou
AU - Kong, Feifei
AU - Zhao, Delong
N1 - Publisher Copyright:
© 2022 The Authors. Advanced Intelligent Systems published by Wiley-VCH GmbH.
PY - 2023/2
Y1 - 2023/2
N2 - Vision-based pose estimation is a basic task in many industrial fields such as bin-picking, autonomous assembly, and augmented reality. One of the most commonly used pose estimation methods first detects the 2D pose keypoints in the input image and then calculates the 6D pose using a pose solver. Recently, deep learning is widely used in pose keypoint detection and performs excellent accuracy and adaptability. However, its over-reliance on sufficient and high-quality samples and supervision is prominent, particularly in the industrial field, leading to high data cost. Based on domain adaptation and computer-aided-design (CAD) models, herein, a virtual-to-real knowledge transfer method for pose keypoint detection to reduce the data cost of deep learning is proposed. To address the disorder of knowledge flow, a viewpoint-driven feature alignment strategy is proposed to simultaneously eliminate interdomain differences and preserve intradomain differences. The shape invariance of rigid objects is then introduced as constraints to address the large assumption space problem in the regressive domain adaptation. The multidimensional experimental results demonstrate the superiority of the method. Without real annotations, the normalized pixel error of keypoint detection is reported as 0.033, and the proportion of pixel errors lower than 0.05 is up to 92.77%.
AB - Vision-based pose estimation is a basic task in many industrial fields such as bin-picking, autonomous assembly, and augmented reality. One of the most commonly used pose estimation methods first detects the 2D pose keypoints in the input image and then calculates the 6D pose using a pose solver. Recently, deep learning is widely used in pose keypoint detection and performs excellent accuracy and adaptability. However, its over-reliance on sufficient and high-quality samples and supervision is prominent, particularly in the industrial field, leading to high data cost. Based on domain adaptation and computer-aided-design (CAD) models, herein, a virtual-to-real knowledge transfer method for pose keypoint detection to reduce the data cost of deep learning is proposed. To address the disorder of knowledge flow, a viewpoint-driven feature alignment strategy is proposed to simultaneously eliminate interdomain differences and preserve intradomain differences. The shape invariance of rigid objects is then introduced as constraints to address the large assumption space problem in the regressive domain adaptation. The multidimensional experimental results demonstrate the superiority of the method. Without real annotations, the normalized pixel error of keypoint detection is reported as 0.033, and the proportion of pixel errors lower than 0.05 is up to 92.77%.
KW - domain adaptations
KW - knowledge transfers
KW - pose keypoint detections
KW - shape self-constraints
KW - viewpoint-driven alignments
UR - https://www.scopus.com/pages/publications/85165813862
U2 - 10.1002/aisy.202200214
DO - 10.1002/aisy.202200214
M3 - 文章
AN - SCOPUS:85165813862
SN - 2640-4567
VL - 5
JO - Advanced Intelligent Systems
JF - Advanced Intelligent Systems
IS - 2
M1 - 2200214
ER -