Skip to main navigation Skip to search Skip to main content

Image segmentation-driven sim-to-real deep reinforcement learning framework for accurate peg-in-hole assembly

  • Ning Zhang
  • , Yongjia Zhao*
  • , Minghao Yang
  • , Shuling Dai
  • *Corresponding author for this work
  • Beihang University
  • University of Chinese Academy of Sciences

Research output: Contribution to journalArticlepeer-review

Abstract

The automation of assembly operations with industrial robots is pivotal in modern manufacturing, particularly for multispecies, low-volume, and customized production. Traditional programing methods are time-consuming and lack adaptability to complex, variable environments. Reinforcement learning-based assembly tasks have shown success in simulation environments, but face challenges like the simulation-to-reality gap and safety concerns when transferred to real-world applications. This article addresses these challenges by proposing a low-cost, image-segmentation-driven deep reinforcement learning strategy tailored for insertion tasks, such as the assembly of peg-in-hole components in satellite manufacturing, which involve extensive contact interactions. Our approach integrates visual and forces feedback into a prior dueling deep Q-network for insertion skill learning, enabling precise alignment of components. To bridge the simulation-to-reality gap, we transform the raw image input space into a canonical space based on image segmentation. Specifically, we employ a segmentation model based on U-net, pretrained in simulation and fine-tuned with real-world data, significantly reducing the need for labor-intensive real image segment labels. To handle the frequent contact inherent in peg-in-hole tasks, we integrated safety protections and impedance control into the training process, providing active compliance and reducing the risk of assembly failures. Our approach was evaluated in both simulated and real robotic environments, demonstrating robust performance in handling camera position errors and varying ambient light intensities and different lighting colors. Finally, the algorithm was validated in a real satellite assembly scenario, achieving a success rate of 15 out of 20 tests.

Original languageEnglish
Pages (from-to)2783-2802
Number of pages20
JournalRobotica
Volume43
Issue number8
DOIs
StatePublished - 1 Aug 2025

Keywords

  • image segmentation
  • impedance control
  • Peg-in-Hole (PiH)
  • reinforcement learning (RL)
  • sim-to-real

Fingerprint

Dive into the research topics of 'Image segmentation-driven sim-to-real deep reinforcement learning framework for accurate peg-in-hole assembly'. Together they form a unique fingerprint.

Cite this