跳到主要导航 跳到搜索 跳到主要内容

Supervised-unsupervised combined transformer for spectral compressive imaging reconstruction

  • Han Zhou
  • , Yusheng Lian*
  • , Jin Li*
  • , Zilong Liu
  • , Xuheng Cao
  • , Chao Ma
  • *此作品的通讯作者
  • Beijing Institute of Graphic Communication
  • National Institute of Metrology China
  • Tongji University

科研成果: 期刊稿件文章同行评审

摘要

To solve the low spatial and/or temporal resolution problem which the conventional hyperspectral cameras often suffer from, spectral compressive imaging systems (SCI) have attracted more attention recently. Recovering a hyperspectral image (HSI) from its corresponding 2D coded image is an ill-posed inverse problem, and learning accurate prior from HSI and 2D coded image is essential to solve this inverse problem. Existing methods only use supervised networks that focus on learning generalized prior from training datasets, or only use unsupervised networks that focus on learning specific prior from 2D coded image, resulting in the inability to learn both generalized and specific priors. Also, when learning the priors, existing methods cannot simultaneously give consideration to both global and local scales, as well as both spatial and spectral dimensions. To cope with this problem, in this paper, we propose a Supervised-Unsupervised Combined Transformer Network (SUCTNet) composed by a supervised Spatio-spectral Transformer network (SSTNet) and an Unsupervised Multi-level Feature Refinement network (UMFRNet). Specifically, we first develop the SSTNet to learn generalized prior and obtain a preliminary HSI. In SSTNet, the proposed spatial encoding and spectral decoding network architecture enables it to simultaneously consider both spatial and spectral dimensions, and a proposed Global and Local Multi head Self Attention block (GL-MSA) enables it simultaneously to consider both global and local scales. Then, the preliminary HSI is fed into the proposed UMFRNet to learn specific prior and obtain the target HSI. In UMFRNet, a proposed multi-level feature refinement mechanism and the physical imaging model of SCI are used to improve reconstruction accuracy and generalization performance. Extensive experiments show that our method significantly outperforms state-of-the-art (SOTA) methods on simulated and real datasets. Codes will be available at https://github.com/Vzhouhan/SUCTNet.

源语言英语
文章编号108030
期刊Optics and Lasers in Engineering
175
DOI
出版状态已出版 - 4月 2024

学术指纹

探究 'Supervised-unsupervised combined transformer for spectral compressive imaging reconstruction' 的科研主题。它们共同构成独一无二的学术指纹。

引用此