Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction

Cai, Yuanhao; Lin, Jing; Hu, Xiaowan; Wang, Haoqian; Yuan, Xin; Zhang, Yulun; Timofte, Radu; Van Gool, Luc

Electrical Engineering and Systems Science > Image and Video Processing

arXiv:2111.07910 (eess)

[Submitted on 15 Nov 2021 (v1), last revised 21 Mar 2022 (this version, v2)]

Title:Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction

Authors:Yuanhao Cai, Jing Lin, Xiaowan Hu, Haoqian Wang, Xin Yuan, Yulun Zhang, Radu Timofte, Luc Van Gool

View PDF

Abstract:Hyperspectral image (HSI) reconstruction aims to recover the 3D spatial-spectral signal from a 2D measurement in the coded aperture snapshot spectral imaging (CASSI) system. The HSI representations are highly similar and correlated across the spectral dimension. Modeling the inter-spectra interactions is beneficial for HSI reconstruction. However, existing CNN-based methods show limitations in capturing spectral-wise similarity and long-range dependencies. Besides, the HSI information is modulated by a coded aperture (physical mask) in CASSI. Nonetheless, current algorithms have not fully explored the guidance effect of the mask for HSI restoration. In this paper, we propose a novel framework, Mask-guided Spectral-wise Transformer (MST), for HSI reconstruction. Specifically, we present a Spectral-wise Multi-head Self-Attention (S-MSA) that treats each spectral feature as a token and calculates self-attention along the spectral dimension. In addition, we customize a Mask-guided Mechanism (MM) that directs S-MSA to pay attention to spatial regions with high-fidelity spectral representations. Extensive experiments show that our MST significantly outperforms state-of-the-art (SOTA) methods on simulation and real HSI datasets while requiring dramatically cheaper computational and memory costs. Code and pre-trained models are available at this https URL

Comments:	CVPR 2022; The first Transformer-based method for snapshot compressive imaging
Subjects:	Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2111.07910 [eess.IV]
	(or arXiv:2111.07910v2 [eess.IV] for this version)
	https://doi.org/10.48550/arXiv.2111.07910

Submission history

From: Yuanhao Cai [view email]
[v1] Mon, 15 Nov 2021 16:59:48 UTC (14,820 KB)
[v2] Mon, 21 Mar 2022 12:56:00 UTC (2,408 KB)

Electrical Engineering and Systems Science > Image and Video Processing

Title:Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science > Image and Video Processing

Title:Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators