Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction


METADATA ONLY
Loading...

Date

2022

Publication Type

Conference Paper

ETH Bibliography

yes

Citations

Altmetric
METADATA ONLY

Data

Rights / License

Abstract

Hyperspectral image (HSI) reconstruction aims to recover the 3D spatial-spectral signal from a 2D measurement in the coded aperture snapshot spectral imaging (CASSI) system. The HSI representations are highly similar and correlated across the spectral dimension. Modeling the inter-spectra interactions is beneficial for HSI reconstruction. However, existing CNN-based methods show limitations in capturing spectral-wise similarity and long-range dependencies. Besides, the HSI information is modulated by a coded aperture (physical mask) in CASSI. Nonetheless, current algorithms have not fully explored the guidance effect of the mask for HSI restoration. In this papa; we propose a novel framework, Mask-guided Spectral-wise Transformer (MST), for HSI reconstruction. Specifically, we present a Spectral-wise Multi-head Self-Attention (S-MSA) that treats each spectral feature as a token and calculates self-attention along the spectral dimension. In addition, we customize a Mask-guided Mechanism (MM) that directs S-MSA to pay attention to spatial regions with high-fidelity spectral representations. Extensive experiments show that our MST significantly outperforms state-of-the-art (SOTA) methods on simulation and real HSI datasets while requiring dramatically cheaper computational and memory costs. https://github.com/caiyuanhao1998/MST/

Publication status

published

Editor

Book title

2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Journal / series

Volume

Pages / Article No.

17481 - 17490

Publisher

IEEE

Event

2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2022)

Edition / version

Methods

Software

Geographic location

Date collected

Date created

Subject

Low-level vision; Computational photography; Photography; Three-dimensional displays; Computational modeling; Memory management; Apertures; Transformers; Pattern recognition

Organisational unit

03514 - Van Gool, Luc (emeritus) / Van Gool, Luc (emeritus) check_circle

Notes

Funding

Related publications and datasets