高光谱成像
计算机科学
人工智能
编码器
特征提取
模式识别(心理学)
空间分析
计算机视觉
遥感
地理
操作系统
作者
Hao Xu,Zhigang Zeng,Wei Yao,Jiayue Lu
标识
DOI:10.1109/lgrs.2023.3321343
摘要
Compared with general optical images, hyperspectral images (HSIs) contain richer spectral information. On one hand, this provides a sufficient basis for ground object recognition. On the other hand, it results in the intermingling of spatial and spectral information. In order to make better use of the rich spatial and spectral information in HSIs, we resort to Vision Transformer (ViT). To be specific, we propose the Cross Spatial–Spectral Dense Transformer (CS2DT) for spatial-spectral feature extracting and feature fusing. For feature extraction, CS2DT employs the Adaptive Dense Encoder (ADE) module, which enables the extraction of multi-scale semantic information. During the features fusion stage, we use the Cross Spatial–Spectral Attention (CS2A) module based on the cross-attention (CA) operation to better integrate spatial and spectral features. We evaluate the classification performance of the proposed CS2DT on three well-known datasets by conducting extensive experiments. Experimental results demonstrate that CS2DT can achieve higher accuracy and higher stability when compared with the state-of-the-art (SOTA) methods. The source code will be made available at https://github.com/shouhengx/CS2DT.
科研通智能强力驱动
Strongly Powered by AbleSci AI