人工智能
计算机视觉
计算机科学
多光谱图像
像素
高光谱成像
保险丝(电气)
图像分辨率
立体视觉
图像配准
图像融合
过程(计算)
融合
图像(数学)
电气工程
工程类
哲学
操作系统
语言学
作者
Yujuan Guo,Xiyou Fu,Meng Xu,Sen Jia
标识
DOI:10.1109/tgrs.2023.3314755
摘要
The necessary prerequisite for effective data fusion is the strict registration of low-resolution hyperspectral images (LR-HSI) and high-resolution multispectral images (HR-MSI). However, registration requires a complex process that takes into account the effects of light, imaging angle, and geometric distortion of the image during acquisition. Therefore, to avoid complex registration, we focused on developing an unregistered HSI and MSI fusion method for pixel shifting, obtaining fused images with high resolution, high signal-to-noise ratio, and feature identifiability. We identified that the unregistered LR-HSI and HR-MSI in the case of pixel shift are very similar to the disparity maps in stereo vision. Inspired by this, we simulate the structure of stereo cameras to propose a stereo cross-attention network (SCANet) to achieve an accurate fusion of unregistered LR-HSI and HR-MSI. Considering the model complexity and computing efficiency, we design a simple and stackable stereo cross-fusion block (SCFBlock) based on a Transformer to simulate the process of light entering the left and right cameras by extracting the abstract features of the images. Moreover, the purpose of cross-convergence fusion self-attention (CCFSA) is to learn cross-complementary attention and collect contextual information in horizontal and vertical directions to fuse unregistered images using multi-directional cross-view information. We have conducted extensive experiments on Pavia University (PaviaU), Chikusei, and PYLake datasets. The results show that SCANet achieves superior or competitive performance in fusing unregistered LR-HSI and HR-MSI in comparison with the other competitors.
科研通智能强力驱动
Strongly Powered by AbleSci AI