人工智能
计算机科学
计算机视觉
恶劣天气
图像融合
模式识别(心理学)
图像处理
图像分割
融合
传感器融合
图像(数学)
图像配准
特征提取
天气预报
遥感
像素
感知
可视化
信息融合
目标检测
支持向量机
作者
Xilai Li,Huichun Liu,Xiaosong Li,Tao Ye,Zhenyu Kuang,Hui Li
标识
DOI:10.1109/tip.2026.3690324
摘要
Multi-modality image fusion (MMIF) in adverse weather aims to address the loss of visual information caused by weather-related degradations, providing clearer scene representations. Although a few studies have attempted to incorporate textual information to improve semantic perception, they often lack effective categorization and thorough analysis of textual content. To address these limitations, we propose AWM-Fuse, a unified fusion framework that handles diverse weather degradations via global and local text perception with shared parameters. In particular, a global text perception module leverages BLIP-generated captions to extract overall scene features and identify primary degradation types, thus promoting generalization across various adverse weather conditions. Complementing this, the local module employs detailed scene descriptions produced by ChatGPT to concentrate on specific degradation effects through concrete textual cues, enabling the recovery of subtle details. Furthermore, textual descriptions are used to constrain the generation of fused images, effectively steering the network learning process toward better alignment with semantic labels, thereby promoting the learning of more meaningful visual features. To facilitate text-guided fusion under adverse weather, we construct AWMM-Text, a large-scale benchmark providing paired global and local annotations for multi-modality image pairs. Extensive experiments demonstrate that AWM-Fuse consistently outperforms state-of-the-art methods under complex weather conditions and on multiple downstream tasks. Our code is available at https://github.com/Feecuin/AWM-Fuse.
科研通智能强力驱动
Strongly Powered by AbleSci AI