融合
红外线的
图像融合
扩散
计算机视觉
人工智能
光学
材料科学
图像处理
传感器融合
计算机科学
可见光谱
目标检测
模式识别(心理学)
特征提取
光学滤波器
作者
Kexin Huang,Zhiyuan Zhang,Chaohua Shi,Lan Luo,Junpeng Shi,Yongxiang Liu
标识
DOI:10.1109/tip.2026.3671618
摘要
Infrared and visible image fusion methods have shown promising results, yet existing approaches either compromise downstream detection performance through independent fusion processes or sacrifice computational efficiency and flexibility by requiring joint training of fusion and detection models. To address these challenges, we propose a detection-driven image fusion network based on diffusion models (termed as DDIF), which optimizes the fused images specifically for object detection tasks. Our method features the following three aspects: 1) we reformulate the image fusion process as an inverse problem solved by a non-differentiable optimization process wherein the fused result preserves the source modality information while conforming to the image prior provided by the diffusion model; 2) we design a Response Guide Learning Module (RGLM) to learn response maps, which determine the contribution of each modality in the fusion process according to the downstream detection task; 3) we establish explicit gradient relationships to ensure compatibility between RGLM training and the non-differentiable optimization process, enabling end-to-end training. Notably, a moderate coupling mechanism is formed in our framework as the subsequent detection model is pre-trained and frozen, enabling flexible integration with various advanced detection networks while maintaining computational efficiency. Extensive experiments indicate that our method achieves superior detection performance compared to SOTA approaches and produces high-quality image fusion results.
科研通智能强力驱动
Strongly Powered by AbleSci AI