光学
图像融合
模态(人机交互)
红外线的
融合
图像配准
计算机视觉
计算机科学
人工智能
物理
语言学
图像(数学)
哲学
作者
Wenqu Zhao,Lingxue Wang,Lian Zhang,Dezhi Zheng,Yi Cai
出处
期刊:Optics Letters
[Optica Publishing Group]
日期:2025-05-16
卷期号:50 (12): 3907-3907
摘要
In this Letter, we propose CLIP-guided multimodal registration and fusion (CGMRF), a semantic understanding-based multimodal image fusion system, for visible and infrared (IR) dual-modality imaging. CGMRF leverages semantic similarity, better aligned with human visual interpretation, to address the challenges of multimodal image registration and fusion. Experimental results across multiple metrics demonstrate the advantages of the proposed CGMRF system.
科研通智能强力驱动
Strongly Powered by AbleSci AI