遥感
计算机科学
机制(生物学)
扩散
目标检测
对象(语法)
人工智能
计算机视觉
模式识别(心理学)
地质学
热力学
认识论
物理
哲学
作者
Chenke Yue,Yin Zhang,Junhua Yan,Zhaolong Luo,Yong Liu,Pengyu Guo
标识
DOI:10.1109/tgrs.2025.3561133
摘要
Multimodal remote sensing images provide complementary information, enhancing the effectiveness of object detection tasks in open-world scenarios. To address the imbalance of information richness between modalities in multimodal object detection, we propose a simple yet effective multi-source image object detection method (DKDNet). Our contributions are twofold: (a) we introduce diffusion deformation convolution (DDConv), which combines deformation convolution with adaptive long-range receptive fields to further enhance the ability to perceive object pose variations and capture distant information. (b) We propose the bidirectional feature distillation and information complementary fusion network (BDFusion), where different modalities exchange information through a knowledge distillation strategy, explicitly enhancing the information interaction between modalities. Finally, we adaptively build spatial domain complementarity between different modalities via self-correction, revealing implicit correlations. Experimental results on the publicly available vehicle detection in aerial imagery (VEDAI) dataset and the optical and SAR ship detection dataset (OSSDD), collected in the Suez Canal region, demonstrate that our proposed method achieves superior performance with acceptable inference time, making it suitable for various realworld scenarios.
科研通智能强力驱动
Strongly Powered by AbleSci AI