计算机科学
目标检测
一致性(知识库)
人工智能
特征提取
关系(数据库)
对象(语法)
探测器
计算机视觉
视觉对象识别的认知神经科学
变压器
遥感
遥感应用
模式识别(心理学)
特征(语言学)
可视化
空间关系
任务分析
上下文图像分类
像素
编码(集合论)
数据挖掘
组分(热力学)
特征学习
空间分析
深度学习
变更检测
情报检索
作者
Peng Sun,Yongbin Zheng,Wanying Xu,Jian Li,Jiansong Yang
标识
DOI:10.1109/tip.2025.3648164
摘要
Recent studies in remote sensing object detection have made excellent progress and shown promising performance. However, most current detectors only explore rotation-invariant feature extraction but disregard the valuable spatial and semantic prior knowledge in remote sensing images (RSIs), which limits the detection performance when encountering blurred or heavy occluded objects. To address this issue, we propose a mask-reconstruction relation learning (MRRL) framework to learn such prior knowledge among objects and a consistency-reasoning transformer over relation proposals (CTRP) to recognize objects with limited visual features via consistency reasoning. Specifically, MRRL framework applies random mask to some objects in the training dataset and performs masked objects reconstruction to guide the network to learn the distribution consistency of objects. CTRP is the core component of the MRRL framework, which models the interaction between spatial and semantic priors, and uses easy detected objects to reason hard detected objects. The trained CTRP can be integrated into the existing detector to improve the ability of object detection with limited visual features in RSIs. Extensive experiments on widely-used datasets for two distinct tasks, namely remote sensing object detection task and occluded object detection task, demonstrate the effectiveness of the proposed method. Source code is available at https://github.com/sunpeng96/CTRP_mmrotate.
科研通智能强力驱动
Strongly Powered by AbleSci AI