计算机科学
棱锥(几何)
特征(语言学)
语义特征
对象(语法)
人工智能
保险丝(电气)
目标检测
代表(政治)
模式识别(心理学)
特征提取
比例(比率)
利用
语义学(计算机科学)
计算机视觉
数学
工程类
计算机安全
政治学
法学
程序设计语言
语言学
几何学
哲学
物理
电气工程
政治
量子力学
作者
Yuqi Chen,Xiangbin Zhu,Yonggang Li,Yuanwang Wei,Lihua Ye
标识
DOI:10.1016/j.image.2023.116919
摘要
Feature-pyramid network-based models, which progressively fuse multi-scale features, have been proven highly effective in object detection. However, these models often learn multi-scale features with ambiguous boundaries, due to small objects with only a few pixels that easily lose information during top-down propagation, which makes multi-scale feature representation less effective. In this work, we propose an efficient Enhanced Semantic Feature Pyramid Network(ES-FPN), which combines semantic information at high-level with contextual information at low-level to improve multi-scale feature learning in small object detection. Specifically, the proposed network first exploits the rich semantic information in lateral connections that enables the features to be more semantic. Then, it excavates the lost information in high-level/low-res feature maps with rich contextual information in low-level/high-res. In this way, the high-level layers suffer the reduced loss of important contextual information during the progressive feature fusion that avoids object disappearance, which is useful to utilize rich semantic information in high-level. Finally, ES-FPN fuses the distributed features of each layer stage-by-stage and the final features are more semantically and better for localizing the object. Extensive experimental results over three widely used object detection benchmarks(MS COCO, VOC and Cityscapes) demonstrate that our network can accurately locate fairly complete objects with clear boundaries and outperforms previous feature pyramid-based methods.
科研通智能强力驱动
Strongly Powered by AbleSci AI