模态(人机交互)
判别式
计算机科学
人工智能
模式识别(心理学)
特征(语言学)
鉴定(生物学)
计算机视觉
语言学
植物
生物
哲学
作者
Haiyun Tao,Yukang Zhang,Yang Lu,Hanzi Wang
标识
DOI:10.1007/978-981-99-8546-3_10
摘要
Visible-infrared person re-identification (VI-ReID) is a challenging cross-modality pedestrian retrieval problem. Due to the significant cross-modality discrepancy, it is difficult to learn discriminative features. Attention-based methods have been widely utilized to extract discriminative features for VI-ReID. However, the existing methods are confined by first-order structures that just exploit simple and coarse information. The existing approach lacks the sufficient capability to learn both modality-irrelevant and modality-relevant features. In this paper, we extract the second-order information from mid-level features to complement the first-order cues. Specifically, we design a flexible second-order module, which considers the correlations between the common features and learns refined feature representations for pedestrian images. Additionally, the visible and infrared modality has a significant gap. Therefore, we propose a plug-and-play mixed intermediate modality module to generate intermediate modality representations to reduce the modality discrepancy between the visible and infrared features. Extensive experimental results on two challenging datasets SYSU-MM01 and RegDB demonstrate that our method considerably achieves competitive performance compared to the state-of-the-art methods.
科研通智能强力驱动
Strongly Powered by AbleSci AI