已入深夜,您辛苦了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!祝你早点完成任务,早点休息,好梦!

Mining Cross-Modality Implicit Semantic Association for Unsupervised Visible-Infrared Person Re-Identification

计算机科学 人工智能 特征学习 语义计算 自然语言处理 卷积神经网络 联想(心理学) 图形 语义学(计算机科学) 语义特征 特征(语言学) 一般化 语义记忆 模态(人机交互) 钥匙(锁) 语义相似性 光学(聚焦) 深度学习 代表(政治) 语义压缩 特征选择 特征向量 模式 语义整合 显式语义分析 语义网络 情报检索 语义映射 监督学习 无监督学习 机器学习 语义空间 语义网格 空格(标点符号) 语义数据模型 特征提取
作者
Bin Yang,Lekai Liu,Wenke Huang,Xiao Wang,Bo Du,Mang Ye
出处
期刊:IEEE Transactions on Information Forensics and Security [Institute of Electrical and Electronics Engineers]
卷期号:21: 697-709
标识
DOI:10.1109/tifs.2025.3645635
摘要

Unsupervised visible-infrared person reidentification (US-VI-ReID) seeks to learn a cross-modality retrieval model without relying on manual annotations, thereby reducing the high cost associated with labeling. Recent large-scale vision-language pre-training models, such as CLIP, have shown significant potential in enhancing pure-vision-based person re-identification. However, existing CLIP-based US-VI-ReID methods focus on independently learning semantic information within the visible and infrared modalities. These methods overlook the mismatch between the pre-training data of CLIP and the downstream cross-modality data, resulting in substantial cross-modal semantic differences. Such inconsistent semantic information, which exhibits modality discrepancies, cannot ensure the accuracy of cross-modality associations and thus hampers the performance of cross-modality learning. To address these challenges and further explore the generalizable semantic representation across modalities in CLIP, we propose a novel framework named Mining Cross-Modality Implicit Semantic Association (MCSA), which focuses on learning a modality-invariant implicit semantic space to enhance cross-modality associations and feature learning. The proposed method comprises two key modules: Modality-invariant Prompt Learning and GCNs-Driven Collaboration Alignment. Specifically, to enable CLIP to learn modality-invariant semantics, we integrate a random color augmentation branch into the visible stream for joint contrastive learning for mining generalizable semantic representations. This ensures the color generalization of the constructed implicit semantic prompts. Moreover, within the cross-modal invariant implicit semantic space, we utilize Graph Convolutional Networks (GCNs) to uncover more reliable cross-modal associations. By integrating information from images and semantic graphs, we jointly refine cross-modal correspondences, enabling the model to perform precise cross-modal feature learning. Extensive experiments conducted on the SYSU-MM01 and RegDB datasets demonstrate the effectiveness of the proposed MCSA. The source code will be released.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
小涂同学发布了新的文献求助10
刚刚
华仔的应助被Jason采纳,获得10
1秒前
邹邹发布了新的文献求助10
1秒前
3秒前
3秒前
小夏饭桶完成签到,获得积分10
4秒前
likun完成签到,获得积分10
4秒前
4秒前
5秒前
Owen的应助被仁爱柠檬采纳,获得10
5秒前
10秒前
SY完成签到,获得积分20
11秒前
秋风的应助被baining采纳,获得10
11秒前
蒙眼过河完成签到,获得积分10
12秒前
今后的应助被彩色的静芙采纳,获得10
14秒前
14秒前
超级苹果完成签到 ,获得积分10
15秒前
16秒前
大白不白发布了新的文献求助10
16秒前
OK发布了新的文献求助100
16秒前
17秒前
17秒前
19秒前
董波波完成签到,获得积分10
20秒前
碧蓝梦寒完成签到,获得积分10
21秒前
Cker发布了新的文献求助10
22秒前
SY发布了新的文献求助10
23秒前
23秒前
xiaolinsang发布了新的文献求助10
24秒前
mm完成签到,获得积分10
24秒前
仁爱柠檬发布了新的文献求助10
24秒前
25秒前
25秒前
大家好完成签到 ,获得积分10
25秒前
26秒前
27秒前
28秒前
斯文败类的应助被整齐颜采纳,获得10
29秒前
11111完成签到,获得积分10
30秒前
30秒前
高分求助中
(应助此贴封号)通过应助OA文献获取积分 10000
Composite Materials Handbook Volume 1 - Revision H 1000
Composite Materials Handbook Volume 3 - Revision H 1000
Rosenblum, Global Change Biology 800
Computational Chemical Reaction Engineering: Modeling, Simulation, and Design with MATLAB 600
Organizational Behavior 510
Management and the Arts 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 内科学 物理 有机化学 化学工程 生物化学 复合材料 光电子学 细胞生物学 心理学 量子力学 催化作用 物理化学 电极
热门帖子
关注 科研通微信公众号,转发送积分 7806400
求助须知:如何正确求助?哪些是违规求助? 9339446
关于积分的说明 20497027
捐赠科研通 7398373
什么是DOI,文献DOI怎么找? 3328053
关于科研通互助平台的介绍 2474779
邀请新用户注册赠送积分活动 2346222