计算机科学
情态动词
知识图
图形
人工智能
推荐系统
图论
机器学习
理论计算机科学
数学
组合数学
化学
高分子化学
作者
Cheng Yan,Qian Xia,Yuhan Hu,Li Li
标识
DOI:10.1109/icftic64248.2024.10913022
摘要
Incorporating multi-modal information into user-item interaction graphs has emerged as a promising strategy to improve recommendation quality by capturing richer semantic relationships. This paper introduces a novel framework that uses bootstrap latent representations improved by contrastive learning to integrate multi-modal knowledge graphs into the recommendation process. Contrastive learning addresses the cold-start problem by maximizing mutual dependencies between item content and collaborative signals [1]. We construct a comprehensive multi-modal knowledge graph by enriching the user-item interaction graph with structured knowledge, images, and textual data. After encoding both user-item interactions and multi-modal information, we propose a self-supervised learning paradigm that eliminates the need for negative sampling and complex data augmentations, thereby reducing computational overhead. Our model incorporates a contrastive view generator and leverages multiple loss functions—including graph reconstruction loss, inter-modality feature alignment loss, and intra-modality feature masking loss—to learn robust and discriminative representations. Our methodology offers improved accuracy and efficiency compared to existing methods, as demonstrated by experimental results on standard recommendation benchmarks. It also outperforms them significantly. This study highlights the effectiveness of combining multi-modal knowledge graphs and contrastive learning to enhance recommendation systems.
科研通智能强力驱动
Strongly Powered by AbleSci AI