计算机科学
像素
人工智能
RGB颜色模型
计算机视觉
目标检测
突出
对象(语法)
模式识别(心理学)
作者
Junbin Yuan,Yiqi Wang,Zhoutao Wang,Qingzhen Xu,Bharadwaj Veeravalli,Xulei Yang
标识
DOI:10.1109/tmm.2025.3535386
摘要
Depth cues are essential for visual perception tasks like Salient Object Detection (SOD). Due to varying depth reliability across scenes, some researchers propose evaluating the overall quality of the depth maps and discarding the less reliable ones to avoid contamination. However, these methods often fail to fully utilize valuable information in depth maps, leading to sub-optimal performance particularly when depth quality is unreliable. Since low-quality depth maps still contain useful information that potentially improves model performance, we propose a Depth Pixel-wise Potential-aware Network to leverage these depth cues effectively. This network includes two novel components designed: 1) A learning strategy for explicitly modeling the confidence of each depth pixel to assist the model in locating valid information in the depth map. 2) A cross-modal adaptive multiple fusion module that fuses features from both RGB and depth modalities. It aims to mitigate the contamination effect of unreliable depth maps and fully exploit the benefits of multiple fusion strategies. Experimental results show that on four publicly available datasets, our method outperforms 17 mainstream methods on various evaluation metrics.
科研通智能强力驱动
Strongly Powered by AbleSci AI