清晨好,您是今天最早来到科研通的研友!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您科研之路漫漫前行!

Breaking Barriers, Localizing Saliency: A Large-Scale Benchmark and Baseline for Condition-Constrained Salient Object Detection

计算机科学 突出 人工智能 目标检测 计算机视觉 水准点(测量) 对象(语法) 解码方法 光学(聚焦) 构造(python库) 模式识别(心理学) 约束(计算机辅助设计) 领域(数学) 先验概率 特征提取 变更检测 视频跟踪 隐马尔可夫模型 姿势 基线(sea) 视觉对象识别的认知神经科学 机器学习 编码 桥(图论)
作者
Runmin Cong,Zhiyang Chen,Hao Fang,Sam Kwong,Wei Zhang
出处
期刊:IEEE Transactions on Pattern Analysis and Machine Intelligence [IEEE Computer Society]
卷期号:48 (4): 4167-4183 被引量:1
标识
DOI:10.1109/tpami.2025.3642893
摘要

Salient Object Detection (SOD) aims to identify and segment the most prominent objects in an image. In real open environments, intelligent systems often encounter complex and challenging scenes, such as low-light, rain, snow, etc., which we call constrained conditions. These real situations pose more severe challenges to existing SOD models. However, there is no comprehensive and in-depth exploration of this field at both the data and model levels, and most of them focus on ideal situations or a single condition. To bridge this gap, we launch a new task, Condition-Constrained Salient Object Detection (CSOD), aimed at robustly and accurately locating salient objects in constrained environments. On the one hand, to compensate for the lack of datasets, we construct the first large-scale condition-constrained salient object detection dataset CSOD10 K, comprising 10,000 pixel-level annotated images and over 100 categories of salient objects. This dataset is oriented towards the real environment and includes 8 real-world constrained scenes under 3 main constraint types, making it extremely challenging. On the other hand, we abandon the paradigm of "restoration before detection" and instead introduce a unified end-to-end framework CSSAM that fully explores scene attributes, eliminating the need for additional ground-truth restored images and reducing computational overhead. Specifically, we design a Scene Prior-Guided Adapter (SPGA), which injects scene priors to enable the foundation model to better adapt to downstream constrained scenes. To automatically decode salient objects, we propose a Hybrid Prompt Decoding Strategy (HPDS), which can effectively integrate multiple types of prompts to achieve adaptation to the SOD task. Extensive experiments show that our model significantly outperforms state-of-the-art methods on both the CSOD10 K dataset and existing standard SOD benchmarks.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
5秒前
一一发布了新的文献求助10
10秒前
笑对人生完成签到 ,获得积分10
13秒前
华仔应助叶潭采纳,获得10
21秒前
FeelingUnreal完成签到,获得积分10
22秒前
Takeda完成签到,获得积分10
24秒前
GHOSTagw完成签到,获得积分10
25秒前
俏皮夏瑶完成签到,获得积分10
28秒前
有魅力以珊完成签到,获得积分10
28秒前
轻舞完成签到,获得积分10
31秒前
34秒前
LMY1470完成签到,获得积分10
35秒前
gtgyh完成签到 ,获得积分10
38秒前
调皮的烤鸡完成签到,获得积分10
38秒前
安详忆梅发布了新的文献求助10
41秒前
HanaTerbush完成签到,获得积分10
42秒前
GinaLundhild06完成签到,获得积分10
45秒前
47秒前
踏实麦片完成签到,获得积分10
48秒前
50秒前
yunsui完成签到,获得积分10
52秒前
叶潭发布了新的文献求助10
52秒前
小小油完成签到,获得积分10
55秒前
1分钟前
xingsixs完成签到,获得积分10
1分钟前
阳光笑颜完成签到,获得积分10
1分钟前
YvesWang完成签到,获得积分10
1分钟前
1分钟前
Orange应助安详忆梅采纳,获得10
1分钟前
深情安青应助孤独太清采纳,获得10
1分钟前
害羞的雁易完成签到 ,获得积分10
1分钟前
传奇3应助竹捷采纳,获得10
1分钟前
1分钟前
1分钟前
1分钟前
孤独太清发布了新的文献求助10
1分钟前
竹捷发布了新的文献求助10
1分钟前
cadcae完成签到,获得积分10
1分钟前
2分钟前
汉堡包应助一一采纳,获得10
2分钟前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Principles of town planning: translating concepts to applications 1000
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
The role of consumer psychology in the marketing strategies of pop mart in Thailand 500
核安全综合知识2024版 500
Photothermal Science and Techniques 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7720773
求助须知:如何正确求助?哪些是违规求助? 9274180
关于积分的说明 20100840
捐赠科研通 7296927
什么是DOI,文献DOI怎么找? 3300250
关于科研通互助平台的介绍 2454141
邀请新用户注册赠送积分活动 2307718