Skip to main navigation Skip to search Skip to main content

Breaking Barriers, Localizing Saliency: A Large-scale Benchmark and Baseline for Condition-Constrained Salient Object Detection

  • Runmin CONG
  • , Zhiyang CHEN
  • , Hao FANG
  • , Sam KWONG
  • , Wei ZHANG

Research output: Journal PublicationsJournal Article (refereed)peer-review

Abstract

Salient Object Detection (SOD) aims to identify and segment the most prominent objects in an image. In real open environments, intelligent systems often encounter complex and challenging scenes, such as low-light, rain, snow, etc., which we call constrained conditions. These real situations pose more severe challenges to existing SOD models. However, there is no comprehensive and in-depth exploration of this field at both the data and model levels, and most of them focus on ideal situations or a single condition. To bridge this gap, we launch a new task, Condition-Constrained Salient Object Detection (CSOD), aimed at robustly and accurately locating salient objects in constrained environments. On the one hand, to compensate for the lack of datasets, we construct the first large-scale condition-constrained salient object detection dataset CSOD10K, comprising 10,000 pixel-level annotated images and over 100 categories of salient objects. This dataset is oriented towards the real environment and includes 8 real-world constrained scenes under 3 main constraint types, making it extremely challenging. On the other hand, we abandon the paradigm of “restoration before detection” and instead introduce a unified end-to-end framework CSSAM that fully explores scene attributes, eliminating the need for additional ground-truth restored images and reducing computational overhead. Specifically, we design a Scene Prior-Guided Adapter (SPGA), which injects scene priors to enable the foundation model to better adapt to downstream constrained scenes. To automatically decode salient objects, we propose a Hybrid Prompt Decoding Strategy (HPDS), which can effectively integrate multiple types of prompts to achieve adaptation to the SOD task. Extensive experiments show that our model significantly outperforms state-of-the-art methods on both the CSOD10K dataset and existing standard SOD benchmarks.
Original languageEnglish
Pages (from-to)4167-4183
Number of pages17
JournalIEEE Transactions on Pattern Analysis and Machine Intelligence
Volume48
Issue number4
Early online date11 Dec 2025
DOIs
Publication statusPublished - Apr 2026

Bibliographical note

Publisher Copyright:
© 1979-2012 IEEE.

Funding

This work was supported in part by the Opening Project of State Key Laboratory of Autonomous Intelligent Unmanned Systems under Grant ZZKF2025-2-8, in part by the National Natural Science Foundation of China under Grant 62471278, in part by the Key R&D Program of Shandong Province under Grant 2025CXGC010111, in part by Hong Kong GRF-RGC General Research Fund under Grant 13200425, and in part by the Research Grants Council of the Hong Kong Special Administrative Region, China, under Grant STG5/E-103/24-R. Recommended for acceptance by Y. Yu.

Keywords

  • Salient Object Detection
  • Constrained Conditions
  • Benchmark Dataset
  • Scene Prior
  • Hybrid Prompt

Fingerprint

Dive into the research topics of 'Breaking Barriers, Localizing Saliency: A Large-scale Benchmark and Baseline for Condition-Constrained Salient Object Detection'. Together they form a unique fingerprint.

Cite this