Rethinking Detecting Salient and Camouflaged Objects in Unconstrained Scenes

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhou, Zhangjun, Li, Yiping, Zhong, Chunlin, Huang, Jianuo, Pei, Jialun, Li, Hua, Tang, He
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913928803516416
author Zhou, Zhangjun
Li, Yiping
Zhong, Chunlin
Huang, Jianuo
Pei, Jialun
Li, Hua
Tang, He
author_facet Zhou, Zhangjun
Li, Yiping
Zhong, Chunlin
Huang, Jianuo
Pei, Jialun
Li, Hua
Tang, He
contents While the human visual system employs distinct mechanisms to perceive salient and camouflaged objects, existing models struggle to disentangle these tasks. Specifically, salient object detection (SOD) models frequently misclassify camouflaged objects as salient, while camouflaged object detection (COD) models conversely misinterpret salient objects as camouflaged. We hypothesize that this can be attributed to two factors: (i) the specific annotation paradigm of current SOD and COD datasets, and (ii) the lack of explicit attribute relationship modeling in current models. Prevalent SOD/COD datasets enforce a mutual exclusivity constraint, assuming scenes contain either salient or camouflaged objects, which poorly aligns with the real world. Furthermore, current SOD/COD methods are primarily designed for these highly constrained datasets and lack explicit modeling of the relationship between salient and camouflaged objects. In this paper, to promote the development of unconstrained salient and camouflaged object detection, we construct a large-scale dataset, USC12K, which features comprehensive labels and four different scenes that cover all possible logical existence scenarios of both salient and camouflaged objects. To explicitly model the relationship between salient and camouflaged objects, we propose a model called USCNet, which introduces two distinct prompt query mechanisms for modeling inter-sample and intra-sample attribute relationships. Additionally, to assess the model's ability to distinguish between salient and camouflaged objects, we design an evaluation metric called CSCS. The proposed method achieves state-of-the-art performance across all scenes in various metrics. The code and dataset will be available at https://github.com/ssecv/USCNet.
format Preprint
id arxiv_https___arxiv_org_abs_2412_10943
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Rethinking Detecting Salient and Camouflaged Objects in Unconstrained Scenes
Zhou, Zhangjun
Li, Yiping
Zhong, Chunlin
Huang, Jianuo
Pei, Jialun
Li, Hua
Tang, He
Computer Vision and Pattern Recognition
While the human visual system employs distinct mechanisms to perceive salient and camouflaged objects, existing models struggle to disentangle these tasks. Specifically, salient object detection (SOD) models frequently misclassify camouflaged objects as salient, while camouflaged object detection (COD) models conversely misinterpret salient objects as camouflaged. We hypothesize that this can be attributed to two factors: (i) the specific annotation paradigm of current SOD and COD datasets, and (ii) the lack of explicit attribute relationship modeling in current models. Prevalent SOD/COD datasets enforce a mutual exclusivity constraint, assuming scenes contain either salient or camouflaged objects, which poorly aligns with the real world. Furthermore, current SOD/COD methods are primarily designed for these highly constrained datasets and lack explicit modeling of the relationship between salient and camouflaged objects. In this paper, to promote the development of unconstrained salient and camouflaged object detection, we construct a large-scale dataset, USC12K, which features comprehensive labels and four different scenes that cover all possible logical existence scenarios of both salient and camouflaged objects. To explicitly model the relationship between salient and camouflaged objects, we propose a model called USCNet, which introduces two distinct prompt query mechanisms for modeling inter-sample and intra-sample attribute relationships. Additionally, to assess the model's ability to distinguish between salient and camouflaged objects, we design an evaluation metric called CSCS. The proposed method achieves state-of-the-art performance across all scenes in various metrics. The code and dataset will be available at https://github.com/ssecv/USCNet.
title Rethinking Detecting Salient and Camouflaged Objects in Unconstrained Scenes
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2412.10943