CPDR: Towards Highly-Efficient Salient Object Detection via Crossed Post-decoder Refinement

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Li, Yijie, Wang, Hewei, Katsaggelos, Aggelos
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912184270848000
author Li, Yijie
Wang, Hewei
Katsaggelos, Aggelos
author_facet Li, Yijie
Wang, Hewei
Katsaggelos, Aggelos
contents Most of the current salient object detection approaches use deeper networks with large backbones to produce more accurate predictions, which results in a significant increase in computational complexity. A great number of network designs follow the pure UNet and Feature Pyramid Network (FPN) architecture which has limited feature extraction and aggregation ability which motivated us to design a lightweight post-decoder refinement module, the crossed post-decoder refinement (CPDR) to enhance the feature representation of a standard FPN or U-Net framework. Specifically, we introduce the Attention Down Sample Fusion (ADF), which employs channel attention mechanisms with attention maps generated by high-level representation to refine the low-level features, and Attention Up Sample Fusion (AUF), leveraging the low-level information to guide the high-level features through spatial attention. Additionally, we proposed the Dual Attention Cross Fusion (DACF) upon ADFs and AUFs, which reduces the number of parameters while maintaining the performance. Experiments on five benchmark datasets demonstrate that our method outperforms previous state-of-the-art approaches.
format Preprint
id arxiv_https___arxiv_org_abs_2501_06441
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle CPDR: Towards Highly-Efficient Salient Object Detection via Crossed Post-decoder Refinement
Li, Yijie
Wang, Hewei
Katsaggelos, Aggelos
Computer Vision and Pattern Recognition
Most of the current salient object detection approaches use deeper networks with large backbones to produce more accurate predictions, which results in a significant increase in computational complexity. A great number of network designs follow the pure UNet and Feature Pyramid Network (FPN) architecture which has limited feature extraction and aggregation ability which motivated us to design a lightweight post-decoder refinement module, the crossed post-decoder refinement (CPDR) to enhance the feature representation of a standard FPN or U-Net framework. Specifically, we introduce the Attention Down Sample Fusion (ADF), which employs channel attention mechanisms with attention maps generated by high-level representation to refine the low-level features, and Attention Up Sample Fusion (AUF), leveraging the low-level information to guide the high-level features through spatial attention. Additionally, we proposed the Dual Attention Cross Fusion (DACF) upon ADFs and AUFs, which reduces the number of parameters while maintaining the performance. Experiments on five benchmark datasets demonstrate that our method outperforms previous state-of-the-art approaches.
title CPDR: Towards Highly-Efficient Salient Object Detection via Crossed Post-decoder Refinement
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2501.06441