CFR-ICL: Cascade-Forward Refinement with Iterative Click Loss for Interactive Image Segmentation

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Sun, Shoukun, Xian, Min, Xu, Fei, Capriotti, Luca, Yao, Tiankai
Format: Preprint
Publié: 2023
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866913253766987776
author Sun, Shoukun
Xian, Min
Xu, Fei
Capriotti, Luca
Yao, Tiankai
author_facet Sun, Shoukun
Xian, Min
Xu, Fei
Capriotti, Luca
Yao, Tiankai
contents The click-based interactive segmentation aims to extract the object of interest from an image with the guidance of user clicks. Recent work has achieved great overall performance by employing feedback from the output. However, in most state-of-the-art approaches, 1) the inference stage involves inflexible heuristic rules and requires a separate refinement model, and 2) the number of user clicks and model performance cannot be balanced. To address the challenges, we propose a click-based and mask-guided interactive image segmentation framework containing three novel components: Cascade-Forward Refinement (CFR), Iterative Click Loss (ICL), and SUEM image augmentation. The CFR offers a unified inference framework to generate segmentation results in a coarse-to-fine manner. The proposed ICL allows model training to improve segmentation and reduce user interactions simultaneously. The proposed SUEM augmentation is a comprehensive way to create large and diverse training sets for interactive image segmentation. Extensive experiments demonstrate the state-of-the-art performance of the proposed approach on five public datasets. Remarkably, our model reduces by 33.2\%, and 15.5\% the number of clicks required to surpass an IoU of 0.95 in the previous state-of-the-art approach on the Berkeley and DAVIS sets, respectively.
format Preprint
id arxiv_https___arxiv_org_abs_2303_05620
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle CFR-ICL: Cascade-Forward Refinement with Iterative Click Loss for Interactive Image Segmentation
Sun, Shoukun
Xian, Min
Xu, Fei
Capriotti, Luca
Yao, Tiankai
Computer Vision and Pattern Recognition
The click-based interactive segmentation aims to extract the object of interest from an image with the guidance of user clicks. Recent work has achieved great overall performance by employing feedback from the output. However, in most state-of-the-art approaches, 1) the inference stage involves inflexible heuristic rules and requires a separate refinement model, and 2) the number of user clicks and model performance cannot be balanced. To address the challenges, we propose a click-based and mask-guided interactive image segmentation framework containing three novel components: Cascade-Forward Refinement (CFR), Iterative Click Loss (ICL), and SUEM image augmentation. The CFR offers a unified inference framework to generate segmentation results in a coarse-to-fine manner. The proposed ICL allows model training to improve segmentation and reduce user interactions simultaneously. The proposed SUEM augmentation is a comprehensive way to create large and diverse training sets for interactive image segmentation. Extensive experiments demonstrate the state-of-the-art performance of the proposed approach on five public datasets. Remarkably, our model reduces by 33.2\%, and 15.5\% the number of clicks required to surpass an IoU of 0.95 in the previous state-of-the-art approach on the Berkeley and DAVIS sets, respectively.
title CFR-ICL: Cascade-Forward Refinement with Iterative Click Loss for Interactive Image Segmentation
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2303.05620