Sequential Amodal Segmentation via Cumulative Occlusion Learning

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Ao, Jiayang, Ke, Qiuhong, Ehinger, Krista A.
Format: Preprint
Veröffentlicht: 2024
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866916240285499392
author Ao, Jiayang
Ke, Qiuhong
Ehinger, Krista A.
author_facet Ao, Jiayang
Ke, Qiuhong
Ehinger, Krista A.
contents To fully understand the 3D context of a single image, a visual system must be able to segment both the visible and occluded regions of objects, while discerning their occlusion order. Ideally, the system should be able to handle any object and not be restricted to segmenting a limited set of object classes, especially in robotic applications. Addressing this need, we introduce a diffusion model with cumulative occlusion learning designed for sequential amodal segmentation of objects with uncertain categories. This model iteratively refines the prediction using the cumulative mask strategy during diffusion, effectively capturing the uncertainty of invisible regions and adeptly reproducing the complex distribution of shapes and occlusion orders of occluded objects. It is akin to the human capability for amodal perception, i.e., to decipher the spatial ordering among objects and accurately predict complete contours for occluded objects in densely layered visual scenes. Experimental results across three amodal datasets show that our method outperforms established baselines.
format Preprint
id arxiv_https___arxiv_org_abs_2405_05791
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Sequential Amodal Segmentation via Cumulative Occlusion Learning
Ao, Jiayang
Ke, Qiuhong
Ehinger, Krista A.
Computer Vision and Pattern Recognition
To fully understand the 3D context of a single image, a visual system must be able to segment both the visible and occluded regions of objects, while discerning their occlusion order. Ideally, the system should be able to handle any object and not be restricted to segmenting a limited set of object classes, especially in robotic applications. Addressing this need, we introduce a diffusion model with cumulative occlusion learning designed for sequential amodal segmentation of objects with uncertain categories. This model iteratively refines the prediction using the cumulative mask strategy during diffusion, effectively capturing the uncertainty of invisible regions and adeptly reproducing the complex distribution of shapes and occlusion orders of occluded objects. It is akin to the human capability for amodal perception, i.e., to decipher the spatial ordering among objects and accurately predict complete contours for occluded objects in densely layered visual scenes. Experimental results across three amodal datasets show that our method outperforms established baselines.
title Sequential Amodal Segmentation via Cumulative Occlusion Learning
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2405.05791