Revisiting Context Aggregation for Image Matting

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Liu, Qinglin, Lv, Xiaoqian, Meng, Quanling, Li, Zonglin, Lan, Xiangyuan, Yang, Shuo, Zhang, Shengping, Nie, Liqiang
Format: Preprint
Veröffentlicht: 2023
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866929343602622464
author Liu, Qinglin
Lv, Xiaoqian
Meng, Quanling
Li, Zonglin
Lan, Xiangyuan
Yang, Shuo
Zhang, Shengping
Nie, Liqiang
author_facet Liu, Qinglin
Lv, Xiaoqian
Meng, Quanling
Li, Zonglin
Lan, Xiangyuan
Yang, Shuo
Zhang, Shengping
Nie, Liqiang
contents Traditional studies emphasize the significance of context information in improving matting performance. Consequently, deep learning-based matting methods delve into designing pooling or affinity-based context aggregation modules to achieve superior results. However, these modules cannot well handle the context scale shift caused by the difference in image size during training and inference, resulting in matting performance degradation. In this paper, we revisit the context aggregation mechanisms of matting networks and find that a basic encoder-decoder network without any context aggregation modules can actually learn more universal context aggregation, thereby achieving higher matting performance compared to existing methods. Building on this insight, we present AEMatter, a matting network that is straightforward yet very effective. AEMatter adopts a Hybrid-Transformer backbone with appearance-enhanced axis-wise learning (AEAL) blocks to build a basic network with strong context aggregation learning capability. Furthermore, AEMatter leverages a large image training strategy to assist the network in learning context aggregation from data. Extensive experiments on five popular matting datasets demonstrate that the proposed AEMatter outperforms state-of-the-art matting methods by a large margin.
format Preprint
id arxiv_https___arxiv_org_abs_2304_01171
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Revisiting Context Aggregation for Image Matting
Liu, Qinglin
Lv, Xiaoqian
Meng, Quanling
Li, Zonglin
Lan, Xiangyuan
Yang, Shuo
Zhang, Shengping
Nie, Liqiang
Computer Vision and Pattern Recognition
Traditional studies emphasize the significance of context information in improving matting performance. Consequently, deep learning-based matting methods delve into designing pooling or affinity-based context aggregation modules to achieve superior results. However, these modules cannot well handle the context scale shift caused by the difference in image size during training and inference, resulting in matting performance degradation. In this paper, we revisit the context aggregation mechanisms of matting networks and find that a basic encoder-decoder network without any context aggregation modules can actually learn more universal context aggregation, thereby achieving higher matting performance compared to existing methods. Building on this insight, we present AEMatter, a matting network that is straightforward yet very effective. AEMatter adopts a Hybrid-Transformer backbone with appearance-enhanced axis-wise learning (AEAL) blocks to build a basic network with strong context aggregation learning capability. Furthermore, AEMatter leverages a large image training strategy to assist the network in learning context aggregation from data. Extensive experiments on five popular matting datasets demonstrate that the proposed AEMatter outperforms state-of-the-art matting methods by a large margin.
title Revisiting Context Aggregation for Image Matting
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2304.01171