Content Adaptive based Motion Alignment Framework for Learned Video Compression

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhang, Tiange, Meng, Xiandong, Ma, Siwei
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918248891547648
author Zhang, Tiange
Meng, Xiandong
Ma, Siwei
author_facet Zhang, Tiange
Meng, Xiandong
Ma, Siwei
contents Recent advances in end-to-end video compression have shown promising results owing to their unified end-to-end learning optimization. However, such generalized frameworks often lack content-specific adaptation, leading to suboptimal compression performance. To address this, this paper proposes a content adaptive based motion alignment framework that improves performance by adapting encoding strategies to diverse content characteristics. Specifically, we first introduce a two-stage flow-guided deformable warping mechanism that refines motion compensation with coarse-to-fine offset prediction and mask modulation, enabling precise feature alignment. Second, we propose a multi-reference quality aware strategy that adjusts distortion weights based on reference quality, and applies it to hierarchical training to reduce error propagation. Third, we integrate a training-free module that downsamples frames by motion magnitude and resolution to obtain smooth motion estimation. Experimental results on standard test datasets demonstrate that our framework CAMA achieves significant improvements over state-of-the-art Neural Video Compression models, achieving a 24.95% BD-rate (PSNR) savings over our baseline model DCVC-TCM, while also outperforming reproduced DCVC-DC and traditional codec HM-16.25.
format Preprint
id arxiv_https___arxiv_org_abs_2512_12936
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Content Adaptive based Motion Alignment Framework for Learned Video Compression
Zhang, Tiange
Meng, Xiandong
Ma, Siwei
Computer Vision and Pattern Recognition
Artificial Intelligence
I.4.2
Recent advances in end-to-end video compression have shown promising results owing to their unified end-to-end learning optimization. However, such generalized frameworks often lack content-specific adaptation, leading to suboptimal compression performance. To address this, this paper proposes a content adaptive based motion alignment framework that improves performance by adapting encoding strategies to diverse content characteristics. Specifically, we first introduce a two-stage flow-guided deformable warping mechanism that refines motion compensation with coarse-to-fine offset prediction and mask modulation, enabling precise feature alignment. Second, we propose a multi-reference quality aware strategy that adjusts distortion weights based on reference quality, and applies it to hierarchical training to reduce error propagation. Third, we integrate a training-free module that downsamples frames by motion magnitude and resolution to obtain smooth motion estimation. Experimental results on standard test datasets demonstrate that our framework CAMA achieves significant improvements over state-of-the-art Neural Video Compression models, achieving a 24.95% BD-rate (PSNR) savings over our baseline model DCVC-TCM, while also outperforming reproduced DCVC-DC and traditional codec HM-16.25.
title Content Adaptive based Motion Alignment Framework for Learned Video Compression
topic Computer Vision and Pattern Recognition
Artificial Intelligence
I.4.2
url https://arxiv.org/abs/2512.12936