RMT-BVQA: Recurrent Memory Transformer-based Blind Video Quality Assessment for Enhanced Video Content

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Peng, Tianhao, Feng, Chen, Danier, Duolikun, Zhang, Fan, Vallade, Benoit, Mackin, Alex, Bull, David
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908399942238208
author Peng, Tianhao
Feng, Chen
Danier, Duolikun
Zhang, Fan
Vallade, Benoit
Mackin, Alex
Bull, David
author_facet Peng, Tianhao
Feng, Chen
Danier, Duolikun
Zhang, Fan
Vallade, Benoit
Mackin, Alex
Bull, David
contents With recent advances in deep learning, numerous algorithms have been developed to enhance video quality, reduce visual artifacts, and improve perceptual quality. However, little research has been reported on the quality assessment of enhanced content - the evaluation of enhancement methods is often based on quality metrics that were designed for compression applications. In this paper, we propose a novel blind deep video quality assessment (VQA) method specifically for enhanced video content. It employs a new Recurrent Memory Transformer (RMT) based network architecture to obtain video quality representations, which is optimized through a novel content-quality-aware contrastive learning strategy based on a new database containing 13K training patches with enhanced content. The extracted quality representations are then combined through linear regression to generate video-level quality indices. The proposed method, RMT-BVQA, has been evaluated on the VDPVE (VQA Dataset for Perceptual Video Enhancement) database through a five-fold cross validation. The results show its superior correlation performance when compared to ten existing no-reference quality metrics.
format Preprint
id arxiv_https___arxiv_org_abs_2405_08621
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle RMT-BVQA: Recurrent Memory Transformer-based Blind Video Quality Assessment for Enhanced Video Content
Peng, Tianhao
Feng, Chen
Danier, Duolikun
Zhang, Fan
Vallade, Benoit
Mackin, Alex
Bull, David
Image and Video Processing
Computer Vision and Pattern Recognition
With recent advances in deep learning, numerous algorithms have been developed to enhance video quality, reduce visual artifacts, and improve perceptual quality. However, little research has been reported on the quality assessment of enhanced content - the evaluation of enhancement methods is often based on quality metrics that were designed for compression applications. In this paper, we propose a novel blind deep video quality assessment (VQA) method specifically for enhanced video content. It employs a new Recurrent Memory Transformer (RMT) based network architecture to obtain video quality representations, which is optimized through a novel content-quality-aware contrastive learning strategy based on a new database containing 13K training patches with enhanced content. The extracted quality representations are then combined through linear regression to generate video-level quality indices. The proposed method, RMT-BVQA, has been evaluated on the VDPVE (VQA Dataset for Perceptual Video Enhancement) database through a five-fold cross validation. The results show its superior correlation performance when compared to ten existing no-reference quality metrics.
title RMT-BVQA: Recurrent Memory Transformer-based Blind Video Quality Assessment for Enhanced Video Content
topic Image and Video Processing
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2405.08621