EyeSim-VQA: A Free-Energy-Guided Eye Simulation Framework for Video Quality Assessment

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Zhaoyang, Lu, Wen, Li, Jie, He, Lihuo, Gong, Maoguo, Gao, Xinbo
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913891690217472
author Wang, Zhaoyang
Lu, Wen
Li, Jie
He, Lihuo
Gong, Maoguo
Gao, Xinbo
author_facet Wang, Zhaoyang
Lu, Wen
Li, Jie
He, Lihuo
Gong, Maoguo
Gao, Xinbo
contents Free-energy-guided self-repair mechanisms have shown promising results in image quality assessment (IQA), but remain under-explored in video quality assessment (VQA), where temporal dynamics and model constraints pose unique challenges. Unlike static images, video content exhibits richer spatiotemporal complexity, making perceptual restoration more difficult. Moreover, VQA systems often rely on pre-trained backbones, which limits the direct integration of enhancement modules without affecting model stability. To address these issues, we propose EyeSimVQA, a novel VQA framework that incorporates free-energy-based self-repair. It adopts a dual-branch architecture, with an aesthetic branch for global perceptual evaluation and a technical branch for fine-grained structural and semantic analysis. Each branch integrates specialized enhancement modules tailored to distinct visual inputs-resized full-frame images and patch-based fragments-to simulate adaptive repair behaviors. We also explore a principled strategy for incorporating high-level visual features without disrupting the original backbone. In addition, we design a biologically inspired prediction head that models sweeping gaze dynamics to better fuse global and local representations for quality prediction. Experiments on five public VQA benchmarks demonstrate that EyeSimVQA achieves competitive or superior performance compared to state-of-the-art methods, while offering improved interpretability through its biologically grounded design.
format Preprint
id arxiv_https___arxiv_org_abs_2506_11549
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle EyeSim-VQA: A Free-Energy-Guided Eye Simulation Framework for Video Quality Assessment
Wang, Zhaoyang
Lu, Wen
Li, Jie
He, Lihuo
Gong, Maoguo
Gao, Xinbo
Computer Vision and Pattern Recognition
Image and Video Processing
Free-energy-guided self-repair mechanisms have shown promising results in image quality assessment (IQA), but remain under-explored in video quality assessment (VQA), where temporal dynamics and model constraints pose unique challenges. Unlike static images, video content exhibits richer spatiotemporal complexity, making perceptual restoration more difficult. Moreover, VQA systems often rely on pre-trained backbones, which limits the direct integration of enhancement modules without affecting model stability. To address these issues, we propose EyeSimVQA, a novel VQA framework that incorporates free-energy-based self-repair. It adopts a dual-branch architecture, with an aesthetic branch for global perceptual evaluation and a technical branch for fine-grained structural and semantic analysis. Each branch integrates specialized enhancement modules tailored to distinct visual inputs-resized full-frame images and patch-based fragments-to simulate adaptive repair behaviors. We also explore a principled strategy for incorporating high-level visual features without disrupting the original backbone. In addition, we design a biologically inspired prediction head that models sweeping gaze dynamics to better fuse global and local representations for quality prediction. Experiments on five public VQA benchmarks demonstrate that EyeSimVQA achieves competitive or superior performance compared to state-of-the-art methods, while offering improved interpretability through its biologically grounded design.
title EyeSim-VQA: A Free-Energy-Guided Eye Simulation Framework for Video Quality Assessment
topic Computer Vision and Pattern Recognition
Image and Video Processing
url https://arxiv.org/abs/2506.11549