Saved in:
Bibliographic Details
Main Authors: Ma, Wei, Chen, Shaowu, Ye, Junjie, Zhang, Peichang, Huang, Lei
Format: Preprint
Published: 2026
Subjects:
Online Access:https://arxiv.org/abs/2601.14568
Tags: Add Tag
No Tags, Be the first to tag this record!
Table of Contents:
  • Existing video inference (VI) enhancement methods typically aim to improve performance by scaling up model sizes and employing sophisticated network architectures. While these approaches demonstrated state-of-the-art performance, they often overlooked the trade-off of resource efficiency and inference effectiveness, leading to inefficient resource utilization and suboptimal inference performance. To address this problem, a fuzzy controller (FC-r) is developed based on key system parameters and inference-related metrics. Guided by the FC-r, a VI enhancement framework is proposed, where the spatiotemporal correlation of targets across adjacent video frames is leveraged. Given the real-time resource conditions of the target device, the framework can dynamically switch between models of varying scales during VI. Experimental results demonstrate that the proposed method effectively achieves a balance between resource utilization and inference performance.