Semi-Supervised State-Space Model with Dynamic Stacking Filter for Real-World Video Deraining

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Sun, Shangquan, Ren, Wenqi, Zhou, Juxiang, Wang, Shu, Gan, Jianhou, Cao, Xiaochun
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866915298706194432
author Sun, Shangquan
Ren, Wenqi
Zhou, Juxiang
Wang, Shu
Gan, Jianhou
Cao, Xiaochun
author_facet Sun, Shangquan
Ren, Wenqi
Zhou, Juxiang
Wang, Shu
Gan, Jianhou
Cao, Xiaochun
contents Significant progress has been made in video restoration under rainy conditions over the past decade, largely propelled by advancements in deep learning. Nevertheless, existing methods that depend on paired data struggle to generalize effectively to real-world scenarios, primarily due to the disparity between synthetic and authentic rain effects. To address these limitations, we propose a dual-branch spatio-temporal state-space model to enhance rain streak removal in video sequences. Specifically, we design spatial and temporal state-space model layers to extract spatial features and incorporate temporal dependencies across frames, respectively. To improve multi-frame feature fusion, we derive a dynamic stacking filter, which adaptively approximates statistical filters for superior pixel-wise feature refinement. Moreover, we develop a median stacking loss to enable semi-supervised learning by generating pseudo-clean patches based on the sparsity prior of rain. To further explore the capacity of deraining models in supporting other vision-based tasks in rainy environments, we introduce a novel real-world benchmark focused on object detection and tracking in rainy conditions. Our method is extensively evaluated across multiple benchmarks containing numerous synthetic and real-world rainy videos, consistently demonstrating its superiority in quantitative metrics, visual quality, efficiency, and its utility for downstream tasks.
format Preprint
id arxiv_https___arxiv_org_abs_2505_16811
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Semi-Supervised State-Space Model with Dynamic Stacking Filter for Real-World Video Deraining
Sun, Shangquan
Ren, Wenqi
Zhou, Juxiang
Wang, Shu
Gan, Jianhou
Cao, Xiaochun
Computer Vision and Pattern Recognition
Significant progress has been made in video restoration under rainy conditions over the past decade, largely propelled by advancements in deep learning. Nevertheless, existing methods that depend on paired data struggle to generalize effectively to real-world scenarios, primarily due to the disparity between synthetic and authentic rain effects. To address these limitations, we propose a dual-branch spatio-temporal state-space model to enhance rain streak removal in video sequences. Specifically, we design spatial and temporal state-space model layers to extract spatial features and incorporate temporal dependencies across frames, respectively. To improve multi-frame feature fusion, we derive a dynamic stacking filter, which adaptively approximates statistical filters for superior pixel-wise feature refinement. Moreover, we develop a median stacking loss to enable semi-supervised learning by generating pseudo-clean patches based on the sparsity prior of rain. To further explore the capacity of deraining models in supporting other vision-based tasks in rainy environments, we introduce a novel real-world benchmark focused on object detection and tracking in rainy conditions. Our method is extensively evaluated across multiple benchmarks containing numerous synthetic and real-world rainy videos, consistently demonstrating its superiority in quantitative metrics, visual quality, efficiency, and its utility for downstream tasks.
title Semi-Supervised State-Space Model with Dynamic Stacking Filter for Real-World Video Deraining
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2505.16811