VFIMamba: Video Frame Interpolation with State Space Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Guozhen, Liu, Chunxu, Cui, Yutao, Zhao, Xiaotong, Ma, Kai, Wang, Limin |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Motion-Aware Generative Frame Interpolation
par: Zhang, Guozhen, et autres
Publié: (2025)
par: Zhang, Guozhen, et autres
Publié: (2025)
Sparse Global Matching for Video Frame Interpolation with Large Motion
par: Liu, Chunxu, et autres
Publié: (2024)
par: Liu, Chunxu, et autres
Publié: (2024)
StableDrag: Stable Dragging for Point-based Image Editing
par: Cui, Yutao, et autres
Publié: (2024)
par: Cui, Yutao, et autres
Publié: (2024)
Dynamic and Compressive Adaptation of Transformers From Images to Videos
par: Zhang, Guozhen, et autres
Publié: (2024)
par: Zhang, Guozhen, et autres
Publié: (2024)
KFS-Bench: Comprehensive Evaluation of Key Frame Sampling in Long Video Understanding
par: Li, Zongyao, et autres
Publié: (2025)
par: Li, Zongyao, et autres
Publié: (2025)
Arbitrary Generative Video Interpolation
par: Zhang, Guozhen, et autres
Publié: (2025)
par: Zhang, Guozhen, et autres
Publié: (2025)
MiVID: Multi-Strategic Self-Supervision for Video Frame Interpolation using Diffusion Model
par: Srivastava, Priyansh, et autres
Publié: (2025)
par: Srivastava, Priyansh, et autres
Publié: (2025)
Mamba-FETrack: Frame-Event Tracking via State Space Model
par: Huang, Ju, et autres
Publié: (2024)
par: Huang, Ju, et autres
Publié: (2024)
DAWN: Dynamic Frame Avatar with Non-autoregressive Diffusion Framework for Talking Head Video Generation
par: Cheng, Hanbo, et autres
Publié: (2024)
par: Cheng, Hanbo, et autres
Publié: (2024)
FrameMind: Frame-Interleaved Video Reasoning via Reinforcement Learning
par: Ge, Haonan, et autres
Publié: (2025)
par: Ge, Haonan, et autres
Publié: (2025)
DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation
par: Ye, Bo, et autres
Publié: (2026)
par: Ye, Bo, et autres
Publié: (2026)
UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban Spaces
par: Zhao, Baining, et autres
Publié: (2025)
par: Zhao, Baining, et autres
Publié: (2025)
UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling
par: Li, Peiming, et autres
Publié: (2025)
par: Li, Peiming, et autres
Publié: (2025)
M-LLM Based Video Frame Selection for Efficient Video Understanding
par: Hu, Kai, et autres
Publié: (2025)
par: Hu, Kai, et autres
Publié: (2025)
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding
par: Zhang, Peng, et autres
Publié: (2026)
par: Zhang, Peng, et autres
Publié: (2026)
Enhancing Long Video Question Answering with Scene-Localized Frame Grouping
par: Yang, Xuyi, et autres
Publié: (2025)
par: Yang, Xuyi, et autres
Publié: (2025)
SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation
par: Zhang, Jiaming, et autres
Publié: (2025)
par: Zhang, Jiaming, et autres
Publié: (2025)
BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models
par: Shi, Fengyuan, et autres
Publié: (2023)
par: Shi, Fengyuan, et autres
Publié: (2023)
VideoAR: Autoregressive Video Generation via Next-Frame & Scale Prediction
par: Ji, Longbin, et autres
Publié: (2026)
par: Ji, Longbin, et autres
Publié: (2026)
Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
par: Jang, Sangwon, et autres
Publié: (2025)
par: Jang, Sangwon, et autres
Publié: (2025)
Focused Forcing: Content-Aware Per-Frame KV Selection for Efficient Autoregressive Video Diffusion
par: Cai, Peiliang, et autres
Publié: (2026)
par: Cai, Peiliang, et autres
Publié: (2026)
Detecting AI-Generated Video via Frame Consistency
par: Ma, Long, et autres
Publié: (2024)
par: Ma, Long, et autres
Publié: (2024)
UniBEVFusion: Unified Radar-Vision BEVFusion for 3D Object Detection
par: Zhao, Haocheng, et autres
Publié: (2024)
par: Zhao, Haocheng, et autres
Publié: (2024)
RichSpace: Enriching Text-to-Video Prompt Space via Text Embedding Interpolation
par: Cao, Yuefan, et autres
Publié: (2025)
par: Cao, Yuefan, et autres
Publié: (2025)
NoiseDiffusion: Correcting Noise for Image Interpolation with Diffusion Models beyond Spherical Linear Interpolation
par: Zheng, PengFei, et autres
Publié: (2024)
par: Zheng, PengFei, et autres
Publié: (2024)
Joint Modeling of Feature, Correspondence, and a Compressed Memory for Video Object Segmentation
par: Zhang, Jiaming, et autres
Publié: (2023)
par: Zhang, Jiaming, et autres
Publié: (2023)
Video Finetuning Improves Reasoning Between Frames
par: Yang, Ruiqi, et autres
Publié: (2025)
par: Yang, Ruiqi, et autres
Publié: (2025)
Mamba-FETrack V2: Revisiting State Space Model for Frame-Event based Visual Object Tracking
par: Wang, Shiao, et autres
Publié: (2025)
par: Wang, Shiao, et autres
Publié: (2025)
MedVSR: Medical Video Super-Resolution with Cross State-Space Propagation
par: Liu, Xinyu, et autres
Publié: (2025)
par: Liu, Xinyu, et autres
Publié: (2025)
UFO: Enhancing Diffusion-Based Video Generation with a Uniform Frame Organizer
par: Liu, Delong, et autres
Publié: (2024)
par: Liu, Delong, et autres
Publié: (2024)
Pack and Force Your Memory: Long-form and Consistent Video Generation
par: Wu, Xiaofei, et autres
Publié: (2025)
par: Wu, Xiaofei, et autres
Publié: (2025)
SSMamba: A Self-Supervised Hybrid State Space Model for Pathological Image Classification
par: Chai, Enhui, et autres
Publié: (2026)
par: Chai, Enhui, et autres
Publié: (2026)
Velocity Disambiguation for Video Frame Interpolation
par: Zhong, Zhihang, et autres
Publié: (2023)
par: Zhong, Zhihang, et autres
Publié: (2023)
DLFR-VAE: Dynamic Latent Frame Rate VAE for Video Generation
par: Yuan, Zhihang, et autres
Publié: (2025)
par: Yuan, Zhihang, et autres
Publié: (2025)
KeyFace: Expressive Audio-Driven Facial Animation for Long Sequences via KeyFrame Interpolation
par: Bigata, Antoni, et autres
Publié: (2025)
par: Bigata, Antoni, et autres
Publié: (2025)
SSM Meets Video Diffusion Models: Efficient Long-Term Video Generation with Structured State Spaces
par: Oshima, Yuta, et autres
Publié: (2024)
par: Oshima, Yuta, et autres
Publié: (2024)
Deformba: Vision State Space Model with Adaptive State Fusion
par: Ke, Hongyu, et autres
Publié: (2026)
par: Ke, Hongyu, et autres
Publié: (2026)
Unified Medical Image Segmentation with State Space Modeling Snake
par: Zhang, Ruicheng, et autres
Publié: (2025)
par: Zhang, Ruicheng, et autres
Publié: (2025)
A Style is Worth One Code: Unlocking Code-to-Style Image Generation with Discrete Style Space
par: Liu, Huijie, et autres
Publié: (2025)
par: Liu, Huijie, et autres
Publié: (2025)
Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models
par: Jeon, Wooseok, et autres
Publié: (2026)
par: Jeon, Wooseok, et autres
Publié: (2026)
Documents similaires
-
Motion-Aware Generative Frame Interpolation
par: Zhang, Guozhen, et autres
Publié: (2025) -
Sparse Global Matching for Video Frame Interpolation with Large Motion
par: Liu, Chunxu, et autres
Publié: (2024) -
StableDrag: Stable Dragging for Point-based Image Editing
par: Cui, Yutao, et autres
Publié: (2024) -
Dynamic and Compressive Adaptation of Transformers From Images to Videos
par: Zhang, Guozhen, et autres
Publié: (2024) -
KFS-Bench: Comprehensive Evaluation of Key Frame Sampling in Long Video Understanding
par: Li, Zongyao, et autres
Publié: (2025)