USV: Unified Sparsification for Accelerating Video Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Xinjian, Wang, Hongmei, Zhou, Yuan, Lu, Qinglin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models
von: Lian, Jiesong, et al.
Veröffentlicht: (2026)
von: Lian, Jiesong, et al.
Veröffentlicht: (2026)
QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention Sparsification
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions
von: Zhang, Guozhen, et al.
Veröffentlicht: (2025)
von: Zhang, Guozhen, et al.
Veröffentlicht: (2025)
USV: Towards Understanding the User-generated Short-form Videos
von: Cheng, Haoyue, et al.
Veröffentlicht: (2026)
von: Cheng, Haoyue, et al.
Veröffentlicht: (2026)
Arbitrary Generative Video Interpolation
von: Zhang, Guozhen, et al.
Veröffentlicht: (2025)
von: Zhang, Guozhen, et al.
Veröffentlicht: (2025)
Phased One-Step Adversarial Equilibrium for Video Diffusion Models
von: Cheng, Jiaxiang, et al.
Veröffentlicht: (2025)
von: Cheng, Jiaxiang, et al.
Veröffentlicht: (2025)
SmoothVideo: Smooth Video Synthesis with Noise Constraints on Diffusion Models for One-shot Video Tuning
von: Peng, Liang, et al.
Veröffentlicht: (2023)
von: Peng, Liang, et al.
Veröffentlicht: (2023)
OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation
von: Liang, Sen, et al.
Veröffentlicht: (2025)
von: Liang, Sen, et al.
Veröffentlicht: (2025)
SparseDiT: Token Sparsification for Efficient Diffusion Transformer
von: Chang, Shuning, et al.
Veröffentlicht: (2024)
von: Chang, Shuning, et al.
Veröffentlicht: (2024)
HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation
von: Huang, Ziyao, et al.
Veröffentlicht: (2025)
von: Huang, Ziyao, et al.
Veröffentlicht: (2025)
Audio-visual Controlled Video Diffusion with Masked Selective State Spaces Modeling for Natural Talking Head Generation
von: Hong, Fa-Ting, et al.
Veröffentlicht: (2025)
von: Hong, Fa-Ting, et al.
Veröffentlicht: (2025)
AccVideo: Accelerating Video Diffusion Model with Synthetic Dataset
von: Zhang, Haiyu, et al.
Veröffentlicht: (2025)
von: Zhang, Haiyu, et al.
Veröffentlicht: (2025)
Video Token Sparsification for Efficient Multimodal LLMs in Autonomous Driving
von: Ma, Yunsheng, et al.
Veröffentlicht: (2024)
von: Ma, Yunsheng, et al.
Veröffentlicht: (2024)
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
Unified Dense Prediction of Video Diffusion
von: Yang, Lehan, et al.
Veröffentlicht: (2025)
von: Yang, Lehan, et al.
Veröffentlicht: (2025)
Rotate Your Character: Revisiting Video Diffusion Models for High-Quality 3D Character Generation
von: Wang, Jin, et al.
Veröffentlicht: (2026)
von: Wang, Jin, et al.
Veröffentlicht: (2026)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
GenRec: Unifying Video Generation and Recognition with Diffusion Models
von: Weng, Zejia, et al.
Veröffentlicht: (2024)
von: Weng, Zejia, et al.
Veröffentlicht: (2024)
Pack and Force Your Memory: Long-form and Consistent Video Generation
von: Wu, Xiaofei, et al.
Veröffentlicht: (2025)
von: Wu, Xiaofei, et al.
Veröffentlicht: (2025)
Unified Multimodal Chain-of-Thought Reward Model through Reinforcement Fine-Tuning
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
Video Generation Models Are Good Latent Reward Models
von: Mi, Xiaoyue, et al.
Veröffentlicht: (2025)
von: Mi, Xiaoyue, et al.
Veröffentlicht: (2025)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
OmniCamera: A Unified Framework for Multi-task Video Generation with Arbitrary Camera Control
von: Wang, Yukun, et al.
Veröffentlicht: (2026)
von: Wang, Yukun, et al.
Veröffentlicht: (2026)
Accelerating Video Diffusion Models via Distribution Matching
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2024)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2024)
Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
Hunyuan-GameCraft: High-dynamic Interactive Game Video Generation with Hybrid History Condition
von: Li, Jiaqi, et al.
Veröffentlicht: (2025)
von: Li, Jiaqi, et al.
Veröffentlicht: (2025)
Harmony: Harmonizing Audio and Video Generation through Cross-Task Synergy
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
von: Chen, Yi, et al.
Veröffentlicht: (2025)
von: Chen, Yi, et al.
Veröffentlicht: (2025)
Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models
von: Guo, Jiayi, et al.
Veröffentlicht: (2026)
von: Guo, Jiayi, et al.
Veröffentlicht: (2026)
Moaw: Unleashing Motion Awareness for Video Diffusion Models
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
Streaming Video Diffusion: Online Video Editing with Diffusion Models
von: Chen, Feng, et al.
Veröffentlicht: (2024)
von: Chen, Feng, et al.
Veröffentlicht: (2024)
Glance: Accelerating Diffusion Models with 1 Sample
von: Dong, Zhuobai, et al.
Veröffentlicht: (2025)
von: Dong, Zhuobai, et al.
Veröffentlicht: (2025)
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
Show Me: Unifying Instructional Image and Video Generation with Diffusion Models
von: Pu, Yujiang, et al.
Veröffentlicht: (2025)
von: Pu, Yujiang, et al.
Veröffentlicht: (2025)
EventDiff: A Unified and Efficient Diffusion Model Framework for Event-based Video Frame Interpolation
von: Zheng, Hanle, et al.
Veröffentlicht: (2025)
von: Zheng, Hanle, et al.
Veröffentlicht: (2025)
UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation
von: Sun, Yang-Tian, et al.
Veröffentlicht: (2025)
von: Sun, Yang-Tian, et al.
Veröffentlicht: (2025)
SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
von: Zhang, Yuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuan, et al.
Veröffentlicht: (2024)
EffectMaker: Unifying Reasoning and Generation for Customized Visual Effect Creation
von: Yang, Shiyuan, et al.
Veröffentlicht: (2026)
von: Yang, Shiyuan, et al.
Veröffentlicht: (2026)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
SoliReward: Mitigating Susceptibility to Reward Hacking and Annotation Noise in Video Generation Reward Models
von: Lian, Jiesong, et al.
Veröffentlicht: (2025)
von: Lian, Jiesong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models
von: Lian, Jiesong, et al.
Veröffentlicht: (2026) -
QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention Sparsification
von: Feng, Weilun, et al.
Veröffentlicht: (2025) -
UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions
von: Zhang, Guozhen, et al.
Veröffentlicht: (2025) -
USV: Towards Understanding the User-generated Short-form Videos
von: Cheng, Haoyue, et al.
Veröffentlicht: (2026) -
Arbitrary Generative Video Interpolation
von: Zhang, Guozhen, et al.
Veröffentlicht: (2025)