COMUNI: Decomposing Common and Unique Video Signals for Diffusion-based Video Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Mingzhen, Wang, Weining, Zhu, Xinxin, Liu, Jing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MM-LDM: Multi-Modal Latent Diffusion Model for Sounding Video Generation
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024)
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024)
AR-Diffusion: Asynchronous Video Generation with Auto-Regressive Diffusion
von: Sun, Mingzhen, et al.
Veröffentlicht: (2025)
von: Sun, Mingzhen, et al.
Veröffentlicht: (2025)
VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image Synthesis
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
ExVideo: Extending Video Diffusion Models via Parameter-Efficient Post-Tuning
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
PhysCorr: Dual-Reward DPO for Physics-Constrained Text-to-Video Generation with Automated Preference Selection
von: Wang, Peiyao, et al.
Veröffentlicht: (2025)
von: Wang, Peiyao, et al.
Veröffentlicht: (2025)
Generative Omnimatte: Learning to Decompose Video into Layers
von: Lee, Yao-Chih, et al.
Veröffentlicht: (2024)
von: Lee, Yao-Chih, et al.
Veröffentlicht: (2024)
Enhancing Motion in Text-to-Video Generation with Decomposed Encoding and Conditioning
von: Ruan, Penghui, et al.
Veröffentlicht: (2024)
von: Ruan, Penghui, et al.
Veröffentlicht: (2024)
Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video Enhancement
von: Zhu, Lingyu, et al.
Veröffentlicht: (2024)
von: Zhu, Lingyu, et al.
Veröffentlicht: (2024)
Mitty: Diffusion-based Human-to-Robot Video Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation
von: Wang, Wenjing, et al.
Veröffentlicht: (2023)
von: Wang, Wenjing, et al.
Veröffentlicht: (2023)
VideoShield: Regulating Diffusion-based Video Generation Models via Watermarking
von: Hu, Runyi, et al.
Veröffentlicht: (2025)
von: Hu, Runyi, et al.
Veröffentlicht: (2025)
Generative Video Diffusion for Unseen Novel Semantic Video Moment Retrieval
von: Luo, Dezhao, et al.
Veröffentlicht: (2024)
von: Luo, Dezhao, et al.
Veröffentlicht: (2024)
Video Decomposition Prior: A Methodology to Decompose Videos into Layers
von: Shrivastava, Gaurav, et al.
Veröffentlicht: (2024)
von: Shrivastava, Gaurav, et al.
Veröffentlicht: (2024)
Decoupling Common and Unique Representations for Multimodal Self-supervised Learning
von: Wang, Yi, et al.
Veröffentlicht: (2023)
von: Wang, Yi, et al.
Veröffentlicht: (2023)
ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
von: Fang, Zixun, et al.
Veröffentlicht: (2025)
von: Fang, Zixun, et al.
Veröffentlicht: (2025)
VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step
von: Wang, Hanyang, et al.
Veröffentlicht: (2025)
von: Wang, Hanyang, et al.
Veröffentlicht: (2025)
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
von: Liang, Jingyun, et al.
Veröffentlicht: (2025)
von: Liang, Jingyun, et al.
Veröffentlicht: (2025)
Dual-Stream Diffusion Net for Text-to-Video Generation
von: Liu, Binhui, et al.
Veröffentlicht: (2023)
von: Liu, Binhui, et al.
Veröffentlicht: (2023)
MagicMirror: ID-Preserved Video Generation in Video Diffusion Transformers
von: Zhang, Yuechen, et al.
Veröffentlicht: (2025)
von: Zhang, Yuechen, et al.
Veröffentlicht: (2025)
Split4D: Decomposed 4D Scene Reconstruction Without Video Segmentation
von: Hu, Yongzhen, et al.
Veröffentlicht: (2025)
von: Hu, Yongzhen, et al.
Veröffentlicht: (2025)
Tora: Trajectory-oriented Diffusion Transformer for Video Generation
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2024)
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2024)
Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval
von: Shlapentokh-Rothman, Michal, et al.
Veröffentlicht: (2026)
von: Shlapentokh-Rothman, Michal, et al.
Veröffentlicht: (2026)
VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
Latte: Latent Diffusion Transformer for Video Generation
von: Ma, Xin, et al.
Veröffentlicht: (2024)
von: Ma, Xin, et al.
Veröffentlicht: (2024)
Bernini: Latent Semantic Planning for Video Diffusion
von: Bernini Team, et al.
Veröffentlicht: (2026)
von: Bernini Team, et al.
Veröffentlicht: (2026)
Infinite Gaze Generation for Videos with Autoregressive Diffusion
von: Kang, Jenna, et al.
Veröffentlicht: (2026)
von: Kang, Jenna, et al.
Veröffentlicht: (2026)
MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling
von: Men, Yifang, et al.
Veröffentlicht: (2024)
von: Men, Yifang, et al.
Veröffentlicht: (2024)
PatchVSR: Breaking Video Diffusion Resolution Limits with Patch-wise Video Super-Resolution
von: Du, Shian, et al.
Veröffentlicht: (2025)
von: Du, Shian, et al.
Veröffentlicht: (2025)
VideoGen-Eval: Agent-based System for Video Generation Evaluation
von: Yang, Yuhang, et al.
Veröffentlicht: (2025)
von: Yang, Yuhang, et al.
Veröffentlicht: (2025)
4Diffusion: Multi-view Video Diffusion Model for 4D Generation
von: Zhang, Haiyu, et al.
Veröffentlicht: (2024)
von: Zhang, Haiyu, et al.
Veröffentlicht: (2024)
AccVideo: Accelerating Video Diffusion Model with Synthetic Dataset
von: Zhang, Haiyu, et al.
Veröffentlicht: (2025)
von: Zhang, Haiyu, et al.
Veröffentlicht: (2025)
Language-driven Description Generation and Common Sense Reasoning for Video Action Recognition
von: Hu, Xiaodan, et al.
Veröffentlicht: (2025)
von: Hu, Xiaodan, et al.
Veröffentlicht: (2025)
Generative Neural Video Compression via Video Diffusion Prior
von: Mao, Qi, et al.
Veröffentlicht: (2025)
von: Mao, Qi, et al.
Veröffentlicht: (2025)
VideoPure: Diffusion-based Adversarial Purification for Video Recognition
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
von: Ren, Yixuan, et al.
Veröffentlicht: (2024)
von: Ren, Yixuan, et al.
Veröffentlicht: (2024)
Uniform Discrete Diffusion with Metric Path for Video Generation
von: Deng, Haoge, et al.
Veröffentlicht: (2025)
von: Deng, Haoge, et al.
Veröffentlicht: (2025)
Omni-Video 2: Scaling MLLM-Conditioned Diffusion for Unified Video Generation and Editing
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
SIGMark: Scalable In-Generation Watermark with Blind Extraction for Video Diffusion
von: Zhu, Xinjie, et al.
Veröffentlicht: (2026)
von: Zhu, Xinjie, et al.
Veröffentlicht: (2026)
SketchVideo: Sketch-based Video Generation and Editing
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2025)
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2025)
Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation
von: Han, Su Ho, et al.
Veröffentlicht: (2025)
von: Han, Su Ho, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MM-LDM: Multi-Modal Latent Diffusion Model for Sounding Video Generation
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024) -
AR-Diffusion: Asynchronous Video Generation with Auto-Regressive Diffusion
von: Sun, Mingzhen, et al.
Veröffentlicht: (2025) -
VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image Synthesis
von: Wu, Shiyu, et al.
Veröffentlicht: (2025) -
ExVideo: Extending Video Diffusion Models via Parameter-Efficient Post-Tuning
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024) -
PhysCorr: Dual-Reward DPO for Physics-Constrained Text-to-Video Generation with Automated Preference Selection
von: Wang, Peiyao, et al.
Veröffentlicht: (2025)