AR-Diffusion: Asynchronous Video Generation with Auto-Regressive Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Mingzhen, Wang, Weining, Li, Gen, Liu, Jiawei, Sun, Jiahui, Feng, Wanquan, Lao, Shanshan, Zhou, SiYu, He, Qian, Liu, Jing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mask$^2$DiT: Dual Mask-based Diffusion Transformer for Multi-Scene Long Video Generation
von: Qi, Tianhao, et al.
Veröffentlicht: (2025)
von: Qi, Tianhao, et al.
Veröffentlicht: (2025)
COMUNI: Decomposing Common and Unique Video Signals for Diffusion-based Video Generation
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024)
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024)
ProAV-DiT: A Projected Latent Diffusion Transformer for Efficient Synchronized Audio-Video Generation
von: Sun, Jiahui, et al.
Veröffentlicht: (2025)
von: Sun, Jiahui, et al.
Veröffentlicht: (2025)
MM-LDM: Multi-Modal Latent Diffusion Model for Sounding Video Generation
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024)
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024)
I2VControl-Camera: Precise Video Camera Control with Adjustable Motion Strength
von: Feng, Wanquan, et al.
Veröffentlicht: (2024)
von: Feng, Wanquan, et al.
Veröffentlicht: (2024)
I2VControl: Disentangled and Unified Video Motion Synthesis Control
von: Feng, Wanquan, et al.
Veröffentlicht: (2024)
von: Feng, Wanquan, et al.
Veröffentlicht: (2024)
NoiseAR: AutoRegressing Initial Noise Prior for Diffusion Models
von: Li, Zeming, et al.
Veröffentlicht: (2025)
von: Li, Zeming, et al.
Veröffentlicht: (2025)
VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image Synthesis
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework
von: Wang, Jing, et al.
Veröffentlicht: (2025)
von: Wang, Jing, et al.
Veröffentlicht: (2025)
ExVideo: Extending Video Diffusion Models via Parameter-Efficient Post-Tuning
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
SDAR: A Synergistic Diffusion-AutoRegression Paradigm for Scalable Sequence Generation
von: Cheng, Shuang, et al.
Veröffentlicht: (2025)
von: Cheng, Shuang, et al.
Veröffentlicht: (2025)
AnyDressing: Customizable Multi-Garment Virtual Dressing via Latent Diffusion Models
von: Li, Xinghui, et al.
Veröffentlicht: (2024)
von: Li, Xinghui, et al.
Veröffentlicht: (2024)
Auto-Regressive Diffusion for Generating 3D Human-Object Interactions
von: Geng, Zichen, et al.
Veröffentlicht: (2025)
von: Geng, Zichen, et al.
Veröffentlicht: (2025)
Personalized Text-to-Image Generation with Auto-Regressive Models
von: Sun, Kaiyue, et al.
Veröffentlicht: (2025)
von: Sun, Kaiyue, et al.
Veröffentlicht: (2025)
Stream-DiffVSR: Low-Latency Streamable Video Super-Resolution via Auto-Regressive Diffusion
von: Shiu, Hau-Shiang, et al.
Veröffentlicht: (2025)
von: Shiu, Hau-Shiang, et al.
Veröffentlicht: (2025)
DreaMontage: Arbitrary Frame-Guided One-Shot Video Generation
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
Breaking AR's Sampling Bottleneck: Provable Acceleration via Diffusion Language Models
von: Li, Gen, et al.
Veröffentlicht: (2025)
von: Li, Gen, et al.
Veröffentlicht: (2025)
Auto-Regressive Masked Diffusion Models
von: Karami, Mahdi, et al.
Veröffentlicht: (2026)
von: Karami, Mahdi, et al.
Veröffentlicht: (2026)
FAR-Drive: Frame-AutoRegressive Video Generation in Closed-Loop Autonomous Driving
von: Li, Yaoru, et al.
Veröffentlicht: (2026)
von: Li, Yaoru, et al.
Veröffentlicht: (2026)
ZipAR: Parallel Auto-regressive Image Generation through Spatial Locality
von: He, Yefei, et al.
Veröffentlicht: (2024)
von: He, Yefei, et al.
Veröffentlicht: (2024)
Versatile Transition Generation with Image-to-Video Diffusion
von: Yang, Zuhao, et al.
Veröffentlicht: (2025)
von: Yang, Zuhao, et al.
Veröffentlicht: (2025)
Bernini: Latent Semantic Planning for Video Diffusion
von: Bernini Team, et al.
Veröffentlicht: (2026)
von: Bernini Team, et al.
Veröffentlicht: (2026)
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
Diffusion-nested Auto-Regressive Synthesis of Heterogeneous Tabular Data
von: Zhang, Hengrui, et al.
Veröffentlicht: (2024)
von: Zhang, Hengrui, et al.
Veröffentlicht: (2024)
Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation
von: Jie, Pengyu, et al.
Veröffentlicht: (2025)
von: Jie, Pengyu, et al.
Veröffentlicht: (2025)
LibraGen: Playing a Balance Game in Subject-Driven Video Generation
von: Zhu, Jiahao, et al.
Veröffentlicht: (2026)
von: Zhu, Jiahao, et al.
Veröffentlicht: (2026)
VideoAR: Autoregressive Video Generation via Next-Frame & Scale Prediction
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
DreamStyle: A Unified Framework for Video Stylization
von: Li, Mengtian, et al.
Veröffentlicht: (2026)
von: Li, Mengtian, et al.
Veröffentlicht: (2026)
VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step
von: Wang, Hanyang, et al.
Veröffentlicht: (2025)
von: Wang, Hanyang, et al.
Veröffentlicht: (2025)
Fluid Composer: Fluid Detail Composition and Rendering Using Video Diffusion Models
von: Duowen Chen, et al.
Veröffentlicht: (2025)
von: Duowen Chen, et al.
Veröffentlicht: (2025)
GeAR: Generation Augmented Retrieval
von: Liu, Haoyu, et al.
Veröffentlicht: (2025)
von: Liu, Haoyu, et al.
Veröffentlicht: (2025)
AsyncDSB: Schedule-Asynchronous Diffusion Schrödinger Bridge for Image Inpainting
von: Han, Zihao, et al.
Veröffentlicht: (2024)
von: Han, Zihao, et al.
Veröffentlicht: (2024)
STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits
von: Papantoniou, Foivos Paraperas, et al.
Veröffentlicht: (2025)
von: Papantoniou, Foivos Paraperas, et al.
Veröffentlicht: (2025)
Generative Edge Detection with Stable Diffusion
von: Zhou, Caixia, et al.
Veröffentlicht: (2024)
von: Zhou, Caixia, et al.
Veröffentlicht: (2024)
READ: Real-time and Efficient Asynchronous Diffusion for Audio-driven Talking Head Generation
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
Expert-Guided Diffusion Planner for Auto-Bidding
von: Peng, Yunshan, et al.
Veröffentlicht: (2025)
von: Peng, Yunshan, et al.
Veröffentlicht: (2025)
Interactive Character Control with Auto-Regressive Motion Diffusion Models
von: Shi, Yi, et al.
Veröffentlicht: (2023)
von: Shi, Yi, et al.
Veröffentlicht: (2023)
Auto-Regressive Moving Diffusion Models for Time Series Forecasting
von: Gao, Jiaxin, et al.
Veröffentlicht: (2024)
von: Gao, Jiaxin, et al.
Veröffentlicht: (2024)
AR-CoPO: Align Autoregressive Video Generation with Contrastive Policy Optimization
von: He, Dailan, et al.
Veröffentlicht: (2026)
von: He, Dailan, et al.
Veröffentlicht: (2026)
HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mask$^2$DiT: Dual Mask-based Diffusion Transformer for Multi-Scene Long Video Generation
von: Qi, Tianhao, et al.
Veröffentlicht: (2025) -
COMUNI: Decomposing Common and Unique Video Signals for Diffusion-based Video Generation
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024) -
ProAV-DiT: A Projected Latent Diffusion Transformer for Efficient Synchronized Audio-Video Generation
von: Sun, Jiahui, et al.
Veröffentlicht: (2025) -
MM-LDM: Multi-Modal Latent Diffusion Model for Sounding Video Generation
von: Sun, Mingzhen, et al.
Veröffentlicht: (2024) -
I2VControl-Camera: Precise Video Camera Control with Adjustable Motion Strength
von: Feng, Wanquan, et al.
Veröffentlicht: (2024)