Streaming Autoregressive Video Generation via Diagonal Distillation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Jinxiu, Liu, Xuanming, Mei, Kangfu, Wen, Yandong, Yang, Ming-Hsuan, Liu, Weiyang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DynamicScaler: Seamless and Scalable Video Generation for Panoramic Scenes
von: Liu, Jinxiu, et al.
Veröffentlicht: (2024)
von: Liu, Jinxiu, et al.
Veröffentlicht: (2024)
Fast Autoregressive Video Generation with Diagonal Decoding
von: Ye, Yang, et al.
Veröffentlicht: (2025)
von: Ye, Yang, et al.
Veröffentlicht: (2025)
DyStream: Streaming Dyadic Talking Heads Generation via Flow Matching-based Autoregressive Model
von: Chen, Bohong, et al.
Veröffentlicht: (2025)
von: Chen, Bohong, et al.
Veröffentlicht: (2025)
Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation
von: Li, Ruibin, et al.
Veröffentlicht: (2026)
von: Li, Ruibin, et al.
Veröffentlicht: (2026)
Context Forcing: Consistent Autoregressive Video Generation with Long Context
von: Chen, Shuo, et al.
Veröffentlicht: (2026)
von: Chen, Shuo, et al.
Veröffentlicht: (2026)
DreamForge: Motion-Aware Autoregressive Video Generation for Multi-View Driving Scenes
von: Mei, Jianbiao, et al.
Veröffentlicht: (2024)
von: Mei, Jianbiao, et al.
Veröffentlicht: (2024)
Speculative Jacobi-Denoising Decoding for Accelerating Autoregressive Text-to-image Generation
von: Teng, Yao, et al.
Veröffentlicht: (2025)
von: Teng, Yao, et al.
Veröffentlicht: (2025)
HiAR: Efficient Autoregressive Long Video Generation via Hierarchical Denoising
von: Zou, Kai, et al.
Veröffentlicht: (2026)
von: Zou, Kai, et al.
Veröffentlicht: (2026)
HiStream: Efficient High-Resolution Video Generation via Redundancy-Eliminated Streaming
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion Generation
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
Dual-Stream Diffusion Net for Text-to-Video Generation
von: Liu, Binhui, et al.
Veröffentlicht: (2023)
von: Liu, Binhui, et al.
Veröffentlicht: (2023)
Direction-Aware Diagonal Autoregressive Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2025)
von: Xu, Yijia, et al.
Veröffentlicht: (2025)
Unleashing the Potential of Large Language Models for Text-to-Image Generation through Autoregressive Representation Alignment
von: Xie, Xing, et al.
Veröffentlicht: (2025)
von: Xie, Xing, et al.
Veröffentlicht: (2025)
Drifting Preference Optimization for One-Step Generative Models
von: Jiang, Zhou, et al.
Veröffentlicht: (2026)
von: Jiang, Zhou, et al.
Veröffentlicht: (2026)
Plenoptic Video Generation
von: Fu, Xiao, et al.
Veröffentlicht: (2026)
von: Fu, Xiao, et al.
Veröffentlicht: (2026)
Stream-R1: Reliability-Perplexity Aware Reward Distillation for Streaming Video Generation
von: Wu, Bin, et al.
Veröffentlicht: (2026)
von: Wu, Bin, et al.
Veröffentlicht: (2026)
Rethinking Preference Alignment for Diffusion Models with Classifier-Free Guidance
von: Jiang, Zhou, et al.
Veröffentlicht: (2026)
von: Jiang, Zhou, et al.
Veröffentlicht: (2026)
MotiMotion: Motion-Controlled Video Generation with Visual Reasoning
von: Hsin-Ying, Lee, et al.
Veröffentlicht: (2026)
von: Hsin-Ying, Lee, et al.
Veröffentlicht: (2026)
Stream-T1: Test-Time Scaling for Streaming Video Generation
von: Tu, Yijing, et al.
Veröffentlicht: (2026)
von: Tu, Yijing, et al.
Veröffentlicht: (2026)
Structure From Tracking: Distilling Structure-Preserving Motion for Video Generation
von: Fei, Yang, et al.
Veröffentlicht: (2025)
von: Fei, Yang, et al.
Veröffentlicht: (2025)
EndoGen: Conditional Autoregressive Endoscopic Video Generation
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
DOLLAR: Few-Step Video Generation via Distillation and Latent Reward Optimization
von: Ding, Zihan, et al.
Veröffentlicht: (2024)
von: Ding, Zihan, et al.
Veröffentlicht: (2024)
Loong: Generating Minute-level Long Videos with Autoregressive Language Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
FaceCam: Portrait Video Camera Control via Scale-Aware Conditioning
von: Lyu, Weijie, et al.
Veröffentlicht: (2026)
von: Lyu, Weijie, et al.
Veröffentlicht: (2026)
Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation
von: Zhao, Min, et al.
Veröffentlicht: (2026)
von: Zhao, Min, et al.
Veröffentlicht: (2026)
Symbolic Graphics Programming with Large Language Models
von: Chen, Yamei, et al.
Veröffentlicht: (2025)
von: Chen, Yamei, et al.
Veröffentlicht: (2025)
VideoAR: Autoregressive Video Generation via Next-Frame & Scale Prediction
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
CoDi: Conditional Diffusion Distillation for Higher-Fidelity and Faster Image Generation
von: Mei, Kangfu, et al.
Veröffentlicht: (2023)
von: Mei, Kangfu, et al.
Veröffentlicht: (2023)
InstructRL4Pix: Training Diffusion for Image Editing by Reinforcement Learning
von: Li, Tiancheng, et al.
Veröffentlicht: (2024)
von: Li, Tiancheng, et al.
Veröffentlicht: (2024)
Towards Variable and Coordinated Holistic Co-Speech Motion Generation
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
MIDAS: Multimodal Interactive Digital-humAn Synthesis via Real-time Autoregressive Video Generation
von: Chen, Ming, et al.
Veröffentlicht: (2025)
von: Chen, Ming, et al.
Veröffentlicht: (2025)
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
von: Luo, Yawen, et al.
Veröffentlicht: (2026)
von: Luo, Yawen, et al.
Veröffentlicht: (2026)
EA3D: Online Open-World 3D Object Extraction from Streaming Videos
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2025)
REST: Diffusion-based Real-time End-to-end Streaming Talking Head Generation via ID-Context Caching and Asynchronous Streaming Distillation
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
von: Lu, Yunhong, et al.
Veröffentlicht: (2025)
von: Lu, Yunhong, et al.
Veröffentlicht: (2025)
Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding
von: Wu, Hang, et al.
Veröffentlicht: (2026)
von: Wu, Hang, et al.
Veröffentlicht: (2026)
Veda: Scalable Video Diffusion via Distilled Sparse Attention
von: Han, Shihao, et al.
Veröffentlicht: (2026)
von: Han, Shihao, et al.
Veröffentlicht: (2026)
VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models
von: Ye, Muchao, et al.
Veröffentlicht: (2024)
von: Ye, Muchao, et al.
Veröffentlicht: (2024)
EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation
von: Xiong, Tianwei, et al.
Veröffentlicht: (2026)
von: Xiong, Tianwei, et al.
Veröffentlicht: (2026)
Layer-Aware Video Composition via Split-then-Merge
von: Kara, Ozgur, et al.
Veröffentlicht: (2025)
von: Kara, Ozgur, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DynamicScaler: Seamless and Scalable Video Generation for Panoramic Scenes
von: Liu, Jinxiu, et al.
Veröffentlicht: (2024) -
Fast Autoregressive Video Generation with Diagonal Decoding
von: Ye, Yang, et al.
Veröffentlicht: (2025) -
DyStream: Streaming Dyadic Talking Heads Generation via Flow Matching-based Autoregressive Model
von: Chen, Bohong, et al.
Veröffentlicht: (2025) -
Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation
von: Li, Ruibin, et al.
Veröffentlicht: (2026) -
Context Forcing: Consistent Autoregressive Video Generation with Long Context
von: Chen, Shuo, et al.
Veröffentlicht: (2026)