Parallelized Autoregressive Visual Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yuqing, Ren, Shuhuai, Lin, Zhijie, Han, Yujin, Guo, Haoyuan, Yang, Zhenheng, Zou, Difan, Feng, Jiashi, Liu, Xihui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation
von: Xiong, Tianwei, et al.
Veröffentlicht: (2026)
von: Xiong, Tianwei, et al.
Veröffentlicht: (2026)
Loong: Generating Minute-level Long Videos with Autoregressive Language Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation
von: Xiong, Tianwei, et al.
Veröffentlicht: (2025)
von: Xiong, Tianwei, et al.
Veröffentlicht: (2025)
Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
LVD-2M: A Long-take Video Dataset with Temporally Dense Captions
von: Xiong, Tianwei, et al.
Veröffentlicht: (2024)
von: Xiong, Tianwei, et al.
Veröffentlicht: (2024)
Speculative Jacobi-Denoising Decoding for Accelerating Autoregressive Text-to-image Generation
von: Teng, Yao, et al.
Veröffentlicht: (2025)
von: Teng, Yao, et al.
Veröffentlicht: (2025)
UVE: Are MLLMs Unified Evaluators for AI-Generated Videos?
von: Liu, Yuanxin, et al.
Veröffentlicht: (2025)
von: Liu, Yuanxin, et al.
Veröffentlicht: (2025)
Understand Before You Generate: Self-Guided Training for Autoregressive Image Generation
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2025)
Next Block Prediction: Video Generation via Semi-Autoregressive Modeling
von: Ren, Shuhuai, et al.
Veröffentlicht: (2025)
von: Ren, Shuhuai, et al.
Veröffentlicht: (2025)
Long Context Tuning for Video Generation
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance
von: Wei, Yujie, et al.
Veröffentlicht: (2025)
von: Wei, Yujie, et al.
Veröffentlicht: (2025)
Can Diffusion Models Learn Hidden Inter-Feature Rules Behind Images?
von: Han, Yujin, et al.
Veröffentlicht: (2025)
von: Han, Yujin, et al.
Veröffentlicht: (2025)
CAR: Controllable Autoregressive Modeling for Visual Generation
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
MACRO: Advancing Multi-Reference Image Generation with Structured Long-Context Data
von: Chen, Zhekai, et al.
Veröffentlicht: (2026)
von: Chen, Zhekai, et al.
Veröffentlicht: (2026)
Autoregressive Image Generation with Randomized Parallel Decoding
von: Li, Haopeng, et al.
Veröffentlicht: (2025)
von: Li, Haopeng, et al.
Veröffentlicht: (2025)
End-to-End Training for Autoregressive Video Diffusion via Self-Resampling
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting
von: Zou, Zhen, et al.
Veröffentlicht: (2026)
von: Zou, Zhen, et al.
Veröffentlicht: (2026)
LaDiC: Are Diffusion Models Really Inferior to Autoregressive Counterparts for Image-to-Text Generation?
von: Wang, Yuchi, et al.
Veröffentlicht: (2024)
von: Wang, Yuchi, et al.
Veröffentlicht: (2024)
How Far is Video Generation from World Model: A Physical Law Perspective
von: Kang, Bingyi, et al.
Veröffentlicht: (2024)
von: Kang, Bingyi, et al.
Veröffentlicht: (2024)
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
von: Xu, Lin, et al.
Veröffentlicht: (2024)
von: Xu, Lin, et al.
Veröffentlicht: (2024)
Head-Aware KV Cache Compression for Efficient Visual Autoregressive Modeling
von: Qin, Ziran, et al.
Veröffentlicht: (2025)
von: Qin, Ziran, et al.
Veröffentlicht: (2025)
Progressive Autoregressive Video Diffusion Models
von: Xie, Desai, et al.
Veröffentlicht: (2024)
von: Xie, Desai, et al.
Veröffentlicht: (2024)
BitDance: Scaling Autoregressive Generative Models with Binary Tokens
von: Ai, Yuang, et al.
Veröffentlicht: (2026)
von: Ai, Yuang, et al.
Veröffentlicht: (2026)
Exploiting Discriminative Codebook Prior for Autoregressive Image Generation
von: Tang, Longxiang, et al.
Veröffentlicht: (2025)
von: Tang, Longxiang, et al.
Veröffentlicht: (2025)
VCE: Safe Autoregressive Image Generation via Visual Contrast Exploitation
von: Han, Feng, et al.
Veröffentlicht: (2025)
von: Han, Feng, et al.
Veröffentlicht: (2025)
Next Patch Prediction for Autoregressive Visual Generation
von: Pang, Yatian, et al.
Veröffentlicht: (2024)
von: Pang, Yatian, et al.
Veröffentlicht: (2024)
Growing Visual Generative Capacity for Pre-Trained MLLMs
von: Wang, Hanyu, et al.
Veröffentlicht: (2025)
von: Wang, Hanyu, et al.
Veröffentlicht: (2025)
Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2025)
From Sequential to Spatial: Reordering Autoregression for Efficient Visual Generation
von: Wang, Siyang, et al.
Veröffentlicht: (2025)
von: Wang, Siyang, et al.
Veröffentlicht: (2025)
VMRNN: Integrating Vision Mamba and LSTM for Efficient and Accurate Spatiotemporal Forecasting
von: Tang, Yujin, et al.
Veröffentlicht: (2024)
von: Tang, Yujin, et al.
Veröffentlicht: (2024)
Image Understanding Makes for A Good Tokenizer for Image Generation
von: Wang, Luting, et al.
Veröffentlicht: (2024)
von: Wang, Luting, et al.
Veröffentlicht: (2024)
Randomized Autoregressive Visual Generation
von: Yu, Qihang, et al.
Veröffentlicht: (2024)
von: Yu, Qihang, et al.
Veröffentlicht: (2024)
AesRM: Improving Video Aesthetics with Expert-Level Feedback
von: Han, Yujin, et al.
Veröffentlicht: (2026)
von: Han, Yujin, et al.
Veröffentlicht: (2026)
Macro-from-Micro Planning for High-Quality and Parallelized Autoregressive Long Video Generation
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
LongStream: Long-Sequence Streaming Autoregressive Visual Geometry
von: Cheng, Chong, et al.
Veröffentlicht: (2026)
von: Cheng, Chong, et al.
Veröffentlicht: (2026)
Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
von: Ma, Fan, et al.
Veröffentlicht: (2023)
von: Ma, Fan, et al.
Veröffentlicht: (2023)
Representation Forcing for Bottleneck-Free Unified Multimodal Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
Mirai: Autoregressive Visual Generation Needs Foresight
von: Yu, Yonghao, et al.
Veröffentlicht: (2026)
von: Yu, Yonghao, et al.
Veröffentlicht: (2026)
REAR: Rethinking Visual Autoregressive Models via Generator-Tokenizer Consistency Regularization
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2025) -
EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation
von: Xiong, Tianwei, et al.
Veröffentlicht: (2026) -
Loong: Generating Minute-level Long Videos with Autoregressive Language Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2024) -
GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation
von: Xiong, Tianwei, et al.
Veröffentlicht: (2025) -
Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)