Taming the Entropy Cliff: Variable Codebook Size Quantization for Autoregressive Visual Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Bowen, Luo, Weijian, Yang, Guang, Zhang, Colin, Hu, Tianyang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning Discrete Autoregressive Priors with Wasserstein Gradient Flow
di: Zheng, Bowen, et al.
Pubblicazione: (2026)
di: Zheng, Bowen, et al.
Pubblicazione: (2026)
Patch-aware Vector Quantized Codebook Learning for Unsupervised Visual Defect Detection
di: Cheng, Qisen, et al.
Pubblicazione: (2025)
di: Cheng, Qisen, et al.
Pubblicazione: (2025)
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
di: Wu, Junyi, et al.
Pubblicazione: (2026)
di: Wu, Junyi, et al.
Pubblicazione: (2026)
Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences
di: Luo, Weijian
Pubblicazione: (2024)
di: Luo, Weijian
Pubblicazione: (2024)
TDM-R1: Reinforcing Few-Step Diffusion Models with Non-Differentiable Reward
di: Luo, Yihong, et al.
Pubblicazione: (2026)
di: Luo, Yihong, et al.
Pubblicazione: (2026)
SpectralAR: Spectral Autoregressive Visual Generation
di: Huang, Yuanhui, et al.
Pubblicazione: (2025)
di: Huang, Yuanhui, et al.
Pubblicazione: (2025)
CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention
di: Lin, Gaojie, et al.
Pubblicazione: (2024)
di: Lin, Gaojie, et al.
Pubblicazione: (2024)
HART: Efficient Visual Generation with Hybrid Autoregressive Transformer
di: Tang, Haotian, et al.
Pubblicazione: (2024)
di: Tang, Haotian, et al.
Pubblicazione: (2024)
David and Goliath: Small One-step Model Beats Large Diffusion with Score Post-training
di: Luo, Weijian, et al.
Pubblicazione: (2024)
di: Luo, Weijian, et al.
Pubblicazione: (2024)
VTBench: Evaluating Visual Tokenizers for Autoregressive Image Generation
di: Lin, Huawei, et al.
Pubblicazione: (2025)
di: Lin, Huawei, et al.
Pubblicazione: (2025)
BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook
di: Gu, Hao, et al.
Pubblicazione: (2025)
di: Gu, Hao, et al.
Pubblicazione: (2025)
AGDC: Autoregressive Generation of Variable-Length Sequences with Joint Discrete and Continuous Spaces
di: Shin, Yeonsang, et al.
Pubblicazione: (2026)
di: Shin, Yeonsang, et al.
Pubblicazione: (2026)
DC-PCN: Point Cloud Completion Network with Dual-Codebook Guided Quantization
di: Wu, Qiuxia, et al.
Pubblicazione: (2025)
di: Wu, Qiuxia, et al.
Pubblicazione: (2025)
Denoising Fisher Training For Neural Implicit Samplers
di: Luo, Weijian, et al.
Pubblicazione: (2024)
di: Luo, Weijian, et al.
Pubblicazione: (2024)
Spherical Leech Quantization for Visual Tokenization and Generation
di: Zhao, Yue, et al.
Pubblicazione: (2025)
di: Zhao, Yue, et al.
Pubblicazione: (2025)
Flow Generator Matching
di: Huang, Zemin, et al.
Pubblicazione: (2024)
di: Huang, Zemin, et al.
Pubblicazione: (2024)
Astra: General Interactive World Model with Autoregressive Denoising
di: Zhu, Yixuan, et al.
Pubblicazione: (2025)
di: Zhu, Yixuan, et al.
Pubblicazione: (2025)
MASC: Boosting Autoregressive Image Generation with a Manifold-Aligned Semantic Clustering
di: He, Lixuan, et al.
Pubblicazione: (2025)
di: He, Lixuan, et al.
Pubblicazione: (2025)
Visual Autoregressive Transformers Must Use $Ω(n^2 d)$ Memory
di: Cao, Yang, et al.
Pubblicazione: (2025)
di: Cao, Yang, et al.
Pubblicazione: (2025)
Spanning Tree Autoregressive Visual Generation
di: Lee, Sangkyu, et al.
Pubblicazione: (2025)
di: Lee, Sangkyu, et al.
Pubblicazione: (2025)
FPQVAR: Floating Point Quantization for Visual Autoregressive Model with FPGA Hardware Co-design
di: Wei, Renjie, et al.
Pubblicazione: (2025)
di: Wei, Renjie, et al.
Pubblicazione: (2025)
$\bf{D^3}$QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection
di: Zhang, Yanran, et al.
Pubblicazione: (2025)
di: Zhang, Yanran, et al.
Pubblicazione: (2025)
Autoregressive Adversarial Post-Training for Real-Time Interactive Video Generation
di: Lin, Shanchuan, et al.
Pubblicazione: (2025)
di: Lin, Shanchuan, et al.
Pubblicazione: (2025)
Frequency Autoregressive Image Generation with Continuous Tokens
di: Yu, Hu, et al.
Pubblicazione: (2025)
di: Yu, Hu, et al.
Pubblicazione: (2025)
Adversarial Error Correction for Visual Autoregressive Generation
di: Bi, Ligong, et al.
Pubblicazione: (2026)
di: Bi, Ligong, et al.
Pubblicazione: (2026)
VariViT: A Vision Transformer for Variable Image Sizes
di: Varma, Aswathi, et al.
Pubblicazione: (2026)
di: Varma, Aswathi, et al.
Pubblicazione: (2026)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
di: Zheng, Anlin, et al.
Pubblicazione: (2025)
di: Zheng, Anlin, et al.
Pubblicazione: (2025)
Efficiently Training A Flat Neural Network Before It has been Quantizated
di: Xia, Peng, et al.
Pubblicazione: (2025)
di: Xia, Peng, et al.
Pubblicazione: (2025)
Taming Outlier Tokens in Diffusion Transformers
di: Wu, Xiaoyu, et al.
Pubblicazione: (2026)
di: Wu, Xiaoyu, et al.
Pubblicazione: (2026)
Universal Approximation of Visual Autoregressive Transformers
di: Chen, Yifang, et al.
Pubblicazione: (2025)
di: Chen, Yifang, et al.
Pubblicazione: (2025)
Speculative Decoding for Autoregressive Video Generation
di: Hu, Yuezhou, et al.
Pubblicazione: (2026)
di: Hu, Yuezhou, et al.
Pubblicazione: (2026)
Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization
di: Gallici, Matteo, et al.
Pubblicazione: (2025)
di: Gallici, Matteo, et al.
Pubblicazione: (2025)
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
di: Lin, Chieh Hubert, et al.
Pubblicazione: (2024)
di: Lin, Chieh Hubert, et al.
Pubblicazione: (2024)
Purrception: Variational Flow Matching for Vector-Quantized Image Generation
di: Matişan, Răzvan-Andrei, et al.
Pubblicazione: (2025)
di: Matişan, Răzvan-Andrei, et al.
Pubblicazione: (2025)
Taming Diffusion for Dataset Distillation with High Representativeness
di: Zhao, Lin, et al.
Pubblicazione: (2025)
di: Zhao, Lin, et al.
Pubblicazione: (2025)
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
di: Gidaris, Spyros, et al.
Pubblicazione: (2023)
di: Gidaris, Spyros, et al.
Pubblicazione: (2023)
JetFormer: An Autoregressive Generative Model of Raw Images and Text
di: Tschannen, Michael, et al.
Pubblicazione: (2024)
di: Tschannen, Michael, et al.
Pubblicazione: (2024)
Annealed Relaxation of Speculative Decoding for Faster Autoregressive Image Generation
di: Li, Xingyao, et al.
Pubblicazione: (2026)
di: Li, Xingyao, et al.
Pubblicazione: (2026)
Visual Generation Without Guidance
di: Chen, Huayu, et al.
Pubblicazione: (2025)
di: Chen, Huayu, et al.
Pubblicazione: (2025)
DocSynthv2: A Practical Autoregressive Modeling for Document Generation
di: Biswas, Sanket, et al.
Pubblicazione: (2024)
di: Biswas, Sanket, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Learning Discrete Autoregressive Priors with Wasserstein Gradient Flow
di: Zheng, Bowen, et al.
Pubblicazione: (2026) -
Patch-aware Vector Quantized Codebook Learning for Unsupervised Visual Defect Detection
di: Cheng, Qisen, et al.
Pubblicazione: (2025) -
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
di: Wu, Junyi, et al.
Pubblicazione: (2026) -
Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences
di: Luo, Weijian
Pubblicazione: (2024) -
TDM-R1: Reinforcing Few-Step Diffusion Models with Non-Differentiable Reward
di: Luo, Yihong, et al.
Pubblicazione: (2026)