LINA: Linear Autoregressive Image Generative Models with Continuous Tokens
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jiahao, Pan, Ting, Deng, Haoge, Han, Dongchen, Wu, Taiqiang, Wang, Xinlong, Luo, Ping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Autoregressive Video Generation without Vector Quantization
von: Deng, Haoge, et al.
Veröffentlicht: (2024)
von: Deng, Haoge, et al.
Veröffentlicht: (2024)
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models
von: Wang, Jiaqi, et al.
Veröffentlicht: (2026)
von: Wang, Jiaqi, et al.
Veröffentlicht: (2026)
Frequency Autoregressive Image Generation with Continuous Tokens
von: Yu, Hu, et al.
Veröffentlicht: (2025)
von: Yu, Hu, et al.
Veröffentlicht: (2025)
Uniform Discrete Diffusion with Metric Path for Video Generation
von: Deng, Haoge, et al.
Veröffentlicht: (2025)
von: Deng, Haoge, et al.
Veröffentlicht: (2025)
Tokenize Anything via Prompting
von: Pan, Ting, et al.
Veröffentlicht: (2023)
von: Pan, Ting, et al.
Veröffentlicht: (2023)
NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale
von: NextStep Team, et al.
Veröffentlicht: (2025)
von: NextStep Team, et al.
Veröffentlicht: (2025)
EVEv2: Improved Baselines for Encoder-Free Vision-Language Models
von: Diao, Haiwen, et al.
Veröffentlicht: (2025)
von: Diao, Haiwen, et al.
Veröffentlicht: (2025)
Hita: Holistic Tokenizer for Autoregressive Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
LiT: Delving into a Simple Linear Diffusion Transformer for Image Generation
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
You See it, You Got it: Learning 3D Creation on Pose-Free Videos at Scale
von: Ma, Baorui, et al.
Veröffentlicht: (2024)
von: Ma, Baorui, et al.
Veröffentlicht: (2024)
Improving Flexible Image Tokenizers for Autoregressive Image Generation
von: Fu, Zixuan, et al.
Veröffentlicht: (2026)
von: Fu, Zixuan, et al.
Veröffentlicht: (2026)
Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis
von: Zheng, Peng, et al.
Veröffentlicht: (2025)
von: Zheng, Peng, et al.
Veröffentlicht: (2025)
D2C: Unlocking the Potential of Continuous Autoregressive Image Generation with Discrete Tokens
von: Wang, Panpan, et al.
Veröffentlicht: (2025)
von: Wang, Panpan, et al.
Veröffentlicht: (2025)
CI-VID: A Coherent Interleaved Text-Video Dataset
von: Ju, Yiming, et al.
Veröffentlicht: (2025)
von: Ju, Yiming, et al.
Veröffentlicht: (2025)
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
Linear-Time Global Visual Modeling without Explicit Attention
von: He, Ruize, et al.
Veröffentlicht: (2026)
von: He, Ruize, et al.
Veröffentlicht: (2026)
DiCoDe: Diffusion-Compressed Deep Tokens for Autoregressive Video Generation with Language Models
von: Li, Yizhuo, et al.
Veröffentlicht: (2024)
von: Li, Yizhuo, et al.
Veröffentlicht: (2024)
ImageFolder: Autoregressive Image Generation with Folded Tokens
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Token-Shuffle: Towards High-Resolution Image Generation with Autoregressive Models
von: Ma, Xu, et al.
Veröffentlicht: (2025)
von: Ma, Xu, et al.
Veröffentlicht: (2025)
Continuous Speculative Decoding for Autoregressive Image Generation
von: Wang, Zili, et al.
Veröffentlicht: (2024)
von: Wang, Zili, et al.
Veröffentlicht: (2024)
VL-Trojan: Multimodal Instruction Backdoor Attacks against Autoregressive Visual Language Models
von: Liang, Jiawei, et al.
Veröffentlicht: (2024)
von: Liang, Jiawei, et al.
Veröffentlicht: (2024)
FastVAR: Linear Visual Autoregressive Modeling via Cached Token Pruning
von: Guo, Hang, et al.
Veröffentlicht: (2025)
von: Guo, Hang, et al.
Veröffentlicht: (2025)
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
von: Sun, Peize, et al.
Veröffentlicht: (2024)
von: Sun, Peize, et al.
Veröffentlicht: (2024)
XQ-GAN: An Open-source Image Tokenization Framework for Autoregressive Generation
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation
von: Wang, Lingfeng, et al.
Veröffentlicht: (2025)
von: Wang, Lingfeng, et al.
Veröffentlicht: (2025)
GSVA: Generalized Segmentation via Multimodal Large Language Models
von: Xia, Zhuofan, et al.
Veröffentlicht: (2023)
von: Xia, Zhuofan, et al.
Veröffentlicht: (2023)
Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens
von: Fan, Lijie, et al.
Veröffentlicht: (2024)
von: Fan, Lijie, et al.
Veröffentlicht: (2024)
Hyperspherical Latents Improve Continuous-Token Autoregressive Generation
von: Ke, Guolin, et al.
Veröffentlicht: (2025)
von: Ke, Guolin, et al.
Veröffentlicht: (2025)
Unified Autoregressive Visual Generation and Understanding with Continuous Tokens
von: Fan, Lijie, et al.
Veröffentlicht: (2025)
von: Fan, Lijie, et al.
Veröffentlicht: (2025)
Autoregressive Models in Vision: A Survey
von: Xiong, Jing, et al.
Veröffentlicht: (2024)
von: Xiong, Jing, et al.
Veröffentlicht: (2024)
Selftok: Discrete Visual Tokens of Autoregression, by Diffusion, and for Reasoning
von: Wang, Bohan, et al.
Veröffentlicht: (2025)
von: Wang, Bohan, et al.
Veröffentlicht: (2025)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
BitDance: Scaling Autoregressive Generative Models with Binary Tokens
von: Ai, Yuang, et al.
Veröffentlicht: (2026)
von: Ai, Yuang, et al.
Veröffentlicht: (2026)
Fractal Autoregressive Depth Estimation with Continuous Token Diffusion
von: Zhang, Jinchang, et al.
Veröffentlicht: (2026)
von: Zhang, Jinchang, et al.
Veröffentlicht: (2026)
Improving Autoregressive Image Generation through Coarse-to-Fine Token Prediction
von: Guo, Ziyao, et al.
Veröffentlicht: (2025)
von: Guo, Ziyao, et al.
Veröffentlicht: (2025)
LlamaSeg: Image Segmentation via Autoregressive Mask Generation
von: Deng, Jiru, et al.
Veröffentlicht: (2025)
von: Deng, Jiru, et al.
Veröffentlicht: (2025)
Adapting LLaMA Decoder to Vision Transformer
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens
von: Li, Zekun, et al.
Veröffentlicht: (2026)
von: Li, Zekun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Autoregressive Video Generation without Vector Quantization
von: Deng, Haoge, et al.
Veröffentlicht: (2024) -
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models
von: Wang, Jiaqi, et al.
Veröffentlicht: (2026) -
Frequency Autoregressive Image Generation with Continuous Tokens
von: Yu, Hu, et al.
Veröffentlicht: (2025) -
Uniform Discrete Diffusion with Metric Path for Video Generation
von: Deng, Haoge, et al.
Veröffentlicht: (2025) -
Tokenize Anything via Prompting
von: Pan, Ting, et al.
Veröffentlicht: (2023)