CAR: Controllable Autoregressive Modeling for Visual Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yao, Ziyu, Li, Jialin, Zhou, Yifeng, Liu, Yong, Jiang, Xi, Wang, Chengjie, Zheng, Feng, Zou, Yuexian, Li, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MMAD: A Comprehensive Benchmark for Multimodal Large Language Models in Industrial Anomaly Detection
von: Jiang, Xi, et al.
Veröffentlicht: (2024)
von: Jiang, Xi, et al.
Veröffentlicht: (2024)
VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language Model
von: Zhuang, Xianwei, et al.
Veröffentlicht: (2025)
von: Zhuang, Xianwei, et al.
Veröffentlicht: (2025)
Decision Boundary-aware Knowledge Consolidation Generates Better Instance-Incremental Learner
von: Nie, Qiang, et al.
Veröffentlicht: (2024)
von: Nie, Qiang, et al.
Veröffentlicht: (2024)
Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting
von: Zou, Zhen, et al.
Veröffentlicht: (2026)
von: Zou, Zhen, et al.
Veröffentlicht: (2026)
Parallelized Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
REAR: Rethinking Visual Autoregressive Models via Generator-Tokenizer Consistency Regularization
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
Toward Multi-class Anomaly Detection: Exploring Class-aware Unified Model against Inter-class Interference
von: Jiang, Xi, et al.
Veröffentlicht: (2024)
von: Jiang, Xi, et al.
Veröffentlicht: (2024)
Visual Implicit Autoregressive Modeling
von: Jiang, Pengfei, et al.
Veröffentlicht: (2026)
von: Jiang, Pengfei, et al.
Veröffentlicht: (2026)
Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
Autoregressive Meta-Actions for Unified Controllable Trajectory Generation
von: Zhao, Jianbo, et al.
Veröffentlicht: (2025)
von: Zhao, Jianbo, et al.
Veröffentlicht: (2025)
DreamVAR: Taming Reinforced Visual Autoregressive Model for High-Fidelity Subject-Driven Image Generation
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding
von: Chen, Zhanpeng, et al.
Veröffentlicht: (2025)
von: Chen, Zhanpeng, et al.
Veröffentlicht: (2025)
LSRS: Latent Scale Rejection Sampling for Visual Autoregressive Modeling
von: Zheng, Hong-Kai, et al.
Veröffentlicht: (2025)
von: Zheng, Hong-Kai, et al.
Veröffentlicht: (2025)
Seg-VAR: Image Segmentation with Visual Autoregressive Modeling
von: Zheng, Rongkun, et al.
Veröffentlicht: (2025)
von: Zheng, Rongkun, et al.
Veröffentlicht: (2025)
VARGPT-v1.1: Improve Visual Autoregressive Large Unified Model via Iterative Instruction Tuning and Reinforcement Learning
von: Zhuang, Xianwei, et al.
Veröffentlicht: (2025)
von: Zhuang, Xianwei, et al.
Veröffentlicht: (2025)
From Sequential to Spatial: Reordering Autoregression for Efficient Visual Generation
von: Wang, Siyang, et al.
Veröffentlicht: (2025)
von: Wang, Siyang, et al.
Veröffentlicht: (2025)
ControlAR: Controllable Image Generation with Autoregressive Models
von: Li, Zongming, et al.
Veröffentlicht: (2024)
von: Li, Zongming, et al.
Veröffentlicht: (2024)
VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsification
von: Zhuang, Xianwei, et al.
Veröffentlicht: (2025)
von: Zhuang, Xianwei, et al.
Veröffentlicht: (2025)
Progressive Supernet Training for Efficient Visual Autoregressive Modeling
von: Chen, Xiaoyue, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyue, et al.
Veröffentlicht: (2025)
StepVAR: Structure-Texture Guided Pruning for Visual Autoregressive Models
von: Liu, Keli, et al.
Veröffentlicht: (2026)
von: Liu, Keli, et al.
Veröffentlicht: (2026)
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
StarGen: A Spatiotemporal Autoregression Framework with Video Diffusion Model for Scalable and Controllable Scene Generation
von: Zhai, Shangjin, et al.
Veröffentlicht: (2025)
von: Zhai, Shangjin, et al.
Veröffentlicht: (2025)
PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training
von: Fu, Weifu, et al.
Veröffentlicht: (2026)
von: Fu, Weifu, et al.
Veröffentlicht: (2026)
MoSa: Motion Generation with Scalable Autoregressive Modeling
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation
von: Liao, Xinyao, et al.
Veröffentlicht: (2026)
von: Liao, Xinyao, et al.
Veröffentlicht: (2026)
ControlVAR: Exploring Controllable Visual Autoregressive Modeling
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Neighboring Autoregressive Modeling for Efficient Visual Generation
von: He, Yefei, et al.
Veröffentlicht: (2025)
von: He, Yefei, et al.
Veröffentlicht: (2025)
VCE: Safe Autoregressive Image Generation via Visual Contrast Exploitation
von: Han, Feng, et al.
Veröffentlicht: (2025)
von: Han, Feng, et al.
Veröffentlicht: (2025)
SoftPatch+: Fully Unsupervised Anomaly Classification and Segmentation
von: Wang, Chengjie, et al.
Veröffentlicht: (2024)
von: Wang, Chengjie, et al.
Veröffentlicht: (2024)
AD-Copilot: A Vision-Language Assistant for Industrial Anomaly Detection via Visual In-context Comparison
von: Jiang, Xi, et al.
Veröffentlicht: (2026)
von: Jiang, Xi, et al.
Veröffentlicht: (2026)
Visual Autoregressive Modeling for Instruction-Guided Image Editing
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
Context-Aware Autoregressive Models for Multi-Conditional Image Generation
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
SCALAR: Scale-wise Controllable Visual Autoregressive Learning
von: Xu, Ryan, et al.
Veröffentlicht: (2025)
von: Xu, Ryan, et al.
Veröffentlicht: (2025)
Mirai: Autoregressive Visual Generation Needs Foresight
von: Yu, Yonghao, et al.
Veröffentlicht: (2026)
von: Yu, Yonghao, et al.
Veröffentlicht: (2026)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
Efficient Conditional Generation on Scale-based Visual Autoregressive Models
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
Autoregressive Image Generation with Masked Bit Modeling
von: Yu, Qihang, et al.
Veröffentlicht: (2026)
von: Yu, Qihang, et al.
Veröffentlicht: (2026)
Depth Adaptive Efficient Visual Autoregressive Modeling
von: Li, Chunliang, et al.
Veröffentlicht: (2026)
von: Li, Chunliang, et al.
Veröffentlicht: (2026)
Speculative Jacobi-Denoising Decoding for Accelerating Autoregressive Text-to-image Generation
von: Teng, Yao, et al.
Veröffentlicht: (2025)
von: Teng, Yao, et al.
Veröffentlicht: (2025)
CPC-VAR:Continual Personalized and Compositional Generation in Visual Autoregressive Models
von: Li, Junhao, et al.
Veröffentlicht: (2026)
von: Li, Junhao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MMAD: A Comprehensive Benchmark for Multimodal Large Language Models in Industrial Anomaly Detection
von: Jiang, Xi, et al.
Veröffentlicht: (2024) -
VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language Model
von: Zhuang, Xianwei, et al.
Veröffentlicht: (2025) -
Decision Boundary-aware Knowledge Consolidation Generates Better Instance-Incremental Learner
von: Nie, Qiang, et al.
Veröffentlicht: (2024) -
Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting
von: Zou, Zhen, et al.
Veröffentlicht: (2026) -
Parallelized Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)