SSG: Scaled Spatial Guidance for Multi-Scale Visual Autoregressive Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shin, Youngwoo, Hur, Jiwan, Kim, Junmo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Neural Deformation Representation for 4D Dynamic Shape Generation
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
Unlocking the Capabilities of Masked Generative Models for Image Synthesis via Self-Guidance
von: Hur, Jiwan, et al.
Veröffentlicht: (2024)
von: Hur, Jiwan, et al.
Veröffentlicht: (2024)
Inlier-Centric Post-Training Quantization for Object Detection Models
von: Kim, Minsu, et al.
Veröffentlicht: (2026)
von: Kim, Minsu, et al.
Veröffentlicht: (2026)
Frequency-Aware Token Reduction for Efficient Vision Transformer
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
B4DL: A Benchmark for 4D LiDAR LLM in Spatio-Temporal Understanding
von: Choi, Changho, et al.
Veröffentlicht: (2025)
von: Choi, Changho, et al.
Veröffentlicht: (2025)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian Conditioning
von: Zhang, Jinhua, et al.
Veröffentlicht: (2025)
von: Zhang, Jinhua, et al.
Veröffentlicht: (2025)
SSG-Dit: A Spatial Signal Guided Framework for Controllable Video Generation
von: Hu, Peng, et al.
Veröffentlicht: (2025)
von: Hu, Peng, et al.
Veröffentlicht: (2025)
Efficient Conditional Generation on Scale-based Visual Autoregressive Models
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation
von: Liao, Xinyao, et al.
Veröffentlicht: (2026)
von: Liao, Xinyao, et al.
Veröffentlicht: (2026)
ScaleMoGen: Autoregressive Next-Scale Prediction for Human Motion Generation
von: Hwang, Inwoo, et al.
Veröffentlicht: (2026)
von: Hwang, Inwoo, et al.
Veröffentlicht: (2026)
Markovian Scale Prediction: A New Era of Visual Autoregressive Generation
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
GenAR: Next-Scale Autoregressive Generation for Spatial Gene Expression Prediction
von: Ouyang, Jiarui, et al.
Veröffentlicht: (2025)
von: Ouyang, Jiarui, et al.
Veröffentlicht: (2025)
AMS-KV: Adaptive KV Caching in Multi-Scale Visual Autoregressive Transformers
von: Xu, Boxun, et al.
Veröffentlicht: (2025)
von: Xu, Boxun, et al.
Veröffentlicht: (2025)
SCALAR: Scale-wise Controllable Visual Autoregressive Learning
von: Xu, Ryan, et al.
Veröffentlicht: (2025)
von: Xu, Ryan, et al.
Veröffentlicht: (2025)
GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation
von: Xiong, Tianwei, et al.
Veröffentlicht: (2025)
von: Xiong, Tianwei, et al.
Veröffentlicht: (2025)
LSRS: Latent Scale Rejection Sampling for Visual Autoregressive Modeling
von: Zheng, Hong-Kai, et al.
Veröffentlicht: (2025)
von: Zheng, Hong-Kai, et al.
Veröffentlicht: (2025)
From Sequential to Spatial: Reordering Autoregression for Efficient Visual Generation
von: Wang, Siyang, et al.
Veröffentlicht: (2025)
von: Wang, Siyang, et al.
Veröffentlicht: (2025)
Next-Scale Autoregressive Models for Text-to-Motion Generation
von: Zheng, Zhiwei, et al.
Veröffentlicht: (2026)
von: Zheng, Zhiwei, et al.
Veröffentlicht: (2026)
Visual Autoregressive Models Beat Diffusion Models on Inference Time Scaling
von: Riise, Erik, et al.
Veröffentlicht: (2025)
von: Riise, Erik, et al.
Veröffentlicht: (2025)
Progress by Pieces: Test-Time Scaling for Autoregressive Image Generation
von: Park, Joonhyung, et al.
Veröffentlicht: (2025)
von: Park, Joonhyung, et al.
Veröffentlicht: (2025)
Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding
von: Kim, Jiwan, et al.
Veröffentlicht: (2026)
von: Kim, Jiwan, et al.
Veröffentlicht: (2026)
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
von: Tian, Keyu, et al.
Veröffentlicht: (2024)
von: Tian, Keyu, et al.
Veröffentlicht: (2024)
DiverseVAR: Balancing Diversity and Quality of Next-Scale Visual Autoregressive Models
von: Park, Mingue, et al.
Veröffentlicht: (2025)
von: Park, Mingue, et al.
Veröffentlicht: (2025)
SoftCFG: Uncertainty-guided Stable Guidance for Visual Autoregressive Model
von: Xu, Dongli, et al.
Veröffentlicht: (2025)
von: Xu, Dongli, et al.
Veröffentlicht: (2025)
Spatial-Temporal Multi-Scale Quantization for Flexible Motion Generation
von: Wang, Zan, et al.
Veröffentlicht: (2025)
von: Wang, Zan, et al.
Veröffentlicht: (2025)
Go with Your Gut: Scaling Confidence for Autoregressive Image Generation
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
A Training-Free Style-aligned Image Generation with Scale-wise Autoregressive Model
von: Park, Jihun, et al.
Veröffentlicht: (2025)
von: Park, Jihun, et al.
Veröffentlicht: (2025)
MAGI-1: Autoregressive Video Generation at Scale
von: ai, Sand., et al.
Veröffentlicht: (2025)
von: ai, Sand., et al.
Veröffentlicht: (2025)
Teaching Metric Distance to Discrete Autoregressive Language Models
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
Parallelized Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
Randomized Autoregressive Visual Generation
von: Yu, Qihang, et al.
Veröffentlicht: (2024)
von: Yu, Qihang, et al.
Veröffentlicht: (2024)
MSC: Multi-Scale Spatio-Temporal Causal Attention for Autoregressive Video Diffusion
von: Xu, Xunnong, et al.
Veröffentlicht: (2024)
von: Xu, Xunnong, et al.
Veröffentlicht: (2024)
Classifier-free Guidance with Adaptive Scaling
von: Malarz, Dawid, et al.
Veröffentlicht: (2025)
von: Malarz, Dawid, et al.
Veröffentlicht: (2025)
Dense Cross-Scale Image Alignment With Fully Spatial Correlation and Just Noticeable Difference Guidance
von: You, Jinkun, et al.
Veröffentlicht: (2025)
von: You, Jinkun, et al.
Veröffentlicht: (2025)
MORPHOS: Autoregressive 4D Generation with Temporal Structured Latents
von: Kwon, Minkyung, et al.
Veröffentlicht: (2026)
von: Kwon, Minkyung, et al.
Veröffentlicht: (2026)
Rethinking Prompt Design for Inference-time Scaling in Text-to-Visual Generation
von: Kim, Subin, et al.
Veröffentlicht: (2025)
von: Kim, Subin, et al.
Veröffentlicht: (2025)
FlowAR: Scale-wise Autoregressive Image Generation Meets Flow Matching
von: Ren, Sucheng, et al.
Veröffentlicht: (2024)
von: Ren, Sucheng, et al.
Veröffentlicht: (2024)
NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale
von: NextStep Team, et al.
Veröffentlicht: (2025)
von: NextStep Team, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Neural Deformation Representation for 4D Dynamic Shape Generation
von: Han, Gyojin, et al.
Veröffentlicht: (2026) -
Unlocking the Capabilities of Masked Generative Models for Image Synthesis via Self-Guidance
von: Hur, Jiwan, et al.
Veröffentlicht: (2024) -
Inlier-Centric Post-Training Quantization for Object Detection Models
von: Kim, Minsu, et al.
Veröffentlicht: (2026) -
Frequency-Aware Token Reduction for Efficient Vision Transformer
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025) -
PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)