MASC: Boosting Autoregressive Image Generation with a Manifold-Aligned Semantic Clustering
Fuente:
arXiv
Saved in:
| Main Authors: | He, Lixuan, Zheng, Shikang, Zhang, Linfeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AMFT: Aligning LLM Reasoners by Meta-Learning the Optimal Imitation-Exploration Balance
by: He, Lixuan, et al.
Published: (2025)
by: He, Lixuan, et al.
Published: (2025)
SMART: Semantic Matching Contrastive Learning for Partially View-Aligned Clustering
by: Peng, Liang, et al.
Published: (2025)
by: Peng, Liang, et al.
Published: (2025)
Annealed Relaxation of Speculative Decoding for Faster Autoregressive Image Generation
by: Li, Xingyao, et al.
Published: (2026)
by: Li, Xingyao, et al.
Published: (2026)
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
by: Kim, Beomsu, et al.
Published: (2025)
by: Kim, Beomsu, et al.
Published: (2025)
VTBench: Evaluating Visual Tokenizers for Autoregressive Image Generation
by: Lin, Huawei, et al.
Published: (2025)
by: Lin, Huawei, et al.
Published: (2025)
Pathways on the Image Manifold: Image Editing via Video Generation
by: Rotstein, Noam, et al.
Published: (2024)
by: Rotstein, Noam, et al.
Published: (2024)
JetFormer: An Autoregressive Generative Model of Raw Images and Text
by: Tschannen, Michael, et al.
Published: (2024)
by: Tschannen, Michael, et al.
Published: (2024)
SpectralAR: Spectral Autoregressive Visual Generation
by: Huang, Yuanhui, et al.
Published: (2025)
by: Huang, Yuanhui, et al.
Published: (2025)
E-CAR: Efficient Continuous Autoregressive Image Generation via Multistage Modeling
by: Yuan, Zhihang, et al.
Published: (2024)
by: Yuan, Zhihang, et al.
Published: (2024)
SAS: Semantic-aware Sampling for Generative Dataset Distillation
by: Li, Mingzhuo, et al.
Published: (2026)
by: Li, Mingzhuo, et al.
Published: (2026)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
by: Liu, Runtao, et al.
Published: (2024)
by: Liu, Runtao, et al.
Published: (2024)
Taming the Entropy Cliff: Variable Codebook Size Quantization for Autoregressive Visual Generation
by: Zheng, Bowen, et al.
Published: (2026)
by: Zheng, Bowen, et al.
Published: (2026)
FreqCa: Accelerating Diffusion Models via Frequency-Aware Caching
by: Liu, Jiacheng, et al.
Published: (2025)
by: Liu, Jiacheng, et al.
Published: (2025)
Watermarking Autoregressive Image Generation
by: Jovanović, Nikola, et al.
Published: (2025)
by: Jovanović, Nikola, et al.
Published: (2025)
Astra: General Interactive World Model with Autoregressive Denoising
by: Zhu, Yixuan, et al.
Published: (2025)
by: Zhu, Yixuan, et al.
Published: (2025)
Geometry-Correct Diffusion Posterior Sampling with Denoiser-Pullback Curvature Guidance and Manifold-Aligned Damping
by: Shin, Seunghyeok, et al.
Published: (2026)
by: Shin, Seunghyeok, et al.
Published: (2026)
PCPO: Proportionate Credit Policy Optimization for Aligning Image Generation Models
by: Lee, Jeongjae, et al.
Published: (2025)
by: Lee, Jeongjae, et al.
Published: (2025)
Autoregressive Adversarial Post-Training for Real-Time Interactive Video Generation
by: Lin, Shanchuan, et al.
Published: (2025)
by: Lin, Shanchuan, et al.
Published: (2025)
HART: Efficient Visual Generation with Hybrid Autoregressive Transformer
by: Tang, Haotian, et al.
Published: (2024)
by: Tang, Haotian, et al.
Published: (2024)
DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness
by: Li, Ruining, et al.
Published: (2025)
by: Li, Ruining, et al.
Published: (2025)
TinyAlign: Boosting Lightweight Vision-Language Models by Mitigating Modal Alignment Bottlenecks
by: Hu, Yuanze, et al.
Published: (2025)
by: Hu, Yuanze, et al.
Published: (2025)
Learning Discrete Autoregressive Priors with Wasserstein Gradient Flow
by: Zheng, Bowen, et al.
Published: (2026)
by: Zheng, Bowen, et al.
Published: (2026)
ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation
by: Chen, Junying, et al.
Published: (2025)
by: Chen, Junying, et al.
Published: (2025)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
by: Mamedov, Timur, et al.
Published: (2026)
by: Mamedov, Timur, et al.
Published: (2026)
Boost Your Human Image Generation Model via Direct Preference Optimization
by: Na, Sanghyeon, et al.
Published: (2024)
by: Na, Sanghyeon, et al.
Published: (2024)
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
by: Mo, Wenyi, et al.
Published: (2024)
by: Mo, Wenyi, et al.
Published: (2024)
Focused Forcing: Content-Aware Per-Frame KV Selection for Efficient Autoregressive Video Diffusion
by: Cai, Peiliang, et al.
Published: (2026)
by: Cai, Peiliang, et al.
Published: (2026)
Camouflaged Image Synthesis Is All You Need to Boost Camouflaged Detection
by: Zhang, Haichao, et al.
Published: (2023)
by: Zhang, Haichao, et al.
Published: (2023)
IPixMatch: Boost Semi-supervised Semantic Segmentation with Inter-Pixel Relation
by: Wu, Kebin, et al.
Published: (2024)
by: Wu, Kebin, et al.
Published: (2024)
DocSynthv2: A Practical Autoregressive Modeling for Document Generation
by: Biswas, Sanket, et al.
Published: (2024)
by: Biswas, Sanket, et al.
Published: (2024)
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
by: Huang, Xun, et al.
Published: (2025)
by: Huang, Xun, et al.
Published: (2025)
SeqSAM: Autoregressive Multiple Hypothesis Prediction for Medical Image Segmentation using SAM
by: Towle, Benjamin, et al.
Published: (2025)
by: Towle, Benjamin, et al.
Published: (2025)
Boosting Single Positive Multi-label Classification with Generalized Robust Loss
by: Chen, Yanxi, et al.
Published: (2024)
by: Chen, Yanxi, et al.
Published: (2024)
AGDC: Autoregressive Generation of Variable-Length Sequences with Joint Discrete and Continuous Spaces
by: Shin, Yeonsang, et al.
Published: (2026)
by: Shin, Yeonsang, et al.
Published: (2026)
Diffuse and Disperse: Image Generation with Representation Regularization
by: Wang, Runqian, et al.
Published: (2025)
by: Wang, Runqian, et al.
Published: (2025)
CanvasMAR: Improving Masked Autoregressive Video Prediction With Canvas
by: Li, Zian, et al.
Published: (2025)
by: Li, Zian, et al.
Published: (2025)
Subimage Overlap Prediction: Task-Aligned Self-Supervised Pretraining For Semantic Segmentation In Remote Sensing Imagery
by: Sharma, Lakshay, et al.
Published: (2026)
by: Sharma, Lakshay, et al.
Published: (2026)
Aligning Text to Image in Diffusion Models is Easier Than You Think
by: Lee, Jaa-Yeon, et al.
Published: (2025)
by: Lee, Jaa-Yeon, et al.
Published: (2025)
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
by: Woo, Byeongju, et al.
Published: (2026)
by: Woo, Byeongju, et al.
Published: (2026)
Controllable Unlearning for Image-to-Image Generative Models via $\varepsilon$-Constrained Optimization
by: Feng, Xiaohua, et al.
Published: (2024)
by: Feng, Xiaohua, et al.
Published: (2024)
Similar Items
-
AMFT: Aligning LLM Reasoners by Meta-Learning the Optimal Imitation-Exploration Balance
by: He, Lixuan, et al.
Published: (2025) -
SMART: Semantic Matching Contrastive Learning for Partially View-Aligned Clustering
by: Peng, Liang, et al.
Published: (2025) -
Annealed Relaxation of Speculative Decoding for Faster Autoregressive Image Generation
by: Li, Xingyao, et al.
Published: (2026) -
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
by: Kim, Beomsu, et al.
Published: (2025) -
VTBench: Evaluating Visual Tokenizers for Autoregressive Image Generation
by: Lin, Huawei, et al.
Published: (2025)