Frequency-Aware Autoregressive Modeling for Efficient High-Resolution Image Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Zhuokun, Fan, Jugang, Yu, Zhuowei, Zhuang, Bohan, Tan, Mingkui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Perception Capabilities of Multimodal LLMs with Training-Free Fusion
by: Chen, Zhuokun, et al.
Published: (2024)
by: Chen, Zhuokun, et al.
Published: (2024)
FlashBlock: Attention Caching for Efficient Long-Context Block Diffusion
by: Chen, Zhuokun, et al.
Published: (2026)
by: Chen, Zhuokun, et al.
Published: (2026)
VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations
by: Patel, Maitreya, et al.
Published: (2026)
by: Patel, Maitreya, et al.
Published: (2026)
MGHF: Multi-Granular High-Frequency Perceptual Loss for Image Super-Resolution
by: Sami, Shoaib Meraj, et al.
Published: (2024)
by: Sami, Shoaib Meraj, et al.
Published: (2024)
Native-Resolution Image Synthesis
by: Wang, Zidong, et al.
Published: (2025)
by: Wang, Zidong, et al.
Published: (2025)
From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation
by: Song, Han, et al.
Published: (2026)
by: Song, Han, et al.
Published: (2026)
LoRAPrune: Structured Pruning Meets Low-Rank Parameter-Efficient Fine-Tuning
by: Zhang, Mingyang, et al.
Published: (2023)
by: Zhang, Mingyang, et al.
Published: (2023)
Efficient High-Resolution Image Editing with Hallucination-Aware Loss and Adaptive Tiling
by: Kwon, Young D., et al.
Published: (2025)
by: Kwon, Young D., et al.
Published: (2025)
High-Resolution Image Synthesis via Next-Token Prediction
by: Chen, Dengsheng, et al.
Published: (2024)
by: Chen, Dengsheng, et al.
Published: (2024)
AQD: Towards Accurate Fully-Quantized Object Detection
by: Chen, Peng, et al.
Published: (2020)
by: Chen, Peng, et al.
Published: (2020)
FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation
by: Zhou, Junkang, et al.
Published: (2026)
by: Zhou, Junkang, et al.
Published: (2026)
Metadata, Wavelet, and Time Aware Diffusion Models for Satellite Image Super Resolution
by: Sigillo, Luigi, et al.
Published: (2025)
by: Sigillo, Luigi, et al.
Published: (2025)
Efficient Stitchable Task Adaptation
by: He, Haoyu, et al.
Published: (2023)
by: He, Haoyu, et al.
Published: (2023)
Progressive Autoregressive Video Diffusion Models
by: Xie, Desai, et al.
Published: (2024)
by: Xie, Desai, et al.
Published: (2024)
Radioactive Watermarks in Diffusion and Autoregressive Image Generative Models
by: Meintz, Michel, et al.
Published: (2025)
by: Meintz, Michel, et al.
Published: (2025)
MixAR: Mixture Autoregressive Image Generation
by: Hu, Jinyuan, et al.
Published: (2025)
by: Hu, Jinyuan, et al.
Published: (2025)
E-CAR: Efficient Continuous Autoregressive Image Generation via Multistage Modeling
by: Yuan, Zhihang, et al.
Published: (2024)
by: Yuan, Zhihang, et al.
Published: (2024)
Prominence-Aware Artifact Detection and Dataset for Image Super-Resolution
by: Molodetskikh, Ivan, et al.
Published: (2025)
by: Molodetskikh, Ivan, et al.
Published: (2025)
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
by: Gu, Youping, et al.
Published: (2025)
by: Gu, Youping, et al.
Published: (2025)
MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation
by: Chen, Junhao, et al.
Published: (2025)
by: Chen, Junhao, et al.
Published: (2025)
The Promise of RL for Autoregressive Image Editing
by: Ahmadi, Saba, et al.
Published: (2025)
by: Ahmadi, Saba, et al.
Published: (2025)
Cost-Aware Routing for Efficient Text-To-Image Generation
by: Li, Qinchan, et al.
Published: (2025)
by: Li, Qinchan, et al.
Published: (2025)
Spatiotemporal Satellite Image Downscaling with Transfer Encoders and Autoregressive Generative Models
by: Xiang, Yang, et al.
Published: (2025)
by: Xiang, Yang, et al.
Published: (2025)
R-Stitch: Dynamic Trajectory Stitching for Efficient Reasoning
by: Chen, Zhuokun, et al.
Published: (2025)
by: Chen, Zhuokun, et al.
Published: (2025)
Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens
by: Fan, Lijie, et al.
Published: (2024)
by: Fan, Lijie, et al.
Published: (2024)
Exploring the Coordination of Frequency and Attention in Masked Image Modeling
by: Gui, Jie, et al.
Published: (2022)
by: Gui, Jie, et al.
Published: (2022)
Fast Autoregressive Models for Continuous Latent Generation
by: Hang, Tiankai, et al.
Published: (2025)
by: Hang, Tiankai, et al.
Published: (2025)
MediQ-GAN: Quantum-Inspired GAN for High Resolution Medical Image Generation
by: Jiao, Qingyue, et al.
Published: (2025)
by: Jiao, Qingyue, et al.
Published: (2025)
Policy-based Tuning of Autoregressive Image Models with Instance- and Distribution-Level Rewards
by: Baran, Orhun Buğra, et al.
Published: (2026)
by: Baran, Orhun Buğra, et al.
Published: (2026)
T-Stitch: Accelerating Sampling in Pre-Trained Diffusion Models with Trajectory Stitching
by: Pan, Zizheng, et al.
Published: (2024)
by: Pan, Zizheng, et al.
Published: (2024)
Missing Fine Details in Images: Last Seen in High Frequencies
by: Medi, Tejaswini, et al.
Published: (2025)
by: Medi, Tejaswini, et al.
Published: (2025)
Pre-trained Visual Dynamics Representations for Efficient Policy Learning
by: Luo, Hao, et al.
Published: (2024)
by: Luo, Hao, et al.
Published: (2024)
Hawk: Leveraging Spatial Context for Faster Autoregressive Text-to-Image Generation
by: Chen, Zhi-Kai, et al.
Published: (2025)
by: Chen, Zhi-Kai, et al.
Published: (2025)
PSA: Pyramid Sparse Attention for Efficient Video Understanding and Generation
by: Li, Xiaolong, et al.
Published: (2025)
by: Li, Xiaolong, et al.
Published: (2025)
Teaching Metric Distance to Discrete Autoregressive Language Models
by: Chung, Jiwan, et al.
Published: (2025)
by: Chung, Jiwan, et al.
Published: (2025)
Scalable High-Resolution Pixel-Space Image Synthesis with Hourglass Diffusion Transformers
by: Crowson, Katherine, et al.
Published: (2024)
by: Crowson, Katherine, et al.
Published: (2024)
CoNav: A Benchmark for Human-Centered Collaborative Navigation
by: Li, Changhao, et al.
Published: (2024)
by: Li, Changhao, et al.
Published: (2024)
ToDo: Token Downsampling for Efficient Generation of High-Resolution Images
by: Smith, Ethan, et al.
Published: (2024)
by: Smith, Ethan, et al.
Published: (2024)
Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection
by: Zhang, Shuhai, et al.
Published: (2025)
by: Zhang, Shuhai, et al.
Published: (2025)
Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis
by: Sigillo, Luigi, et al.
Published: (2025)
by: Sigillo, Luigi, et al.
Published: (2025)
Similar Items
-
Enhancing Perception Capabilities of Multimodal LLMs with Training-Free Fusion
by: Chen, Zhuokun, et al.
Published: (2024) -
FlashBlock: Attention Caching for Efficient Long-Context Block Diffusion
by: Chen, Zhuokun, et al.
Published: (2026) -
VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations
by: Patel, Maitreya, et al.
Published: (2026) -
MGHF: Multi-Granular High-Frequency Perceptual Loss for Image Super-Resolution
by: Sami, Shoaib Meraj, et al.
Published: (2024) -
Native-Resolution Image Synthesis
by: Wang, Zidong, et al.
Published: (2025)