Hawk: Leveraging Spatial Context for Faster Autoregressive Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Zhi-Kai, Jiang, Jun-Peng, Ye, Han-Jia, Zhan, De-Chuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs
von: Chen, Zhi-Kai, et al.
Veröffentlicht: (2026)
von: Chen, Zhi-Kai, et al.
Veröffentlicht: (2026)
TV100: A TV Series Dataset that Pre-Trained CLIP Has Not Seen
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
Annealed Relaxation of Speculative Decoding for Faster Autoregressive Image Generation
von: Li, Xingyao, et al.
Veröffentlicht: (2026)
von: Li, Xingyao, et al.
Veröffentlicht: (2026)
Adaptive Adapter Routing for Long-Tailed Class-Incremental Learning
von: Qi, Zhi-Hong, et al.
Veröffentlicht: (2024)
von: Qi, Zhi-Hong, et al.
Veröffentlicht: (2024)
Addressing Imbalanced Domain-Incremental Learning through Dual-Balance Collaborative Experts
von: Li, Lan, et al.
Veröffentlicht: (2025)
von: Li, Lan, et al.
Veröffentlicht: (2025)
Bridge the Modality and Capability Gaps in Vision-Language Model Selection
von: Yi, Chao, et al.
Veröffentlicht: (2024)
von: Yi, Chao, et al.
Veröffentlicht: (2024)
Context-Aware Autoregressive Models for Multi-Conditional Image Generation
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
Leveraging Cross-Modal Neighbor Representation for Improved CLIP Classification
von: Yi, Chao, et al.
Veröffentlicht: (2024)
von: Yi, Chao, et al.
Veröffentlicht: (2024)
MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation
von: Chen, Junhao, et al.
Veröffentlicht: (2025)
von: Chen, Junhao, et al.
Veröffentlicht: (2025)
Expandable Subspace Ensemble for Pre-Trained Model-Based Class-Incremental Learning
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
PILOT: A Pre-Trained Model-Based Continual Learning Toolbox
von: Sun, Hai-Long, et al.
Veröffentlicht: (2023)
von: Sun, Hai-Long, et al.
Veröffentlicht: (2023)
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
External Knowledge Injection for CLIP-Based Class-Incremental Learning
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2025)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2025)
Class-Incremental Learning: A Survey
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2023)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2023)
BOFA: Bridge-Layer Orthogonal Low-Rank Fusion for CLIP-Based Class-Incremental Learning
von: Li, Lan, et al.
Veröffentlicht: (2025)
von: Li, Lan, et al.
Veröffentlicht: (2025)
Expressive Text-to-Image Generation with Rich Text
von: Ge, Songwei, et al.
Veröffentlicht: (2023)
von: Ge, Songwei, et al.
Veröffentlicht: (2023)
Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads
von: Kou, Siqi, et al.
Veröffentlicht: (2024)
von: Kou, Siqi, et al.
Veröffentlicht: (2024)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
Revisiting Class-Incremental Learning with Pre-Trained Models: Generalizability and Adaptivity are All You Need
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2023)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2023)
Continual Learning with Pre-Trained Models: A Survey
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
Dual Consolidation for Pre-Trained Model-Based Domain-Incremental Learning
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
MixAR: Mixture Autoregressive Image Generation
von: Hu, Jinyuan, et al.
Veröffentlicht: (2025)
von: Hu, Jinyuan, et al.
Veröffentlicht: (2025)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
JetFormer: An Autoregressive Generative Model of Raw Images and Text
von: Tschannen, Michael, et al.
Veröffentlicht: (2024)
von: Tschannen, Michael, et al.
Veröffentlicht: (2024)
Towards Better & Faster Autoregressive Image Generation: From the Perspective of Entropy
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2025)
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2025)
MOS: Model Surgery for Pre-Trained Model-Based Class-Incremental Learning
von: Sun, Hai-Long, et al.
Veröffentlicht: (2024)
von: Sun, Hai-Long, et al.
Veröffentlicht: (2024)
Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens
von: Fan, Lijie, et al.
Veröffentlicht: (2024)
von: Fan, Lijie, et al.
Veröffentlicht: (2024)
Unleashing the Potential of Large Language Models for Text-to-Image Generation through Autoregressive Representation Alignment
von: Xie, Xing, et al.
Veröffentlicht: (2025)
von: Xie, Xing, et al.
Veröffentlicht: (2025)
Context Forcing: Consistent Autoregressive Video Generation with Long Context
von: Chen, Shuo, et al.
Veröffentlicht: (2026)
von: Chen, Shuo, et al.
Veröffentlicht: (2026)
End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer
von: Chu, Wenda, et al.
Veröffentlicht: (2026)
von: Chu, Wenda, et al.
Veröffentlicht: (2026)
Learning without Forgetting for Vision-Language Models
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2023)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2023)
From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation
von: Song, Han, et al.
Veröffentlicht: (2026)
von: Song, Han, et al.
Veröffentlicht: (2026)
Rethinking Diffusion for Text-Driven Human Motion Generation: Redundant Representations, Evaluation, and Masked Autoregression
von: Meng, Zichong, et al.
Veröffentlicht: (2024)
von: Meng, Zichong, et al.
Veröffentlicht: (2024)
Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better
von: Lai, Zihang, et al.
Veröffentlicht: (2025)
von: Lai, Zihang, et al.
Veröffentlicht: (2025)
Radioactive Watermarks in Diffusion and Autoregressive Image Generative Models
von: Meintz, Michel, et al.
Veröffentlicht: (2025)
von: Meintz, Michel, et al.
Veröffentlicht: (2025)
FlashMesh: Faster and Better Autoregressive Mesh Synthesis via Structured Speculation
von: Shen, Tingrui, et al.
Veröffentlicht: (2025)
von: Shen, Tingrui, et al.
Veröffentlicht: (2025)
ACDC: Autoregressive Coherent Multimodal Generation using Diffusion Correction
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM Encoders
von: Kou, Siqi, et al.
Veröffentlicht: (2026)
von: Kou, Siqi, et al.
Veröffentlicht: (2026)
Progressive Compositionality in Text-to-Image Generative Models
von: Han, Evans Xu, et al.
Veröffentlicht: (2024)
von: Han, Evans Xu, et al.
Veröffentlicht: (2024)
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models
von: Li, Senmao, et al.
Veröffentlicht: (2025)
von: Li, Senmao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs
von: Chen, Zhi-Kai, et al.
Veröffentlicht: (2026) -
TV100: A TV Series Dataset that Pre-Trained CLIP Has Not Seen
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024) -
Annealed Relaxation of Speculative Decoding for Faster Autoregressive Image Generation
von: Li, Xingyao, et al.
Veröffentlicht: (2026) -
Adaptive Adapter Routing for Long-Tailed Class-Incremental Learning
von: Qi, Zhi-Hong, et al.
Veröffentlicht: (2024) -
Addressing Imbalanced Domain-Incremental Learning through Dual-Balance Collaborative Experts
von: Li, Lan, et al.
Veröffentlicht: (2025)