Direction-Aware Diagonal Autoregressive Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Yijia, Ju, Jianzhong, Luan, Jian, Cui, Jinshi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
LLaVA-SG: Leveraging Scene Graphs as Visual Semantic Expression in Vision-Language Models
von: Wang, Jingyi, et al.
Veröffentlicht: (2024)
von: Wang, Jingyi, et al.
Veröffentlicht: (2024)
Image-Feature Weak-to-Strong Consistency: An Enhanced Paradigm for Semi-Supervised Learning
von: Wu, Zhiyu, et al.
Veröffentlicht: (2024)
von: Wu, Zhiyu, et al.
Veröffentlicht: (2024)
Fast Autoregressive Video Generation with Diagonal Decoding
von: Ye, Yang, et al.
Veröffentlicht: (2025)
von: Ye, Yang, et al.
Veröffentlicht: (2025)
Visual Para-Thinker: Divide-and-Conquer Reasoning for Visual Comprehension
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
TimeViper: A Hybrid Mamba-Transformer Vision-Language Model for Efficient Long Video Understanding
von: Xu, Boshen, et al.
Veröffentlicht: (2025)
von: Xu, Boshen, et al.
Veröffentlicht: (2025)
Motion-Aware Caching for Efficient Autoregressive Video Generation
von: Xu, Jing, et al.
Veröffentlicht: (2026)
von: Xu, Jing, et al.
Veröffentlicht: (2026)
DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation
von: Ye, Bo, et al.
Veröffentlicht: (2026)
von: Ye, Bo, et al.
Veröffentlicht: (2026)
StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video
von: Li, Ao, et al.
Veröffentlicht: (2026)
von: Li, Ao, et al.
Veröffentlicht: (2026)
Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective
von: Mao, Yuxin, et al.
Veröffentlicht: (2025)
von: Mao, Yuxin, et al.
Veröffentlicht: (2025)
Frequency Autoregressive Image Generation with Continuous Tokens
von: Yu, Hu, et al.
Veröffentlicht: (2025)
von: Yu, Hu, et al.
Veröffentlicht: (2025)
Hita: Holistic Tokenizer for Autoregressive Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
RandAR: Decoder-only Autoregressive Visual Generation in Random Orders
von: Pang, Ziqi, et al.
Veröffentlicht: (2024)
von: Pang, Ziqi, et al.
Veröffentlicht: (2024)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
von: Zheng, Zirui, et al.
Veröffentlicht: (2025)
von: Zheng, Zirui, et al.
Veröffentlicht: (2025)
Make It Efficient: Dynamic Sparse Attention for Autoregressive Image Generation
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
BitMark: Watermarking Bitwise Autoregressive Image Generative Models
von: Kerner, Louis, et al.
Veröffentlicht: (2025)
von: Kerner, Louis, et al.
Veröffentlicht: (2025)
Progress by Pieces: Test-Time Scaling for Autoregressive Image Generation
von: Park, Joonhyung, et al.
Veröffentlicht: (2025)
von: Park, Joonhyung, et al.
Veröffentlicht: (2025)
Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2025)
Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads
von: Kou, Siqi, et al.
Veröffentlicht: (2024)
von: Kou, Siqi, et al.
Veröffentlicht: (2024)
One-Forcing: Towards Stable One-Step Autoregressive Video Generation
von: Feng, Jiaqi, et al.
Veröffentlicht: (2026)
von: Feng, Jiaqi, et al.
Veröffentlicht: (2026)
Improving Chain-of-Thought Efficiency for Autoregressive Image Generation
von: Gu, Zeqi, et al.
Veröffentlicht: (2025)
von: Gu, Zeqi, et al.
Veröffentlicht: (2025)
Adversarial Error Correction for Visual Autoregressive Generation
von: Bi, Ligong, et al.
Veröffentlicht: (2026)
von: Bi, Ligong, et al.
Veröffentlicht: (2026)
Training-Free Text-Guided Image Editing with Visual Autoregressive Model
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
VTBench: Evaluating Visual Tokenizers for Autoregressive Image Generation
von: Lin, Huawei, et al.
Veröffentlicht: (2025)
von: Lin, Huawei, et al.
Veröffentlicht: (2025)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
RestoreVAR: Visual Autoregressive Generation for All-in-One Image Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2025)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2025)
Nexus-Gen: Unified Image Understanding, Generation, and Editing via Prefilled Autoregression in Shared Embedding Space
von: Zhang, Hong, et al.
Veröffentlicht: (2025)
von: Zhang, Hong, et al.
Veröffentlicht: (2025)
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
von: Tian, Keyu, et al.
Veröffentlicht: (2024)
von: Tian, Keyu, et al.
Veröffentlicht: (2024)
On the Robustness of Watermarking for Autoregressive Image Generation
von: Müller, Andreas, et al.
Veröffentlicht: (2026)
von: Müller, Andreas, et al.
Veröffentlicht: (2026)
Think-Clip-Sample: Slow-Fast Frame Selection for Video Understanding
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective
von: Zhu, Yongxin, et al.
Veröffentlicht: (2024)
von: Zhu, Yongxin, et al.
Veröffentlicht: (2024)
Moving Beyond Diffusion: Hierarchy-to-Hierarchy Autoregression for fMRI-to-Image Reconstruction
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
MAGI-1: Autoregressive Video Generation at Scale
von: ai, Sand., et al.
Veröffentlicht: (2025)
von: ai, Sand., et al.
Veröffentlicht: (2025)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
CtrlNeRF: The Generative Neural Radiation Fields for the Controllable Synthesis of High-fidelity 3D-Aware Images
von: Liu, Jian, et al.
Veröffentlicht: (2024)
von: Liu, Jian, et al.
Veröffentlicht: (2024)
Autoregressive Omni-Aware Outpainting for Open-Vocabulary 360-Degree Image Generation
von: Lu, Zhuqiang, et al.
Veröffentlicht: (2023)
von: Lu, Zhuqiang, et al.
Veröffentlicht: (2023)
Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
Next Block Prediction: Video Generation via Semi-Autoregressive Modeling
von: Ren, Shuhuai, et al.
Veröffentlicht: (2025)
von: Ren, Shuhuai, et al.
Veröffentlicht: (2025)
Spanning Tree Autoregressive Visual Generation
von: Lee, Sangkyu, et al.
Veröffentlicht: (2025)
von: Lee, Sangkyu, et al.
Veröffentlicht: (2025)
Speculative Decoding for Autoregressive Video Generation
von: Hu, Yuezhou, et al.
Veröffentlicht: (2026)
von: Hu, Yuezhou, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026) -
LLaVA-SG: Leveraging Scene Graphs as Visual Semantic Expression in Vision-Language Models
von: Wang, Jingyi, et al.
Veröffentlicht: (2024) -
Image-Feature Weak-to-Strong Consistency: An Enhanced Paradigm for Semi-Supervised Learning
von: Wu, Zhiyu, et al.
Veröffentlicht: (2024) -
Fast Autoregressive Video Generation with Diagonal Decoding
von: Ye, Yang, et al.
Veröffentlicht: (2025) -
Visual Para-Thinker: Divide-and-Conquer Reasoning for Visual Comprehension
von: Xu, Haoran, et al.
Veröffentlicht: (2026)