Many-to-many Image Generation with Auto-regressive Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Ying, Zhang, Yizhe, Zhai, Shuangfei, Huang, Lifu, Susskind, Joshua M., Gu, Jiatao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kaleido Diffusion: Improving Conditional Diffusion Models with Autoregressive Latent Modeling
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
Matryoshka Diffusion Models
von: Gu, Jiatao, et al.
Veröffentlicht: (2023)
von: Gu, Jiatao, et al.
Veröffentlicht: (2023)
Improving GFlowNets for Text-to-Image Diffusion Alignment
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
von: Sampaio, Georgia Gabriela, et al.
Veröffentlicht: (2024)
von: Sampaio, Georgia Gabriela, et al.
Veröffentlicht: (2024)
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation
von: Shen, Ying, et al.
Veröffentlicht: (2026)
von: Shen, Ying, et al.
Veröffentlicht: (2026)
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
World-consistent Video Diffusion with Explicit 3D Modeling
von: Zhang, Qihang, et al.
Veröffentlicht: (2024)
von: Zhang, Qihang, et al.
Veröffentlicht: (2024)
Normalizing Trajectory Models
von: Gu, Jiatao, et al.
Veröffentlicht: (2026)
von: Gu, Jiatao, et al.
Veröffentlicht: (2026)
Normalizing Flows with Iterative Denoising
von: Chen, Tianrong, et al.
Veröffentlicht: (2026)
von: Chen, Tianrong, et al.
Veröffentlicht: (2026)
How Far Are We from Intelligent Visual Deductive Reasoning?
von: Zhang, Yizhe, et al.
Veröffentlicht: (2024)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2024)
STARFlow-V: End-to-End Video Generative Modeling with Normalizing Flows
von: Gu, Jiatao, et al.
Veröffentlicht: (2025)
von: Gu, Jiatao, et al.
Veröffentlicht: (2025)
Normalizing Flows are Capable Generative Models
von: Zhai, Shuangfei, et al.
Veröffentlicht: (2024)
von: Zhai, Shuangfei, et al.
Veröffentlicht: (2024)
Scalable Pre-training of Large Autoregressive Image Models
von: El-Nouby, Alaaeldin, et al.
Veröffentlicht: (2024)
von: El-Nouby, Alaaeldin, et al.
Veröffentlicht: (2024)
Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation
von: Xu, Zhiyang, et al.
Veröffentlicht: (2025)
von: Xu, Zhiyang, et al.
Veröffentlicht: (2025)
The Coupling Within: Flow Matching via Distilled Normalizing Flows
von: Berthelot, David, et al.
Veröffentlicht: (2026)
von: Berthelot, David, et al.
Veröffentlicht: (2026)
STARFlow: Scaling Latent Normalizing Flows for High-resolution Image Synthesis
von: Gu, Jiatao, et al.
Veröffentlicht: (2025)
von: Gu, Jiatao, et al.
Veröffentlicht: (2025)
PLANNER: Generating Diversified Paragraph via Latent Language Diffusion Model
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
Pushing Auto-regressive Models for 3D Shape Generation at Capacity and Scalability
von: Qian, Xuelin, et al.
Veröffentlicht: (2024)
von: Qian, Xuelin, et al.
Veröffentlicht: (2024)
Towards Holistic Modeling for Video Frame Interpolation with Auto-regressive Diffusion Transformers
von: Peng, Xinyu, et al.
Veröffentlicht: (2026)
von: Peng, Xinyu, et al.
Veröffentlicht: (2026)
AAMDM: Accelerated Auto-regressive Motion Diffusion Model
von: Li, Tianyu, et al.
Veröffentlicht: (2023)
von: Li, Tianyu, et al.
Veröffentlicht: (2023)
SPARTUN3D: Situated Spatial Understanding of 3D World in Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2024)
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2024)
Scaling Down Text Encoders of Text-to-Image Diffusion Models
von: Wang, Lifu, et al.
Veröffentlicht: (2025)
von: Wang, Lifu, et al.
Veröffentlicht: (2025)
iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation
von: Fu, Zhoujie, et al.
Veröffentlicht: (2025)
von: Fu, Zhoujie, et al.
Veröffentlicht: (2025)
Zero-1-to-G: Taming Pretrained 2D Diffusion Model for Direct 3D Generation
von: Meng, Xuyi, et al.
Veröffentlicht: (2025)
von: Meng, Xuyi, et al.
Veröffentlicht: (2025)
GECO: Generative Image-to-3D within a SECOnd
von: Wang, Chen, et al.
Veröffentlicht: (2024)
von: Wang, Chen, et al.
Veröffentlicht: (2024)
AutoDIR: Automatic All-in-One Image Restoration with Latent Diffusion
von: Jiang, Yitong, et al.
Veröffentlicht: (2023)
von: Jiang, Yitong, et al.
Veröffentlicht: (2023)
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
AR-RAG: Autoregressive Retrieval Augmentation for Image Generation
von: Qi, Jingyuan, et al.
Veröffentlicht: (2025)
von: Qi, Jingyuan, et al.
Veröffentlicht: (2025)
BAD: Bidirectional Auto-regressive Diffusion for Text-to-Motion Generation
von: Hosseyni, S. Rohollah, et al.
Veröffentlicht: (2024)
von: Hosseyni, S. Rohollah, et al.
Veröffentlicht: (2024)
Accelerating Auto-regressive Text-to-Image Generation with Training-free Speculative Jacobi Decoding
von: Teng, Yao, et al.
Veröffentlicht: (2024)
von: Teng, Yao, et al.
Veröffentlicht: (2024)
One Layer Is Enough: Adapting Pretrained Visual Encoders for Image Generation
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
ZipAR: Parallel Auto-regressive Image Generation through Spatial Locality
von: He, Yefei, et al.
Veröffentlicht: (2024)
von: He, Yefei, et al.
Veröffentlicht: (2024)
SuperFlow: Training Flow Matching Models with RL on the Fly
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
A Watermark for Auto-Regressive Image Generation Models
von: Wu, Yihan, et al.
Veröffentlicht: (2025)
von: Wu, Yihan, et al.
Veröffentlicht: (2025)
Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation
von: Xin, Yi, et al.
Veröffentlicht: (2025)
von: Xin, Yi, et al.
Veröffentlicht: (2025)
GenDoP: Auto-regressive Camera Trajectory Generation as a Director of Photography
von: Zhang, Mengchen, et al.
Veröffentlicht: (2025)
von: Zhang, Mengchen, et al.
Veröffentlicht: (2025)
SJD++: Improved Speculative Jacobi Decoding for Training-free Acceleration of Discrete Auto-regressive Text-to-Image Generation
von: Teng, Yao, et al.
Veröffentlicht: (2025)
von: Teng, Yao, et al.
Veröffentlicht: (2025)
SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation
von: Huang, Kaiwen, et al.
Veröffentlicht: (2026)
von: Huang, Kaiwen, et al.
Veröffentlicht: (2026)
Pre-trained Language Models Do Not Help Auto-regressive Text-to-Image Generation
von: Zhang, Yuhui, et al.
Veröffentlicht: (2023)
von: Zhang, Yuhui, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Kaleido Diffusion: Improving Conditional Diffusion Models with Autoregressive Latent Modeling
von: Gu, Jiatao, et al.
Veröffentlicht: (2024) -
Matryoshka Diffusion Models
von: Gu, Jiatao, et al.
Veröffentlicht: (2023) -
Improving GFlowNets for Text-to-Image Diffusion Alignment
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024) -
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
von: Sampaio, Georgia Gabriela, et al.
Veröffentlicht: (2024) -
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation
von: Shen, Ying, et al.
Veröffentlicht: (2026)