Imagine Flash: Accelerating Emu Diffusion Models with Backward Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Kohler, Jonas, Pumarola, Albert, Schönfeld, Edgar, Sanakoyeu, Artsiom, Sumbaly, Roshan, Vajda, Peter, Thabet, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Autoregressive Distillation of Diffusion Transformers
by: Kim, Yeongmin, et al.
Published: (2025)
by: Kim, Yeongmin, et al.
Published: (2025)
FlexiDiT: Your Diffusion Transformer Can Easily Generate High-Quality Samples with Less Compute
by: Anagnostidis, Sotiris, et al.
Published: (2025)
by: Anagnostidis, Sotiris, et al.
Published: (2025)
Cache Me if You Can: Accelerating Diffusion Models through Block Caching
by: Wimbauer, Felix, et al.
Published: (2023)
by: Wimbauer, Felix, et al.
Published: (2023)
INGeo: Accelerating Instant Neural Scene Reconstruction with Noisy Geometry Priors
by: Li, Chaojian, et al.
Published: (2022)
by: Li, Chaojian, et al.
Published: (2022)
Judge Decoding: Faster Speculative Sampling Requires Going Beyond Model Alignment
by: Bachmann, Gregor, et al.
Published: (2025)
by: Bachmann, Gregor, et al.
Published: (2025)
Bespoke Non-Stationary Solvers for Fast Sampling of Diffusion and Flow Models
by: Shaul, Neta, et al.
Published: (2024)
by: Shaul, Neta, et al.
Published: (2024)
Animated Stickers: Bringing Stickers to Life with Video Diffusion
by: Yan, David, et al.
Published: (2024)
by: Yan, David, et al.
Published: (2024)
Imagine yourself: Tuning-Free Personalized Image Generation
by: He, Zecheng, et al.
Published: (2024)
by: He, Zecheng, et al.
Published: (2024)
Using Motion Cues to Supervise Single-Frame Body Pose and Shape Estimation in Low Data Regimes
by: Davydov, Andrey, et al.
Published: (2024)
by: Davydov, Andrey, et al.
Published: (2024)
Emu: Generative Pretraining in Multimodality
by: Sun, Quan, et al.
Published: (2023)
by: Sun, Quan, et al.
Published: (2023)
SneakPeek: Future-Guided Instructional Streaming Video Generation
by: Hong, Cheeun, et al.
Published: (2025)
by: Hong, Cheeun, et al.
Published: (2025)
Emu3.5: Native Multimodal Models are World Learners
by: Cui, Yufeng, et al.
Published: (2025)
by: Cui, Yufeng, et al.
Published: (2025)
Self-supervised Depth Denoising Using Lower- and Higher-quality RGB-D sensors
by: Shabanov, Akhmedkhan, et al.
Published: (2020)
by: Shabanov, Akhmedkhan, et al.
Published: (2020)
XR-MBT: Multi-modal Full Body Tracking for XR through Self-Supervision with Learned Depth Point Cloud Registration
by: Rozumnyi, Denys, et al.
Published: (2024)
by: Rozumnyi, Denys, et al.
Published: (2024)
SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation
by: Liu, Hongjian, et al.
Published: (2024)
by: Liu, Hongjian, et al.
Published: (2024)
Emu3: Next-Token Prediction is All You Need
by: Wang, Xinlong, et al.
Published: (2024)
by: Wang, Xinlong, et al.
Published: (2024)
Video Self-Stitching Graph Network for Temporal Action Localization
by: Zhao, Chen, et al.
Published: (2020)
by: Zhao, Chen, et al.
Published: (2020)
AVID: Any-Length Video Inpainting with Diffusion Model
by: Zhang, Zhixing, et al.
Published: (2023)
by: Zhang, Zhixing, et al.
Published: (2023)
Flash Diffusion: Accelerating Any Conditional Diffusion Model for Few Steps Image Generation
by: Chadebec, Clément, et al.
Published: (2024)
by: Chadebec, Clément, et al.
Published: (2024)
Forward-Backward Knowledge Distillation for Continual Clustering
by: Sadeghi, Mohammadreza, et al.
Published: (2024)
by: Sadeghi, Mohammadreza, et al.
Published: (2024)
Accelerating Diffusion Models with One-to-Many Knowledge Distillation
by: Zhang, Linfeng, et al.
Published: (2024)
by: Zhang, Linfeng, et al.
Published: (2024)
Development and Enhancement of Text-to-Image Diffusion Models
by: Sahu, Rajdeep Roshan
Published: (2025)
by: Sahu, Rajdeep Roshan
Published: (2025)
Token-Shuffle: Towards High-Resolution Image Generation with Autoregressive Models
by: Ma, Xu, et al.
Published: (2025)
by: Ma, Xu, et al.
Published: (2025)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
by: Zhai, Yuanhao, et al.
Published: (2024)
by: Zhai, Yuanhao, et al.
Published: (2024)
World2VLM: Distilling World Model Imagination into VLMs for Dynamic Spatial Reasoning
by: Zhang, Wanyue, et al.
Published: (2026)
by: Zhang, Wanyue, et al.
Published: (2026)
MixRT: Mixed Neural Representations For Real-Time NeRF Rendering
by: Li, Chaojian, et al.
Published: (2023)
by: Li, Chaojian, et al.
Published: (2023)
Imagine with the Teacher: Complete Shape in a Multi-View Distillation Way
by: Luo, Zhanpeng, et al.
Published: (2025)
by: Luo, Zhanpeng, et al.
Published: (2025)
FVGen: Accelerating Novel-View Synthesis with Adversarial Video Diffusion Distillation
by: Teng, Wenbin, et al.
Published: (2025)
by: Teng, Wenbin, et al.
Published: (2025)
Flash-Unified: A Training-Free and Task-Aware Acceleration Framework for Native Unified Models
by: Ke, Junlong, et al.
Published: (2026)
by: Ke, Junlong, et al.
Published: (2026)
FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation
by: Zhou, Junkang, et al.
Published: (2026)
by: Zhou, Junkang, et al.
Published: (2026)
Layout Control and Semantic Guidance with Attention Loss Backward for T2I Diffusion Model
by: Li, Guandong
Published: (2024)
by: Li, Guandong
Published: (2024)
Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation
by: Li, Badi, et al.
Published: (2025)
by: Li, Badi, et al.
Published: (2025)
Warm Starts Accelerate Conditional Diffusion
by: Scholz, Jonas, et al.
Published: (2025)
by: Scholz, Jonas, et al.
Published: (2025)
FlashEval: Towards Fast and Accurate Evaluation of Text-to-image Diffusion Generative Models
by: Zhao, Lin, et al.
Published: (2024)
by: Zhao, Lin, et al.
Published: (2024)
Accelerating Diffusion Decoders via Multi-Scale Sampling and One-Step Distillation
by: Wang, Chuhan, et al.
Published: (2026)
by: Wang, Chuhan, et al.
Published: (2026)
Flash-Split: 2D Reflection Removal with Flash Cues and Latent Diffusion Separation
by: Wang, Tianfu, et al.
Published: (2024)
by: Wang, Tianfu, et al.
Published: (2024)
BAgger: Backwards Aggregation for Mitigating Drift in Autoregressive Video Diffusion Models
by: Po, Ryan, et al.
Published: (2025)
by: Po, Ryan, et al.
Published: (2025)
SAM-Lightening: A Lightweight Segment Anything Model with Dilated Flash Attention to Achieve 30 times Acceleration
by: Song, Yanfei, et al.
Published: (2024)
by: Song, Yanfei, et al.
Published: (2024)
Scale-wise Distillation of Diffusion Models
by: Starodubcev, Nikita, et al.
Published: (2025)
by: Starodubcev, Nikita, et al.
Published: (2025)
Simple and Fast Distillation of Diffusion Models
by: Zhou, Zhenyu, et al.
Published: (2024)
by: Zhou, Zhenyu, et al.
Published: (2024)
Similar Items
-
Autoregressive Distillation of Diffusion Transformers
by: Kim, Yeongmin, et al.
Published: (2025) -
FlexiDiT: Your Diffusion Transformer Can Easily Generate High-Quality Samples with Less Compute
by: Anagnostidis, Sotiris, et al.
Published: (2025) -
Cache Me if You Can: Accelerating Diffusion Models through Block Caching
by: Wimbauer, Felix, et al.
Published: (2023) -
INGeo: Accelerating Instant Neural Scene Reconstruction with Noisy Geometry Priors
by: Li, Chaojian, et al.
Published: (2022) -
Judge Decoding: Faster Speculative Sampling Requires Going Beyond Model Alignment
by: Bachmann, Gregor, et al.
Published: (2025)