Faster Diffusion: Rethinking the Role of the Encoder for Diffusion Model Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Senmao, Hu, Taihang, van de Weijer, Joost, Khan, Fahad Shahbaz, Liu, Tao, Li, Linxuan, Yang, Shiqi, Wang, Yaxing, Cheng, Ming-Ming, Yang, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
StyleDiffusion: Prompt-Embedding Inversion for Text-Based Editing
by: Li, Senmao, et al.
Published: (2023)
by: Li, Senmao, et al.
Published: (2023)
Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
by: Li, Senmao, et al.
Published: (2024)
by: Li, Senmao, et al.
Published: (2024)
One-Way Ticket:Time-Independent Unified Encoder for Distilling Text-to-Image Diffusion Models
by: Li, Senmao, et al.
Published: (2025)
by: Li, Senmao, et al.
Published: (2025)
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
by: Hu, Taihang, et al.
Published: (2024)
by: Hu, Taihang, et al.
Published: (2024)
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
InterLCM: Low-Quality Images as Intermediate States of Latent Consistency Models for Effective Blind Face Restoration
by: Li, Senmao, et al.
Published: (2025)
by: Li, Senmao, et al.
Published: (2025)
Adversarial Concept Distillation for One-Step Diffusion Personalization
by: Yang, Yixiong, et al.
Published: (2025)
by: Yang, Yixiong, et al.
Published: (2025)
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models
by: Li, Senmao, et al.
Published: (2025)
by: Li, Senmao, et al.
Published: (2025)
Training-free image inversion for one-step diffusion models
by: Wu, Tao, et al.
Published: (2026)
by: Wu, Tao, et al.
Published: (2026)
Anchor Token Matching: Implicit Structure Locking for Training-free AR Image Editing
by: Hu, Taihang, et al.
Published: (2025)
by: Hu, Taihang, et al.
Published: (2025)
Free-Lunch Color-Texture Disentanglement for Stylized Image Generation
by: Qin, Jiang, et al.
Published: (2025)
by: Qin, Jiang, et al.
Published: (2025)
Diversity Has Always Been There in Your Visual Autoregressive Models
by: Wang, Tong, et al.
Published: (2025)
by: Wang, Tong, et al.
Published: (2025)
Not All Parameters Matter: Masking Diffusion Models for Enhancing Generation Ability
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Continuous-Time Distribution Matching for Few-Step Diffusion Distillation
by: Liu, Tao, et al.
Published: (2026)
by: Liu, Tao, et al.
Published: (2026)
WaDi: Weight Direction-aware Distillation for One-step Image Synthesis
by: Wang, Lei, et al.
Published: (2026)
by: Wang, Lei, et al.
Published: (2026)
Enhancing Perceptual Quality in Video Super-Resolution through Temporally-Consistent Detail Synthesis using Diffusion Models
by: Rota, Claudio, et al.
Published: (2023)
by: Rota, Claudio, et al.
Published: (2023)
Video-GroundingDINO: Towards Open-Vocabulary Spatio-Temporal Video Grounding
by: Wasim, Syed Talal, et al.
Published: (2023)
by: Wasim, Syed Talal, et al.
Published: (2023)
RainDiff: End-to-end Precipitation Nowcasting Via Token-wise Attention Diffusion
by: Nguyen, Thao, et al.
Published: (2025)
by: Nguyen, Thao, et al.
Published: (2025)
ColorPeel: Color Prompt Learning with Diffusion Models via Color and Shape Disentanglement
by: Butt, Muhammad Atif, et al.
Published: (2024)
by: Butt, Muhammad Atif, et al.
Published: (2024)
SED: A Simple Encoder-Decoder for Open-Vocabulary Semantic Segmentation
by: Xie, Bin, et al.
Published: (2023)
by: Xie, Bin, et al.
Published: (2023)
From Cradle to Cane: A Two-Pass Framework for High-Fidelity Lifespan Face Aging
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
Video-CoM: Interactive Video Reasoning via Chain of Manipulations
by: Rasheed, Hanoona, et al.
Published: (2025)
by: Rasheed, Hanoona, et al.
Published: (2025)
UNETR++: Delving into Efficient and Accurate 3D Medical Image Segmentation
by: Shaker, Abdelrahman, et al.
Published: (2022)
by: Shaker, Abdelrahman, et al.
Published: (2022)
How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?
by: Dong, Jiahua, et al.
Published: (2024)
by: Dong, Jiahua, et al.
Published: (2024)
EoCD: Encoder only Remote Sensing Change Detection
by: Noman, Mubashir, et al.
Published: (2026)
by: Noman, Mubashir, et al.
Published: (2026)
Discriminative Image Generation with Diffusion Models for Zero-Shot Learning
by: Fu, Dingjie, et al.
Published: (2024)
by: Fu, Dingjie, et al.
Published: (2024)
Faster Diffusion Action Segmentation
by: Wang, Shuaibing, et al.
Published: (2024)
by: Wang, Shuaibing, et al.
Published: (2024)
Dynamic Pre-training: Towards Efficient and Scalable All-in-One Image Restoration
by: Dudhane, Akshay, et al.
Published: (2024)
by: Dudhane, Akshay, et al.
Published: (2024)
Efficient Video Object Segmentation via Modulated Cross-Attention Memory
by: Shaker, Abdelrahman, et al.
Published: (2024)
by: Shaker, Abdelrahman, et al.
Published: (2024)
CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
Rethinking Transformers Pre-training for Multi-Spectral Satellite Imagery
by: Noman, Mubashir, et al.
Published: (2024)
by: Noman, Mubashir, et al.
Published: (2024)
Exploiting the Semantic Knowledge of Pre-trained Text-Encoders for Continual Learning
by: Yu, Lu, et al.
Published: (2024)
by: Yu, Lu, et al.
Published: (2024)
Multi-Class Textual-Inversion Secretly Yields a Semantic-Agnostic Classifier
by: Wang, Kai, et al.
Published: (2024)
by: Wang, Kai, et al.
Published: (2024)
LocInv: Localization-aware Inversion for Text-Guided Image Editing
by: Tang, Chuanming, et al.
Published: (2024)
by: Tang, Chuanming, et al.
Published: (2024)
VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Videos
by: Rasheed, Hanoona, et al.
Published: (2025)
by: Rasheed, Hanoona, et al.
Published: (2025)
Calibrating Higher-Order Statistics for Few-Shot Class-Incremental Learning with Pre-trained Vision Transformers
by: Goswami, Dipam, et al.
Published: (2024)
by: Goswami, Dipam, et al.
Published: (2024)
The Art of Deception: Color Visual Illusions and Diffusion Models
by: Gomez-Villa, Alex, et al.
Published: (2024)
by: Gomez-Villa, Alex, et al.
Published: (2024)
Distribution Backtracking Builds A Faster Convergence Trajectory for Diffusion Distillation
by: Zhang, Shengyuan, et al.
Published: (2024)
by: Zhang, Shengyuan, et al.
Published: (2024)
IterInv: Iterative Inversion for Pixel-Level T2I Models
by: Tang, Chuanming, et al.
Published: (2023)
by: Tang, Chuanming, et al.
Published: (2023)
Self-Distilled Masked Auto-Encoders are Efficient Video Anomaly Detectors
by: Ristea, Nicolae-Catalin, et al.
Published: (2023)
by: Ristea, Nicolae-Catalin, et al.
Published: (2023)
Similar Items
-
StyleDiffusion: Prompt-Embedding Inversion for Text-Based Editing
by: Li, Senmao, et al.
Published: (2023) -
Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
by: Li, Senmao, et al.
Published: (2024) -
One-Way Ticket:Time-Independent Unified Encoder for Distilling Text-to-Image Diffusion Models
by: Li, Senmao, et al.
Published: (2025) -
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
by: Hu, Taihang, et al.
Published: (2024) -
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
by: Liu, Tao, et al.
Published: (2025)