StyleDiffusion: Prompt-Embedding Inversion for Text-Based Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Senmao, van de Weijer, Joost, Hu, Taihang, Khan, Fahad Shahbaz, Hou, Qibin, Wang, Yaxing, Yang, Jian, Cheng, Ming-Ming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
von: Li, Senmao, et al.
Veröffentlicht: (2024)
von: Li, Senmao, et al.
Veröffentlicht: (2024)
Faster Diffusion: Rethinking the Role of the Encoder for Diffusion Model Inference
von: Li, Senmao, et al.
Veröffentlicht: (2023)
von: Li, Senmao, et al.
Veröffentlicht: (2023)
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
von: Liu, Tao, et al.
Veröffentlicht: (2025)
von: Liu, Tao, et al.
Veröffentlicht: (2025)
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
InterLCM: Low-Quality Images as Intermediate States of Latent Consistency Models for Effective Blind Face Restoration
von: Li, Senmao, et al.
Veröffentlicht: (2025)
von: Li, Senmao, et al.
Veröffentlicht: (2025)
One-Way Ticket:Time-Independent Unified Encoder for Distilling Text-to-Image Diffusion Models
von: Li, Senmao, et al.
Veröffentlicht: (2025)
von: Li, Senmao, et al.
Veröffentlicht: (2025)
Anchor Token Matching: Implicit Structure Locking for Training-free AR Image Editing
von: Hu, Taihang, et al.
Veröffentlicht: (2025)
von: Hu, Taihang, et al.
Veröffentlicht: (2025)
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models
von: Li, Senmao, et al.
Veröffentlicht: (2025)
von: Li, Senmao, et al.
Veröffentlicht: (2025)
LocInv: Localization-aware Inversion for Text-Guided Image Editing
von: Tang, Chuanming, et al.
Veröffentlicht: (2024)
von: Tang, Chuanming, et al.
Veröffentlicht: (2024)
Adversarial Concept Distillation for One-Step Diffusion Personalization
von: Yang, Yixiong, et al.
Veröffentlicht: (2025)
von: Yang, Yixiong, et al.
Veröffentlicht: (2025)
Training-free image inversion for one-step diffusion models
von: Wu, Tao, et al.
Veröffentlicht: (2026)
von: Wu, Tao, et al.
Veröffentlicht: (2026)
Free-Lunch Color-Texture Disentanglement for Stylized Image Generation
von: Qin, Jiang, et al.
Veröffentlicht: (2025)
von: Qin, Jiang, et al.
Veröffentlicht: (2025)
SimInversion: A Simple Framework for Inversion-Based Text-to-Image Editing
von: Qian, Qi, et al.
Veröffentlicht: (2024)
von: Qian, Qi, et al.
Veröffentlicht: (2024)
Diversity Has Always Been There in Your Visual Autoregressive Models
von: Wang, Tong, et al.
Veröffentlicht: (2025)
von: Wang, Tong, et al.
Veröffentlicht: (2025)
IterInv: Iterative Inversion for Pixel-Level T2I Models
von: Tang, Chuanming, et al.
Veröffentlicht: (2023)
von: Tang, Chuanming, et al.
Veröffentlicht: (2023)
Multi-Class Textual-Inversion Secretly Yields a Semantic-Agnostic Classifier
von: Wang, Kai, et al.
Veröffentlicht: (2024)
von: Wang, Kai, et al.
Veröffentlicht: (2024)
ColorPeel: Color Prompt Learning with Diffusion Models via Color and Shape Disentanglement
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2024)
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2024)
WaDi: Weight Direction-aware Distillation for One-step Image Synthesis
von: Wang, Lei, et al.
Veröffentlicht: (2026)
von: Wang, Lei, et al.
Veröffentlicht: (2026)
ProEdit: Inversion-based Editing From Prompts Done Right
von: Ouyang, Zhi, et al.
Veröffentlicht: (2025)
von: Ouyang, Zhi, et al.
Veröffentlicht: (2025)
Continuous-Time Distribution Matching for Few-Step Diffusion Distillation
von: Liu, Tao, et al.
Veröffentlicht: (2026)
von: Liu, Tao, et al.
Veröffentlicht: (2026)
Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation
von: Li, Yunheng, et al.
Veröffentlicht: (2024)
von: Li, Yunheng, et al.
Veröffentlicht: (2024)
StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
von: Zhou, Yupeng, et al.
Veröffentlicht: (2024)
von: Zhou, Yupeng, et al.
Veröffentlicht: (2024)
Mixture of Style Experts for Diverse Image Stylization
von: Zhu, Shihao, et al.
Veröffentlicht: (2026)
von: Zhu, Shihao, et al.
Veröffentlicht: (2026)
Enhancing Perceptual Quality in Video Super-Resolution through Temporally-Consistent Detail Synthesis using Diffusion Models
von: Rota, Claudio, et al.
Veröffentlicht: (2023)
von: Rota, Claudio, et al.
Veröffentlicht: (2023)
Query Drift Compensation: Enabling Compatibility in Continual Learning of Retrieval Embedding Models
von: Goswami, Dipam, et al.
Veröffentlicht: (2025)
von: Goswami, Dipam, et al.
Veröffentlicht: (2025)
Progressive Semantic-Guided Vision Transformer for Zero-Shot Learning
von: Chen, Shiming, et al.
Veröffentlicht: (2024)
von: Chen, Shiming, et al.
Veröffentlicht: (2024)
LumiCtrl : Learning Illuminant Prompts for Lighting Control in Personalized Text-to-Image Models
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
Not All Parameters Matter: Masking Diffusion Models for Enhancing Generation Ability
von: Wang, Lei, et al.
Veröffentlicht: (2025)
von: Wang, Lei, et al.
Veröffentlicht: (2025)
Video-GroundingDINO: Towards Open-Vocabulary Spatio-Temporal Video Grounding
von: Wasim, Syed Talal, et al.
Veröffentlicht: (2023)
von: Wasim, Syed Talal, et al.
Veröffentlicht: (2023)
Wavelet-Guided Acceleration of Text Inversion in Diffusion-Based Image Editing
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
HiStyle: Hierarchical Style Embedding Predictor for Text-Prompt-Guided Controllable Speech Synthesis
von: Zhang, Ziyu, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyu, et al.
Veröffentlicht: (2025)
KAC: Kolmogorov-Arnold Classifier for Continual Learning
von: Hu, Yusong, et al.
Veröffentlicht: (2025)
von: Hu, Yusong, et al.
Veröffentlicht: (2025)
How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?
von: Dong, Jiahua, et al.
Veröffentlicht: (2024)
von: Dong, Jiahua, et al.
Veröffentlicht: (2024)
Video-CoM: Interactive Video Reasoning via Chain of Manipulations
von: Rasheed, Hanoona, et al.
Veröffentlicht: (2025)
von: Rasheed, Hanoona, et al.
Veröffentlicht: (2025)
UNETR++: Delving into Efficient and Accurate 3D Medical Image Segmentation
von: Shaker, Abdelrahman, et al.
Veröffentlicht: (2022)
von: Shaker, Abdelrahman, et al.
Veröffentlicht: (2022)
Discriminative Image Generation with Diffusion Models for Zero-Shot Learning
von: Fu, Dingjie, et al.
Veröffentlicht: (2024)
von: Fu, Dingjie, et al.
Veröffentlicht: (2024)
Prompt-Softbox-Prompt: A Free-Text Embedding Control for Image Editing
von: Yang, Yitong, et al.
Veröffentlicht: (2024)
von: Yang, Yitong, et al.
Veröffentlicht: (2024)
Task-Oriented Diffusion Inversion for High-Fidelity Text-based Editing
von: Xu, Yangyang, et al.
Veröffentlicht: (2024)
von: Xu, Yangyang, et al.
Veröffentlicht: (2024)
BAPLe: Backdoor Attacks on Medical Foundational Models using Prompt Learning
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
MaTe3D: Mask-guided Text-based 3D-aware Portrait Editing
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
von: Li, Senmao, et al.
Veröffentlicht: (2024) -
Faster Diffusion: Rethinking the Role of the Encoder for Diffusion Model Inference
von: Li, Senmao, et al.
Veröffentlicht: (2023) -
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
von: Liu, Tao, et al.
Veröffentlicht: (2025) -
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
von: Hu, Taihang, et al.
Veröffentlicht: (2024) -
InterLCM: Low-Quality Images as Intermediate States of Latent Consistency Models for Effective Blind Face Restoration
von: Li, Senmao, et al.
Veröffentlicht: (2025)