DeltaSpace: A Semantic-aligned Feature Space for Flexible Text-guided Image Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Lyu, Yueming, Zhao, Kang, Peng, Bo, Chen, Huafeng, Jiang, Yue, Zhang, Yingya, Dong, Jing, Shan, Caifeng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Adversarial Transferability between Kolmogorov-arnold Networks
by: Wang, Songping, et al.
Published: (2025)
by: Wang, Songping, et al.
Published: (2025)
DeltaEdit: Exploring Text-free Training for Text-Driven Image Manipulation
by: Lyu, Yueming, et al.
Published: (2023)
by: Lyu, Yueming, et al.
Published: (2023)
An Effective End-to-End Solution for Multimodal Action Recognition
by: Wang, Songping, et al.
Published: (2025)
by: Wang, Songping, et al.
Published: (2025)
Anti-Aesthetics: Protecting Facial Privacy against Customized Text-to-Image Synthesis
by: Wang, Songping, et al.
Published: (2025)
by: Wang, Songping, et al.
Published: (2025)
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing
by: Cai, Lingling, et al.
Published: (2025)
by: Cai, Lingling, et al.
Published: (2025)
Towards Scalable Human-aligned Benchmark for Text-guided Image Editing
by: Ryu, Suho, et al.
Published: (2025)
by: Ryu, Suho, et al.
Published: (2025)
RADAR: Defending RAG Dynamically against Retrieval Corruption
by: Chen, Ziyuan, et al.
Published: (2026)
by: Chen, Ziyuan, et al.
Published: (2026)
Towards Real-Time Fake News Detection under Evidence Scarcity
by: Wei, Guangyu, et al.
Published: (2025)
by: Wei, Guangyu, et al.
Published: (2025)
Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval
by: Liu, Delong, et al.
Published: (2023)
by: Liu, Delong, et al.
Published: (2023)
Exposing and Defending the Achilles' Heel of Video Mixture-of-Experts
by: Wang, Songping, et al.
Published: (2026)
by: Wang, Songping, et al.
Published: (2026)
RunawayEvil: Jailbreaking the Image-to-Video Generative Models
by: Wang, Songping, et al.
Published: (2025)
by: Wang, Songping, et al.
Published: (2025)
S^3D-NeRF: Single-Shot Speech-Driven Neural Radiance Field for High Fidelity Talking Head Synthesis
by: Li, Dongze, et al.
Published: (2024)
by: Li, Dongze, et al.
Published: (2024)
Instant Preference Alignment for Text-to-Image Diffusion Models
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Delta Rectified Flow Sampling for Text-to-Image Editing
by: Beaudouin, Gaspard, et al.
Published: (2025)
by: Beaudouin, Gaspard, et al.
Published: (2025)
SemanticAudio: Audio Generation and Editing in Semantic Space
by: Dai, Zheqi, et al.
Published: (2026)
by: Dai, Zheqi, et al.
Published: (2026)
Latent Watermark: Inject and Detect Watermarks in Latent Diffusion Space
by: Meng, Zheling, et al.
Published: (2024)
by: Meng, Zheling, et al.
Published: (2024)
Semantic Mismatch and Perceptual Degradation: A New Perspective on Image Editing Immunity
by: Dong, Shuai, et al.
Published: (2025)
by: Dong, Shuai, et al.
Published: (2025)
FreeMask: Rethinking the Importance of Attention Masks for Zero-Shot Video Editing
by: Cai, Lingling, et al.
Published: (2024)
by: Cai, Lingling, et al.
Published: (2024)
InstaFace: Identity-Preserving Facial Editing with Single Image Inference
by: Khan, MD Wahiduzzaman, et al.
Published: (2025)
by: Khan, MD Wahiduzzaman, et al.
Published: (2025)
Fast Adversarial Training with Weak-to-Strong Spatial-Temporal Consistency in the Frequency Domain on Videos
by: Wang, Songping, et al.
Published: (2025)
by: Wang, Songping, et al.
Published: (2025)
S$^2$Edit: Text-Guided Image Editing with Precise Semantic and Spatial Control
by: Liu, Xudong, et al.
Published: (2025)
by: Liu, Xudong, et al.
Published: (2025)
Exploring Text-Guided Single Image Editing for Remote Sensing Images
by: Han, Fangzhou, et al.
Published: (2024)
by: Han, Fangzhou, et al.
Published: (2024)
NOVA: Sparse Control, Dense Synthesis for Pair-Free Video Editing
by: Pan, Tianlin, et al.
Published: (2026)
by: Pan, Tianlin, et al.
Published: (2026)
Defense against Unauthorized Distillation in Image Restoration via Feature Space Perturbation
by: Hu, Han, et al.
Published: (2025)
by: Hu, Han, et al.
Published: (2025)
SGEdit: Bridging LLM with Text2Image Generative Model for Scene Graph-based Image Editing
by: Zhang, Zhiyuan, et al.
Published: (2024)
by: Zhang, Zhiyuan, et al.
Published: (2024)
Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation
by: Ko, Jungmin, et al.
Published: (2026)
by: Ko, Jungmin, et al.
Published: (2026)
SAM-COD: SAM-guided Unified Framework for Weakly-Supervised Camouflaged Object Detection
by: Chen, Huafeng, et al.
Published: (2024)
by: Chen, Huafeng, et al.
Published: (2024)
FluxSpace: Disentangled Semantic Editing in Rectified Flow Transformers
by: Dalva, Yusuf, et al.
Published: (2024)
by: Dalva, Yusuf, et al.
Published: (2024)
Image Inpainting Models are Effective Tools for Instruction-guided Image Editing
by: Ju, Xuan, et al.
Published: (2024)
by: Ju, Xuan, et al.
Published: (2024)
Concept Corrector: Erase concepts on the fly for text-to-image diffusion models
by: Meng, Zheling, et al.
Published: (2025)
by: Meng, Zheling, et al.
Published: (2025)
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
by: Nam, Hyelin, et al.
Published: (2023)
by: Nam, Hyelin, et al.
Published: (2023)
Artifact Feature Purification for Cross-domain Detection of AI-generated Images
by: Meng, Zheling, et al.
Published: (2024)
by: Meng, Zheling, et al.
Published: (2024)
Visual Lexicon: Rich Image Features in Language Space
by: Wang, XuDong, et al.
Published: (2024)
by: Wang, XuDong, et al.
Published: (2024)
AvatarBack: Back-Head Generation for Complete 3D Avatars from Front-View Images
by: Xin, Shiqi, et al.
Published: (2025)
by: Xin, Shiqi, et al.
Published: (2025)
COLLIE: Guiding Skill Discovery in Semantically Coherent Latent Space
by: Luan, Yao, et al.
Published: (2026)
by: Luan, Yao, et al.
Published: (2026)
RealignDiff: Boosting Text-to-Image Diffusion Model with Coarse-to-fine Semantic Re-alignment
by: Jiang, Zutao, et al.
Published: (2023)
by: Jiang, Zutao, et al.
Published: (2023)
O-Mamba: O-shape State-Space Model for Underwater Image Enhancement
by: Dong, Chenyu, et al.
Published: (2024)
by: Dong, Chenyu, et al.
Published: (2024)
Editing Massive Concepts in Text-to-Image Diffusion Models
by: Xiong, Tianwei, et al.
Published: (2024)
by: Xiong, Tianwei, et al.
Published: (2024)
DiffEditor: Boosting Accuracy and Flexibility on Diffusion-based Image Editing
by: Mou, Chong, et al.
Published: (2024)
by: Mou, Chong, et al.
Published: (2024)
Semantic Structure of Feature Space in Large Language Models
by: Kozlowski, Austin C., et al.
Published: (2026)
by: Kozlowski, Austin C., et al.
Published: (2026)
Similar Items
-
Exploring Adversarial Transferability between Kolmogorov-arnold Networks
by: Wang, Songping, et al.
Published: (2025) -
DeltaEdit: Exploring Text-free Training for Text-Driven Image Manipulation
by: Lyu, Yueming, et al.
Published: (2023) -
An Effective End-to-End Solution for Multimodal Action Recognition
by: Wang, Songping, et al.
Published: (2025) -
Anti-Aesthetics: Protecting Facial Privacy against Customized Text-to-Image Synthesis
by: Wang, Songping, et al.
Published: (2025) -
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing
by: Cai, Lingling, et al.
Published: (2025)