Zero-shot Image Editing with Reference Imitation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xi, Feng, Yutong, Chen, Mengting, Wang, Yiyang, Zhang, Shilong, Liu, Yu, Shen, Yujun, Zhao, Hengshuang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AnyDoor: Zero-shot Object-level Image Customization
von: Chen, Xi, et al.
Veröffentlicht: (2023)
von: Chen, Xi, et al.
Veröffentlicht: (2023)
DiffDoctor: Diagnosing Image Diffusion Models Before Treating
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
DiffCamera: Arbitrary Refocusing on Images
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
FlashFace: Human Image Personalization with High-fidelity Identity Preservation
von: Zhang, Shilong, et al.
Veröffentlicht: (2024)
von: Zhang, Shilong, et al.
Veröffentlicht: (2024)
FashionComposer: Compositional Fashion Image Generation
von: Ji, Sihui, et al.
Veröffentlicht: (2024)
von: Ji, Sihui, et al.
Veröffentlicht: (2024)
GDRO: Group-level Reward Post-training Suitable for Diffusion Models
von: Wang, Yiyang, et al.
Veröffentlicht: (2026)
von: Wang, Yiyang, et al.
Veröffentlicht: (2026)
FocalClick-XL: Towards Unified and High-quality Interactive Segmentation
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
Exploring Iterative Manifold Constraint for Zero-shot Image Editing
von: Li, Maomao, et al.
Veröffentlicht: (2025)
von: Li, Maomao, et al.
Veröffentlicht: (2025)
CoDance: An Unbind-Rebind Paradigm for Robust Multi-Subject Animation
von: Tan, Shuai, et al.
Veröffentlicht: (2026)
von: Tan, Shuai, et al.
Veröffentlicht: (2026)
Ranni: Taming Text-to-Image Diffusion for Accurate Instruction Following
von: Feng, Yutong, et al.
Veröffentlicht: (2023)
von: Feng, Yutong, et al.
Veröffentlicht: (2023)
Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
von: Yuan, Zhihao, et al.
Veröffentlicht: (2023)
von: Yuan, Zhihao, et al.
Veröffentlicht: (2023)
LayerFlow: A Unified Model for Layer-aware Video Generation
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
FREE-Edit: Using Editing-aware Injection in Rectified Flow Models for Zero-shot Image-Driven Video Editing
von: Li, Maomao, et al.
Veröffentlicht: (2026)
von: Li, Maomao, et al.
Veröffentlicht: (2026)
MagicQuill: An Intelligent Interactive Image Editing System
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
Zero-shot Face Editing via ID-Attribute Decoupled Inversion
von: Hou, Yang, et al.
Veröffentlicht: (2025)
von: Hou, Yang, et al.
Veröffentlicht: (2025)
UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics
von: Chen, Xi, et al.
Veröffentlicht: (2024)
von: Chen, Xi, et al.
Veröffentlicht: (2024)
Data-Efficient Generalization for Zero-shot Composed Image Retrieval
von: Chen, Zining, et al.
Veröffentlicht: (2025)
von: Chen, Zining, et al.
Veröffentlicht: (2025)
Free-Editor: Zero-shot Text-driven 3D Scene Editing
von: Karim, Nazmul, et al.
Veröffentlicht: (2023)
von: Karim, Nazmul, et al.
Veröffentlicht: (2023)
MeaCap: Memory-Augmented Zero-shot Image Captioning
von: Zeng, Zequn, et al.
Veröffentlicht: (2024)
von: Zeng, Zequn, et al.
Veröffentlicht: (2024)
Seg-VAR: Image Segmentation with Visual Autoregressive Modeling
von: Zheng, Rongkun, et al.
Veröffentlicht: (2025)
von: Zheng, Rongkun, et al.
Veröffentlicht: (2025)
Early Semantic Grounding in Image Editing Models for Zero-Shot Referring Image Segmentation
von: He, Jingxuan, et al.
Veröffentlicht: (2026)
von: He, Jingxuan, et al.
Veröffentlicht: (2026)
Edicho: Consistent Image Editing in the Wild
von: Bai, Qingyan, et al.
Veröffentlicht: (2024)
von: Bai, Qingyan, et al.
Veröffentlicht: (2024)
Zero-shot Composed Text-Image Retrieval
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models
von: Wang, Wen, et al.
Veröffentlicht: (2023)
von: Wang, Wen, et al.
Veröffentlicht: (2023)
IMAGHarmony: Controllable Image Editing with Consistent Object Quantity and Layout
von: Shen, Fei, et al.
Veröffentlicht: (2025)
von: Shen, Fei, et al.
Veröffentlicht: (2025)
FLORA: Formal Language Model Enables Robust Training-free Zero-shot Object Referring Analysis
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing
von: Cai, Lingling, et al.
Veröffentlicht: (2025)
von: Cai, Lingling, et al.
Veröffentlicht: (2025)
Latent Space Disentanglement in Diffusion Transformers Enables Precise Zero-shot Semantic Editing
von: Shuai, Zitao, et al.
Veröffentlicht: (2024)
von: Shuai, Zitao, et al.
Veröffentlicht: (2024)
MangaNinja: Line Art Colorization with Precise Reference Following
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
ZeroStereo: Zero-shot Stereo Matching from Single Images
von: Wang, Xianqi, et al.
Veröffentlicht: (2025)
von: Wang, Xianqi, et al.
Veröffentlicht: (2025)
Generative Editing in the Joint Vision-Language Space for Zero-Shot Composed Image Retrieval
von: Wang, Xin, et al.
Veröffentlicht: (2025)
von: Wang, Xin, et al.
Veröffentlicht: (2025)
LogoSticker: Inserting Logos into Diffusion Models for Customized Generation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
Latent Space Disentanglement in Diffusion Transformers Enables Zero-shot Fine-grained Semantic Editing
von: Shuai, Zitao, et al.
Veröffentlicht: (2024)
von: Shuai, Zitao, et al.
Veröffentlicht: (2024)
Zero-shot Referring Expression Comprehension via Structural Similarity Between Images and Captions
von: Han, Zeyu, et al.
Veröffentlicht: (2023)
von: Han, Zeyu, et al.
Veröffentlicht: (2023)
A Reference-Based 3D Semantic-Aware Framework for Accurate Local Facial Attribute Editing
von: Huang, Yu-Kai, et al.
Veröffentlicht: (2024)
von: Huang, Yu-Kai, et al.
Veröffentlicht: (2024)
IteRPrimE: Zero-shot Referring Image Segmentation with Iterative Grad-CAM Refinement and Primary Word Emphasis
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
Modular Customization of Diffusion Models via Blockwise-Parameterized Low-Rank Adaptation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
DreamMask: Boosting Open-vocabulary Panoptic Segmentation with Synthetic Data
von: Tu, Yuanpeng, et al.
Veröffentlicht: (2025)
von: Tu, Yuanpeng, et al.
Veröffentlicht: (2025)
PLANA3R: Zero-shot Metric Planar 3D Reconstruction via Feed-Forward Planar Splatting
von: Liu, Changkun, et al.
Veröffentlicht: (2025)
von: Liu, Changkun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AnyDoor: Zero-shot Object-level Image Customization
von: Chen, Xi, et al.
Veröffentlicht: (2023) -
DiffDoctor: Diagnosing Image Diffusion Models Before Treating
von: Wang, Yiyang, et al.
Veröffentlicht: (2025) -
DiffCamera: Arbitrary Refocusing on Images
von: Wang, Yiyang, et al.
Veröffentlicht: (2025) -
FlashFace: Human Image Personalization with High-fidelity Identity Preservation
von: Zhang, Shilong, et al.
Veröffentlicht: (2024) -
FashionComposer: Compositional Fashion Image Generation
von: Ji, Sihui, et al.
Veröffentlicht: (2024)