Training-Free Text-Guided Color Editing with Multi-Modal Diffusion Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yin, Zixin, Dai, Xili, Chen, Ling-Hao, Zhou, Deyu, Wang, Jianan, Wang, Duomin, Yu, Gang, Ni, Lionel M., Zhang, Lei, Shum, Heung-Yeung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LazyDrag: Enabling Stable Drag-Based Editing on Multi-Modal Diffusion Transformers via Explicit Correspondence
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
ConsistEdit: Highly Consistent and Precise Training-free Visual Editing
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
SQuadGen: Generating Simple Quad Layouts via Chart Distance Fields
von: Kong, Youkang, et al.
Veröffentlicht: (2026)
von: Kong, Youkang, et al.
Veröffentlicht: (2026)
Controllable Segmentation-Based Text-Guided Style Editing
von: Li, Jingwen, et al.
Veröffentlicht: (2025)
von: Li, Jingwen, et al.
Veröffentlicht: (2025)
LGTM: Training-Free Light-Guided Text-to-Image Diffusion Model via Initial Noise Manipulation
von: Morita, Ryugo, et al.
Veröffentlicht: (2026)
von: Morita, Ryugo, et al.
Veröffentlicht: (2026)
Image Editing with Diffusion Models: A Survey
von: Wang, Jia, et al.
Veröffentlicht: (2025)
von: Wang, Jia, et al.
Veröffentlicht: (2025)
Garment Animation NeRF with Color Editing
von: Renke Wang, et al.
Veröffentlicht: (2024)
von: Renke Wang, et al.
Veröffentlicht: (2024)
Screen, Cache, and Match: A Training-Free Causality-Consistent Reference Frame Framework for Human Animation
von: Wang, Jianan, et al.
Veröffentlicht: (2025)
von: Wang, Jianan, et al.
Veröffentlicht: (2025)
From Rigging to Waving: 3D-Guided Diffusion for Natural Animation of Hand-Drawn Characters
von: Zhou, Jie, et al.
Veröffentlicht: (2025)
von: Zhou, Jie, et al.
Veröffentlicht: (2025)
SVGBuilder: Component-Based Colored SVG Generation with Text-Guided Autoregressive Transformers
von: Chen, Zehao, et al.
Veröffentlicht: (2024)
von: Chen, Zehao, et al.
Veröffentlicht: (2024)
SketchDream: Sketch-based Text-to-3D Generation and Editing
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2024)
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2024)
iShapEditing: Intelligent Shape Editing with Diffusion Models
von: Jing Li, et al.
Veröffentlicht: (2024)
von: Jing Li, et al.
Veröffentlicht: (2024)
Blended Point Cloud Diffusion for Localized Text-guided Shape Editing
von: Sella, Etai, et al.
Veröffentlicht: (2025)
von: Sella, Etai, et al.
Veröffentlicht: (2025)
DreamTime: An Improved Optimization Strategy for Diffusion-Guided 3D Generation
von: Huang, Yukun, et al.
Veröffentlicht: (2023)
von: Huang, Yukun, et al.
Veröffentlicht: (2023)
Real-time and Controllable Reactive Motion Synthesis via Intention Guidance
von: Zhang, Xiaotang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaotang, et al.
Veröffentlicht: (2025)
Lazy Diffusion Transformer for Interactive Image Editing
von: Nitzan, Yotam, et al.
Veröffentlicht: (2024)
von: Nitzan, Yotam, et al.
Veröffentlicht: (2024)
Text‐Guided Diffusion with Spectral Convolution for 3D Human Pose Estimation
von: Liyuan Shi, et al.
Veröffentlicht: (2025)
von: Liyuan Shi, et al.
Veröffentlicht: (2025)
SemanticGarment: Semantic-Controlled Generation and Editing of 3D Gaussian Garments
von: Wang, Ruiyan, et al.
Veröffentlicht: (2025)
von: Wang, Ruiyan, et al.
Veröffentlicht: (2025)
JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
Stable Flow: Vital Layers for Training-Free Image Editing
von: Avrahami, Omri, et al.
Veröffentlicht: (2024)
von: Avrahami, Omri, et al.
Veröffentlicht: (2024)
TurboEdit: Text-Based Image Editing Using Few-Step Diffusion Models
von: Deutch, Gilad, et al.
Veröffentlicht: (2024)
von: Deutch, Gilad, et al.
Veröffentlicht: (2024)
Colorful-Noise: Training-Free Low-Frequency Noise Manipulation for Color-Based Conditional Image Generation
von: Cohen, Nadav Z., et al.
Veröffentlicht: (2026)
von: Cohen, Nadav Z., et al.
Veröffentlicht: (2026)
LUSD: Localized Update Score Distillation for Text-Guided Image Editing
von: Chinchuthakun, Worameth, et al.
Veröffentlicht: (2025)
von: Chinchuthakun, Worameth, et al.
Veröffentlicht: (2025)
Physics-Based Motion Tracking of Contact-Rich Interacting Characters
von: Zhang, Xiaotang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiaotang, et al.
Veröffentlicht: (2026)
Representing Flow Fields with Divergence-Free Kernels for Reconstruction
von: Ni, Xingyu, et al.
Veröffentlicht: (2025)
von: Ni, Xingyu, et al.
Veröffentlicht: (2025)
Guiding-Based Importance Sampling for Walk on Stars
von: Huang, Tianyu, et al.
Veröffentlicht: (2024)
von: Huang, Tianyu, et al.
Veröffentlicht: (2024)
DreamWaltz-G: Expressive 3D Gaussian Avatars from Skeleton-Guided 2D Diffusion
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
TextFlux: An OCR‐Free DiT Model for High‐Fidelity Multilingual Scene Text Synthesis
von: Yu Xie, et al.
Veröffentlicht: (2026)
von: Yu Xie, et al.
Veröffentlicht: (2026)
GaussianEditor: Editing 3D Gaussians Delicately with Text Instructions
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
A Hessian-Based Field Deformer for Real-Time Topology-Aware Shape Editing
von: Zhang, Yunxiao, et al.
Veröffentlicht: (2024)
von: Zhang, Yunxiao, et al.
Veröffentlicht: (2024)
PartMotionEdit: Fine-Grained Text-Driven 3D Human Motion Editing via Part-Level Modulation
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
Region-Aware Color Smudging
von: Jiang, Ying, et al.
Veröffentlicht: (2024)
von: Jiang, Ying, et al.
Veröffentlicht: (2024)
SpeakerVid-5M: A Large-Scale High-Quality Dataset for Audio-Visual Dyadic Interactive Human Generation
von: Zhang, Youliang, et al.
Veröffentlicht: (2025)
von: Zhang, Youliang, et al.
Veröffentlicht: (2025)
Improved 3D Scene Stylization via Text-Guided Generative Image Editing with Region-Based Control
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2025)
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2025)
LGCC: Enhancing Flow Matching Based Text-Guided Image Editing with Local Gaussian Coupling and Context Consistency
von: Liu, Fangbing, et al.
Veröffentlicht: (2025)
von: Liu, Fangbing, et al.
Veröffentlicht: (2025)
MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning
von: Zhang, Yi-Yang, et al.
Veröffentlicht: (2025)
von: Zhang, Yi-Yang, et al.
Veröffentlicht: (2025)
Leveraging Semantic Attribute Binding for Free-Lunch Color Control in Diffusion Models
von: Laria, Héctor, et al.
Veröffentlicht: (2025)
von: Laria, Héctor, et al.
Veröffentlicht: (2025)
StyleBlend: Enhancing Style‐Specific Content Creation in Text‐to‐Image Diffusion Models
von: Zichong Chen, et al.
Veröffentlicht: (2025)
von: Zichong Chen, et al.
Veröffentlicht: (2025)
Curve-based slicer for multi-axis DLP 3D printing
von: Dai, Chengkai, et al.
Veröffentlicht: (2025)
von: Dai, Chengkai, et al.
Veröffentlicht: (2025)
On Copyright Risks of Text-to-Image Diffusion Models
von: Zhang, Yang, et al.
Veröffentlicht: (2023)
von: Zhang, Yang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LazyDrag: Enabling Stable Drag-Based Editing on Multi-Modal Diffusion Transformers via Explicit Correspondence
von: Yin, Zixin, et al.
Veröffentlicht: (2025) -
ConsistEdit: Highly Consistent and Precise Training-free Visual Editing
von: Yin, Zixin, et al.
Veröffentlicht: (2025) -
SQuadGen: Generating Simple Quad Layouts via Chart Distance Fields
von: Kong, Youkang, et al.
Veröffentlicht: (2026) -
Controllable Segmentation-Based Text-Guided Style Editing
von: Li, Jingwen, et al.
Veröffentlicht: (2025) -
LGTM: Training-Free Light-Guided Text-to-Image Diffusion Model via Initial Noise Manipulation
von: Morita, Ryugo, et al.
Veröffentlicht: (2026)