EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yuxuan, Yuan, Yirui, Song, Yiren, Wang, Haofan, Liu, Jiaming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation
von: Wang, Cong, et al.
Veröffentlicht: (2024)
von: Wang, Cong, et al.
Veröffentlicht: (2024)
Stable-Makeup: When Real-World Makeup Transfer Meets Diffusion Model
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
Unlocking the Latent Canvas: Eliciting and Benchmarking Symbolic Visual Expression in LLMs
von: Zheng, Yiren, et al.
Veröffentlicht: (2026)
von: Zheng, Yiren, et al.
Veröffentlicht: (2026)
OmniPSD: Layered PSD Generation with Diffusion Transformer
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
Loom: Diffusion-Transformer for Interleaved Generation
von: Ye, Mingcheng, et al.
Veröffentlicht: (2025)
von: Ye, Mingcheng, et al.
Veröffentlicht: (2025)
VISTA: Triplet-Supervised Video Style Transfer with Diffusion Transformers
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
WMAdapter: Adding WaterMark Control to Latent Diffusion Models
von: Ci, Hai, et al.
Veröffentlicht: (2024)
von: Ci, Hai, et al.
Veröffentlicht: (2024)
Stable-Hair: Real-World Hair Transfer via Diffusion Model
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
SIGMA: Selective-Interleaved Generation with Multi-Attribute Tokens
von: Zhang, Xiaoyan, et al.
Veröffentlicht: (2026)
von: Zhang, Xiaoyan, et al.
Veröffentlicht: (2026)
WordCon: Word-level Typography Control in Scene Text Rendering
von: Shi, Wenda, et al.
Veröffentlicht: (2025)
von: Shi, Wenda, et al.
Veröffentlicht: (2025)
Adding Additional Control to One-Step Diffusion with Joint Distribution Matching
von: Luo, Yihong, et al.
Veröffentlicht: (2025)
von: Luo, Yihong, et al.
Veröffentlicht: (2025)
MakeAnything: Harnessing Diffusion Transformers for Multi-Domain Procedural Sequence Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
FonTS: Text Rendering with Typography and Style Controls
von: Shi, Wenda, et al.
Veröffentlicht: (2024)
von: Shi, Wenda, et al.
Veröffentlicht: (2024)
Image Watermarks are Removable Using Controllable Regeneration from Clean Noise
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
PhotoDoodle: Learning Artistic Image Editing from Few-Shot Pairwise Data
von: Huang, Shijie, et al.
Veröffentlicht: (2025)
von: Huang, Shijie, et al.
Veröffentlicht: (2025)
RelaCtrl: Relevance-Guided Efficient Control for Diffusion Transformers
von: Cao, Ke, et al.
Veröffentlicht: (2025)
von: Cao, Ke, et al.
Veröffentlicht: (2025)
LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
DisControlFace: Adding Disentangled Control to Diffusion Autoencoder for One-shot Explicit Facial Image Editing
von: Jia, Haozhe, et al.
Veröffentlicht: (2023)
von: Jia, Haozhe, et al.
Veröffentlicht: (2023)
RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers
von: Gong, Yan, et al.
Veröffentlicht: (2025)
von: Gong, Yan, et al.
Veröffentlicht: (2025)
Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
NanoControl: A Lightweight Framework for Precise and Efficient Control in Diffusion Transformer
von: Liu, Shanyuan, et al.
Veröffentlicht: (2025)
von: Liu, Shanyuan, et al.
Veröffentlicht: (2025)
Training-free Regional Prompting for Diffusion Transformers
von: Chen, Anthony, et al.
Veröffentlicht: (2024)
von: Chen, Anthony, et al.
Veröffentlicht: (2024)
DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
OminiControl2: Efficient Conditioning for Diffusion Transformers
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2025)
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2025)
GRE Suite: Geo-localization Inference via Fine-Tuned Vision-Language Models and Enhanced Reasoning Chains
von: Wang, Chun, et al.
Veröffentlicht: (2025)
von: Wang, Chun, et al.
Veröffentlicht: (2025)
In-Context Audio Control of Video Diffusion Transformers
von: Liu, Wenze, et al.
Veröffentlicht: (2025)
von: Liu, Wenze, et al.
Veröffentlicht: (2025)
Any2AnyTryon: Leveraging Adaptive Position Embeddings for Versatile Virtual Clothing Tasks
von: Guo, Hailong, et al.
Veröffentlicht: (2025)
von: Guo, Hailong, et al.
Veröffentlicht: (2025)
MegActor-$Σ$: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
von: Yang, Shurong, et al.
Veröffentlicht: (2024)
von: Yang, Shurong, et al.
Veröffentlicht: (2024)
COME: Adding Scene-Centric Forecasting Control to Occupancy World Model
von: Shi, Yining, et al.
Veröffentlicht: (2025)
von: Shi, Yining, et al.
Veröffentlicht: (2025)
OmniRefiner: Reinforcement-Guided Local Diffusion Refinement
von: Liu, Yaoli, et al.
Veröffentlicht: (2025)
von: Liu, Yaoli, et al.
Veröffentlicht: (2025)
NumeriKontrol: Adding Numeric Control to Diffusion Transformers for Instruction-based Image Editing
von: Xu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2025)
TransAnimate: Taming Layer Diffusion to Generate RGBA Video
von: Chen, Xuewei, et al.
Veröffentlicht: (2025)
von: Chen, Xuewei, et al.
Veröffentlicht: (2025)
TrackGo: A Flexible and Efficient Method for Controllable Video Generation
von: Zhou, Haitao, et al.
Veröffentlicht: (2024)
von: Zhou, Haitao, et al.
Veröffentlicht: (2024)
DiffDesign: Controllable Diffusion with Meta Prior for Efficient Interior Design Generation
von: Yang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuxuan, et al.
Veröffentlicht: (2024)
Fast Personalized Text-to-Image Syntheses With Attention Injection
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
Mitty: Diffusion-based Human-to-Robot Video Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers
von: Chen, Ruidong, et al.
Veröffentlicht: (2026)
von: Chen, Ruidong, et al.
Veröffentlicht: (2026)
ObjectAdd: Adding Objects into Image via a Training-Free Diffusion Modification Fashion
von: Zhang, Ziyue, et al.
Veröffentlicht: (2024)
von: Zhang, Ziyue, et al.
Veröffentlicht: (2024)
FiT: Flexible Vision Transformer for Diffusion Model
von: Lu, Zeyu, et al.
Veröffentlicht: (2024)
von: Lu, Zeyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering
von: Lu, Runnan, et al.
Veröffentlicht: (2025) -
EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation
von: Wang, Cong, et al.
Veröffentlicht: (2024) -
Stable-Makeup: When Real-World Makeup Transfer Meets Diffusion Model
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024) -
Unlocking the Latent Canvas: Eliciting and Benchmarking Symbolic Visual Expression in LLMs
von: Zheng, Yiren, et al.
Veröffentlicht: (2026) -
OmniPSD: Layered PSD Generation with Diffusion Transformer
von: Liu, Cheng, et al.
Veröffentlicht: (2025)