Controlling Text-to-Image Diffusion by Orthogonal Finetuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiu, Zeju, Liu, Weiyang, Feng, Haiwen, Xue, Yuxuan, Feng, Yao, Liu, Zhen, Zhang, Dan, Weller, Adrian, Schölkopf, Bernhard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Orthogonal Finetuning Made Scalable
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
Parameter-Efficient Orthogonal Finetuning via Butterfly Factorization
von: Liu, Weiyang, et al.
Veröffentlicht: (2023)
von: Liu, Weiyang, et al.
Veröffentlicht: (2023)
Ghost on the Shell: An Expressive Representation of General 3D Shapes
von: Liu, Zhen, et al.
Veröffentlicht: (2023)
von: Liu, Zhen, et al.
Veröffentlicht: (2023)
GraphDreamer: Compositional 3D Scene Synthesis from Scene Graphs
von: Gao, Gege, et al.
Veröffentlicht: (2023)
von: Gao, Gege, et al.
Veröffentlicht: (2023)
JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
GenLit: Reformulating Single-Image Relighting as Video Generation
von: Bharadwaj, Shrisha, et al.
Veröffentlicht: (2024)
von: Bharadwaj, Shrisha, et al.
Veröffentlicht: (2024)
Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models
von: Fortes, Armando, et al.
Veröffentlicht: (2025)
von: Fortes, Armando, et al.
Veröffentlicht: (2025)
InteractDiffusion: Interaction Control in Text-to-Image Diffusion Models
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2023)
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2023)
LGTM: Training-Free Light-Guided Text-to-Image Diffusion Model via Initial Noise Manipulation
von: Morita, Ryugo, et al.
Veröffentlicht: (2026)
von: Morita, Ryugo, et al.
Veröffentlicht: (2026)
DreamBlend: Advancing Personalized Fine-tuning of Text-to-Image Diffusion Models
von: Ram, Shwetha, et al.
Veröffentlicht: (2024)
von: Ram, Shwetha, et al.
Veröffentlicht: (2024)
Style Customization of Text-to-Vector Generation with Image Diffusion Priors
von: Zhang, Peiying, et al.
Veröffentlicht: (2025)
von: Zhang, Peiying, et al.
Veröffentlicht: (2025)
Controllable Weather Synthesis and Removal with Video Diffusion Models
von: Lin, Chih-Hao, et al.
Veröffentlicht: (2025)
von: Lin, Chih-Hao, et al.
Veröffentlicht: (2025)
Generating Human Interaction Motions in Scenes with Text Control
von: Yi, Hongwei, et al.
Veröffentlicht: (2024)
von: Yi, Hongwei, et al.
Veröffentlicht: (2024)
Can Large Language Models Understand Symbolic Graphics Programs?
von: Qiu, Zeju, et al.
Veröffentlicht: (2024)
von: Qiu, Zeju, et al.
Veröffentlicht: (2024)
Geometry Image Diffusion: Fast and Data-Efficient Text-to-3D with Image-Based Surface Representation
von: Elizarov, Slava, et al.
Veröffentlicht: (2024)
von: Elizarov, Slava, et al.
Veröffentlicht: (2024)
Agentic Design of Compositional Machines
von: Zhang, Wenqian, et al.
Veröffentlicht: (2025)
von: Zhang, Wenqian, et al.
Veröffentlicht: (2025)
Taming Real-World Space-Time Video Super-Resolution with One-Step Diffusion
von: Wei, Shuoyan, et al.
Veröffentlicht: (2026)
von: Wei, Shuoyan, et al.
Veröffentlicht: (2026)
CtrlVDiff: Controllable Video Generation via Unified Multimodal Video Diffusion
von: Xi, Dianbing, et al.
Veröffentlicht: (2025)
von: Xi, Dianbing, et al.
Veröffentlicht: (2025)
DartControl: A Diffusion-Based Autoregressive Motion Model for Real-Time Text-Driven Motion Control
von: Zhao, Kaifeng, et al.
Veröffentlicht: (2024)
von: Zhao, Kaifeng, et al.
Veröffentlicht: (2024)
TurboEdit: Text-Based Image Editing Using Few-Step Diffusion Models
von: Deutch, Gilad, et al.
Veröffentlicht: (2024)
von: Deutch, Gilad, et al.
Veröffentlicht: (2024)
Controlling Avatar Diffusion with Learnable Gaussian Embedding
von: Gao, Xuan, et al.
Veröffentlicht: (2025)
von: Gao, Xuan, et al.
Veröffentlicht: (2025)
LightLab: Controlling Light Sources in Images with Diffusion Models
von: Magar, Nadav, et al.
Veröffentlicht: (2025)
von: Magar, Nadav, et al.
Veröffentlicht: (2025)
Endoscopic Depth Estimation Based on Deep Learning: A Survey
von: Niu, Ke, et al.
Veröffentlicht: (2025)
von: Niu, Ke, et al.
Veröffentlicht: (2025)
In-Context LoRA for Diffusion Transformers
von: Huang, Lianghua, et al.
Veröffentlicht: (2024)
von: Huang, Lianghua, et al.
Veröffentlicht: (2024)
ETCH: Generalizing Body Fitting to Clothed Humans via Equivariant Tightness
von: Li, Boqian, et al.
Veröffentlicht: (2025)
von: Li, Boqian, et al.
Veröffentlicht: (2025)
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
DressCode: Autoregressively Sewing and Generating Garments from Text Guidance
von: He, Kai, et al.
Veröffentlicht: (2024)
von: He, Kai, et al.
Veröffentlicht: (2024)
TextToon: Real-Time Text Toonify Head Avatar from Single Video
von: Song, Luchuan, et al.
Veröffentlicht: (2024)
von: Song, Luchuan, et al.
Veröffentlicht: (2024)
DiLightNet: Fine-grained Lighting Control for Diffusion-based Image Generation
von: Zeng, Chong, et al.
Veröffentlicht: (2024)
von: Zeng, Chong, et al.
Veröffentlicht: (2024)
HPR3D: Hierarchical Proxy Representation for High-Fidelity 3D Reconstruction and Controllable Editing
von: Wang, Tielong, et al.
Veröffentlicht: (2025)
von: Wang, Tielong, et al.
Veröffentlicht: (2025)
TexControl: Sketch-Based Two-Stage Fashion Image Generation Using Diffusion Model
von: Zhang, Yongming, et al.
Veröffentlicht: (2024)
von: Zhang, Yongming, et al.
Veröffentlicht: (2024)
Improved 3D Scene Stylization via Text-Guided Generative Image Editing with Region-Based Control
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2025)
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2025)
Image Sculpting: Precise Object Editing with 3D Geometry Control
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2024)
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2024)
Shape Conditioned Human Motion Generation with Diffusion Model
von: Xue, Kebing, et al.
Veröffentlicht: (2024)
von: Xue, Kebing, et al.
Veröffentlicht: (2024)
Human-VDM: Learning Single-Image 3D Human Gaussian Splatting from Video Diffusion Models
von: Liu, Zhibin, et al.
Veröffentlicht: (2024)
von: Liu, Zhibin, et al.
Veröffentlicht: (2024)
Generic 3D Diffusion Adapter Using Controlled Multi-View Editing
von: Chen, Hansheng, et al.
Veröffentlicht: (2024)
von: Chen, Hansheng, et al.
Veröffentlicht: (2024)
MIND: Multi-Scale Intent Diffusion for Text-Driven Physics-Based Humanoid Control
von: Li, Bin, et al.
Veröffentlicht: (2026)
von: Li, Bin, et al.
Veröffentlicht: (2026)
Surf3R: Rapid Surface Reconstruction from Sparse RGB Views in Seconds
von: Zhu, Haodong, et al.
Veröffentlicht: (2025)
von: Zhu, Haodong, et al.
Veröffentlicht: (2025)
Controllable Video Generation: A Survey
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models
von: Yi, Taoran, et al.
Veröffentlicht: (2023)
von: Yi, Taoran, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Orthogonal Finetuning Made Scalable
von: Qiu, Zeju, et al.
Veröffentlicht: (2025) -
Parameter-Efficient Orthogonal Finetuning via Butterfly Factorization
von: Liu, Weiyang, et al.
Veröffentlicht: (2023) -
Ghost on the Shell: An Expressive Representation of General 3D Shapes
von: Liu, Zhen, et al.
Veröffentlicht: (2023) -
GraphDreamer: Compositional 3D Scene Synthesis from Scene Graphs
von: Gao, Gege, et al.
Veröffentlicht: (2023) -
JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
von: Zeng, Yu, et al.
Veröffentlicht: (2024)