From Text to Pose to Image: Improving Diffusion Model Control and Quality
Fuente:
arXiv
Saved in:
| Main Authors: | Bonnet, Clément, Lee, Ariel N., Wertel, Franck, Tamano, Antoine, Cizain, Tanguy, Ducru, Pablo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Controlling Human Shape and Pose in Text-to-Image Diffusion Models via Domain Adaptation
by: Buchheim, Benito, et al.
Published: (2024)
by: Buchheim, Benito, et al.
Published: (2024)
VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control
by: Wu, Shaojin, et al.
Published: (2024)
by: Wu, Shaojin, et al.
Published: (2024)
Grounding Text-to-Image Diffusion Models for Controlled High-Quality Image Generation
by: Süleyman, Ahmad, et al.
Published: (2025)
by: Süleyman, Ahmad, et al.
Published: (2025)
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
by: Wang, Zhendong, et al.
Published: (2025)
by: Wang, Zhendong, et al.
Published: (2025)
Guidance-base Diffusion Models for Improving Photoacoustic Image Quality
by: Eguchi, Tatsuhiro, et al.
Published: (2025)
by: Eguchi, Tatsuhiro, et al.
Published: (2025)
Analyzing and Improving Fast Sampling of Text-to-Image Diffusion Models
by: Zhou, Zhenyu, et al.
Published: (2026)
by: Zhou, Zhenyu, et al.
Published: (2026)
Contrastive Prompts Improve Disentanglement in Text-to-Image Diffusion Models
by: Wu, Chen, et al.
Published: (2024)
by: Wu, Chen, et al.
Published: (2024)
ECNet: Effective Controllable Text-to-Image Diffusion Models
by: Li, Sicheng, et al.
Published: (2024)
by: Li, Sicheng, et al.
Published: (2024)
Local Conditional Controlling for Text-to-Image Diffusion Models
by: Zhao, Yibo, et al.
Published: (2023)
by: Zhao, Yibo, et al.
Published: (2023)
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model
by: Lee, Donghwna, et al.
Published: (2024)
by: Lee, Donghwna, et al.
Published: (2024)
Dual Recursive Feedback on Generation and Appearance Latents for Pose-Robust Text-to-Image Diffusion
by: Kim, Jiwon, et al.
Published: (2025)
by: Kim, Jiwon, et al.
Published: (2025)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
by: Cao, Pu, et al.
Published: (2024)
by: Cao, Pu, et al.
Published: (2024)
PiCo: Enhancing Text-Image Alignment with Improved Noise Selection and Precise Mask Control in Diffusion Models
by: Xie, Chang, et al.
Published: (2025)
by: Xie, Chang, et al.
Published: (2025)
Stable-Pose: Leveraging Transformers for Pose-Guided Text-to-Image Generation
by: Wang, Jiajun, et al.
Published: (2024)
by: Wang, Jiajun, et al.
Published: (2024)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
by: Gao, Xiang, et al.
Published: (2024)
by: Gao, Xiang, et al.
Published: (2024)
PoseTraj: Pose-Aware Trajectory Control in Video Diffusion
by: Ji, Longbin, et al.
Published: (2025)
by: Ji, Longbin, et al.
Published: (2025)
From Text to Mask: Localizing Entities Using the Attention of Text-to-Image Diffusion Models
by: Xiao, Changming, et al.
Published: (2023)
by: Xiao, Changming, et al.
Published: (2023)
Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models
by: Fortes, Armando, et al.
Published: (2025)
by: Fortes, Armando, et al.
Published: (2025)
PriorFormer: A Transformer for Real-time Monocular 3D Human Pose Estimation with Versatile Geometric Priors
by: Adjel, Mohamed, et al.
Published: (2025)
by: Adjel, Mohamed, et al.
Published: (2025)
Toward Early Quality Assessment of Text-to-Image Diffusion Models
by: Guo, Huanlei, et al.
Published: (2026)
by: Guo, Huanlei, et al.
Published: (2026)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
by: Liu, Luping, et al.
Published: (2024)
by: Liu, Luping, et al.
Published: (2024)
Pose-dIVE: Pose-Diversified Augmentation with Diffusion Model for Person Re-Identification
by: Kim, Inès Hyeonsu, et al.
Published: (2024)
by: Kim, Inès Hyeonsu, et al.
Published: (2024)
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
by: Zhang, Yabo, et al.
Published: (2024)
by: Zhang, Yabo, et al.
Published: (2024)
InteractDiffusion: Interaction Control in Text-to-Image Diffusion Models
by: Hoe, Jiun Tian, et al.
Published: (2023)
by: Hoe, Jiun Tian, et al.
Published: (2023)
FashionPose: Text to Pose to Relight Image Generation for Personalized Fashion Visualization
by: Shi, Chuancheng, et al.
Published: (2025)
by: Shi, Chuancheng, et al.
Published: (2025)
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
by: Rahman, Tanzila, et al.
Published: (2024)
by: Rahman, Tanzila, et al.
Published: (2024)
ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems
by: Zavadski, Denis, et al.
Published: (2023)
by: Zavadski, Denis, et al.
Published: (2023)
PreciseControl: Enhancing Text-To-Image Diffusion Models with Fine-Grained Attribute Control
by: Parihar, Rishubh, et al.
Published: (2024)
by: Parihar, Rishubh, et al.
Published: (2024)
Semantic Guidance Tuning for Text-To-Image Diffusion Models
by: Kang, Hyun, et al.
Published: (2023)
by: Kang, Hyun, et al.
Published: (2023)
FBSDiff++: Improved Frequency Band Substitution of Diffusion Features for Efficient and Highly Controllable Text-Driven Image-to-Image Translation
by: Gao, Xiang, et al.
Published: (2026)
by: Gao, Xiang, et al.
Published: (2026)
ARTIST: Improving the Generation of Text-rich Images with Disentangled Diffusion Models and Large Language Models
by: Zhang, Jianyi, et al.
Published: (2024)
by: Zhang, Jianyi, et al.
Published: (2024)
Wuerstchen: An Efficient Architecture for Large-Scale Text-to-Image Diffusion Models
by: Pernias, Pablo, et al.
Published: (2023)
by: Pernias, Pablo, et al.
Published: (2023)
Ultrasound Image Enhancement with the Variance of Diffusion Models
by: Zhang, Yuxin, et al.
Published: (2024)
by: Zhang, Yuxin, et al.
Published: (2024)
Direct Consistency Optimization for Robust Customization of Text-to-Image Diffusion Models
by: Lee, Kyungmin, et al.
Published: (2024)
by: Lee, Kyungmin, et al.
Published: (2024)
Adapting Diffusion Models for Improved Prompt Compliance and Controllable Image Synthesis
by: Sridhar, Deepak, et al.
Published: (2024)
by: Sridhar, Deepak, et al.
Published: (2024)
Extreme Two-View Geometry From Object Poses with Diffusion Models
by: Sun, Yujing, et al.
Published: (2024)
by: Sun, Yujing, et al.
Published: (2024)
Customizing Text-to-Image Diffusion with Object Viewpoint Control
by: Kumari, Nupur, et al.
Published: (2024)
by: Kumari, Nupur, et al.
Published: (2024)
Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models
by: Shen, Fei, et al.
Published: (2023)
by: Shen, Fei, et al.
Published: (2023)
Debiasing Text-to-Image Diffusion Models
by: He, Ruifei, et al.
Published: (2024)
by: He, Ruifei, et al.
Published: (2024)
Adjusting Initial Noise to Mitigate Memorization in Text-to-Image Diffusion Models
by: Han, Hyeonggeun, et al.
Published: (2025)
by: Han, Hyeonggeun, et al.
Published: (2025)
Similar Items
-
Controlling Human Shape and Pose in Text-to-Image Diffusion Models via Domain Adaptation
by: Buchheim, Benito, et al.
Published: (2024) -
VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control
by: Wu, Shaojin, et al.
Published: (2024) -
Grounding Text-to-Image Diffusion Models for Controlled High-Quality Image Generation
by: Süleyman, Ahmad, et al.
Published: (2025) -
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
by: Wang, Zhendong, et al.
Published: (2025) -
Guidance-base Diffusion Models for Improving Photoacoustic Image Quality
by: Eguchi, Tatsuhiro, et al.
Published: (2025)