Expressive Text-to-Image Generation with Rich Text
Fuente:
arXiv
Saved in:
| Main Authors: | Ge, Songwei, Park, Taesung, Zhu, Jun-Yan, Huang, Jia-Bin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
One-Step Image Translation with Text-to-Image Models
by: Parmar, Gaurav, et al.
Published: (2024)
by: Parmar, Gaurav, et al.
Published: (2024)
On the Content Bias in Fréchet Video Distance
by: Ge, Songwei, et al.
Published: (2024)
by: Ge, Songwei, et al.
Published: (2024)
Generating Multi-Image Synthetic Data for Text-to-Image Customization
by: Kumari, Nupur, et al.
Published: (2025)
by: Kumari, Nupur, et al.
Published: (2025)
Customizing Text-to-Image Models with a Single Image Pair
by: Jones, Maxwell, et al.
Published: (2024)
by: Jones, Maxwell, et al.
Published: (2024)
Rethinking Score Distillation as a Bridge Between Image Distributions
by: McAllister, David, et al.
Published: (2024)
by: McAllister, David, et al.
Published: (2024)
Distilling Diffusion Models into Conditional GANs
by: Kang, Minguk, et al.
Published: (2024)
by: Kang, Minguk, et al.
Published: (2024)
Preserve Your Own Correlation: A Noise Prior for Video Diffusion Models
by: Ge, Songwei, et al.
Published: (2023)
by: Ge, Songwei, et al.
Published: (2023)
Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation
by: Petrovich, Mathis, et al.
Published: (2024)
by: Petrovich, Mathis, et al.
Published: (2024)
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
by: Avrahami, Omri, et al.
Published: (2023)
by: Avrahami, Omri, et al.
Published: (2023)
LumiX: Structured and Coherent Text-to-Intrinsic Generation
by: Han, Xu, et al.
Published: (2025)
by: Han, Xu, et al.
Published: (2025)
Training-Free Consistent Text-to-Image Generation
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
LUSD: Localized Update Score Distillation for Text-Guided Image Editing
by: Chinchuthakun, Worameth, et al.
Published: (2025)
by: Chinchuthakun, Worameth, et al.
Published: (2025)
CAD-Coder:Text-Guided CAD Files Code Generation
by: He, Changqi, et al.
Published: (2025)
by: He, Changqi, et al.
Published: (2025)
Ghost on the Shell: An Expressive Representation of General 3D Shapes
by: Liu, Zhen, et al.
Published: (2023)
by: Liu, Zhen, et al.
Published: (2023)
DreamBlend: Advancing Personalized Fine-tuning of Text-to-Image Diffusion Models
by: Ram, Shwetha, et al.
Published: (2024)
by: Ram, Shwetha, et al.
Published: (2024)
GaussEdit: Adaptive 3D Scene Editing with Text and Image Prompts
by: Shu, Zhenyu, et al.
Published: (2025)
by: Shu, Zhenyu, et al.
Published: (2025)
DreamWaltz-G: Expressive 3D Gaussian Avatars from Skeleton-Guided 2D Diffusion
by: Huang, Yukun, et al.
Published: (2024)
by: Huang, Yukun, et al.
Published: (2024)
Story2Board: A Training-Free Approach for Expressive Storyboard Generation
by: Dinkevich, David, et al.
Published: (2025)
by: Dinkevich, David, et al.
Published: (2025)
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
by: Guo, Yuwei, et al.
Published: (2023)
by: Guo, Yuwei, et al.
Published: (2023)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
Diverse Text-to-Image Generation via Contrastive Noise Optimization
by: Kim, Byungjun, et al.
Published: (2025)
by: Kim, Byungjun, et al.
Published: (2025)
CookingDiffusion: Cooking Procedural Image Generation with Stable Diffusion
by: Wang, Yuan, et al.
Published: (2025)
by: Wang, Yuan, et al.
Published: (2025)
MESA: Text-Driven Terrain Generation Using Latent Diffusion and Global Copernicus Data
by: Borne--Pons, Paul, et al.
Published: (2025)
by: Borne--Pons, Paul, et al.
Published: (2025)
InseRF: Text-Driven Generative Object Insertion in Neural 3D Scenes
by: Shahbazi, Mohamad, et al.
Published: (2024)
by: Shahbazi, Mohamad, et al.
Published: (2024)
Text-to-Image Generation for Vocabulary Learning Using the Keyword Method
by: Attygalle, Nuwan T., et al.
Published: (2025)
by: Attygalle, Nuwan T., et al.
Published: (2025)
PALP: Prompt Aligned Personalization of Text-to-Image Models
by: Arar, Moab, et al.
Published: (2024)
by: Arar, Moab, et al.
Published: (2024)
From Competition to Synergy: Unlocking Reinforcement Learning for Subject-Driven Image Generation
by: Huang, Ziwei, et al.
Published: (2025)
by: Huang, Ziwei, et al.
Published: (2025)
Consolidating Attention Features for Multi-view Image Editing
by: Patashnik, Or, et al.
Published: (2024)
by: Patashnik, Or, et al.
Published: (2024)
Instant3D: Instant Text-to-3D Generation
by: Li, Ming, et al.
Published: (2023)
by: Li, Ming, et al.
Published: (2023)
LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR
by: Fekri, Pedram, et al.
Published: (2026)
by: Fekri, Pedram, et al.
Published: (2026)
Controlling Text-to-Image Diffusion by Orthogonal Finetuning
by: Qiu, Zeju, et al.
Published: (2023)
by: Qiu, Zeju, et al.
Published: (2023)
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
by: Wei, Yuanxin, et al.
Published: (2025)
by: Wei, Yuanxin, et al.
Published: (2025)
Scaling Group Inference for Diverse and High-Quality Generation
by: Parmar, Gaurav, et al.
Published: (2025)
by: Parmar, Gaurav, et al.
Published: (2025)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
by: Girdhar, Rohit, et al.
Published: (2023)
by: Girdhar, Rohit, et al.
Published: (2023)
Hollowed Net for On-Device Personalization of Text-to-Image Diffusion Models
by: Cho, Wonguk, et al.
Published: (2024)
by: Cho, Wonguk, et al.
Published: (2024)
Animus3D: Text-driven 3D Animation via Motion Score Distillation
by: Sun, Qi, et al.
Published: (2025)
by: Sun, Qi, et al.
Published: (2025)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
by: Kim, Bumsoo, et al.
Published: (2024)
by: Kim, Bumsoo, et al.
Published: (2024)
Detection-Driven Object Count Optimization for Text-to-Image Diffusion Models
by: Zafar, Oz, et al.
Published: (2024)
by: Zafar, Oz, et al.
Published: (2024)
Depth-supervised NeRF: Fewer Views and Faster Training for Free
by: Deng, Kangle, et al.
Published: (2021)
by: Deng, Kangle, et al.
Published: (2021)
DebiasPI: Inference-time Debiasing by Prompt Iteration of a Text-to-Image Generative Model
by: Bonna, Sarah, et al.
Published: (2025)
by: Bonna, Sarah, et al.
Published: (2025)
Similar Items
-
One-Step Image Translation with Text-to-Image Models
by: Parmar, Gaurav, et al.
Published: (2024) -
On the Content Bias in Fréchet Video Distance
by: Ge, Songwei, et al.
Published: (2024) -
Generating Multi-Image Synthetic Data for Text-to-Image Customization
by: Kumari, Nupur, et al.
Published: (2025) -
Customizing Text-to-Image Models with a Single Image Pair
by: Jones, Maxwell, et al.
Published: (2024) -
Rethinking Score Distillation as a Bridge Between Image Distributions
by: McAllister, David, et al.
Published: (2024)