Creative Image Generation with Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Kunpeng, Elgammal, Ahmed |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoMA: Multimodal LLM Adapter for Fast Personalized Image Generation
by: Song, Kunpeng, et al.
Published: (2024)
by: Song, Kunpeng, et al.
Published: (2024)
Modality Agnostic Efficient Long Range Encoder
by: Parag, Toufiq, et al.
Published: (2025)
by: Parag, Toufiq, et al.
Published: (2025)
Enhancing Creative Generation on Stable Diffusion-based Models
by: Han, Jiyeon, et al.
Published: (2025)
by: Han, Jiyeon, et al.
Published: (2025)
CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation
by: Zhang, Hui, et al.
Published: (2024)
by: Zhang, Hui, et al.
Published: (2024)
Noise-Consistent Siamese-Diffusion for Medical Image Synthesis and Segmentation
by: Qiu, Kunpeng, et al.
Published: (2025)
by: Qiu, Kunpeng, et al.
Published: (2025)
Llama Learns to Direct: DirectorLLM for Human-Centric Video Generation
by: Song, Kunpeng, et al.
Published: (2024)
by: Song, Kunpeng, et al.
Published: (2024)
MultiShadow: Multi-Object Shadow Generation for Image Compositing via Diffusion Model
by: Ahmed, Waqas, et al.
Published: (2026)
by: Ahmed, Waqas, et al.
Published: (2026)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
by: Cao, Pu, et al.
Published: (2024)
by: Cao, Pu, et al.
Published: (2024)
HiDiffusion: Unlocking Higher-Resolution Creativity and Efficiency in Pretrained Diffusion Models
by: Zhang, Shen, et al.
Published: (2023)
by: Zhang, Shen, et al.
Published: (2023)
ProCreate, Don't Reproduce! Propulsive Energy Diffusion for Creative Generation
by: Lu, Jack, et al.
Published: (2024)
by: Lu, Jack, et al.
Published: (2024)
CCEdit: Creative and Controllable Video Editing via Diffusion Models
by: Feng, Ruoyu, et al.
Published: (2023)
by: Feng, Ruoyu, et al.
Published: (2023)
AI Art Neural Constellation: Revealing the Collective and Contrastive State of AI-Generated and Human Art
by: Khan, Faizan Farooq, et al.
Published: (2024)
by: Khan, Faizan Farooq, et al.
Published: (2024)
Learn From Zoom: Decoupled Supervised Contrastive Learning For WCE Image Classification
by: Qiu, Kunpeng, et al.
Published: (2024)
by: Qiu, Kunpeng, et al.
Published: (2024)
Versatile Transition Generation with Image-to-Video Diffusion
by: Yang, Zuhao, et al.
Published: (2025)
by: Yang, Zuhao, et al.
Published: (2025)
Adaptively Distilled ControlNet: Accelerated Training and Superior Sampling for Medical Image Synthesis
by: Qiu, Kunpeng, et al.
Published: (2025)
by: Qiu, Kunpeng, et al.
Published: (2025)
CAP: Evaluation of Persuasive and Creative Image Generation
by: Aghazadeh, Aysan, et al.
Published: (2024)
by: Aghazadeh, Aysan, et al.
Published: (2024)
Origins of Creativity in Attention-Based Diffusion Models
by: Finn, Emma, et al.
Published: (2025)
by: Finn, Emma, et al.
Published: (2025)
CREA: A Collaborative Multi-Agent Framework for Creative Image Editing and Generation
by: Venkatesh, Kavana, et al.
Published: (2025)
by: Venkatesh, Kavana, et al.
Published: (2025)
Disciplined Diffusion: Text-to-Image Diffusion Model against NSFW Generation
by: Zhang, Chi, et al.
Published: (2026)
by: Zhang, Chi, et al.
Published: (2026)
Debiasing Text-to-Image Diffusion Models
by: He, Ruifei, et al.
Published: (2024)
by: He, Ruifei, et al.
Published: (2024)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
by: Hu, Teng, et al.
Published: (2023)
by: Hu, Teng, et al.
Published: (2023)
Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation
by: Ng, Kam Woh, et al.
Published: (2025)
by: Ng, Kam Woh, et al.
Published: (2025)
Creative4U: MLLMs-based Advertising Creative Image Selector with Comparative Reasoning
by: Lin, Yukang, et al.
Published: (2025)
by: Lin, Yukang, et al.
Published: (2025)
Image-to-Brain Signal Generation for Visual Prosthesis with CLIP Guided Multimodal Diffusion Models
by: Xu, Ganxi, et al.
Published: (2025)
by: Xu, Ganxi, et al.
Published: (2025)
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
by: Rahman, Tanzila, et al.
Published: (2024)
by: Rahman, Tanzila, et al.
Published: (2024)
Reflection Generation for Composite Image Using Diffusion Model
by: Zhao, Haonan, et al.
Published: (2026)
by: Zhao, Haonan, et al.
Published: (2026)
Diffusion Models Need Visual Priors for Image Generation
by: Yue, Xiaoyu, et al.
Published: (2024)
by: Yue, Xiaoyu, et al.
Published: (2024)
Paragraph-to-Image Generation with Information-Enriched Diffusion Model
by: Wu, Weijia, et al.
Published: (2023)
by: Wu, Weijia, et al.
Published: (2023)
TwinDiffusion: Enhancing Coherence and Efficiency in Panoramic Image Generation with Diffusion Models
by: Zhou, Teng, et al.
Published: (2024)
by: Zhou, Teng, et al.
Published: (2024)
STAY Diffusion: Styled Layout Diffusion Model for Diverse Layout-to-Image Generation
by: Wang, Ruyu, et al.
Published: (2025)
by: Wang, Ruyu, et al.
Published: (2025)
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
by: Wang, Zhendong, et al.
Published: (2025)
by: Wang, Zhendong, et al.
Published: (2025)
SonicDiffusion: Audio-Driven Image Generation and Editing with Pretrained Diffusion Models
by: Biner, Burak Can, et al.
Published: (2024)
by: Biner, Burak Can, et al.
Published: (2024)
EasyRef: Omni-Generalized Group Image Reference for Diffusion Models via Multimodal LLM
by: Zong, Zhuofan, et al.
Published: (2024)
by: Zong, Zhuofan, et al.
Published: (2024)
EmoGen: Emotional Image Content Generation with Text-to-Image Diffusion Models
by: Yang, Jingyuan, et al.
Published: (2024)
by: Yang, Jingyuan, et al.
Published: (2024)
Multi-Object Advertisement Creative Generation
by: Gao, Jialu, et al.
Published: (2026)
by: Gao, Jialu, et al.
Published: (2026)
ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices
by: Du, Kunpeng, et al.
Published: (2026)
by: Du, Kunpeng, et al.
Published: (2026)
RAW-Diffusion: RGB-Guided Diffusion Models for High-Fidelity RAW Image Generation
by: Reinders, Christoph, et al.
Published: (2024)
by: Reinders, Christoph, et al.
Published: (2024)
Discriminative Image Generation with Diffusion Models for Zero-Shot Learning
by: Fu, Dingjie, et al.
Published: (2024)
by: Fu, Dingjie, et al.
Published: (2024)
PID: Physics-Informed Diffusion Model for Infrared Image Generation
by: Mao, Fangyuan, et al.
Published: (2024)
by: Mao, Fangyuan, et al.
Published: (2024)
Similar Items
-
MoMA: Multimodal LLM Adapter for Fast Personalized Image Generation
by: Song, Kunpeng, et al.
Published: (2024) -
Modality Agnostic Efficient Long Range Encoder
by: Parag, Toufiq, et al.
Published: (2025) -
Enhancing Creative Generation on Stable Diffusion-based Models
by: Han, Jiyeon, et al.
Published: (2025) -
CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation
by: Zhang, Hui, et al.
Published: (2024) -
Noise-Consistent Siamese-Diffusion for Medical Image Synthesis and Segmentation
by: Qiu, Kunpeng, et al.
Published: (2025)