Diverse Text-to-Image Generation via Contrastive Noise Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Byungjun, Um, Soobin, Ye, Jong Chul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Minority-Focused Text-to-Image Generation via Prompt Optimization
by: Um, Soobin, et al.
Published: (2024)
by: Um, Soobin, et al.
Published: (2024)
MotionCFG: Boosting Motion Dynamics via Stochastic Concept Perturbation
by: Kim, Byungjun, et al.
Published: (2026)
by: Kim, Byungjun, et al.
Published: (2026)
Self-Guided Generation of Minority Samples Using Diffusion Models
by: Um, Soobin, et al.
Published: (2024)
by: Um, Soobin, et al.
Published: (2024)
Boost-and-Skip: A Simple Guidance-Free Diffusion for Minority Generation
by: Um, Soobin, et al.
Published: (2025)
by: Um, Soobin, et al.
Published: (2025)
Don't Play Favorites: Minority Guidance for Diffusion Models
by: Um, Soobin, et al.
Published: (2023)
by: Um, Soobin, et al.
Published: (2023)
Training-Free Consistent Text-to-Image Generation
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
Detection-Driven Object Count Optimization for Text-to-Image Diffusion Models
by: Zafar, Oz, et al.
Published: (2024)
by: Zafar, Oz, et al.
Published: (2024)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion
by: Park, Sol, et al.
Published: (2026)
by: Park, Sol, et al.
Published: (2026)
Hollowed Net for On-Device Personalization of Text-to-Image Diffusion Models
by: Cho, Wonguk, et al.
Published: (2024)
by: Cho, Wonguk, et al.
Published: (2024)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
by: Kim, Bumsoo, et al.
Published: (2024)
by: Kim, Bumsoo, et al.
Published: (2024)
Gradient-Free Noise Optimization for Reward Alignment in Generative Models
by: Kim, Jeongsol, et al.
Published: (2026)
by: Kim, Jeongsol, et al.
Published: (2026)
Φ-Noise: Training-Free Temporal Video Conditioning via Phase-Based Noise Manipulation
by: Abramovich, Ofir, et al.
Published: (2026)
by: Abramovich, Ofir, et al.
Published: (2026)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
by: Dahary, Omer, et al.
Published: (2025)
by: Dahary, Omer, et al.
Published: (2025)
Stylized Text-to-Motion Generation via Hypernetwork-Driven Low-Rank Adaptation
by: Jeon, Junhyuk, et al.
Published: (2026)
by: Jeon, Junhyuk, et al.
Published: (2026)
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
by: Hong, Seokhyeon, et al.
Published: (2025)
by: Hong, Seokhyeon, et al.
Published: (2025)
Controlling Text-to-Image Diffusion by Orthogonal Finetuning
by: Qiu, Zeju, et al.
Published: (2023)
by: Qiu, Zeju, et al.
Published: (2023)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
by: Girdhar, Rohit, et al.
Published: (2023)
by: Girdhar, Rohit, et al.
Published: (2023)
HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models
by: Ruiz, Nataniel, et al.
Published: (2023)
by: Ruiz, Nataniel, et al.
Published: (2023)
ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts
by: Petrov, Dmitry, et al.
Published: (2024)
by: Petrov, Dmitry, et al.
Published: (2024)
Navigating Text-To-Image Customization: From LyCORIS Fine-Tuning to Model Evaluation
by: Yeh, Shih-Ying, et al.
Published: (2023)
by: Yeh, Shih-Ying, et al.
Published: (2023)
PCPO: Proportionate Credit Policy Optimization for Aligning Image Generation Models
by: Lee, Jeongjae, et al.
Published: (2025)
by: Lee, Jeongjae, et al.
Published: (2025)
RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion
by: Shriram, Jaidev, et al.
Published: (2024)
by: Shriram, Jaidev, et al.
Published: (2024)
Infinite-Resolution Integral Noise Warping for Diffusion Models
by: Deng, Yitong, et al.
Published: (2024)
by: Deng, Yitong, et al.
Published: (2024)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
by: Nam, Hyelin, et al.
Published: (2023)
by: Nam, Hyelin, et al.
Published: (2023)
TAUE: Training-free Noise Transplant and Cultivation Diffusion Model
by: Nagai, Daichi, et al.
Published: (2025)
by: Nagai, Daichi, et al.
Published: (2025)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
by: Cai, Shengqu, et al.
Published: (2024)
by: Cai, Shengqu, et al.
Published: (2024)
RealFill: Reference-Driven Generation for Authentic Image Completion
by: Tang, Luming, et al.
Published: (2023)
by: Tang, Luming, et al.
Published: (2023)
Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters
by: Chiu, Pin-Yen, et al.
Published: (2025)
by: Chiu, Pin-Yen, et al.
Published: (2025)
Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM
by: Kim, Jaemin, et al.
Published: (2024)
by: Kim, Jaemin, et al.
Published: (2024)
On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers
by: Dahary, Omer, et al.
Published: (2026)
by: Dahary, Omer, et al.
Published: (2026)
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation
by: Tan, Xianfeng, et al.
Published: (2024)
by: Tan, Xianfeng, et al.
Published: (2024)
Instant3D: Instant Text-to-3D Generation
by: Li, Ming, et al.
Published: (2023)
by: Li, Ming, et al.
Published: (2023)
ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
DreamCatalyst: Fast and High-Quality 3D Editing via Controlling Editability and Identity Preservation
by: Kim, Jiwook, et al.
Published: (2024)
by: Kim, Jiwook, et al.
Published: (2024)
Generalized Consistency Trajectory Models for Image Manipulation
by: Kim, Beomsu, et al.
Published: (2024)
by: Kim, Beomsu, et al.
Published: (2024)
Unpaired Image-to-Image Translation via Neural Schrödinger Bridge
by: Kim, Beomsu, et al.
Published: (2023)
by: Kim, Beomsu, et al.
Published: (2023)
Walk Before You Dance: High-fidelity and Editable Dance Synthesis via Generative Masked Motion Prior
by: Shah, Foram N, et al.
Published: (2025)
by: Shah, Foram N, et al.
Published: (2025)
Extreme Blind Image Restoration via Prompt-Conditioned Information Bottleneck
by: Kim, Hongeun, et al.
Published: (2025)
by: Kim, Hongeun, et al.
Published: (2025)
Similar Items
-
Minority-Focused Text-to-Image Generation via Prompt Optimization
by: Um, Soobin, et al.
Published: (2024) -
MotionCFG: Boosting Motion Dynamics via Stochastic Concept Perturbation
by: Kim, Byungjun, et al.
Published: (2026) -
Self-Guided Generation of Minority Samples Using Diffusion Models
by: Um, Soobin, et al.
Published: (2024) -
Boost-and-Skip: A Simple Guidance-Free Diffusion for Minority Generation
by: Um, Soobin, et al.
Published: (2025) -
Don't Play Favorites: Minority Guidance for Diffusion Models
by: Um, Soobin, et al.
Published: (2023)