Long-Text-to-Image Generation via Compositional Prompt Decomposition
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Jen-Yuan, Lin, Tong, Du, Yilun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?
di: Jiao, Qirui, et al.
Pubblicazione: (2025)
di: Jiao, Qirui, et al.
Pubblicazione: (2025)
MEPG:Multi-Expert Planning and Generation for Compositionally-Rich Image Generation
di: Zhao, Yuan, et al.
Pubblicazione: (2025)
di: Zhao, Yuan, et al.
Pubblicazione: (2025)
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
Compositional Generative Modeling: A Single Model is Not All You Need
di: Du, Yilun, et al.
Pubblicazione: (2024)
di: Du, Yilun, et al.
Pubblicazione: (2024)
Interpreting CLIP's Image Representation via Text-Based Decomposition
di: Gandelsman, Yossi, et al.
Pubblicazione: (2023)
di: Gandelsman, Yossi, et al.
Pubblicazione: (2023)
Optimizing Few-Step Sampler for Diffusion Probabilistic Model
di: Huang, Jen-Yuan
Pubblicazione: (2024)
di: Huang, Jen-Yuan
Pubblicazione: (2024)
Progressive Image Restoration via Text-Conditioned Video Generation
di: Kang, Peng, et al.
Pubblicazione: (2025)
di: Kang, Peng, et al.
Pubblicazione: (2025)
Reverse Prompt: Cracking the Recipe Inside Text-to-Image Generation
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
VersusDebias: Universal Zero-Shot Debiasing for Text-to-Image Models via SLM-Based Prompt Engineering and Generative Adversary
di: Luo, Hanjun, et al.
Pubblicazione: (2024)
di: Luo, Hanjun, et al.
Pubblicazione: (2024)
PhyPrompt: RL-based Prompt Refinement for Physically Plausible Text-to-Video Generation
di: Wu, Shang, et al.
Pubblicazione: (2026)
di: Wu, Shang, et al.
Pubblicazione: (2026)
FairQueue: Rethinking Prompt Learning for Fair Text-to-Image Generation
di: Teo, Christopher T. H, et al.
Pubblicazione: (2024)
di: Teo, Christopher T. H, et al.
Pubblicazione: (2024)
Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
di: Dai, Dawei, et al.
Pubblicazione: (2025)
di: Dai, Dawei, et al.
Pubblicazione: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
di: Park, Sangha, et al.
Pubblicazione: (2025)
di: Park, Sangha, et al.
Pubblicazione: (2025)
Progressive Prompt Detailing for Improved Alignment in Text-to-Image Generative Models
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
Minority-Focused Text-to-Image Generation via Prompt Optimization
di: Um, Soobin, et al.
Pubblicazione: (2024)
di: Um, Soobin, et al.
Pubblicazione: (2024)
Memory-Inspired Temporal Prompt Interaction for Text-Image Classification
di: Yu, Xinyao, et al.
Pubblicazione: (2024)
di: Yu, Xinyao, et al.
Pubblicazione: (2024)
Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
di: Chen, Wenting, et al.
Pubblicazione: (2024)
di: Chen, Wenting, et al.
Pubblicazione: (2024)
Interactive Visual Assessment for Text-to-Image Generation Models
di: Mi, Xiaoyue, et al.
Pubblicazione: (2024)
di: Mi, Xiaoyue, et al.
Pubblicazione: (2024)
Object-level Visual Prompts for Compositional Image Generation
di: Parmar, Gaurav, et al.
Pubblicazione: (2025)
di: Parmar, Gaurav, et al.
Pubblicazione: (2025)
DiffDecompose: Layer-Wise Decomposition of Alpha-Composited Images via Diffusion Transformers
di: Wang, Zitong, et al.
Pubblicazione: (2025)
di: Wang, Zitong, et al.
Pubblicazione: (2025)
SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation
di: Ren, Tianfei, et al.
Pubblicazione: (2026)
di: Ren, Tianfei, et al.
Pubblicazione: (2026)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
ArtCrafter: Text-Image Aligning Style Transfer via Embedding Reframing
di: Huang, Nisha, et al.
Pubblicazione: (2025)
di: Huang, Nisha, et al.
Pubblicazione: (2025)
Right Looks, Wrong Reasons: Compositional Fidelity in Text-to-Image Generation
di: Vatsa, Mayank, et al.
Pubblicazione: (2025)
di: Vatsa, Mayank, et al.
Pubblicazione: (2025)
InstantIR: Blind Image Restoration with Instant Generative Reference
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2024)
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2024)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
di: Zheng, Zirui, et al.
Pubblicazione: (2025)
di: Zheng, Zirui, et al.
Pubblicazione: (2025)
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
di: Wang, Yuran, et al.
Pubblicazione: (2025)
di: Wang, Yuran, et al.
Pubblicazione: (2025)
MetaLogic: Robustness Evaluation of Text-to-Image Models via Logically Equivalent Prompts
di: Shen, Yifan, et al.
Pubblicazione: (2025)
di: Shen, Yifan, et al.
Pubblicazione: (2025)
Region Prompt Tuning: Fine-grained Scene Text Detection Utilizing Region Text Prompt
di: Lin, Xingtao, et al.
Pubblicazione: (2024)
di: Lin, Xingtao, et al.
Pubblicazione: (2024)
Equilibrium Matching: Generative Modeling with Implicit Energy-Based Models
di: Wang, Runqian, et al.
Pubblicazione: (2025)
di: Wang, Runqian, et al.
Pubblicazione: (2025)
Prompt Decoupling for Text-to-Image Person Re-identification
di: Li, Weihao, et al.
Pubblicazione: (2024)
di: Li, Weihao, et al.
Pubblicazione: (2024)
PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models
di: Yuan, Lingzhi, et al.
Pubblicazione: (2025)
di: Yuan, Lingzhi, et al.
Pubblicazione: (2025)
Ctrl-VI: Controllable Video Synthesis via Variational Inference
di: Duan, Haoyi, et al.
Pubblicazione: (2025)
di: Duan, Haoyi, et al.
Pubblicazione: (2025)
How Long Can Unified Multimodal Models Generate Images Reliably? Taming Long-Horizon Interleaved Image Generation via Context Curation
di: Chen, Haoyu, et al.
Pubblicazione: (2026)
di: Chen, Haoyu, et al.
Pubblicazione: (2026)
Position: Towards Implicit Prompt For Text-To-Image Models
di: Yang, Yue, et al.
Pubblicazione: (2024)
di: Yang, Yue, et al.
Pubblicazione: (2024)
TGC-Net: A Structure-Aware and Semantically-Aligned Framework for Text-Guided Medical Image Segmentation
di: Lin, Gaoren, et al.
Pubblicazione: (2025)
di: Lin, Gaoren, et al.
Pubblicazione: (2025)
How Bias Binds: Measuring Hidden Associations for Bias Control in Text-to-Image Compositions
di: Li, Jeng-Lin, et al.
Pubblicazione: (2025)
di: Li, Jeng-Lin, et al.
Pubblicazione: (2025)
Batch-Instructed Gradient for Prompt Evolution:Systematic Prompt Optimization for Enhanced Text-to-Image Synthesis
di: Yang, Xinrui, et al.
Pubblicazione: (2024)
di: Yang, Xinrui, et al.
Pubblicazione: (2024)
Documenti analoghi
-
DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?
di: Jiao, Qirui, et al.
Pubblicazione: (2025) -
MEPG:Multi-Expert Planning and Generation for Compositionally-Rich Image Generation
di: Zhao, Yuan, et al.
Pubblicazione: (2025) -
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
di: Wu, Mingrui, et al.
Pubblicazione: (2025) -
Dynamic Prompt Optimizing for Text-to-Image Generation
di: Mo, Wenyi, et al.
Pubblicazione: (2024) -
Compositional Generative Modeling: A Single Model is Not All You Need
di: Du, Yilun, et al.
Pubblicazione: (2024)