Dynamic Prompt Optimizing for Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mo, Wenyi, Zhang, Tianyu, Bai, Yalong, Su, Bing, Wen, Ji-Rong, Yang, Qing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2025)
von: Ba, Ying, et al.
Veröffentlicht: (2025)
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
Learning User Preferences for Image Generation Model
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Pareto-Guided Optimal Transport for Multi-Reward Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2026)
von: Ba, Ying, et al.
Veröffentlicht: (2026)
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
von: Wu, Mingrui, et al.
Veröffentlicht: (2025)
von: Wu, Mingrui, et al.
Veröffentlicht: (2025)
Batch-Instructed Gradient for Prompt Evolution:Systematic Prompt Optimization for Enhanced Text-to-Image Synthesis
von: Yang, Xinrui, et al.
Veröffentlicht: (2024)
von: Yang, Xinrui, et al.
Veröffentlicht: (2024)
IFAdapter: Instance Feature Control for Grounded Text-to-Image Generation
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
GenPilot: A Multi-Agent System for Test-Time Prompt Optimization in Image Generation
von: Ye, Wen, et al.
Veröffentlicht: (2025)
von: Ye, Wen, et al.
Veröffentlicht: (2025)
Minority-Focused Text-to-Image Generation via Prompt Optimization
von: Um, Soobin, et al.
Veröffentlicht: (2024)
von: Um, Soobin, et al.
Veröffentlicht: (2024)
Memory-Inspired Temporal Prompt Interaction for Text-Image Classification
von: Yu, Xinyao, et al.
Veröffentlicht: (2024)
von: Yu, Xinyao, et al.
Veröffentlicht: (2024)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
Rethinking Prompt Design for Inference-time Scaling in Text-to-Visual Generation
von: Kim, Subin, et al.
Veröffentlicht: (2025)
von: Kim, Subin, et al.
Veröffentlicht: (2025)
Reverse Prompt: Cracking the Recipe Inside Text-to-Image Generation
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
Long-Text-to-Image Generation via Compositional Prompt Decomposition
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
Human-Guided Image Generation for Expanding Small-Scale Training Image Datasets
von: Chen, Changjian, et al.
Veröffentlicht: (2024)
von: Chen, Changjian, et al.
Veröffentlicht: (2024)
Optimizing Negative Prompts for Enhanced Aesthetics and Fidelity in Text-To-Image Generation
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
Reflective Human-Machine Co-adaptation for Enhanced Text-to-Image Generation Dialogue System
von: Feng, Yuheng, et al.
Veröffentlicht: (2024)
von: Feng, Yuheng, et al.
Veröffentlicht: (2024)
Supporting Vision-Language Model Inference with Confounder-pruning Knowledge Prompt
von: Li, Jiangmeng, et al.
Veröffentlicht: (2022)
von: Li, Jiangmeng, et al.
Veröffentlicht: (2022)
FairQueue: Rethinking Prompt Learning for Fair Text-to-Image Generation
von: Teo, Christopher T. H, et al.
Veröffentlicht: (2024)
von: Teo, Christopher T. H, et al.
Veröffentlicht: (2024)
Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
von: Park, Sangha, et al.
Veröffentlicht: (2025)
von: Park, Sangha, et al.
Veröffentlicht: (2025)
Progressive Prompt Detailing for Improved Alignment in Text-to-Image Generative Models
von: Saichandran, Ketan Suhaas, et al.
Veröffentlicht: (2025)
von: Saichandran, Ketan Suhaas, et al.
Veröffentlicht: (2025)
Prompt Decoupling for Text-to-Image Person Re-identification
von: Li, Weihao, et al.
Veröffentlicht: (2024)
von: Li, Weihao, et al.
Veröffentlicht: (2024)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
Guidance Matters: Rethinking the Evaluation Pitfall for Text-to-Image Generation
von: Xie, Dian, et al.
Veröffentlicht: (2026)
von: Xie, Dian, et al.
Veröffentlicht: (2026)
StyleInject: Parameter Efficient Tuning of Text-to-Image Diffusion Models
von: Zhou, Mohan, et al.
Veröffentlicht: (2024)
von: Zhou, Mohan, et al.
Veröffentlicht: (2024)
Enhancing Text-to-Image Diffusion Transformer via Split-Text Conditioning
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
Position: Towards Implicit Prompt For Text-To-Image Models
von: Yang, Yue, et al.
Veröffentlicht: (2024)
von: Yang, Yue, et al.
Veröffentlicht: (2024)
Hierarchical Prompt Learning for Image- and Text-Based Person Re-Identification
von: Zhou, Linhan, et al.
Veröffentlicht: (2025)
von: Zhou, Linhan, et al.
Veröffentlicht: (2025)
Retrieval, Refinement, and Ranking for Text-to-Video Generation via Prompt Optimization and Test-Time Scaling
von: Rahman, Zillur, et al.
Veröffentlicht: (2026)
von: Rahman, Zillur, et al.
Veröffentlicht: (2026)
PriorCLIP: Visual Prior Guided Vision-Language Model for Remote Sensing Image-Text Retrieval
von: Pan, Jiancheng, et al.
Veröffentlicht: (2024)
von: Pan, Jiancheng, et al.
Veröffentlicht: (2024)
InstructAvatar: Text-Guided Emotion and Motion Control for Avatar Generation
von: Wang, Yuchi, et al.
Veröffentlicht: (2024)
von: Wang, Yuchi, et al.
Veröffentlicht: (2024)
Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
von: Peng, Xingkai, et al.
Veröffentlicht: (2025)
von: Peng, Xingkai, et al.
Veröffentlicht: (2025)
A User-Friendly Framework for Generating Model-Preferred Prompts in Text-to-Image Synthesis
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
GreenStableYolo: Optimizing Inference Time and Image Quality of Text-to-Image Generation
von: Gong, Jingzhi, et al.
Veröffentlicht: (2024)
von: Gong, Jingzhi, et al.
Veröffentlicht: (2024)
Compress3D: a Compressed Latent Space for 3D Generation from a Single Image
von: Zhang, Bowen, et al.
Veröffentlicht: (2024)
von: Zhang, Bowen, et al.
Veröffentlicht: (2024)
Optimizing LVLMs with On-Policy Data for Effective Hallucination Mitigation
von: Yu, Chengzhi, et al.
Veröffentlicht: (2025)
von: Yu, Chengzhi, et al.
Veröffentlicht: (2025)
Curriculum Group Policy Optimization: Adaptive Sampling for Unleashing the Potential of Text-to-Image Generation
von: Li, Baoteng, et al.
Veröffentlicht: (2026)
von: Li, Baoteng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
von: Mo, Wenyi, et al.
Veröffentlicht: (2024) -
Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2025) -
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025) -
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025) -
Learning User Preferences for Image Generation Model
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)