Reverse Prompt: Cracking the Recipe Inside Text-to-Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Ren, Zhiyao, Zhan, Yibing, Yu, Baosheng, Tao, Dacheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis
di: Zhao, Shuxian, et al.
Pubblicazione: (2026)
di: Zhao, Shuxian, et al.
Pubblicazione: (2026)
Image Captions are Natural Prompts for Text-to-Image Models
di: Lei, Shiye, et al.
Pubblicazione: (2023)
di: Lei, Shiye, et al.
Pubblicazione: (2023)
Dynamic Prompt Optimizing for Text-to-Image Generation
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
Prompt-Consistency Image Generation (PCIG): A Unified Framework Integrating LLMs, Knowledge Graphs, and Controllable Diffusion Models
di: Sun, Yichen, et al.
Pubblicazione: (2024)
di: Sun, Yichen, et al.
Pubblicazione: (2024)
Self-Supervised Learning for Detecting AI-Generated Faces as Anomalies
di: Zou, Mian, et al.
Pubblicazione: (2025)
di: Zou, Mian, et al.
Pubblicazione: (2025)
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
CookAnything: A Framework for Flexible and Consistent Multi-Step Recipe Image Generation
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models
di: Ma, Sihan, et al.
Pubblicazione: (2026)
di: Ma, Sihan, et al.
Pubblicazione: (2026)
Long-Text-to-Image Generation via Compositional Prompt Decomposition
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2026)
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2026)
Capability-aware Prompt Reformulation Learning for Text-to-Image Generation
di: Zhan, Jingtao, et al.
Pubblicazione: (2024)
di: Zhan, Jingtao, et al.
Pubblicazione: (2024)
Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
di: Dai, Dawei, et al.
Pubblicazione: (2025)
di: Dai, Dawei, et al.
Pubblicazione: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
di: Park, Sangha, et al.
Pubblicazione: (2025)
di: Park, Sangha, et al.
Pubblicazione: (2025)
Progressive Prompt Detailing for Improved Alignment in Text-to-Image Generative Models
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
FairQueue: Rethinking Prompt Learning for Fair Text-to-Image Generation
di: Teo, Christopher T. H, et al.
Pubblicazione: (2024)
di: Teo, Christopher T. H, et al.
Pubblicazione: (2024)
Memory-Inspired Temporal Prompt Interaction for Text-Image Classification
di: Yu, Xinyao, et al.
Pubblicazione: (2024)
di: Yu, Xinyao, et al.
Pubblicazione: (2024)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
Contact-aware Human Motion Generation from Textual Descriptions
di: Ma, Sihan, et al.
Pubblicazione: (2024)
di: Ma, Sihan, et al.
Pubblicazione: (2024)
Position: Towards Implicit Prompt For Text-To-Image Models
di: Yang, Yue, et al.
Pubblicazione: (2024)
di: Yang, Yue, et al.
Pubblicazione: (2024)
Prompt Decoupling for Text-to-Image Person Re-identification
di: Li, Weihao, et al.
Pubblicazione: (2024)
di: Li, Weihao, et al.
Pubblicazione: (2024)
Audio Visual Segmentation Through Text Embeddings
di: Lee, Kyungbok, et al.
Pubblicazione: (2025)
di: Lee, Kyungbok, et al.
Pubblicazione: (2025)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
di: Zhao, Yu, et al.
Pubblicazione: (2024)
di: Zhao, Yu, et al.
Pubblicazione: (2024)
LARGO: Low-Rank Regulated Gradient Projection for Robust Parameter Efficient Fine-Tuning
di: Zhang, Haotian, et al.
Pubblicazione: (2025)
di: Zhang, Haotian, et al.
Pubblicazione: (2025)
Improving Post-Earthquake Crack Detection using Semi-Synthetic Generated Images
di: Dondi, Piercarlo, et al.
Pubblicazione: (2024)
di: Dondi, Piercarlo, et al.
Pubblicazione: (2024)
Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
di: Chen, Wenting, et al.
Pubblicazione: (2024)
di: Chen, Wenting, et al.
Pubblicazione: (2024)
Bi-Level Optimization for Self-Supervised AI-Generated Face Detection
di: Zou, Mian, et al.
Pubblicazione: (2025)
di: Zou, Mian, et al.
Pubblicazione: (2025)
Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
di: Wei, Yibing, et al.
Pubblicazione: (2024)
di: Wei, Yibing, et al.
Pubblicazione: (2024)
MASTER: Multimodal Segmentation with Text Prompts
di: Liu, Fuyang, et al.
Pubblicazione: (2025)
di: Liu, Fuyang, et al.
Pubblicazione: (2025)
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
di: Liu, Tao, et al.
Pubblicazione: (2025)
di: Liu, Tao, et al.
Pubblicazione: (2025)
Batch-Instructed Gradient for Prompt Evolution:Systematic Prompt Optimization for Enhanced Text-to-Image Synthesis
di: Yang, Xinrui, et al.
Pubblicazione: (2024)
di: Yang, Xinrui, et al.
Pubblicazione: (2024)
PhyPrompt: RL-based Prompt Refinement for Physically Plausible Text-to-Video Generation
di: Wu, Shang, et al.
Pubblicazione: (2026)
di: Wu, Shang, et al.
Pubblicazione: (2026)
PaintScene4D: Consistent 4D Scene Generation from Text Prompts
di: Gupta, Vinayak, et al.
Pubblicazione: (2024)
di: Gupta, Vinayak, et al.
Pubblicazione: (2024)
Review of Hallucination Understanding in Large Language and Vision Models
di: Ho, Zhengyi, et al.
Pubblicazione: (2025)
di: Ho, Zhengyi, et al.
Pubblicazione: (2025)
Minority-Focused Text-to-Image Generation via Prompt Optimization
di: Um, Soobin, et al.
Pubblicazione: (2024)
di: Um, Soobin, et al.
Pubblicazione: (2024)
Reinforcement Learning-Based Prompt Template Stealing for Text-to-Image Models
di: Zou, Xiaotian
Pubblicazione: (2025)
di: Zou, Xiaotian
Pubblicazione: (2025)
Hierarchical Prompt Learning for Image- and Text-Based Person Re-Identification
di: Zhou, Linhan, et al.
Pubblicazione: (2025)
di: Zhou, Linhan, et al.
Pubblicazione: (2025)
PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models
di: Yuan, Lingzhi, et al.
Pubblicazione: (2025)
di: Yuan, Lingzhi, et al.
Pubblicazione: (2025)
Automated Prompt Generation for Creative and Counterfactual Text-to-image Synthesis
di: Jelaca, Aleksa, et al.
Pubblicazione: (2025)
di: Jelaca, Aleksa, et al.
Pubblicazione: (2025)
A User-Friendly Framework for Generating Model-Preferred Prompts in Text-to-Image Synthesis
di: Hei, Nailei, et al.
Pubblicazione: (2024)
di: Hei, Nailei, et al.
Pubblicazione: (2024)
Progressive Image Restoration via Text-Conditioned Video Generation
di: Kang, Peng, et al.
Pubblicazione: (2025)
di: Kang, Peng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis
di: Zhao, Shuxian, et al.
Pubblicazione: (2026) -
Image Captions are Natural Prompts for Text-to-Image Models
di: Lei, Shiye, et al.
Pubblicazione: (2023) -
Dynamic Prompt Optimizing for Text-to-Image Generation
di: Mo, Wenyi, et al.
Pubblicazione: (2024) -
Prompt-Consistency Image Generation (PCIG): A Unified Framework Integrating LLMs, Knowledge Graphs, and Controllable Diffusion Models
di: Sun, Yichen, et al.
Pubblicazione: (2024) -
Self-Supervised Learning for Detecting AI-Generated Faces as Anomalies
di: Zou, Mian, et al.
Pubblicazione: (2025)