Universal Prompt Optimizer for Safe Text-to-Image Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Zongyu, Gao, Hongcheng, Wang, Yueze, Zhang, Xiang, Wang, Suhang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Devil is in the Prompts: Retrieval-Augmented Prompt Optimization for Text-to-Video Generation
por: Gao, Bingjie, et al.
Publicado: (2025)
por: Gao, Bingjie, et al.
Publicado: (2025)
Optimizing Prompts for Text-to-Image Generation
por: Hao, Yaru, et al.
Publicado: (2022)
por: Hao, Yaru, et al.
Publicado: (2022)
Fast Prompt Alignment for Text-to-Image Generation
por: Mrini, Khalil, et al.
Publicado: (2024)
por: Mrini, Khalil, et al.
Publicado: (2024)
TextTIGER: Text-based Intelligent Generation with Entity Prompt Refinement for Text-to-Image Generation
por: Ozaki, Shintaro, et al.
Publicado: (2025)
por: Ozaki, Shintaro, et al.
Publicado: (2025)
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
por: Jing, Zonglei, et al.
Publicado: (2025)
por: Jing, Zonglei, et al.
Publicado: (2025)
Improving Text-to-Image Consistency via Automatic Prompt Optimization
por: Mañas, Oscar, et al.
Publicado: (2024)
por: Mañas, Oscar, et al.
Publicado: (2024)
Image Corruption-Inspired Membership Inference Attacks against Large Vision-Language Models
por: Wu, Zongyu, et al.
Publicado: (2025)
por: Wu, Zongyu, et al.
Publicado: (2025)
MegaPairs: Massive Data Synthesis For Universal Multimodal Retrieval
por: Zhou, Junjie, et al.
Publicado: (2024)
por: Zhou, Junjie, et al.
Publicado: (2024)
VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
por: Cheng, Jiale, et al.
Publicado: (2025)
por: Cheng, Jiale, et al.
Publicado: (2025)
ComfyGen: Prompt-Adaptive Workflows for Text-to-Image Generation
por: Gal, Rinon, et al.
Publicado: (2024)
por: Gal, Rinon, et al.
Publicado: (2024)
Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
por: Chen, Wenting, et al.
Publicado: (2024)
por: Chen, Wenting, et al.
Publicado: (2024)
TextInVision: Text and Prompt Complexity Driven Visual Text Generation Benchmark
por: Fallah, Forouzan, et al.
Publicado: (2025)
por: Fallah, Forouzan, et al.
Publicado: (2025)
Can Prompt Modifiers Control Bias? A Comparative Analysis of Text-to-Image Generative Models
por: Shin, Philip Wootaek, et al.
Publicado: (2024)
por: Shin, Philip Wootaek, et al.
Publicado: (2024)
Prompting4Debugging: Red-Teaming Text-to-Image Diffusion Models by Finding Problematic Prompts
por: Chin, Zhi-Yi, et al.
Publicado: (2023)
por: Chin, Zhi-Yi, et al.
Publicado: (2023)
NSFW-Classifier Guided Prompt Sanitization for Safe Text-to-Image Generation
por: Xie, Yu, et al.
Publicado: (2025)
por: Xie, Yu, et al.
Publicado: (2025)
LADR: Locality-Aware Dynamic Rescue for Efficient Text-to-Image Generation with Diffusion Large Language Models
por: Wang, Chenglin, et al.
Publicado: (2026)
por: Wang, Chenglin, et al.
Publicado: (2026)
Visual-RAG: Benchmarking Text-to-Image Retrieval Augmented Generation for Visual Knowledge Intensive Queries
por: Wu, Yin, et al.
Publicado: (2025)
por: Wu, Yin, et al.
Publicado: (2025)
AnchorOPT: Towards Optimizing Dynamic Anchors for Adaptive Prompt Learning
por: Li, Zheng, et al.
Publicado: (2025)
por: Li, Zheng, et al.
Publicado: (2025)
Emu: Generative Pretraining in Multimodality
por: Sun, Quan, et al.
Publicado: (2023)
por: Sun, Quan, et al.
Publicado: (2023)
Training-Free Generation of Diverse and High-Fidelity Images via Prompt Semantic Space Optimization
por: Meng, Debin, et al.
Publicado: (2025)
por: Meng, Debin, et al.
Publicado: (2025)
Efficient Universal Goal Hijacking with Semantics-guided Prompt Organization
por: Huang, Yihao, et al.
Publicado: (2024)
por: Huang, Yihao, et al.
Publicado: (2024)
SafeCFG: Controlling Harmful Features with Dynamic Safe Guidance for Safe Generation
por: Pan, Jiadong, et al.
Publicado: (2024)
por: Pan, Jiadong, et al.
Publicado: (2024)
Image2Text2Image: A Novel Framework for Label-Free Evaluation of Image-to-Text Generation with Text-to-Image Diffusion Models
por: Huang, Jia-Hong, et al.
Publicado: (2024)
por: Huang, Jia-Hong, et al.
Publicado: (2024)
Image Over Text: Transforming Formula Recognition Evaluation with Character Detection Matching
por: Wang, Bin, et al.
Publicado: (2024)
por: Wang, Bin, et al.
Publicado: (2024)
PromptTA: Prompt-driven Text Adapter for Source-free Domain Generalization
por: Zhang, Haoran, et al.
Publicado: (2024)
por: Zhang, Haoran, et al.
Publicado: (2024)
Interleaved Scene Graphs for Interleaved Text-and-Image Generation Assessment
por: Chen, Dongping, et al.
Publicado: (2024)
por: Chen, Dongping, et al.
Publicado: (2024)
MM-SAP: A Comprehensive Benchmark for Assessing Self-Awareness of Multimodal Large Language Models in Perception
por: Wang, Yuhao, et al.
Publicado: (2024)
por: Wang, Yuhao, et al.
Publicado: (2024)
Beyond Filtering: Adaptive Image-Text Quality Enhancement for MLLM Pretraining
por: Huang, Han, et al.
Publicado: (2024)
por: Huang, Han, et al.
Publicado: (2024)
LanP: Rethinking the Impact of Language Priors in Large Vision-Language Models
por: Wu, Zongyu, et al.
Publicado: (2025)
por: Wu, Zongyu, et al.
Publicado: (2025)
VidProM: A Million-scale Real Prompt-Gallery Dataset for Text-to-Video Diffusion Models
por: Wang, Wenhao, et al.
Publicado: (2024)
por: Wang, Wenhao, et al.
Publicado: (2024)
DiffChat: Learning to Chat with Text-to-Image Synthesis Models for Interactive Image Creation
por: Wang, Jiapeng, et al.
Publicado: (2024)
por: Wang, Jiapeng, et al.
Publicado: (2024)
SegHist: A General Segmentation-based Framework for Chinese Historical Document Text Line Detection
por: Hu, Xingjian, et al.
Publicado: (2024)
por: Hu, Xingjian, et al.
Publicado: (2024)
UniVS: Unified and Universal Video Segmentation with Prompts as Queries
por: Li, Minghan, et al.
Publicado: (2024)
por: Li, Minghan, et al.
Publicado: (2024)
MM-Interleaved: Interleaved Image-Text Generative Modeling via Multi-modal Feature Synchronizer
por: Tian, Changyao, et al.
Publicado: (2024)
por: Tian, Changyao, et al.
Publicado: (2024)
Text Prompt Injection of Vision Language Models
por: Zhu, Ruizhe
Publicado: (2025)
por: Zhu, Ruizhe
Publicado: (2025)
UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs
por: Jiang, Houcheng, et al.
Publicado: (2026)
por: Jiang, Houcheng, et al.
Publicado: (2026)
Evaluating the Evaluators: Metrics for Compositional Text-to-Image Generation
por: Kasaei, Seyed Amir, et al.
Publicado: (2025)
por: Kasaei, Seyed Amir, et al.
Publicado: (2025)
Pseudo-Prompt Generating in Pre-trained Vision-Language Models for Multi-Label Medical Image Classification
por: Ye, Yaoqin, et al.
Publicado: (2024)
por: Ye, Yaoqin, et al.
Publicado: (2024)
Text-only Synthesis for Image Captioning
por: Zhou, Qing, et al.
Publicado: (2024)
por: Zhou, Qing, et al.
Publicado: (2024)
Drawing the Line: Enhancing Trustworthiness of MLLMs Through the Power of Refusal
por: Wang, Yuhao, et al.
Publicado: (2024)
por: Wang, Yuhao, et al.
Publicado: (2024)
Ejemplares similares
-
The Devil is in the Prompts: Retrieval-Augmented Prompt Optimization for Text-to-Video Generation
por: Gao, Bingjie, et al.
Publicado: (2025) -
Optimizing Prompts for Text-to-Image Generation
por: Hao, Yaru, et al.
Publicado: (2022) -
Fast Prompt Alignment for Text-to-Image Generation
por: Mrini, Khalil, et al.
Publicado: (2024) -
TextTIGER: Text-based Intelligent Generation with Entity Prompt Refinement for Text-to-Image Generation
por: Ozaki, Shintaro, et al.
Publicado: (2025) -
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
por: Jing, Zonglei, et al.
Publicado: (2025)