Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Xingkai, Jiang, Jun, Tong, Meng, Li, Shuai, Zhang, Weiming, Yu, Nenghai, Chen, Kejiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
T2SMark: Balancing Robustness and Diversity in Noise-as-Watermark for Diffusion Models
von: Yang, Jindong, et al.
Veröffentlicht: (2025)
von: Yang, Jindong, et al.
Veröffentlicht: (2025)
Clean Image May be Dangerous: Data Poisoning Attacks Against Deep Hashing
von: Li, Shuai, et al.
Veröffentlicht: (2025)
von: Li, Shuai, et al.
Veröffentlicht: (2025)
WMVLM: Evaluating Diffusion Model Image Watermarking via Vision-Language Models
von: Yang, Zijin, et al.
Veröffentlicht: (2026)
von: Yang, Zijin, et al.
Veröffentlicht: (2026)
M3-AGIQA: Multimodal, Multi-Round, Multi-Aspect AI-Generated Image Quality Assessment
von: Cui, Chuan, et al.
Veröffentlicht: (2025)
von: Cui, Chuan, et al.
Veröffentlicht: (2025)
©Plug-in Authorization for Human Content Copyright Protection in Text-to-Image Model
von: Zhou, Chao, et al.
Veröffentlicht: (2024)
von: Zhou, Chao, et al.
Veröffentlicht: (2024)
Gaussian Shading: Provable Performance-Lossless Image Watermarking for Diffusion Models
von: Yang, Zijin, et al.
Veröffentlicht: (2024)
von: Yang, Zijin, et al.
Veröffentlicht: (2024)
SafeGuider: Robust and Practical Content Safety Control for Text-to-Image Models
von: Qi, Peigui, et al.
Veröffentlicht: (2025)
von: Qi, Peigui, et al.
Veröffentlicht: (2025)
Natias: Neuron Attribution based Transferable Image Adversarial Steganography
von: Fan, Zexin, et al.
Veröffentlicht: (2024)
von: Fan, Zexin, et al.
Veröffentlicht: (2024)
SemBind: Binding Diffusion Watermarks to Semantics Against Black-Box Forgery Attacks
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
Prompt Decoupling for Text-to-Image Person Re-identification
von: Li, Weihao, et al.
Veröffentlicht: (2024)
von: Li, Weihao, et al.
Veröffentlicht: (2024)
T2I-RiskyPrompt: A Benchmark for Safety Evaluation, Attack, and Defense on Text-to-Image Model
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
LiteUpdate: A Lightweight Framework for Updating AI-Generated Image Detectors
von: Lu, Jiajie, et al.
Veröffentlicht: (2025)
von: Lu, Jiajie, et al.
Veröffentlicht: (2025)
Gaussian Shading++: Rethinking the Realistic Deployment Challenge of Performance-Lossless Image Watermark for Diffusion Models
von: Yang, Zijin, et al.
Veröffentlicht: (2025)
von: Yang, Zijin, et al.
Veröffentlicht: (2025)
Adversarial Prompt Injection Attack on Multimodal Large Language Models
von: Ding, Meiwen, et al.
Veröffentlicht: (2026)
von: Ding, Meiwen, et al.
Veröffentlicht: (2026)
Unified Prompt Attack Against Text-to-Image Generation Models
von: Peng, Duo, et al.
Veröffentlicht: (2025)
von: Peng, Duo, et al.
Veröffentlicht: (2025)
SWIFT: Sliding Window Reconstruction for Few-Shot Training-Free Generated Video Attribution
von: Wang, Chao, et al.
Veröffentlicht: (2026)
von: Wang, Chao, et al.
Veröffentlicht: (2026)
UPAM: Unified Prompt Attack in Text-to-Image Generation Models Against Both Textual Filters and Visual Checkers
von: Peng, Duo, et al.
Veröffentlicht: (2024)
von: Peng, Duo, et al.
Veröffentlicht: (2024)
SurrogatePrompt: Bypassing the Safety Filter of Text-to-Image Models via Substitution
von: Ba, Zhongjie, et al.
Veröffentlicht: (2023)
von: Ba, Zhongjie, et al.
Veröffentlicht: (2023)
MASTER: Multimodal Segmentation with Text Prompts
von: Liu, Fuyang, et al.
Veröffentlicht: (2025)
von: Liu, Fuyang, et al.
Veröffentlicht: (2025)
Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation
von: Wu, Chengyue, et al.
Veröffentlicht: (2024)
von: Wu, Chengyue, et al.
Veröffentlicht: (2024)
Low-Effort Jailbreak Attacks Against Text-to-Image Safety Filters
von: Mustafa, Ahmed B, et al.
Veröffentlicht: (2026)
von: Mustafa, Ahmed B, et al.
Veröffentlicht: (2026)
PLA: Prompt Learning Attack against Text-to-Image Generative Models
von: Lyu, Xinqi, et al.
Veröffentlicht: (2025)
von: Lyu, Xinqi, et al.
Veröffentlicht: (2025)
Long-Text-to-Image Generation via Compositional Prompt Decomposition
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
Personalized Safety Alignment for Text-to-Image Diffusion Models
von: Lei, Yu, et al.
Veröffentlicht: (2025)
von: Lei, Yu, et al.
Veröffentlicht: (2025)
Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning
von: Liu, Jinlong, et al.
Veröffentlicht: (2026)
von: Liu, Jinlong, et al.
Veröffentlicht: (2026)
ConceptGuard: Proactive Safety in Text-and-Image-to-Video Generation through Multimodal Risk Detection
von: Ma, Ruize, et al.
Veröffentlicht: (2025)
von: Ma, Ruize, et al.
Veröffentlicht: (2025)
Prompt-Based Safety Guidance Is Ineffective for Unlearned Text-to-Image Diffusion Models
von: Shin, Jiwoo, et al.
Veröffentlicht: (2025)
von: Shin, Jiwoo, et al.
Veröffentlicht: (2025)
Implicit Bias Injection Attacks against Text-to-Image Diffusion Models
von: Huang, Huayang, et al.
Veröffentlicht: (2025)
von: Huang, Huayang, et al.
Veröffentlicht: (2025)
Ask, Attend, Attack: A Effective Decision-Based Black-Box Targeted Attack for Image-to-Text Models
von: Zeng, Qingyuan, et al.
Veröffentlicht: (2024)
von: Zeng, Qingyuan, et al.
Veröffentlicht: (2024)
SAPL: Semantic-Agnostic Prompt Learning in CLIP for Weakly Supervised Image Manipulation Localization
von: Wang, Xinghao, et al.
Veröffentlicht: (2026)
von: Wang, Xinghao, et al.
Veröffentlicht: (2026)
Position: Towards Implicit Prompt For Text-To-Image Models
von: Yang, Yue, et al.
Veröffentlicht: (2024)
von: Yang, Yue, et al.
Veröffentlicht: (2024)
2AFC Prompting of Large Multimodal Models for Image Quality Assessment
von: Zhu, Hanwei, et al.
Veröffentlicht: (2024)
von: Zhu, Hanwei, et al.
Veröffentlicht: (2024)
Safety of Multimodal Large Language Models on Images and Texts
von: Liu, Xin, et al.
Veröffentlicht: (2024)
von: Liu, Xin, et al.
Veröffentlicht: (2024)
Image Captions are Natural Prompts for Text-to-Image Models
von: Lei, Shiye, et al.
Veröffentlicht: (2023)
von: Lei, Shiye, et al.
Veröffentlicht: (2023)
Memory-Inspired Temporal Prompt Interaction for Text-Image Classification
von: Yu, Xinyao, et al.
Veröffentlicht: (2024)
von: Yu, Xinyao, et al.
Veröffentlicht: (2024)
Runge-Kutta Approximation and Decoupled Attention for Rectified Flow Inversion and Semantic Editing
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
MedIQA: A Scalable Foundation Model for Prompt-Driven Medical Image Quality Assessment
von: Xun, Siyi, et al.
Veröffentlicht: (2025)
von: Xun, Siyi, et al.
Veröffentlicht: (2025)
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
von: Shuai, Xincheng, et al.
Veröffentlicht: (2024)
von: Shuai, Xincheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
T2SMark: Balancing Robustness and Diversity in Noise-as-Watermark for Diffusion Models
von: Yang, Jindong, et al.
Veröffentlicht: (2025) -
Clean Image May be Dangerous: Data Poisoning Attacks Against Deep Hashing
von: Li, Shuai, et al.
Veröffentlicht: (2025) -
WMVLM: Evaluating Diffusion Model Image Watermarking via Vision-Language Models
von: Yang, Zijin, et al.
Veröffentlicht: (2026) -
M3-AGIQA: Multimodal, Multi-Round, Multi-Aspect AI-Generated Image Quality Assessment
von: Cui, Chuan, et al.
Veröffentlicht: (2025) -
©Plug-in Authorization for Human Content Copyright Protection in Text-to-Image Model
von: Zhou, Chao, et al.
Veröffentlicht: (2024)