Perception-guided Jailbreak against Text-to-Image Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Yihao, Liang, Le, Li, Tianlin, Jia, Xiaojun, Wang, Run, Miao, Weikai, Pu, Geguang, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scale-Invariant Adversarial Attack against Arbitrary-scale Super-resolution
von: Huang, Yihao, et al.
Veröffentlicht: (2025)
von: Huang, Yihao, et al.
Veröffentlicht: (2025)
Text Modality Oriented Image Feature Extraction for Detecting Diffusion-based DeepFake
von: Yang, Di, et al.
Veröffentlicht: (2024)
von: Yang, Di, et al.
Veröffentlicht: (2024)
Improving Robustness of LiDAR-Camera Fusion Model against Weather Corruption from Fusion Strategy Perspective
von: Huang, Yihao, et al.
Veröffentlicht: (2024)
von: Huang, Yihao, et al.
Veröffentlicht: (2024)
Verify Claimed Text-to-Image Models via Boundary-Aware Prompt Optimization
von: Zhao, Zidong, et al.
Veröffentlicht: (2026)
von: Zhao, Zidong, et al.
Veröffentlicht: (2026)
Efficient Universal Goal Hijacking with Semantics-guided Prompt Organization
von: Huang, Yihao, et al.
Veröffentlicht: (2024)
von: Huang, Yihao, et al.
Veröffentlicht: (2024)
Privacy Protection Against Personalized Text-to-Image Synthesis via Cross-image Consistency Constraints
von: Wang, Guanyu, et al.
Veröffentlicht: (2025)
von: Wang, Guanyu, et al.
Veröffentlicht: (2025)
HTS-Attack: Heuristic Token Search for Jailbreaking Text-to-Image Models
von: Gao, Sensen, et al.
Veröffentlicht: (2024)
von: Gao, Sensen, et al.
Veröffentlicht: (2024)
Efficient and Effective Universal Adversarial Attack against Vision-Language Pre-training Models
von: Yang, Fan, et al.
Veröffentlicht: (2024)
von: Yang, Fan, et al.
Veröffentlicht: (2024)
Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models
von: Yang, Fan, et al.
Veröffentlicht: (2025)
von: Yang, Fan, et al.
Veröffentlicht: (2025)
CosalPure: Learning Concept from Group Images for Robust Co-Saliency Detection
von: Zhu, Jiayi, et al.
Veröffentlicht: (2024)
von: Zhu, Jiayi, et al.
Veröffentlicht: (2024)
Texture Re-scalable Universal Adversarial Perturbation
von: Huang, Yihao, et al.
Veröffentlicht: (2024)
von: Huang, Yihao, et al.
Veröffentlicht: (2024)
Beyond Pixels: Semantic-aware Typographic Attack for Geo-Privacy Protection
von: Zhu, Jiayi, et al.
Veröffentlicht: (2025)
von: Zhu, Jiayi, et al.
Veröffentlicht: (2025)
Concept Guided Co-salient Object Detection
von: Zhu, Jiayi, et al.
Veröffentlicht: (2024)
von: Zhu, Jiayi, et al.
Veröffentlicht: (2024)
Dodging DeepFake Detection via Implicit Spatial-Domain Notch Filtering
von: Huang, Yihao, et al.
Veröffentlicht: (2020)
von: Huang, Yihao, et al.
Veröffentlicht: (2020)
OmniSafeBench-MM: A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack-Defense Evaluation
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
When Memory Becomes a Vulnerability: Towards Multi-turn Jailbreak Attacks against Text-to-Image Generation Systems
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
PLA: Prompt Learning Attack against Text-to-Image Generative Models
von: Lyu, Xinqi, et al.
Veröffentlicht: (2025)
von: Lyu, Xinqi, et al.
Veröffentlicht: (2025)
IRAD: Implicit Representation-driven Image Resampling against Adversarial Attacks
von: Cao, Yue, et al.
Veröffentlicht: (2023)
von: Cao, Yue, et al.
Veröffentlicht: (2023)
PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models
von: Yuan, Lingzhi, et al.
Veröffentlicht: (2025)
von: Yuan, Lingzhi, et al.
Veröffentlicht: (2025)
Reading Between the Pixels: An Inscriptive Jailbreak Attack on Text-to-Image Models
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
PATFinger: Prompt-Adapted Transferable Fingerprinting against Unauthorized Multimodal Dataset Usage
von: Zhang, Wenyi, et al.
Veröffentlicht: (2025)
von: Zhang, Wenyi, et al.
Veröffentlicht: (2025)
GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models
von: Hu, Xiangdong, et al.
Veröffentlicht: (2026)
von: Hu, Xiangdong, et al.
Veröffentlicht: (2026)
ALA: Naturalness-aware Adversarial Lightness Attack
von: Huang, Yihao, et al.
Veröffentlicht: (2022)
von: Huang, Yihao, et al.
Veröffentlicht: (2022)
Implicit Bias Injection Attacks against Text-to-Image Diffusion Models
von: Huang, Huayang, et al.
Veröffentlicht: (2025)
von: Huang, Huayang, et al.
Veröffentlicht: (2025)
Universally Unfiltered and Unseen:Input-Agnostic Multimodal Jailbreaks against Text-to-Image Model Safeguards
von: Yan, Song, et al.
Veröffentlicht: (2025)
von: Yan, Song, et al.
Veröffentlicht: (2025)
Text-guided Foundation Model Adaptation for Long-Tailed Medical Image Classification
von: Li, Sirui, et al.
Veröffentlicht: (2024)
von: Li, Sirui, et al.
Veröffentlicht: (2024)
Metaphor-based Jailbreak Attacks on Text-to-Image Models
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
Progressive Compositionality in Text-to-Image Generative Models
von: Han, Evans Xu, et al.
Veröffentlicht: (2024)
von: Han, Evans Xu, et al.
Veröffentlicht: (2024)
Jailbreak Attacks and Defenses against Multimodal Generative Models: A Survey
von: Liu, Xuannan, et al.
Veröffentlicht: (2024)
von: Liu, Xuannan, et al.
Veröffentlicht: (2024)
Exploring Fine-Grained Image-Text Alignment for Referring Remote Sensing Image Segmentation
von: Lei, Sen, et al.
Veröffentlicht: (2024)
von: Lei, Sen, et al.
Veröffentlicht: (2024)
Low-Effort Jailbreak Attacks Against Text-to-Image Safety Filters
von: Mustafa, Ahmed B, et al.
Veröffentlicht: (2026)
von: Mustafa, Ahmed B, et al.
Veröffentlicht: (2026)
MIDAS: Multi-Image Dispersion and Semantic Reconstruction for Jailbreaking MLLMs
von: Liu, Yilian, et al.
Veröffentlicht: (2026)
von: Liu, Yilian, et al.
Veröffentlicht: (2026)
Disciplined Diffusion: Text-to-Image Diffusion Model against NSFW Generation
von: Zhang, Chi, et al.
Veröffentlicht: (2026)
von: Zhang, Chi, et al.
Veröffentlicht: (2026)
Token-Level Constraint Boundary Search for Jailbreaking Text-to-Image Models
von: Liu, Jiangtao, et al.
Veröffentlicht: (2025)
von: Liu, Jiangtao, et al.
Veröffentlicht: (2025)
Evolution-based Region Adversarial Prompt Learning for Robustness Enhancement in Vision-Language Models
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
DreamCom: Finetuning Text-guided Inpainting Model for Image Composition
von: Lu, Lingxiao, et al.
Veröffentlicht: (2023)
von: Lu, Lingxiao, et al.
Veröffentlicht: (2023)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
von: Cao, Pu, et al.
Veröffentlicht: (2024)
von: Cao, Pu, et al.
Veröffentlicht: (2024)
Text-guided Controllable Diffusion for Realistic Camouflage Images Generation
von: Qian, Yuhang, et al.
Veröffentlicht: (2025)
von: Qian, Yuhang, et al.
Veröffentlicht: (2025)
AutoPrompt: Automated Red-Teaming of Text-to-Image Models via LLM-Driven Adversarial Prompts
von: Liu, Yufan, et al.
Veröffentlicht: (2025)
von: Liu, Yufan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Scale-Invariant Adversarial Attack against Arbitrary-scale Super-resolution
von: Huang, Yihao, et al.
Veröffentlicht: (2025) -
Text Modality Oriented Image Feature Extraction for Detecting Diffusion-based DeepFake
von: Yang, Di, et al.
Veröffentlicht: (2024) -
Improving Robustness of LiDAR-Camera Fusion Model against Weather Corruption from Fusion Strategy Perspective
von: Huang, Yihao, et al.
Veröffentlicht: (2024) -
Verify Claimed Text-to-Image Models via Boundary-Aware Prompt Optimization
von: Zhao, Zidong, et al.
Veröffentlicht: (2026) -
Efficient Universal Goal Hijacking with Semantics-guided Prompt Organization
von: Huang, Yihao, et al.
Veröffentlicht: (2024)