ToxicTextCLIP: Text-Based Poisoning and Backdoor Attacks on CLIP Pre-training
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Xin, Zhao, Haiyang, Chen, Yimin, Guo, Jiawei, Huang, Kecheng, Zhao, Ming |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
by: Yang, Wenhan, et al.
Published: (2023)
by: Yang, Wenhan, et al.
Published: (2023)
Backdooring CLIP through Concept Confusion
by: Hu, Lijie, et al.
Published: (2025)
by: Hu, Lijie, et al.
Published: (2025)
REDEditing: Relationship-Driven Precise Backdoor Poisoning on Text-to-Image Diffusion Models
by: Guo, Chongye, et al.
Published: (2025)
by: Guo, Chongye, et al.
Published: (2025)
Poisoning-based Backdoor Attacks for Arbitrary Target Label with Positive Triggers
by: Huang, Binxiao, et al.
Published: (2024)
by: Huang, Binxiao, et al.
Published: (2024)
Model Supply Chain Poisoning: Backdooring Pre-trained Models via Embedding Indistinguishability
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
A Proxy Attack-Free Strategy for Practically Improving the Poisoning Efficiency in Backdoor Attacks
by: Li, Ziqiang, et al.
Published: (2023)
by: Li, Ziqiang, et al.
Published: (2023)
RPP: A Certified Poisoned-Sample Detection Framework for Backdoor Attacks under Dataset Imbalance
by: Lin, Miao, et al.
Published: (2026)
by: Lin, Miao, et al.
Published: (2026)
Perturb and Recover: Fine-tuning for Effective Backdoor Removal from CLIP
by: Singh, Naman Deep, et al.
Published: (2024)
by: Singh, Naman Deep, et al.
Published: (2024)
BadBlocks: Low-Cost and Stealthy Backdoor Attacks Tailored for Text-to-Image Diffusion Models
by: Wu, Jia, et al.
Published: (2025)
by: Wu, Jia, et al.
Published: (2025)
CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion
by: Jindal, Akshit, et al.
Published: (2026)
by: Jindal, Akshit, et al.
Published: (2026)
IPA-NeRF: Illusory Poisoning Attack Against Neural Radiance Fields
by: Jiang, Wenxiang, et al.
Published: (2024)
by: Jiang, Wenxiang, et al.
Published: (2024)
Silent Branding Attack: Trigger-free Data Poisoning Attack on Text-to-Image Diffusion Models
by: Jang, Sangwon, et al.
Published: (2025)
by: Jang, Sangwon, et al.
Published: (2025)
CIS-BA: Continuous Interaction Space Based Backdoor Attack for Object Detection in the Real-World
by: Zhao, Shuxin, et al.
Published: (2025)
by: Zhao, Shuxin, et al.
Published: (2025)
Wicked Oddities: Selectively Poisoning for Effective Clean-Label Backdoor Attacks
by: Nguyen, Quang H., et al.
Published: (2024)
by: Nguyen, Quang H., et al.
Published: (2024)
CorruptEncoder: Data Poisoning based Backdoor Attacks to Contrastive Learning
by: Zhang, Jinghuai, et al.
Published: (2022)
by: Zhang, Jinghuai, et al.
Published: (2022)
IPBA: Imperceptible Perturbation Backdoor Attack in Federated Self-Supervised Learning
by: Wang, Jiayao, et al.
Published: (2025)
by: Wang, Jiayao, et al.
Published: (2025)
CLIP-Flow: A Universal Discriminator for AI-Generated Images Inspired by Anomaly Detection
by: Yuan, Zhipeng, et al.
Published: (2025)
by: Yuan, Zhipeng, et al.
Published: (2025)
Checkerboard: A Simple, Effective, Efficient and Learning-free Clean Label Backdoor Attack with Low Poisoning Budget
by: Yang, Yi, et al.
Published: (2026)
by: Yang, Yi, et al.
Published: (2026)
Not All Prompts Are Secure: A Switchable Backdoor Attack Against Pre-trained Vision Transformers
by: Yang, Sheng, et al.
Published: (2024)
by: Yang, Sheng, et al.
Published: (2024)
HTS-Attack: Heuristic Token Search for Jailbreaking Text-to-Image Models
by: Gao, Sensen, et al.
Published: (2024)
by: Gao, Sensen, et al.
Published: (2024)
Backdoor Federated Learning by Poisoning Backdoor-Critical Layers
by: Zhuang, Haomin, et al.
Published: (2023)
by: Zhuang, Haomin, et al.
Published: (2023)
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
by: Wei, Shaokui, et al.
Published: (2024)
by: Wei, Shaokui, et al.
Published: (2024)
Event Trojan: Asynchronous Event-based Backdoor Attacks
by: Wang, Ruofei, et al.
Published: (2024)
by: Wang, Ruofei, et al.
Published: (2024)
SATBA: An Invisible Backdoor Attack Based On Spatial Attention
by: Zhou, Huasong, et al.
Published: (2023)
by: Zhou, Huasong, et al.
Published: (2023)
CAPAA: Classifier-Agnostic Projector-Based Adversarial Attack
by: Li, Zhan, et al.
Published: (2025)
by: Li, Zhan, et al.
Published: (2025)
DeeCLIP: A Robust and Generalizable Transformer-Based Framework for Detecting AI-Generated Images
by: Keita, Mamadou, et al.
Published: (2025)
by: Keita, Mamadou, et al.
Published: (2025)
Clean-image Backdoor Attacks
by: Rong, Dazhong, et al.
Published: (2024)
by: Rong, Dazhong, et al.
Published: (2024)
VidLeaks: Membership Inference Attacks Against Text-to-Video Models
by: Wang, Li, et al.
Published: (2026)
by: Wang, Li, et al.
Published: (2026)
Multi-Target Federated Backdoor Attack Based on Feature Aggregation
by: Hao, Lingguag, et al.
Published: (2025)
by: Hao, Lingguag, et al.
Published: (2025)
Does CLIP Know My Face?
by: Hintersdorf, Dominik, et al.
Published: (2022)
by: Hintersdorf, Dominik, et al.
Published: (2022)
FedPoisonTTP: A Threat Model and Poisoning Attack for Federated Test-Time Personalization
by: Iftee, Md Akil Raihan, et al.
Published: (2025)
by: Iftee, Md Akil Raihan, et al.
Published: (2025)
Backdoor Attacks on Prompt-Driven Video Segmentation Foundation Models
by: Zhang, Zongmin, et al.
Published: (2025)
by: Zhang, Zongmin, et al.
Published: (2025)
Exposing Functional Fusion: A New Class of Strategic Backdoor in Dynamic Prompt Architectures
by: Liu, Zeyao, et al.
Published: (2026)
by: Liu, Zeyao, et al.
Published: (2026)
Backdoor Attack with Mode Mixture Latent Modification
by: Zhang, Hongwei, et al.
Published: (2024)
by: Zhang, Hongwei, et al.
Published: (2024)
INK: Inheritable Natural Backdoor Attack Against Model Distillation
by: Liu, Xiaolei, et al.
Published: (2023)
by: Liu, Xiaolei, et al.
Published: (2023)
One Perturbation is Enough: On Generating Universal Adversarial Perturbations against Vision-Language Pre-training Models
by: Fang, Hao, et al.
Published: (2024)
by: Fang, Hao, et al.
Published: (2024)
VLATTACK: Multimodal Adversarial Attacks on Vision-Language Tasks via Pre-trained Models
by: Yin, Ziyi, et al.
Published: (2023)
by: Yin, Ziyi, et al.
Published: (2023)
Data Free Backdoor Attacks
by: Cao, Bochuan, et al.
Published: (2024)
by: Cao, Bochuan, et al.
Published: (2024)
Tracing Back the Malicious Clients in Poisoning Attacks to Federated Learning
by: Jia, Yuqi, et al.
Published: (2024)
by: Jia, Yuqi, et al.
Published: (2024)
BadDet+: Robust Backdoor Attacks for Object Detection
by: Dunnett, Kealan, et al.
Published: (2026)
by: Dunnett, Kealan, et al.
Published: (2026)
Similar Items
-
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
by: Yang, Wenhan, et al.
Published: (2023) -
Backdooring CLIP through Concept Confusion
by: Hu, Lijie, et al.
Published: (2025) -
REDEditing: Relationship-Driven Precise Backdoor Poisoning on Text-to-Image Diffusion Models
by: Guo, Chongye, et al.
Published: (2025) -
Poisoning-based Backdoor Attacks for Arbitrary Target Label with Positive Triggers
by: Huang, Binxiao, et al.
Published: (2024) -
Model Supply Chain Poisoning: Backdooring Pre-trained Models via Embedding Indistinguishability
by: Wang, Hao, et al.
Published: (2024)