Mudjacking: Patching Backdoor Vulnerabilities in Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Hongbin, Reiter, Michael K., Gong, Neil Zhenqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CorruptEncoder: Data Poisoning based Backdoor Attacks to Contrastive Learning
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2022)
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2022)
Robustness of Vision Foundation Models to Common Perturbations
von: Liu, Hongbin, et al.
Veröffentlicht: (2026)
von: Liu, Hongbin, et al.
Veröffentlicht: (2026)
Refusing Safe Prompts for Multi-modal Large Language Models
von: Shao, Zedian, et al.
Veröffentlicht: (2024)
von: Shao, Zedian, et al.
Veröffentlicht: (2024)
Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection
von: Shao, Zedian, et al.
Veröffentlicht: (2026)
von: Shao, Zedian, et al.
Veröffentlicht: (2026)
Unveiling and Mitigating Backdoor Vulnerabilities based on Unlearning Weight Changes and Backdoor Activeness
von: Lin, Weilin, et al.
Veröffentlicht: (2024)
von: Lin, Weilin, et al.
Veröffentlicht: (2024)
Certifiably Robust Image Watermark
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2024)
Securing Visually-Aware Recommender Systems: An Adversarial Image Reconstruction and Detection Framework
von: Yin, Minglei, et al.
Veröffentlicht: (2023)
von: Yin, Minglei, et al.
Veröffentlicht: (2023)
Tracing Back the Malicious Clients in Poisoning Attacks to Federated Learning
von: Jia, Yuqi, et al.
Veröffentlicht: (2024)
von: Jia, Yuqi, et al.
Veröffentlicht: (2024)
EditTrack: Detecting and Attributing AI-assisted Image Editing
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
Revealing Vulnerabilities of Neural Networks in Parameter Learning and Defense Against Explanation-Aware Backdoors
von: Kadir, Md Abdul, et al.
Veröffentlicht: (2024)
von: Kadir, Md Abdul, et al.
Veröffentlicht: (2024)
Backdoor Attack with Sparse and Invisible Trigger
von: Gao, Yinghua, et al.
Veröffentlicht: (2023)
von: Gao, Yinghua, et al.
Veröffentlicht: (2023)
SafeText: Safe Text-to-image Models via Aligning the Text Encoder
von: Hu, Yuepeng, et al.
Veröffentlicht: (2025)
von: Hu, Yuepeng, et al.
Veröffentlicht: (2025)
How to Backdoor Consistency Models?
von: Wang, Chengen, et al.
Veröffentlicht: (2024)
von: Wang, Chengen, et al.
Veröffentlicht: (2024)
Invisible Backdoor Attacks on Diffusion Models
von: Li, Sen, et al.
Veröffentlicht: (2024)
von: Li, Sen, et al.
Veröffentlicht: (2024)
Under-confidence Backdoors Are Resilient and Stealthy Backdoors
von: Peng, Minlong, et al.
Veröffentlicht: (2022)
von: Peng, Minlong, et al.
Veröffentlicht: (2022)
Backdoor Federated Learning by Poisoning Backdoor-Critical Layers
von: Zhuang, Haomin, et al.
Veröffentlicht: (2023)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2023)
On the Multi-modal Vulnerability of Diffusion Models
von: Yang, Dingcheng, et al.
Veröffentlicht: (2024)
von: Yang, Dingcheng, et al.
Veröffentlicht: (2024)
Model Supply Chain Poisoning: Backdooring Pre-trained Models via Embedding Indistinguishability
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
Beating Backdoor Attack at Its Own Game
von: Liu, Min, et al.
Veröffentlicht: (2023)
von: Liu, Min, et al.
Veröffentlicht: (2023)
Universal Backdoor Attacks
von: Schneider, Benjamin, et al.
Veröffentlicht: (2023)
von: Schneider, Benjamin, et al.
Veröffentlicht: (2023)
DisDet: Exploring Detectability of Backdoor Attack on Diffusion Models
von: Sui, Yang, et al.
Veröffentlicht: (2024)
von: Sui, Yang, et al.
Veröffentlicht: (2024)
A General Framework for Data-Use Auditing of ML Models
von: Huang, Zonghao, et al.
Veröffentlicht: (2024)
von: Huang, Zonghao, et al.
Veröffentlicht: (2024)
Instance-Level Data-Use Auditing of Visual ML Models
von: Huang, Zonghao, et al.
Veröffentlicht: (2025)
von: Huang, Zonghao, et al.
Veröffentlicht: (2025)
Unsupervised Backdoor Detection and Mitigation for Spiking Neural Networks
von: Li, Jiachen, et al.
Veröffentlicht: (2025)
von: Li, Jiachen, et al.
Veröffentlicht: (2025)
Watermark-based Attribution of AI-Generated Content
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2024)
VillanDiffusion: A Unified Backdoor Attack Framework for Diffusion Models
von: Chou, Sheng-Yen, et al.
Veröffentlicht: (2023)
von: Chou, Sheng-Yen, et al.
Veröffentlicht: (2023)
Memory Backdoor Attacks on Neural Networks
von: Luzon, Eden, et al.
Veröffentlicht: (2024)
von: Luzon, Eden, et al.
Veröffentlicht: (2024)
IU: Imperceptible Universal Backdoor Attack
von: Lin, Hsin, et al.
Veröffentlicht: (2026)
von: Lin, Hsin, et al.
Veröffentlicht: (2026)
UFID: A Unified Framework for Input-level Backdoor Detection on Diffusion Models
von: Guan, Zihan, et al.
Veröffentlicht: (2024)
von: Guan, Zihan, et al.
Veröffentlicht: (2024)
BB-Patch: BlackBox Adversarial Patch-Attack using Zeroth-Order Optimization
von: Kumar, Satyadwyoom, et al.
Veröffentlicht: (2024)
von: Kumar, Satyadwyoom, et al.
Veröffentlicht: (2024)
PatchDEMUX: A Certifiably Robust Framework for Multi-label Classifiers Against Adversarial Patches
von: Jacob, Dennis, et al.
Veröffentlicht: (2025)
von: Jacob, Dennis, et al.
Veröffentlicht: (2025)
ImageNet-Patch: A Dataset for Benchmarking Machine Learning Robustness against Adversarial Patches
von: Pintor, Maura, et al.
Veröffentlicht: (2022)
von: Pintor, Maura, et al.
Veröffentlicht: (2022)
Goal-oriented Backdoor Attack against Vision-Language-Action Models via Physical Objects
von: Zhou, Zirun, et al.
Veröffentlicht: (2025)
von: Zhou, Zirun, et al.
Veröffentlicht: (2025)
VideoMarkBench: Benchmarking Robustness of Video Watermarking
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
Benchmarking Adversarial Patch Selection and Location
von: Kimhi, Shai, et al.
Veröffentlicht: (2025)
von: Kimhi, Shai, et al.
Veröffentlicht: (2025)
When Think-with-Image Meets Safety: What Determines Multimodal Jailbreak Robustness?
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
Generating Potent Poisons and Backdoors from Scratch with Guided Diffusion
von: Souri, Hossein, et al.
Veröffentlicht: (2024)
von: Souri, Hossein, et al.
Veröffentlicht: (2024)
Identifying Physically Realizable Triggers for Backdoored Face Recognition Networks
von: Raj, Ankita, et al.
Veröffentlicht: (2025)
von: Raj, Ankita, et al.
Veröffentlicht: (2025)
Deferred Poisoning: Making the Model More Vulnerable via Hessian Singularization
von: He, Yuhao, et al.
Veröffentlicht: (2024)
von: He, Yuhao, et al.
Veröffentlicht: (2024)
Wicked Oddities: Selectively Poisoning for Effective Clean-Label Backdoor Attacks
von: Nguyen, Quang H., et al.
Veröffentlicht: (2024)
von: Nguyen, Quang H., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CorruptEncoder: Data Poisoning based Backdoor Attacks to Contrastive Learning
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2022) -
Robustness of Vision Foundation Models to Common Perturbations
von: Liu, Hongbin, et al.
Veröffentlicht: (2026) -
Refusing Safe Prompts for Multi-modal Large Language Models
von: Shao, Zedian, et al.
Veröffentlicht: (2024) -
Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection
von: Shao, Zedian, et al.
Veröffentlicht: (2026) -
Unveiling and Mitigating Backdoor Vulnerabilities based on Unlearning Weight Changes and Backdoor Activeness
von: Lin, Weilin, et al.
Veröffentlicht: (2024)