Rethinking Pruning for Backdoor Mitigation: An Optimization Perspective
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Nan, Yu, Haiyang, Yi, Ping |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Magnitude-based Neuron Pruning for Backdoor Defens
di: Li, Nan, et al.
Pubblicazione: (2024)
di: Li, Nan, et al.
Pubblicazione: (2024)
OCGEC: One-class Graph Embedding Classification for DNN Backdoor Detection
di: Jiang, Haoyu, et al.
Pubblicazione: (2023)
di: Jiang, Haoyu, et al.
Pubblicazione: (2023)
Uncovering, Explaining, and Mitigating the Superficial Safety of Backdoor Defense
di: Min, Rui, et al.
Pubblicazione: (2024)
di: Min, Rui, et al.
Pubblicazione: (2024)
Mutual Information Guided Backdoor Mitigation for Pre-trained Encoders
di: Han, Tingxu, et al.
Pubblicazione: (2024)
di: Han, Tingxu, et al.
Pubblicazione: (2024)
ICLShield: Exploring and Mitigating In-Context Learning Backdoor Attacks
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
Mitigating Deep Reinforcement Learning Backdoors in the Neural Activation Space
di: Vyas, Sanyam, et al.
Pubblicazione: (2024)
di: Vyas, Sanyam, et al.
Pubblicazione: (2024)
Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs
di: Wang, Yifei, et al.
Pubblicazione: (2026)
di: Wang, Yifei, et al.
Pubblicazione: (2026)
Backdoor Secrets Unveiled: Identifying Backdoor Data with Optimized Scaled Prediction Consistency
di: Pal, Soumyadeep, et al.
Pubblicazione: (2024)
di: Pal, Soumyadeep, et al.
Pubblicazione: (2024)
BackdoorMBTI: A Backdoor Learning Multimodal Benchmark Tool Kit for Backdoor Defense Evaluation
di: Yu, Haiyang, et al.
Pubblicazione: (2024)
di: Yu, Haiyang, et al.
Pubblicazione: (2024)
PSBD: Prediction Shift Uncertainty Unlocks Backdoor Detection
di: Li, Wei, et al.
Pubblicazione: (2024)
di: Li, Wei, et al.
Pubblicazione: (2024)
Backdoor Graph Condensation
di: Wu, Jiahao, et al.
Pubblicazione: (2024)
di: Wu, Jiahao, et al.
Pubblicazione: (2024)
Fusing Pruned and Backdoored Models: Optimal Transport-based Data-free Backdoor Mitigation
di: Lin, Weilin, et al.
Pubblicazione: (2024)
di: Lin, Weilin, et al.
Pubblicazione: (2024)
Heterogeneous Graph Backdoor Attack
di: Chen, Jiawei, et al.
Pubblicazione: (2025)
di: Chen, Jiawei, et al.
Pubblicazione: (2025)
Concealing Backdoor Model Updates in Federated Learning by Trigger-Optimized Data Poisoning
di: Zhang, Yujie, et al.
Pubblicazione: (2024)
di: Zhang, Yujie, et al.
Pubblicazione: (2024)
Backdoor Vectors: a Task Arithmetic View on Backdoor Attacks and Defenses
di: Pawlak, Stanisław, et al.
Pubblicazione: (2025)
di: Pawlak, Stanisław, et al.
Pubblicazione: (2025)
Elijah: Eliminating Backdoors Injected in Diffusion Models via Distribution Shift
di: An, Shengwei, et al.
Pubblicazione: (2023)
di: An, Shengwei, et al.
Pubblicazione: (2023)
Unlearn to Relearn Backdoors: Deferred Backdoor Functionality Attacks on Deep Learning Models
di: Shin, Jeongjin, et al.
Pubblicazione: (2024)
di: Shin, Jeongjin, et al.
Pubblicazione: (2024)
Graph Neural Backdoor: Fundamentals, Methodologies, Applications, and Future Directions
di: Yang, Xiao, et al.
Pubblicazione: (2024)
di: Yang, Xiao, et al.
Pubblicazione: (2024)
A Cryptographic Perspective on Mitigation vs. Detection in Machine Learning
di: Gluch, Greg, et al.
Pubblicazione: (2025)
di: Gluch, Greg, et al.
Pubblicazione: (2025)
Backdoor defense, learnability and obfuscation
di: Christiano, Paul, et al.
Pubblicazione: (2024)
di: Christiano, Paul, et al.
Pubblicazione: (2024)
How to Backdoor the Knowledge Distillation
di: Wu, Chen, et al.
Pubblicazione: (2025)
di: Wu, Chen, et al.
Pubblicazione: (2025)
BadImplant: Injection-based Multi-Targeted Graph Backdoor Attack
di: Khan, Md Nabi Newaz, et al.
Pubblicazione: (2026)
di: Khan, Md Nabi Newaz, et al.
Pubblicazione: (2026)
Compromising Embodied Agents with Contextual Backdoor Attacks
di: Liu, Aishan, et al.
Pubblicazione: (2024)
di: Liu, Aishan, et al.
Pubblicazione: (2024)
Fast and Lightweight Backdoor Detection via Head Random Probing
di: Yu, Yinbo, et al.
Pubblicazione: (2026)
di: Yu, Yinbo, et al.
Pubblicazione: (2026)
TERD: A Unified Framework for Safeguarding Diffusion Models Against Backdoors
di: Mo, Yichuan, et al.
Pubblicazione: (2024)
di: Mo, Yichuan, et al.
Pubblicazione: (2024)
Your Agent Can Defend Itself against Backdoor Attacks
di: Changjiang, Li, et al.
Pubblicazione: (2025)
di: Changjiang, Li, et al.
Pubblicazione: (2025)
Revisiting Backdoor Attacks on Time Series Classification in the Frequency Domain
di: Huang, Yuanmin, et al.
Pubblicazione: (2025)
di: Huang, Yuanmin, et al.
Pubblicazione: (2025)
BAFFLE: Hiding Backdoors in Offline Reinforcement Learning Datasets
di: Gong, Chen, et al.
Pubblicazione: (2022)
di: Gong, Chen, et al.
Pubblicazione: (2022)
Purifying Generative LLMs from Backdoors without Prior Knowledge or Clean Reference
di: Li, Jianwei, et al.
Pubblicazione: (2026)
di: Li, Jianwei, et al.
Pubblicazione: (2026)
How to Craft Backdoors with Unlabeled Data Alone?
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
Injecting Universal Jailbreak Backdoors into LLMs in Minutes
di: Chen, Zhuowei, et al.
Pubblicazione: (2025)
di: Chen, Zhuowei, et al.
Pubblicazione: (2025)
LoBAM: LoRA-Based Backdoor Attack on Model Merging
di: Yin, Ming, et al.
Pubblicazione: (2024)
di: Yin, Ming, et al.
Pubblicazione: (2024)
Position: Retire the "Positive Backdoor" Label -- Secret Alignment Requires Strict and Systematic Evaluation
di: Li, Jianwei, et al.
Pubblicazione: (2026)
di: Li, Jianwei, et al.
Pubblicazione: (2026)
PBP: Post-training Backdoor Purification for Malware Classifiers
di: Nguyen, Dung Thuy, et al.
Pubblicazione: (2024)
di: Nguyen, Dung Thuy, et al.
Pubblicazione: (2024)
On the (In)feasibility of ML Backdoor Detection as an Hypothesis Testing Problem
di: Pichler, Georg, et al.
Pubblicazione: (2024)
di: Pichler, Georg, et al.
Pubblicazione: (2024)
BACKTIME: Backdoor Attacks on Multivariate Time Series Forecasting
di: Lin, Xiao, et al.
Pubblicazione: (2024)
di: Lin, Xiao, et al.
Pubblicazione: (2024)
Backdooring Bias ($B^2$) into Stable Diffusion Models
di: Naseh, Ali, et al.
Pubblicazione: (2024)
di: Naseh, Ali, et al.
Pubblicazione: (2024)
Flatness-aware Sequential Learning Generates Resilient Backdoors
di: Pham, Hoang, et al.
Pubblicazione: (2024)
di: Pham, Hoang, et al.
Pubblicazione: (2024)
Invisible Backdoor Attack Through Singular Value Decomposition
di: Chen, Wenmin, et al.
Pubblicazione: (2024)
di: Chen, Wenmin, et al.
Pubblicazione: (2024)
Pay Attention to the Triggers: Constructing Backdoors That Survive Distillation
di: De Muri, Giovanni, et al.
Pubblicazione: (2025)
di: De Muri, Giovanni, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Magnitude-based Neuron Pruning for Backdoor Defens
di: Li, Nan, et al.
Pubblicazione: (2024) -
OCGEC: One-class Graph Embedding Classification for DNN Backdoor Detection
di: Jiang, Haoyu, et al.
Pubblicazione: (2023) -
Uncovering, Explaining, and Mitigating the Superficial Safety of Backdoor Defense
di: Min, Rui, et al.
Pubblicazione: (2024) -
Mutual Information Guided Backdoor Mitigation for Pre-trained Encoders
di: Han, Tingxu, et al.
Pubblicazione: (2024) -
ICLShield: Exploring and Mitigating In-Context Learning Backdoor Attacks
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)