Class-Conditional Neural Polarizer: A Lightweight and Effective Backdoor Defense by Purifying Poisoned Features
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Mingli, Wei, Shaokui, Zha, Hongyuan, Wu, Baoyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
von: Wei, Shaokui, et al.
Veröffentlicht: (2024)
von: Wei, Shaokui, et al.
Veröffentlicht: (2024)
Backdoor Mitigation by Distance-Driven Detoxification
von: Wei, Shaokui, et al.
Veröffentlicht: (2024)
von: Wei, Shaokui, et al.
Veröffentlicht: (2024)
Activation Gradient based Poisoned Sample Detection Against Backdoor Attacks
von: Yuan, Danni, et al.
Veröffentlicht: (2023)
von: Yuan, Danni, et al.
Veröffentlicht: (2023)
Revisiting the Auxiliary Data in Backdoor Purification
von: Wei, Shaokui, et al.
Veröffentlicht: (2025)
von: Wei, Shaokui, et al.
Veröffentlicht: (2025)
P2P: A Poison-to-Poison Remedy for Reliable Backdoor Defense in LLMs
von: Zhao, Shuai, et al.
Veröffentlicht: (2025)
von: Zhao, Shuai, et al.
Veröffentlicht: (2025)
Evolutionary Trigger Detection and Lightweight Model Repair Based Backdoor Defense
von: Zhou, Qi, et al.
Veröffentlicht: (2024)
von: Zhou, Qi, et al.
Veröffentlicht: (2024)
BackdoorBench: A Comprehensive Benchmark and Analysis of Backdoor Learning
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
BackdoorMBTI: A Backdoor Learning Multimodal Benchmark Tool Kit for Backdoor Defense Evaluation
von: Yu, Haiyang, et al.
Veröffentlicht: (2024)
von: Yu, Haiyang, et al.
Veröffentlicht: (2024)
MARS: A Malignity-Aware Backdoor Defense in Federated Learning
von: Wan, Wei, et al.
Veröffentlicht: (2025)
von: Wan, Wei, et al.
Veröffentlicht: (2025)
Lightweight and Fast Backdoor Model Detection
von: Yu, Yinbo, et al.
Veröffentlicht: (2026)
von: Yu, Yinbo, et al.
Veröffentlicht: (2026)
A Set of Generalized Components to Achieve Effective Poison-only Clean-label Backdoor Attacks with Collaborative Sample Selection and Triggers
von: Wu, Zhixiao, et al.
Veröffentlicht: (2025)
von: Wu, Zhixiao, et al.
Veröffentlicht: (2025)
SoK: The Last Line of Defense: On Backdoor Defense Evaluation
von: Abad, Gorka, et al.
Veröffentlicht: (2025)
von: Abad, Gorka, et al.
Veröffentlicht: (2025)
Toward Polymorphic Backdoor against Semantic Communication via Intensity-Based Poisoning
von: Yang, Xiao, et al.
Veröffentlicht: (2026)
von: Yang, Xiao, et al.
Veröffentlicht: (2026)
From Poisoned to Aware: Fostering Backdoor Self-Awareness in LLMs
von: Shen, Guangyu, et al.
Veröffentlicht: (2025)
von: Shen, Guangyu, et al.
Veröffentlicht: (2025)
CBPF: Filtering Poisoned Data Based on Composite Backdoor Attack
von: Xia, Hanfeng, et al.
Veröffentlicht: (2024)
von: Xia, Hanfeng, et al.
Veröffentlicht: (2024)
A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations
von: Zhou, Yihe, et al.
Veröffentlicht: (2025)
von: Zhou, Yihe, et al.
Veröffentlicht: (2025)
Exploring Backdoor Attack and Defense for LLM-empowered Recommendations
von: Ning, Liangbo, et al.
Veröffentlicht: (2025)
von: Ning, Liangbo, et al.
Veröffentlicht: (2025)
Purifying Generative LLMs from Backdoors without Prior Knowledge or Clean Reference
von: Li, Jianwei, et al.
Veröffentlicht: (2026)
von: Li, Jianwei, et al.
Veröffentlicht: (2026)
TED-LaST: Towards Robust Backdoor Defense Against Adaptive Attacks
von: Mo, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Mo, Xiaoxing, et al.
Veröffentlicht: (2025)
BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
BELT: Old-School Backdoor Attacks can Evade the State-of-the-Art Defense with Backdoor Exclusivity Lifting
von: Qiu, Huming, et al.
Veröffentlicht: (2023)
von: Qiu, Huming, et al.
Veröffentlicht: (2023)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
Dual Defense: Enhancing Privacy and Mitigating Poisoning Attacks in Federated Learning
von: Xu, Runhua, et al.
Veröffentlicht: (2025)
von: Xu, Runhua, et al.
Veröffentlicht: (2025)
WPDA: Frequency-based Backdoor Attack with Wavelet Packet Decomposition
von: Song, Zhengyao, et al.
Veröffentlicht: (2024)
von: Song, Zhengyao, et al.
Veröffentlicht: (2024)
Poison Once, Control Anywhere: Clean-Text Visual Backdoors in VLM-based Mobile Agents
von: Wang, Xuan, et al.
Veröffentlicht: (2025)
von: Wang, Xuan, et al.
Veröffentlicht: (2025)
Data Poisoning in Deep Learning: A Survey
von: Zhao, Pinlong, et al.
Veröffentlicht: (2025)
von: Zhao, Pinlong, et al.
Veröffentlicht: (2025)
SoK: Benchmarking Poisoning Attacks and Defenses in Federated Learning
von: Zhang, Heyi, et al.
Veröffentlicht: (2025)
von: Zhang, Heyi, et al.
Veröffentlicht: (2025)
Backdoor Vectors: a Task Arithmetic View on Backdoor Attacks and Defenses
von: Pawlak, Stanisław, et al.
Veröffentlicht: (2025)
von: Pawlak, Stanisław, et al.
Veröffentlicht: (2025)
CUBA: Controlled Untargeted Backdoor Attack against Deep Neural Networks
von: Wu, Yinghao, et al.
Veröffentlicht: (2025)
von: Wu, Yinghao, et al.
Veröffentlicht: (2025)
FFCBA: Feature-based Full-target Clean-label Backdoor Attacks
von: Yin, Yangxu, et al.
Veröffentlicht: (2025)
von: Yin, Yangxu, et al.
Veröffentlicht: (2025)
Acquiring Clean Language Models from Backdoor Poisoned Datasets by Downscaling Frequency Space
von: Wu, Zongru, et al.
Veröffentlicht: (2024)
von: Wu, Zongru, et al.
Veröffentlicht: (2024)
The Stronger the Diffusion Model, the Easier the Backdoor: Data Poisoning to Induce Copyright Breaches Without Adjusting Finetuning Pipeline
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
A Survey of Recent Backdoor Attacks and Defenses in Large Language Models
von: Zhao, Shuai, et al.
Veröffentlicht: (2024)
von: Zhao, Shuai, et al.
Veröffentlicht: (2024)
PAD-FT: A Lightweight Defense for Backdoor Attacks via Data Purification and Fine-Tuning
von: Xu, Yukai, et al.
Veröffentlicht: (2024)
von: Xu, Yukai, et al.
Veröffentlicht: (2024)
Impart: An Imperceptible and Effective Label-Specific Backdoor Attack
von: Zhao, Jingke, et al.
Veröffentlicht: (2024)
von: Zhao, Jingke, et al.
Veröffentlicht: (2024)
Purify Once, Edit Freely: Breaking Image Protections under Model Mismatch
von: Zhao, Qichen, et al.
Veröffentlicht: (2026)
von: Zhao, Qichen, et al.
Veröffentlicht: (2026)
Stealthy Poisoning Attacks Bypass Defenses in Regression Settings
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
Enhancing All-to-X Backdoor Attacks with Optimized Target Class Mapping
von: Wang, Lei, et al.
Veröffentlicht: (2025)
von: Wang, Lei, et al.
Veröffentlicht: (2025)
Fast and Lightweight Backdoor Detection via Head Random Probing
von: Yu, Yinbo, et al.
Veröffentlicht: (2026)
von: Yu, Yinbo, et al.
Veröffentlicht: (2026)
Critical-CoT: A Robust Defense Framework against Reasoning-Level Backdoor Attacks in Large Language Models
von: Truong, Vu Tuan, et al.
Veröffentlicht: (2026)
von: Truong, Vu Tuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
von: Wei, Shaokui, et al.
Veröffentlicht: (2024) -
Backdoor Mitigation by Distance-Driven Detoxification
von: Wei, Shaokui, et al.
Veröffentlicht: (2024) -
Activation Gradient based Poisoned Sample Detection Against Backdoor Attacks
von: Yuan, Danni, et al.
Veröffentlicht: (2023) -
Revisiting the Auxiliary Data in Backdoor Purification
von: Wei, Shaokui, et al.
Veröffentlicht: (2025) -
P2P: A Poison-to-Poison Remedy for Reliable Backdoor Defense in LLMs
von: Zhao, Shuai, et al.
Veröffentlicht: (2025)