BeDKD: Backdoor Defense Based on Directional Mapping Module and Adversarial Knowledge Distillation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Zhengxian, Wen, Juan, Peng, Wanli, Zhou, Yinghan, dou, Changtong, Xue, Yiming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SLIP: Soft Label Mechanism and Key-Extraction-Guided CoT-based Defense Against Instruction Backdoor in APIs
por: Wu, Zhengxian, et al.
Publicado: (2025)
por: Wu, Zhengxian, et al.
Publicado: (2025)
BadApex: Backdoor Attack Based on Adaptive Optimization Mechanism of Black-box Large Language Models
por: Wu, Zhengxian, et al.
Publicado: (2025)
por: Wu, Zhengxian, et al.
Publicado: (2025)
Inhibitory Attacks on Backdoor-based Fingerprinting for Large Language Models
por: Fu, Hang, et al.
Publicado: (2026)
por: Fu, Hang, et al.
Publicado: (2026)
Is Your Writing Being Mimicked by AI? Unveiling Imitation with Invisible Watermarks in Creative Writing
por: Zhang, Ziwei, et al.
Publicado: (2025)
por: Zhang, Ziwei, et al.
Publicado: (2025)
Self-Disguise Attack: Induce the LLM to disguise itself for AIGT detection evasion
por: Zhou, Yinghan, et al.
Publicado: (2025)
por: Zhou, Yinghan, et al.
Publicado: (2025)
GTSD: Generative Text Steganography Based on Diffusion Model
por: Wu, Zhengxian, et al.
Publicado: (2025)
por: Wu, Zhengxian, et al.
Publicado: (2025)
EditMF: Drawing an Invisible Fingerprint for Your Large Language Models
por: Wu, Jiaxuan, et al.
Publicado: (2025)
por: Wu, Jiaxuan, et al.
Publicado: (2025)
BDPFL: Backdoor Defense for Personalized Federated Learning via Explainable Distillation
por: Zhu, Chengcheng, et al.
Publicado: (2025)
por: Zhu, Chengcheng, et al.
Publicado: (2025)
Transferring Backdoors between Large Language Models by Knowledge Distillation
por: Cheng, Pengzhou, et al.
Publicado: (2024)
por: Cheng, Pengzhou, et al.
Publicado: (2024)
BURN: Backdoor Unlearning via Adversarial Boundary Analysis
por: Su, Yanghao, et al.
Publicado: (2025)
por: Su, Yanghao, et al.
Publicado: (2025)
How to Backdoor the Knowledge Distillation
por: Wu, Chen, et al.
Publicado: (2025)
por: Wu, Chen, et al.
Publicado: (2025)
FedBAP: Backdoor Defense via Benign Adversarial Perturbation in Federated Learning
por: Yan, Xinhai, et al.
Publicado: (2025)
por: Yan, Xinhai, et al.
Publicado: (2025)
BadActs: A Universal Backdoor Defense in the Activation Space
por: Yi, Biao, et al.
Publicado: (2024)
por: Yi, Biao, et al.
Publicado: (2024)
TRAP: Hijacking VLA CoT-Reasoning via Adversarial Patches
por: Huang, Zhengxian, et al.
Publicado: (2026)
por: Huang, Zhengxian, et al.
Publicado: (2026)
Retrieval-Confused Generation is a Good Defender for Privacy Violation Attack of Large Language Models
por: Peng, Wanli, et al.
Publicado: (2025)
por: Peng, Wanli, et al.
Publicado: (2025)
Nearest is Not Dearest: Towards Practical Defense against Quantization-conditioned Backdoor Attacks
por: Li, Boheng, et al.
Publicado: (2024)
por: Li, Boheng, et al.
Publicado: (2024)
Robustness Inspired Graph Backdoor Defense
por: Zhang, Zhiwei, et al.
Publicado: (2024)
por: Zhang, Zhiwei, et al.
Publicado: (2024)
BackdoorMBTI: A Backdoor Learning Multimodal Benchmark Tool Kit for Backdoor Defense Evaluation
por: Yu, Haiyang, et al.
Publicado: (2024)
por: Yu, Haiyang, et al.
Publicado: (2024)
Evolutionary Trigger Detection and Lightweight Model Repair Based Backdoor Defense
por: Zhou, Qi, et al.
Publicado: (2024)
por: Zhou, Qi, et al.
Publicado: (2024)
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
por: Wei, Shaokui, et al.
Publicado: (2024)
por: Wei, Shaokui, et al.
Publicado: (2024)
ProtoGuard-SL: Prototype Consistency Based Backdoor Defense for Vertical Split Learning
por: Shui, Yuhan, et al.
Publicado: (2026)
por: Shui, Yuhan, et al.
Publicado: (2026)
Robust Knowledge Distillation in Federated Learning: Counteracting Backdoor Attacks
por: Alharbi, Ebtisaam, et al.
Publicado: (2025)
por: Alharbi, Ebtisaam, et al.
Publicado: (2025)
Backdoor-Powered Prompt Injection Attacks Nullify Defense Methods
por: Chen, Yulin, et al.
Publicado: (2025)
por: Chen, Yulin, et al.
Publicado: (2025)
Backdoor Attacks and Defenses in Computer Vision Domain: A Survey
por: Abbasi, Bilal Hussain, et al.
Publicado: (2025)
por: Abbasi, Bilal Hussain, et al.
Publicado: (2025)
Distributed Backdoor Attacks on Federated Graph Learning and Certified Defenses
por: Yang, Yuxin, et al.
Publicado: (2024)
por: Yang, Yuxin, et al.
Publicado: (2024)
SoK: The Last Line of Defense: On Backdoor Defense Evaluation
por: Abad, Gorka, et al.
Publicado: (2025)
por: Abad, Gorka, et al.
Publicado: (2025)
Seal Your Backdoor with Variational Defense
por: Sabolić, Ivan, et al.
Publicado: (2025)
por: Sabolić, Ivan, et al.
Publicado: (2025)
FilterFL: Knowledge Filtering-based Data-Free Backdoor Defense for Federated Learning
por: Yang, Yanxin, et al.
Publicado: (2023)
por: Yang, Yanxin, et al.
Publicado: (2023)
Architectural Backdoors in Deep Learning: A Survey of Vulnerabilities, Detection, and Defense
por: Childress, Victoria, et al.
Publicado: (2025)
por: Childress, Victoria, et al.
Publicado: (2025)
Attack as Defense: Run-time Backdoor Implantation for Image Content Protection
por: Zhang, Haichuan, et al.
Publicado: (2024)
por: Zhang, Haichuan, et al.
Publicado: (2024)
BDFirewall: Towards Effective and Expeditiously Black-Box Backdoor Defense in MLaaS
por: Li, Ye, et al.
Publicado: (2025)
por: Li, Ye, et al.
Publicado: (2025)
Physical Backdoor Attack Against Deep Learning-Based Modulation Classification
por: Salmi, Younes, et al.
Publicado: (2026)
por: Salmi, Younes, et al.
Publicado: (2026)
MARS: A Malignity-Aware Backdoor Defense in Federated Learning
por: Wan, Wei, et al.
Publicado: (2025)
por: Wan, Wei, et al.
Publicado: (2025)
Dark Distillation: Backdooring Distilled Datasets without Accessing Raw Data
por: Yang, Ziyuan, et al.
Publicado: (2025)
por: Yang, Ziyuan, et al.
Publicado: (2025)
Unlearning Backdoor Attacks for LLMs with Weak-to-Strong Knowledge Distillation
por: Zhao, Shuai, et al.
Publicado: (2024)
por: Zhao, Shuai, et al.
Publicado: (2024)
Stateful Agent Backdoor
por: Dai, Zhengchunmin, et al.
Publicado: (2026)
por: Dai, Zhengchunmin, et al.
Publicado: (2026)
Diffusion-Guided Adversarial Perturbation Injection for Generalizable Defense Against Facial Manipulations
por: Li, Yue, et al.
Publicado: (2026)
por: Li, Yue, et al.
Publicado: (2026)
Towards Imperceptible Adversarial Defense: A Gradient-Driven Shield against Facial Manipulations
por: Li, Yue, et al.
Publicado: (2025)
por: Li, Yue, et al.
Publicado: (2025)
Plato's Form: Toward Backdoor Defense-as-a-Service for LLMs with Prototype Representations
por: Chen, Chen, et al.
Publicado: (2026)
por: Chen, Chen, et al.
Publicado: (2026)
DistillGuard: Evaluating Defenses Against LLM Knowledge Distillation
por: Jiang, Bo
Publicado: (2026)
por: Jiang, Bo
Publicado: (2026)
Ejemplares similares
-
SLIP: Soft Label Mechanism and Key-Extraction-Guided CoT-based Defense Against Instruction Backdoor in APIs
por: Wu, Zhengxian, et al.
Publicado: (2025) -
BadApex: Backdoor Attack Based on Adaptive Optimization Mechanism of Black-box Large Language Models
por: Wu, Zhengxian, et al.
Publicado: (2025) -
Inhibitory Attacks on Backdoor-based Fingerprinting for Large Language Models
por: Fu, Hang, et al.
Publicado: (2026) -
Is Your Writing Being Mimicked by AI? Unveiling Imitation with Invisible Watermarks in Creative Writing
por: Zhang, Ziwei, et al.
Publicado: (2025) -
Self-Disguise Attack: Induce the LLM to disguise itself for AIGT detection evasion
por: Zhou, Yinghan, et al.
Publicado: (2025)