NBA: defensive distillation for backdoor removal via neural behavior alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ying, Zonghao, Wu, Bin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DLP: towards active defense against backdoor attacks with decoupled learning process
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
ImpNet: Imperceptible and blackbox-undetectable backdoors in compiled neural networks
von: Clifford, Eleanor, et al.
Veröffentlicht: (2022)
von: Clifford, Eleanor, et al.
Veröffentlicht: (2022)
FL-CLEANER: byzantine and backdoor defense by CLustering Errors of Activation maps in Non-iid fedErated leaRning
von: Ghali, Mehdi Ben, et al.
Veröffentlicht: (2025)
von: Ghali, Mehdi Ben, et al.
Veröffentlicht: (2025)
Effective backdoor attack on graph neural networks in link prediction tasks
von: Dai, Jiazhu, et al.
Veröffentlicht: (2024)
von: Dai, Jiazhu, et al.
Veröffentlicht: (2024)
From base cases to backdoors: An Empirical Study of Unnatural Crypto-API Misuse
von: Olaiya, Victor, et al.
Veröffentlicht: (2025)
von: Olaiya, Victor, et al.
Veröffentlicht: (2025)
Attacks on the neural network and defense methods
von: Korenev, A., et al.
Veröffentlicht: (2024)
von: Korenev, A., et al.
Veröffentlicht: (2024)
On the critical path to implant backdoors and the effectiveness of potential mitigation techniques: Early learnings from XZ
von: Lins, Mario, et al.
Veröffentlicht: (2024)
von: Lins, Mario, et al.
Veröffentlicht: (2024)
Model-agnostic clean-label backdoor mitigation in cybersecurity environments
von: Severi, Giorgio, et al.
Veröffentlicht: (2024)
von: Severi, Giorgio, et al.
Veröffentlicht: (2024)
Vul-LMGNNs: Fusing language models and online-distilled graph neural networks for code vulnerability detection
von: Liu, Ruitong, et al.
Veröffentlicht: (2024)
von: Liu, Ruitong, et al.
Veröffentlicht: (2024)
A general approach to enhance the survivability of backdoor attacks by decision path coupling
von: Zhao, Yufei, et al.
Veröffentlicht: (2024)
von: Zhao, Yufei, et al.
Veröffentlicht: (2024)
A clean-label graph backdoor attack method in node classification task
von: Xing, Xiaogang, et al.
Veröffentlicht: (2023)
von: Xing, Xiaogang, et al.
Veröffentlicht: (2023)
Non-omniscient backdoor injection with one poison sample: Proving the one-poison hypothesis for linear regression, linear classification, and 2-layer ReLU neural networks
von: Peinemann, Thorsten, et al.
Veröffentlicht: (2025)
von: Peinemann, Thorsten, et al.
Veröffentlicht: (2025)
CAT: Concept-level backdoor ATtacks for Concept Bottleneck Models
von: Lai, Songning, et al.
Veröffentlicht: (2024)
von: Lai, Songning, et al.
Veröffentlicht: (2024)
AgentVisor: Defending LLM Agents Against Prompt Injection via Semantic Virtualization
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
The Impact of Exposed Passwords on Honeyword Efficacy
von: Huang, Zonghao, et al.
Veröffentlicht: (2023)
von: Huang, Zonghao, et al.
Veröffentlicht: (2023)
Are aligned neural networks adversarially aligned?
von: Carlini, Nicholas, et al.
Veröffentlicht: (2023)
von: Carlini, Nicholas, et al.
Veröffentlicht: (2023)
The last Dance : Robust backdoor attack via diffusion models and bayesian approach
von: Mengara, Orson
Veröffentlicht: (2024)
von: Mengara, Orson
Veröffentlicht: (2024)
Trading Devil: Robust backdoor attack via Stochastic investment models and Bayesian approach
von: Mengara, Orson
Veröffentlicht: (2024)
von: Mengara, Orson
Veröffentlicht: (2024)
SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
Active Sybil attack and efficient defense strategy in IPFS DHT
von: Netto, V. H. de Moura, et al.
Veröffentlicht: (2025)
von: Netto, V. H. de Moura, et al.
Veröffentlicht: (2025)
Uncovering Security Threats and Architecting Defenses in Autonomous Agents: A Case Study of OpenClaw
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
DefTesPY: Cyber defense model with enhanced data modeling and analysis for Tesla company via Python Language
von: Kshetri, Naresh, et al.
Veröffentlicht: (2024)
von: Kshetri, Naresh, et al.
Veröffentlicht: (2024)
A limited technical background is sufficient for attack-defense tree acceptability
von: Schiele, Nathan Daniel, et al.
Veröffentlicht: (2025)
von: Schiele, Nathan Daniel, et al.
Veröffentlicht: (2025)
Evolving Deception: When Agents Evolve, Deception Wins
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
Reasoning-Oriented Programming: Chaining Semantic Gadgets to Jailbreak Large Vision Language Models
von: Zou, Quanchen, et al.
Veröffentlicht: (2026)
von: Zou, Quanchen, et al.
Veröffentlicht: (2026)
Proactive security defense: cyber threat intelligence modeling for connected autonomous vehicles
von: Wang, Yinghui, et al.
Veröffentlicht: (2024)
von: Wang, Yinghui, et al.
Veröffentlicht: (2024)
Unveiling the Safety of GPT-4o: An Empirical Study using Jailbreak Attacks
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
SlowBA: An efficiency backdoor attack towards VLM-based GUI agents
von: Li, Junxian, et al.
Veröffentlicht: (2026)
von: Li, Junxian, et al.
Veröffentlicht: (2026)
Sequential Comics for Jailbreaking Multimodal Large Language Models via Structured Visual Storytelling
von: Zhang, Deyue, et al.
Veröffentlicht: (2025)
von: Zhang, Deyue, et al.
Veröffentlicht: (2025)
CS-Eval: A Comprehensive Large Language Model Benchmark for CyberSecurity
von: Yu, Zhengmin, et al.
Veröffentlicht: (2024)
von: Yu, Zhengmin, et al.
Veröffentlicht: (2024)
HackCar: a test platform for attacks and defenses on a cost-contained automotive architecture
von: Stabili, Dario, et al.
Veröffentlicht: (2024)
von: Stabili, Dario, et al.
Veröffentlicht: (2024)
Proactive defense against LLM Jailbreak
von: Zhao, Weiliang, et al.
Veröffentlicht: (2025)
von: Zhao, Weiliang, et al.
Veröffentlicht: (2025)
FreeTalk:A plug-and-play and black-box defense against speech synthesis attacks
von: Pu, Yuwen, et al.
Veröffentlicht: (2025)
von: Pu, Yuwen, et al.
Veröffentlicht: (2025)
TeleAI-Safety: A comprehensive LLM jailbreaking benchmark towards attacks, defenses, and evaluations
von: Chen, Xiuyuan, et al.
Veröffentlicht: (2025)
von: Chen, Xiuyuan, et al.
Veröffentlicht: (2025)
Enhancing cybersecurity defenses: a multicriteria decision-making approach to MITRE ATT&CK mitigation strategy
von: Mohamed, Ihab, et al.
Veröffentlicht: (2024)
von: Mohamed, Ihab, et al.
Veröffentlicht: (2024)
A General Framework for Data-Use Auditing of ML Models
von: Huang, Zonghao, et al.
Veröffentlicht: (2024)
von: Huang, Zonghao, et al.
Veröffentlicht: (2024)
Instance-Level Data-Use Auditing of Visual ML Models
von: Huang, Zonghao, et al.
Veröffentlicht: (2025)
von: Huang, Zonghao, et al.
Veröffentlicht: (2025)
DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs
von: Xu, Wenzhuo, et al.
Veröffentlicht: (2026)
von: Xu, Wenzhuo, et al.
Veröffentlicht: (2026)
Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
Enhanced cast-128 with adaptive s-box optimization via neural networks for image protection
von: Fadhil, Fadhil Abbas, et al.
Veröffentlicht: (2025)
von: Fadhil, Fadhil Abbas, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DLP: towards active defense against backdoor attacks with decoupled learning process
von: Ying, Zonghao, et al.
Veröffentlicht: (2024) -
ImpNet: Imperceptible and blackbox-undetectable backdoors in compiled neural networks
von: Clifford, Eleanor, et al.
Veröffentlicht: (2022) -
FL-CLEANER: byzantine and backdoor defense by CLustering Errors of Activation maps in Non-iid fedErated leaRning
von: Ghali, Mehdi Ben, et al.
Veröffentlicht: (2025) -
Effective backdoor attack on graph neural networks in link prediction tasks
von: Dai, Jiazhu, et al.
Veröffentlicht: (2024) -
From base cases to backdoors: An Empirical Study of Unnatural Crypto-API Misuse
von: Olaiya, Victor, et al.
Veröffentlicht: (2025)