Model-agnostic clean-label backdoor mitigation in cybersecurity environments
Fuente:
arXiv
Saved in:
| Main Authors: | Severi, Giorgio, Boboila, Simona, Holodnak, John, Kratkiewicz, Kendra, Izmailov, Rauf, De Lucia, Michael J., Oprea, Alina |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quantitative Resilience Modeling for Autonomous Cyber Defense
by: Cadet, Xavier, et al.
Published: (2025)
by: Cadet, Xavier, et al.
Published: (2025)
Backdoor Attacks in Peer-to-Peer Federated Learning
by: Syros, Georgios, et al.
Published: (2023)
by: Syros, Georgios, et al.
Published: (2023)
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
by: Cadet, Xavier, et al.
Published: (2025)
by: Cadet, Xavier, et al.
Published: (2025)
A clean-label graph backdoor attack method in node classification task
by: Xing, Xiaogang, et al.
Published: (2023)
by: Xing, Xiaogang, et al.
Published: (2023)
Hierarchical Multi-agent Reinforcement Learning for Cyber Network Defense
by: Singh, Aditya Vikram, et al.
Published: (2024)
by: Singh, Aditya Vikram, et al.
Published: (2024)
On the critical path to implant backdoors and the effectiveness of potential mitigation techniques: Early learnings from XZ
by: Lins, Mario, et al.
Published: (2024)
by: Lins, Mario, et al.
Published: (2024)
Retrieval-Augmented LLMs for Security Incident Analysis
by: Cadet, Xavier, et al.
Published: (2026)
by: Cadet, Xavier, et al.
Published: (2026)
Enhancing cybersecurity defenses: a multicriteria decision-making approach to MITRE ATT&CK mitigation strategy
by: Mohamed, Ihab, et al.
Published: (2024)
by: Mohamed, Ihab, et al.
Published: (2024)
Reconstruction of Personally Identifiable Information from Supervised Finetuned Models
by: Furukawa, Sae, et al.
Published: (2026)
by: Furukawa, Sae, et al.
Published: (2026)
TMI! Finetuned Models Leak Private Information from their Pretraining Data
by: Abascal, John, et al.
Published: (2023)
by: Abascal, John, et al.
Published: (2023)
Synthesizing Tight Privacy and Accuracy Bounds via Weighted Model Counting
by: Oakley, Lisa, et al.
Published: (2024)
by: Oakley, Lisa, et al.
Published: (2024)
Measuring likelihood in cybersecurity
by: Corona-Fraga, Pablo, et al.
Published: (2025)
by: Corona-Fraga, Pablo, et al.
Published: (2025)
Toward a Principled Framework for Agent Safety Measurement
by: Lin, Shuyi, et al.
Published: (2026)
by: Lin, Shuyi, et al.
Published: (2026)
Adversarial Inception Backdoor Attacks against Reinforcement Learning
by: Rathbun, Ethan, et al.
Published: (2024)
by: Rathbun, Ethan, et al.
Published: (2024)
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
by: Rathbun, Ethan, et al.
Published: (2024)
by: Rathbun, Ethan, et al.
Published: (2024)
Navigating the road to automotive cybersecurity compliance
by: Oberti, Franco, et al.
Published: (2024)
by: Oberti, Franco, et al.
Published: (2024)
NBA: defensive distillation for backdoor removal via neural behavior alignment
by: Ying, Zonghao, et al.
Published: (2024)
by: Ying, Zonghao, et al.
Published: (2024)
Phantom: General Backdoor Attacks on Retrieval Augmented Language Generation
by: Chaudhari, Harsh, et al.
Published: (2024)
by: Chaudhari, Harsh, et al.
Published: (2024)
From base cases to backdoors: An Empirical Study of Unnatural Crypto-API Misuse
by: Olaiya, Victor, et al.
Published: (2025)
by: Olaiya, Victor, et al.
Published: (2025)
DLP: towards active defense against backdoor attacks with decoupled learning process
by: Ying, Zonghao, et al.
Published: (2024)
by: Ying, Zonghao, et al.
Published: (2024)
Artificial intelligence and cybersecurity in banking sector: opportunities and risks
by: Kovacevic, Ana, et al.
Published: (2024)
by: Kovacevic, Ana, et al.
Published: (2024)
Coherence-driven inference for cybersecurity
by: Huntsman, Steve
Published: (2025)
by: Huntsman, Steve
Published: (2025)
Syntax- and Compilation-Preserving Evasion of LLM Vulnerability Detectors
by: Sun, Luze, et al.
Published: (2026)
by: Sun, Luze, et al.
Published: (2026)
Towards Cybersecurity SuperIntelligence (CSI): What's the best harness for cybersecurity?
by: Mayoral-Vilches, Víctor, et al.
Published: (2026)
by: Mayoral-Vilches, Víctor, et al.
Published: (2026)
CAT: Concept-level backdoor ATtacks for Concept Bottleneck Models
by: Lai, Songning, et al.
Published: (2024)
by: Lai, Songning, et al.
Published: (2024)
ImpNet: Imperceptible and blackbox-undetectable backdoors in compiled neural networks
by: Clifford, Eleanor, et al.
Published: (2022)
by: Clifford, Eleanor, et al.
Published: (2022)
Exploiting Leaderboards for Large-Scale Distribution of Malicious Models
by: Suri, Anshuman, et al.
Published: (2025)
by: Suri, Anshuman, et al.
Published: (2025)
CYGENT: A cybersecurity conversational agent with log summarization powered by GPT-3
by: Balasubramanian, Prasasthy, et al.
Published: (2024)
by: Balasubramanian, Prasasthy, et al.
Published: (2024)
Showcasing standards and approaches for cybersecurity, safety, and privacy issues in connected and autonomous vehicles
by: Czekster, Ricardo M.
Published: (2025)
by: Czekster, Ricardo M.
Published: (2025)
Black-Box Privacy Attacks on Shared Representations in Multitask Learning
by: Abascal, John, et al.
Published: (2025)
by: Abascal, John, et al.
Published: (2025)
Cascading Adversarial Bias from Injection to Distillation in Language Models
by: Chaudhari, Harsh, et al.
Published: (2025)
by: Chaudhari, Harsh, et al.
Published: (2025)
Text-to-Image Models Leave Identifiable Signatures: Implications for Leaderboard Security
by: Naseh, Ali, et al.
Published: (2025)
by: Naseh, Ali, et al.
Published: (2025)
Interactive cybersecurity training system based on simulation environments
by: Tymoshchuk, Dmytro, et al.
Published: (2024)
by: Tymoshchuk, Dmytro, et al.
Published: (2024)
A general approach to enhance the survivability of backdoor attacks by decision path coupling
by: Zhao, Yufei, et al.
Published: (2024)
by: Zhao, Yufei, et al.
Published: (2024)
ChatGPT, is this real? The influence of generative AI on writing style in top-tier cybersecurity papers
by: Vansteenhuyse, Daan
Published: (2026)
by: Vansteenhuyse, Daan
Published: (2026)
Beware Untrusted Simulators -- Reward-Free Backdoor Attacks in Reinforcement Learning
by: Rathbun, Ethan, et al.
Published: (2026)
by: Rathbun, Ethan, et al.
Published: (2026)
Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem
by: Lin, Shuyi, et al.
Published: (2025)
by: Lin, Shuyi, et al.
Published: (2025)
Attacks and Mitigations for Distributed Governance of Agentic AI under Byzantine Adversaries
by: Laws, Matthew D., et al.
Published: (2026)
by: Laws, Matthew D., et al.
Published: (2026)
We need to aim at the top: Factors associated with cybersecurity awareness of cyber and information security decision-makers
by: Vrhovec, Simon, et al.
Published: (2024)
by: Vrhovec, Simon, et al.
Published: (2024)
A Bayesian Approach to Membership Inference for Statistical Release
by: Oakley, Lisa, et al.
Published: (2026)
by: Oakley, Lisa, et al.
Published: (2026)
Similar Items
-
Quantitative Resilience Modeling for Autonomous Cyber Defense
by: Cadet, Xavier, et al.
Published: (2025) -
Backdoor Attacks in Peer-to-Peer Federated Learning
by: Syros, Georgios, et al.
Published: (2023) -
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
by: Cadet, Xavier, et al.
Published: (2025) -
A clean-label graph backdoor attack method in node classification task
by: Xing, Xiaogang, et al.
Published: (2023) -
Hierarchical Multi-agent Reinforcement Learning for Cyber Network Defense
by: Singh, Aditya Vikram, et al.
Published: (2024)