Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
Fuente:
arXiv
Guardado en:
| Autores principales: | Mia, Maraz, Pritom, Mir Mehedi A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Can Features for Phishing URL Detection Be Trusted Across Diverse Datasets? A Case Study with Explainable AI
por: Mia, Maraz, et al.
Publicado: (2024)
por: Mia, Maraz, et al.
Publicado: (2024)
Characterizing Event-themed Malicious Web Campaigns: A Case Study on War-themed Websites
por: Mia, Maraz, et al.
Publicado: (2025)
por: Mia, Maraz, et al.
Publicado: (2025)
Visually Analyze SHAP Plots to Diagnose Misclassifications in ML-based Intrusion Detection
por: Mia, Maraz, et al.
Publicado: (2024)
por: Mia, Maraz, et al.
Publicado: (2024)
AbuseGPT: Abuse of Generative AI ChatBots to Create Smishing Campaigns
por: Shibli, Ashfak Md, et al.
Publicado: (2024)
por: Shibli, Ashfak Md, et al.
Publicado: (2024)
Cloud-based XAI Services for Assessing Open Repository Models Under Adversarial Attacks
por: Wang, Zerui, et al.
Publicado: (2024)
por: Wang, Zerui, et al.
Publicado: (2024)
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities
por: Ferrag, Mohamed Amine, et al.
Publicado: (2024)
por: Ferrag, Mohamed Amine, et al.
Publicado: (2024)
Empirical Analysis of Adversarial Robustness and Explainability Drift in Cybersecurity Classifiers
por: Rajhans, Mona, et al.
Publicado: (2026)
por: Rajhans, Mona, et al.
Publicado: (2026)
Explainable Artificial Intelligence (XAI) for Malware Analysis: A Survey of Techniques, Applications, and Open Challenges
por: Manthena, Harikha, et al.
Publicado: (2024)
por: Manthena, Harikha, et al.
Publicado: (2024)
Vulnerability Disclosure through Adaptive Black-Box Adversarial Attacks on NIDS
por: Ennaji, Sabrine, et al.
Publicado: (2025)
por: Ennaji, Sabrine, et al.
Publicado: (2025)
Large Language Models in Cybersecurity: Applications, Vulnerabilities, and Defense Techniques
por: Jaffal, Niveen O., et al.
Publicado: (2025)
por: Jaffal, Niveen O., et al.
Publicado: (2025)
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
por: Hasan, Md. Mehedi, et al.
Publicado: (2025)
por: Hasan, Md. Mehedi, et al.
Publicado: (2025)
XAI-CF -- Examining the Role of Explainable Artificial Intelligence in Cyber Forensics
por: Alam, Shahid, et al.
Publicado: (2024)
por: Alam, Shahid, et al.
Publicado: (2024)
Adversarial Defense in Cybersecurity: A Systematic Review of GANs for Threat Detection and Mitigation
por: Ndayipfukamiye, Tharcisse, et al.
Publicado: (2025)
por: Ndayipfukamiye, Tharcisse, et al.
Publicado: (2025)
Are Frontier LLMs Ready for Cybersecurity? Evidence for Vertical Foundation Models from Dual-Mode Vulnerability Benchmarks
por: Dahiya, Vivek, et al.
Publicado: (2026)
por: Dahiya, Vivek, et al.
Publicado: (2026)
On the interplay of Explainability, Privacy and Predictive Performance with Explanation-assisted Model Extraction
por: Ezzeddine, Fatima, et al.
Publicado: (2025)
por: Ezzeddine, Fatima, et al.
Publicado: (2025)
Evaluating Adversarial Vulnerabilities in Modern Large Language Models
por: Perel, Tom
Publicado: (2025)
por: Perel, Tom
Publicado: (2025)
LENS-XAI: Redefining Lightweight and Explainable Network Security through Knowledge Distillation and Variational Autoencoders for Scalable Intrusion Detection in Cybersecurity
por: Yagiz, Muhammet Anil, et al.
Publicado: (2025)
por: Yagiz, Muhammet Anil, et al.
Publicado: (2025)
VADER: A Human-Evaluated Benchmark for Vulnerability Assessment, Detection, Explanation, and Remediation
por: Liu, Ethan TS., et al.
Publicado: (2025)
por: Liu, Ethan TS., et al.
Publicado: (2025)
Weaponizing Language Models for Cybersecurity Offensive Operations: Automating Vulnerability Assessment Report Validation; A Review Paper
por: Almuhaidib, Abdulrahman S, et al.
Publicado: (2025)
por: Almuhaidib, Abdulrahman S, et al.
Publicado: (2025)
Conflicts Make Large Reasoning Models Vulnerable to Attacks
por: Liu, Honghao, et al.
Publicado: (2026)
por: Liu, Honghao, et al.
Publicado: (2026)
DiffAttack: Evasion Attacks Against Diffusion-Based Adversarial Purification
por: Kang, Mintong, et al.
Publicado: (2023)
por: Kang, Mintong, et al.
Publicado: (2023)
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models
por: Ma, Jiachen, et al.
Publicado: (2024)
por: Ma, Jiachen, et al.
Publicado: (2024)
Selection-Based Vulnerabilities: Clean-Label Backdoor Attacks in Active Learning
por: Zhi, Yuhan, et al.
Publicado: (2025)
por: Zhi, Yuhan, et al.
Publicado: (2025)
From Similarity to Vulnerability: Key Collision Attack on LLM Semantic Caching
por: Zhang, Zhixiang, et al.
Publicado: (2026)
por: Zhang, Zhixiang, et al.
Publicado: (2026)
Vulnerabilities in AI Code Generators: Exploring Targeted Data Poisoning Attacks
por: Cotroneo, Domenico, et al.
Publicado: (2023)
por: Cotroneo, Domenico, et al.
Publicado: (2023)
Integrated Simulation Framework for Adversarial Attacks on Autonomous Vehicles
por: Anagnostopoulos, Christos, et al.
Publicado: (2025)
por: Anagnostopoulos, Christos, et al.
Publicado: (2025)
Adversarial Machine Learning: Attacks, Defenses, and Open Challenges
por: Jha, Pranav K
Publicado: (2025)
por: Jha, Pranav K
Publicado: (2025)
Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate
por: Qi, Senmao, et al.
Publicado: (2025)
por: Qi, Senmao, et al.
Publicado: (2025)
Exploring the Vulnerabilities of Federated Learning: A Deep Dive into Gradient Inversion Attacks
por: Guo, Pengxin, et al.
Publicado: (2025)
por: Guo, Pengxin, et al.
Publicado: (2025)
SEASONED: Semantic-Enhanced Self-Counterfactual Explainable Detection of Adversarial Exploiter Contracts
por: Ai, Xng, et al.
Publicado: (2025)
por: Ai, Xng, et al.
Publicado: (2025)
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations
por: Ge, Huaizhi, et al.
Publicado: (2024)
por: Ge, Huaizhi, et al.
Publicado: (2024)
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning
por: Domico, Kyle, et al.
Publicado: (2025)
por: Domico, Kyle, et al.
Publicado: (2025)
Special-Character Adversarial Attacks on Open-Source Language Model
por: Sarabamoun, Ephraiem
Publicado: (2025)
por: Sarabamoun, Ephraiem
Publicado: (2025)
Enhancing TinyML Security: Study of Adversarial Attack Transferability
por: Shah, Parin, et al.
Publicado: (2024)
por: Shah, Parin, et al.
Publicado: (2024)
Attention Masks Help Adversarial Attacks to Bypass Safety Detectors
por: Shi, Yunfan
Publicado: (2024)
por: Shi, Yunfan
Publicado: (2024)
Integrative Approaches in Cybersecurity and AI
por: Omar, Marwan
Publicado: (2024)
por: Omar, Marwan
Publicado: (2024)
When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output
por: Zhang, Shuoming, et al.
Publicado: (2025)
por: Zhang, Shuoming, et al.
Publicado: (2025)
GUARD-SLM: Token Activation-Based Defense Against Jailbreak Attacks for Small Language Models
por: Mia, Md Jueal, et al.
Publicado: (2026)
por: Mia, Md Jueal, et al.
Publicado: (2026)
Exploiting Vulnerabilities in Speech Translation Systems through Targeted Adversarial Attacks
por: Liu, Chang, et al.
Publicado: (2025)
por: Liu, Chang, et al.
Publicado: (2025)
Medical Multimodal Model Stealing Attacks via Adversarial Domain Alignment
por: Shen, Yaling, et al.
Publicado: (2025)
por: Shen, Yaling, et al.
Publicado: (2025)
Ejemplares similares
-
Can Features for Phishing URL Detection Be Trusted Across Diverse Datasets? A Case Study with Explainable AI
por: Mia, Maraz, et al.
Publicado: (2024) -
Characterizing Event-themed Malicious Web Campaigns: A Case Study on War-themed Websites
por: Mia, Maraz, et al.
Publicado: (2025) -
Visually Analyze SHAP Plots to Diagnose Misclassifications in ML-based Intrusion Detection
por: Mia, Maraz, et al.
Publicado: (2024) -
AbuseGPT: Abuse of Generative AI ChatBots to Create Smishing Campaigns
por: Shibli, Ashfak Md, et al.
Publicado: (2024) -
Cloud-based XAI Services for Assessing Open Repository Models Under Adversarial Attacks
por: Wang, Zerui, et al.
Publicado: (2024)