Cloud-based XAI Services for Assessing Open Repository Models Under Adversarial Attacks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zerui, Liu, Yan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
von: Mia, Maraz, et al.
Veröffentlicht: (2025)
von: Mia, Maraz, et al.
Veröffentlicht: (2025)
Special-Character Adversarial Attacks on Open-Source Language Model
von: Sarabamoun, Ephraiem
Veröffentlicht: (2025)
von: Sarabamoun, Ephraiem
Veröffentlicht: (2025)
Adversarial Machine Learning: Attacks, Defenses, and Open Challenges
von: Jha, Pranav K
Veröffentlicht: (2025)
von: Jha, Pranav K
Veröffentlicht: (2025)
Multi-task Adversarial Attacks against Black-box Model with Few-shot Queries
von: Wang, Wenqiang, et al.
Veröffentlicht: (2025)
von: Wang, Wenqiang, et al.
Veröffentlicht: (2025)
DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems
von: Ou, Haoran, et al.
Veröffentlicht: (2026)
von: Ou, Haoran, et al.
Veröffentlicht: (2026)
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models
von: Ma, Jiachen, et al.
Veröffentlicht: (2024)
von: Ma, Jiachen, et al.
Veröffentlicht: (2024)
Rethinking Latency Denial-of-Service: Attacking the LLM Serving Framework, Not the Model
von: Wang, Tianyi, et al.
Veröffentlicht: (2026)
von: Wang, Tianyi, et al.
Veröffentlicht: (2026)
MaPPing Your Model: Assessing the Impact of Adversarial Attacks on LLM-based Programming Assistants
von: Heibel, John, et al.
Veröffentlicht: (2024)
von: Heibel, John, et al.
Veröffentlicht: (2024)
Explainable Artificial Intelligence (XAI) for Malware Analysis: A Survey of Techniques, Applications, and Open Challenges
von: Manthena, Harikha, et al.
Veröffentlicht: (2024)
von: Manthena, Harikha, et al.
Veröffentlicht: (2024)
RAG-targeted Adversarial Attack on LLM-based Threat Detection and Mitigation Framework
von: Ikbarieh, Seif, et al.
Veröffentlicht: (2025)
von: Ikbarieh, Seif, et al.
Veröffentlicht: (2025)
Adversarial Attacks on Multimodal Large Language Models: A Comprehensive Survey
von: Jain, Bhavuk, et al.
Veröffentlicht: (2026)
von: Jain, Bhavuk, et al.
Veröffentlicht: (2026)
Medical Multimodal Model Stealing Attacks via Adversarial Domain Alignment
von: Shen, Yaling, et al.
Veröffentlicht: (2025)
von: Shen, Yaling, et al.
Veröffentlicht: (2025)
DiffAttack: Evasion Attacks Against Diffusion-Based Adversarial Purification
von: Kang, Mintong, et al.
Veröffentlicht: (2023)
von: Kang, Mintong, et al.
Veröffentlicht: (2023)
XAI-based Feature Selection for Improved Network Intrusion Detection Systems
von: Arreche, Osvaldo, et al.
Veröffentlicht: (2024)
von: Arreche, Osvaldo, et al.
Veröffentlicht: (2024)
VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models
von: Liu, Pang, et al.
Veröffentlicht: (2026)
von: Liu, Pang, et al.
Veröffentlicht: (2026)
An Attack Method for Medical Insurance Claim Fraud Detection based on Generative Adversarial Network
von: Pang, Yining, et al.
Veröffentlicht: (2025)
von: Pang, Yining, et al.
Veröffentlicht: (2025)
TEAM: Temporal Adversarial Examples Attack Model against Network Intrusion Detection System Applied to RNN
von: Liu, Ziyi, et al.
Veröffentlicht: (2024)
von: Liu, Ziyi, et al.
Veröffentlicht: (2024)
Unraveling Attacks in Machine Learning-based IoT Ecosystems: A Survey and the Open Libraries Behind Them
von: Liu, Chao, et al.
Veröffentlicht: (2024)
von: Liu, Chao, et al.
Veröffentlicht: (2024)
Disttack: Graph Adversarial Attacks Toward Distributed GNN Training
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2024)
Integrated Simulation Framework for Adversarial Attacks on Autonomous Vehicles
von: Anagnostopoulos, Christos, et al.
Veröffentlicht: (2025)
von: Anagnostopoulos, Christos, et al.
Veröffentlicht: (2025)
Explainer-guided Targeted Adversarial Attacks against Binary Code Similarity Detection Models
von: Chen, Mingjie, et al.
Veröffentlicht: (2025)
von: Chen, Mingjie, et al.
Veröffentlicht: (2025)
Scam Shield: Multi-Model Voting and Fine-Tuned LLMs Against Adversarial Attacks
von: Chang, Chen-Wei, et al.
Veröffentlicht: (2025)
von: Chang, Chen-Wei, et al.
Veröffentlicht: (2025)
ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathologically Long Reasoning in Large Reasoning Models
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2026)
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2026)
Doppelganger Method: Breaking Role Consistency in LLM Agent via Prompt-based Transferable Adversarial Attack
von: Kang, Daewon, et al.
Veröffentlicht: (2025)
von: Kang, Daewon, et al.
Veröffentlicht: (2025)
A Rusty Link in the AI Supply Chain: Detecting Evil Configurations in Model Repositories
von: Ding, Ziqi, et al.
Veröffentlicht: (2025)
von: Ding, Ziqi, et al.
Veröffentlicht: (2025)
GenoArmory: A Unified Evaluation Framework for Adversarial Attacks on Genomic Foundation Models
von: Luo, Haozheng, et al.
Veröffentlicht: (2025)
von: Luo, Haozheng, et al.
Veröffentlicht: (2025)
OML: A Primitive for Reconciling Open Access with Owner Control in AI Model Distribution
von: Cheng, Zerui, et al.
Veröffentlicht: (2024)
von: Cheng, Zerui, et al.
Veröffentlicht: (2024)
Enhancing TinyML Security: Study of Adversarial Attack Transferability
von: Shah, Parin, et al.
Veröffentlicht: (2024)
von: Shah, Parin, et al.
Veröffentlicht: (2024)
Attention Masks Help Adversarial Attacks to Bypass Safety Detectors
von: Shi, Yunfan
Veröffentlicht: (2024)
von: Shi, Yunfan
Veröffentlicht: (2024)
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning
von: Domico, Kyle, et al.
Veröffentlicht: (2025)
von: Domico, Kyle, et al.
Veröffentlicht: (2025)
Frequency-Domain Regularized Adversarial Alignment for Transferable Attacks against Closed-Source MLLMs
von: Yuan, Leitao, et al.
Veröffentlicht: (2026)
von: Yuan, Leitao, et al.
Veröffentlicht: (2026)
Causal-Guided Detoxify Backdoor Attack of Open-Weight LoRA Models
von: Chen, Linzhi, et al.
Veröffentlicht: (2025)
von: Chen, Linzhi, et al.
Veröffentlicht: (2025)
AdaDoS: Adaptive DoS Attack via Deep Adversarial Reinforcement Learning in SDN
von: Shao, Wei, et al.
Veröffentlicht: (2025)
von: Shao, Wei, et al.
Veröffentlicht: (2025)
A General Black-box Adversarial Attack on Graph-based Fake News Detectors
von: Zhu, Peican, et al.
Veröffentlicht: (2024)
von: Zhu, Peican, et al.
Veröffentlicht: (2024)
MF-CLIP: Leveraging CLIP as Surrogate Models for No-box Adversarial Attacks
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
AttackPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
Vulnerability Disclosure through Adaptive Black-Box Adversarial Attacks on NIDS
von: Ennaji, Sabrine, et al.
Veröffentlicht: (2025)
von: Ennaji, Sabrine, et al.
Veröffentlicht: (2025)
Foe for Fraud: Transferable Adversarial Attacks in Credit Card Fraud Detection
von: Fok, Jan Lum, et al.
Veröffentlicht: (2025)
von: Fok, Jan Lum, et al.
Veröffentlicht: (2025)
Energy-Latency Attacks: A New Adversarial Threat to Deep Learning
von: Meftah, Hanene F. Z. Brachemi, et al.
Veröffentlicht: (2025)
von: Meftah, Hanene F. Z. Brachemi, et al.
Veröffentlicht: (2025)
LLM-Driven Feature-Level Adversarial Attacks on Android Malware Detectors
von: Lan, Tianwei, et al.
Veröffentlicht: (2025)
von: Lan, Tianwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
von: Mia, Maraz, et al.
Veröffentlicht: (2025) -
Special-Character Adversarial Attacks on Open-Source Language Model
von: Sarabamoun, Ephraiem
Veröffentlicht: (2025) -
Adversarial Machine Learning: Attacks, Defenses, and Open Challenges
von: Jha, Pranav K
Veröffentlicht: (2025) -
Multi-task Adversarial Attacks against Black-box Model with Few-shot Queries
von: Wang, Wenqiang, et al.
Veröffentlicht: (2025) -
DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems
von: Ou, Haoran, et al.
Veröffentlicht: (2026)