On the interplay of Explainability, Privacy and Predictive Performance with Explanation-assisted Model Extraction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ezzeddine, Fatima, Akel, Rinad, Sbeity, Ihab, Giordano, Silvia, Langheinrich, Marc, Ayoub, Omran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Knowledge Distillation-Based Model Extraction Attack using GAN-based Private Counterfactual Explanations
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2024)
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2024)
Differential Privacy for Anomaly Detection: Analyzing the Trade-off Between Privacy and Explainability
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2024)
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2024)
Privacy Implications of Explainable AI in Data-Driven Systems
von: Ezzeddine, Fatima
Veröffentlicht: (2024)
von: Ezzeddine, Fatima
Veröffentlicht: (2024)
Fair Recourse for All: Ensuring Individual and Group Fairness in Counterfactual Explanations
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2026)
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2026)
Explainable Machine Learning-Based Security and Privacy Protection Framework for Internet of Medical Things Systems
von: Si-ahmed, Ayoub, et al.
Veröffentlicht: (2024)
von: Si-ahmed, Ayoub, et al.
Veröffentlicht: (2024)
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
von: Mia, Maraz, et al.
Veröffentlicht: (2025)
von: Mia, Maraz, et al.
Veröffentlicht: (2025)
Privacy-Preserving Decentralized Federated Learning via Explainable Adaptive Differential Privacy
von: Piran, Fardin Jalil, et al.
Veröffentlicht: (2025)
von: Piran, Fardin Jalil, et al.
Veröffentlicht: (2025)
Privacy Risks and Preservation Methods in Explainable Artificial Intelligence: A Scoping Review
von: Allana, Sonal, et al.
Veröffentlicht: (2025)
von: Allana, Sonal, et al.
Veröffentlicht: (2025)
INVARLLM: LLM-assisted Physical Invariant Extraction for Cyber-Physical Systems Anomaly Detection
von: Abshari, Danial, et al.
Veröffentlicht: (2024)
von: Abshari, Danial, et al.
Veröffentlicht: (2024)
Small Language Models for Phishing Website Detection: Cost, Performance, and Privacy Trade-Offs
von: Goldenits, Georg, et al.
Veröffentlicht: (2025)
von: Goldenits, Georg, et al.
Veröffentlicht: (2025)
PrivacyRestore: Privacy-Preserving Inference in Large Language Models via Privacy Removal and Restoration
von: Zeng, Ziqian, et al.
Veröffentlicht: (2024)
von: Zeng, Ziqian, et al.
Veröffentlicht: (2024)
Towards a scalable AI-driven framework for data-independent Cyber Threat Intelligence Information Extraction
von: Sorokoletova, Olga, et al.
Veröffentlicht: (2025)
von: Sorokoletova, Olga, et al.
Veröffentlicht: (2025)
A Survey of Privacy-Preserving Model Explanations: Privacy Risks, Attacks, and Countermeasures
von: Nguyen, Thanh Tam, et al.
Veröffentlicht: (2024)
von: Nguyen, Thanh Tam, et al.
Veröffentlicht: (2024)
Explanations Leak: Membership Inference with Differential Privacy and Active Learning Defense
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2026)
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2026)
Routing-Aware Explanations for Mixture of Experts Graph Models in Malware Detection
von: Shokouhinejad, Hossein, et al.
Veröffentlicht: (2026)
von: Shokouhinejad, Hossein, et al.
Veröffentlicht: (2026)
Fair-FLIP: Fair Deepfake Detection with Fairness-Oriented Final Layer Input Prioritising
von: Szandala, Tomasz, et al.
Veröffentlicht: (2025)
von: Szandala, Tomasz, et al.
Veröffentlicht: (2025)
QUEEN: Query Unlearning against Model Extraction
von: Chen, Huajie, et al.
Veröffentlicht: (2024)
von: Chen, Huajie, et al.
Veröffentlicht: (2024)
MISLEADER: Defending against Model Extraction with Ensembles of Distilled Models
von: Cheng, Xueqi, et al.
Veröffentlicht: (2025)
von: Cheng, Xueqi, et al.
Veröffentlicht: (2025)
Enhancing Privacy in Federated Learning: Secure Aggregation for Real-World Healthcare Applications
von: Taiello, Riccardo, et al.
Veröffentlicht: (2024)
von: Taiello, Riccardo, et al.
Veröffentlicht: (2024)
User Behavior Analysis in Privacy Protection with Large Language Models: A Study on Privacy Preferences with Limited Data
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
When Better Features Mean Greater Risks: The Performance-Privacy Trade-Off in Contrastive Learning
von: Sun, Ruining, et al.
Veröffentlicht: (2025)
von: Sun, Ruining, et al.
Veröffentlicht: (2025)
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations
von: Ge, Huaizhi, et al.
Veröffentlicht: (2024)
von: Ge, Huaizhi, et al.
Veröffentlicht: (2024)
Privacy-Preserving Diffusion Model Using Homomorphic Encryption
von: Chen, Yaojian, et al.
Veröffentlicht: (2024)
von: Chen, Yaojian, et al.
Veröffentlicht: (2024)
SoK: Semantic Privacy in Large Language Models
von: Ma, Baihe, et al.
Veröffentlicht: (2025)
von: Ma, Baihe, et al.
Veröffentlicht: (2025)
LLMAC: A Global and Explainable Access Control Framework with Large Language Model
von: Zisad, Sharif Noor, et al.
Veröffentlicht: (2026)
von: Zisad, Sharif Noor, et al.
Veröffentlicht: (2026)
Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models
von: Biswas, Umesh, et al.
Veröffentlicht: (2026)
von: Biswas, Umesh, et al.
Veröffentlicht: (2026)
LLM-PBE: Assessing Data Privacy in Large Language Models
von: Li, Qinbin, et al.
Veröffentlicht: (2024)
von: Li, Qinbin, et al.
Veröffentlicht: (2024)
RAG Security and Privacy: Formalizing the Threat Model and Attack Surface
von: Arzanipour, Atousa, et al.
Veröffentlicht: (2025)
von: Arzanipour, Atousa, et al.
Veröffentlicht: (2025)
Eclectic Rule Extraction for Explainability of Deep Neural Network based Intrusion Detection Systems
von: Ables, Jesse, et al.
Veröffentlicht: (2024)
von: Ables, Jesse, et al.
Veröffentlicht: (2024)
Prompt Inject Detection with Generative Explanation as an Investigative Tool
von: Pan, Jonathan, et al.
Veröffentlicht: (2025)
von: Pan, Jonathan, et al.
Veröffentlicht: (2025)
A Survey: Towards Privacy and Security in Mobile Large Language Models
von: Xu, Honghui, et al.
Veröffentlicht: (2025)
von: Xu, Honghui, et al.
Veröffentlicht: (2025)
CompLeak: Deep Learning Model Compression Exacerbates Privacy Leakage
von: Li, Na, et al.
Veröffentlicht: (2025)
von: Li, Na, et al.
Veröffentlicht: (2025)
PPBFL: A Privacy Protected Blockchain-based Federated Learning Model
von: Li, Yang, et al.
Veröffentlicht: (2024)
von: Li, Yang, et al.
Veröffentlicht: (2024)
Privacy-Preserving Large Language Models: Mechanisms, Applications, and Future Directions
von: Zhao, Guoshenghui, et al.
Veröffentlicht: (2024)
von: Zhao, Guoshenghui, et al.
Veröffentlicht: (2024)
Privacy in Action: Towards Realistic Privacy Mitigation and Evaluation for LLM-Powered Agents
von: Wang, Shouju, et al.
Veröffentlicht: (2025)
von: Wang, Shouju, et al.
Veröffentlicht: (2025)
Neural Honeytrace: Plug&Play Watermarking Framework against Model Extraction Attacks
von: Xu, Yixiao, et al.
Veröffentlicht: (2025)
von: Xu, Yixiao, et al.
Veröffentlicht: (2025)
MARVEL: Multi-Agent RTL Vulnerability Extraction using Large Language Models
von: Collini, Luca, et al.
Veröffentlicht: (2025)
von: Collini, Luca, et al.
Veröffentlicht: (2025)
Fault Injection and Safe-Error Attack for Extraction of Embedded Neural Network Models
von: Hector, Kevin, et al.
Veröffentlicht: (2023)
von: Hector, Kevin, et al.
Veröffentlicht: (2023)
Watermarking Graph Neural Networks via Explanations for Ownership Protection
von: Downer, Jane, et al.
Veröffentlicht: (2025)
von: Downer, Jane, et al.
Veröffentlicht: (2025)
ExplainableGuard: Interpretable Adversarial Defense for Large Language Models Using Chain-of-Thought Reasoning
von: Guan, Shaowei, et al.
Veröffentlicht: (2025)
von: Guan, Shaowei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Knowledge Distillation-Based Model Extraction Attack using GAN-based Private Counterfactual Explanations
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2024) -
Differential Privacy for Anomaly Detection: Analyzing the Trade-off Between Privacy and Explainability
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2024) -
Privacy Implications of Explainable AI in Data-Driven Systems
von: Ezzeddine, Fatima
Veröffentlicht: (2024) -
Fair Recourse for All: Ensuring Individual and Group Fairness in Counterfactual Explanations
von: Ezzeddine, Fatima, et al.
Veröffentlicht: (2026) -
Explainable Machine Learning-Based Security and Privacy Protection Framework for Internet of Medical Things Systems
von: Si-ahmed, Ayoub, et al.
Veröffentlicht: (2024)