Knowledge Distillation-Based Model Extraction Attack using GAN-based Private Counterfactual Explanations
Fuente:
arXiv
Saved in:
| Main Authors: | Ezzeddine, Fatima, Ayoub, Omran, Giordano, Silvia |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the interplay of Explainability, Privacy and Predictive Performance with Explanation-assisted Model Extraction
by: Ezzeddine, Fatima, et al.
Published: (2025)
by: Ezzeddine, Fatima, et al.
Published: (2025)
Fair Recourse for All: Ensuring Individual and Group Fairness in Counterfactual Explanations
by: Ezzeddine, Fatima, et al.
Published: (2026)
by: Ezzeddine, Fatima, et al.
Published: (2026)
Privacy Implications of Explainable AI in Data-Driven Systems
by: Ezzeddine, Fatima
Published: (2024)
by: Ezzeddine, Fatima
Published: (2024)
A Survey of Privacy-Preserving Model Explanations: Privacy Risks, Attacks, and Countermeasures
by: Nguyen, Thanh Tam, et al.
Published: (2024)
by: Nguyen, Thanh Tam, et al.
Published: (2024)
On Membership Inference Attacks in Knowledge Distillation
by: Cui, Ziyao, et al.
Published: (2025)
by: Cui, Ziyao, et al.
Published: (2025)
VISION: Robust and Interpretable Code Vulnerability Detection Leveraging Counterfactual Augmentation
by: Egea, David, et al.
Published: (2025)
by: Egea, David, et al.
Published: (2025)
Differentially Private Data Release on Graphs: Inefficiencies and Unfairness
by: Fioretto, Ferdinando, et al.
Published: (2024)
by: Fioretto, Ferdinando, et al.
Published: (2024)
FAIRPLAI: A Human-in-the-Loop Approach to Fair and Private Machine Learning
by: Sanchez Jr., David, et al.
Published: (2025)
by: Sanchez Jr., David, et al.
Published: (2025)
A Survey on Model Extraction Attacks and Defenses for Large Language Models
by: Zhao, Kaixiang, et al.
Published: (2025)
by: Zhao, Kaixiang, et al.
Published: (2025)
Fair-FLIP: Fair Deepfake Detection with Fairness-Oriented Final Layer Input Prioritising
by: Szandala, Tomasz, et al.
Published: (2025)
by: Szandala, Tomasz, et al.
Published: (2025)
Machine Unlearning Fails to Remove Data Poisoning Attacks
by: Pawelczyk, Martin, et al.
Published: (2024)
by: Pawelczyk, Martin, et al.
Published: (2024)
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods
by: Jang, Yeonwoo, et al.
Published: (2025)
by: Jang, Yeonwoo, et al.
Published: (2025)
A Public Theory of Distillation Resistance via Constraint-Coupled Reasoning Architectures
by: Wei, Peng, et al.
Published: (2026)
by: Wei, Peng, et al.
Published: (2026)
Hidden Poison: Machine Unlearning Enables Camouflaged Poisoning Attacks
by: Di, Jimmy Z., et al.
Published: (2022)
by: Di, Jimmy Z., et al.
Published: (2022)
A Survey of Model Extraction Attacks and Defenses in Distributed Computing Environments
by: Zhao, Kaixiang, et al.
Published: (2025)
by: Zhao, Kaixiang, et al.
Published: (2025)
How to Backdoor the Knowledge Distillation
by: Wu, Chen, et al.
Published: (2025)
by: Wu, Chen, et al.
Published: (2025)
A Systematic Survey of Model Extraction Attacks and Defenses: State-of-the-Art and Perspectives
by: Zhao, Kaixiang, et al.
Published: (2025)
by: Zhao, Kaixiang, et al.
Published: (2025)
Differentially Private Deep Model-Based Reinforcement Learning
by: Rio, Alexandre, et al.
Published: (2024)
by: Rio, Alexandre, et al.
Published: (2024)
Privacy Bias in Language Models: A Contextual Integrity-based Auditing Metric
by: Shvartzshnaider, Yan, et al.
Published: (2024)
by: Shvartzshnaider, Yan, et al.
Published: (2024)
Precise Extraction of Deep Learning Models via Side-Channel Attacks on Edge/Endpoint Devices
by: Lee, Younghan, et al.
Published: (2024)
by: Lee, Younghan, et al.
Published: (2024)
Can Differentially Private Fine-tuning LLMs Protect Against Privacy Attacks?
by: Du, Hao, et al.
Published: (2025)
by: Du, Hao, et al.
Published: (2025)
Urania: Differentially Private Insights into AI Use
by: Liu, Daogao, et al.
Published: (2025)
by: Liu, Daogao, et al.
Published: (2025)
Untargeted Adversarial Attack on Knowledge Graph Embeddings
by: Zhao, Tianzhe, et al.
Published: (2024)
by: Zhao, Tianzhe, et al.
Published: (2024)
Decentralized autonomous organization and blockchain-based incentivization framework for community-based facilities management
by: Ly, Reachsak, et al.
Published: (2026)
by: Ly, Reachsak, et al.
Published: (2026)
Differentially Private Model Merging
by: Yin, Qichuan, et al.
Published: (2026)
by: Yin, Qichuan, et al.
Published: (2026)
EGAN: Evolutional GAN for Ransomware Evasion
by: Commey, Daniel, et al.
Published: (2024)
by: Commey, Daniel, et al.
Published: (2024)
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
by: Ngong, Ivoline C., et al.
Published: (2024)
by: Ngong, Ivoline C., et al.
Published: (2024)
A GAN-based data poisoning framework against anomaly detection in vertical federated learning
by: Chen, Xiaolin, et al.
Published: (2024)
by: Chen, Xiaolin, et al.
Published: (2024)
The Application of Transformer-Based Models for Predicting Consequences of Cyber Attacks
by: Chhetri, Bipin, et al.
Published: (2025)
by: Chhetri, Bipin, et al.
Published: (2025)
Privacy Assessment of Federated Learning using Private Personalized Layers
by: Jourdan, Théo, et al.
Published: (2021)
by: Jourdan, Théo, et al.
Published: (2021)
Trustless Audits without Revealing Data or Models
by: Waiwitlikhit, Suppakit, et al.
Published: (2024)
by: Waiwitlikhit, Suppakit, et al.
Published: (2024)
Robust Federated Learning with Confidence-Weighted Filtering and GAN-Based Completion under Noisy and Incomplete Data
by: Gokcen, Alpaslan, et al.
Published: (2025)
by: Gokcen, Alpaslan, et al.
Published: (2025)
ExpProof : Operationalizing Explanations for Confidential Models with ZKPs
by: Yadav, Chhavi, et al.
Published: (2025)
by: Yadav, Chhavi, et al.
Published: (2025)
Confidential Guardian: Cryptographically Prohibiting the Abuse of Model Abstention
by: Rabanser, Stephan, et al.
Published: (2025)
by: Rabanser, Stephan, et al.
Published: (2025)
SecGenAI: Enhancing Security of Cloud-based Generative AI Applications within Australian Critical Technologies of National Interest
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
LoBAM: LoRA-Based Backdoor Attack on Model Merging
by: Yin, Ming, et al.
Published: (2024)
by: Yin, Ming, et al.
Published: (2024)
Robust Safety Monitoring of Language Models via Activation Watermarking
by: Aremu, Toluwani, et al.
Published: (2026)
by: Aremu, Toluwani, et al.
Published: (2026)
Making AI-Assisted Grant Evaluation Auditable without Exposing the Model
by: Bicakci, Kemal
Published: (2026)
by: Bicakci, Kemal
Published: (2026)
Towards Robust Stability Prediction in Smart Grids: GAN-based Approach under Data Constraints and Adversarial Challenges
by: Efatinasab, Emad, et al.
Published: (2025)
by: Efatinasab, Emad, et al.
Published: (2025)
Explanations Leak: Membership Inference with Differential Privacy and Active Learning Defense
by: Ezzeddine, Fatima, et al.
Published: (2026)
by: Ezzeddine, Fatima, et al.
Published: (2026)
Similar Items
-
On the interplay of Explainability, Privacy and Predictive Performance with Explanation-assisted Model Extraction
by: Ezzeddine, Fatima, et al.
Published: (2025) -
Fair Recourse for All: Ensuring Individual and Group Fairness in Counterfactual Explanations
by: Ezzeddine, Fatima, et al.
Published: (2026) -
Privacy Implications of Explainable AI in Data-Driven Systems
by: Ezzeddine, Fatima
Published: (2024) -
A Survey of Privacy-Preserving Model Explanations: Privacy Risks, Attacks, and Countermeasures
by: Nguyen, Thanh Tam, et al.
Published: (2024) -
On Membership Inference Attacks in Knowledge Distillation
by: Cui, Ziyao, et al.
Published: (2025)