Model Mimic Attack: Knowledge Distillation for Provably Transferable Adversarial Examples
Fuente:
arXiv
Saved in:
| Main Authors: | Lukyanov, Kirill, Perminov, Andrew, Turdakov, Denis, Pautov, Mikhail |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contract And Conquer: How to Provably Compute Adversarial Examples for a Black-Box Model?
by: Chistyakova, Anna, et al.
Published: (2026)
by: Chistyakova, Anna, et al.
Published: (2026)
GLiRA: Black-Box Membership Inference Attack via Knowledge Distillation
by: Galichin, Andrey V., et al.
Published: (2024)
by: Galichin, Andrey V., et al.
Published: (2024)
Graph Neural Network for Crawling Target Nodes in Social Networks
by: Lukyanov, Kirill, et al.
Published: (2024)
by: Lukyanov, Kirill, et al.
Published: (2024)
Improving the Transferability of Adversarial Examples by Inverse Knowledge Distillation
by: Wu, Wenyuan, et al.
Published: (2025)
by: Wu, Wenyuan, et al.
Published: (2025)
Framework GNN-AID: Graph Neural Network Analysis Interpretation and Defense
by: Lukyanov, Kirill, et al.
Published: (2025)
by: Lukyanov, Kirill, et al.
Published: (2025)
Towards Interpretable Adversarial Examples via Sparse Adversarial Attack
by: Lin, Fudong, et al.
Published: (2025)
by: Lin, Fudong, et al.
Published: (2025)
Teach Me to Trick: Exploring Adversarial Transferability via Knowledge Distillation
by: Pradhan, Siddhartha, et al.
Published: (2025)
by: Pradhan, Siddhartha, et al.
Published: (2025)
Robustness questions the interpretability of graph neural networks: what to do?
by: Lukyanov, Kirill, et al.
Published: (2025)
by: Lukyanov, Kirill, et al.
Published: (2025)
Understanding Model Ensemble in Transferable Adversarial Attack
by: Yao, Wei, et al.
Published: (2024)
by: Yao, Wei, et al.
Published: (2024)
Rethinking Adversarial Policies: A Generalized Attack Formulation and Provable Defense in RL
by: Liu, Xiangyu, et al.
Published: (2023)
by: Liu, Xiangyu, et al.
Published: (2023)
Rethinking the Intermediate Features in Adversarial Attacks: Misleading Robotic Models via Adversarial Distillation
by: Zhao, Ke, et al.
Published: (2024)
by: Zhao, Ke, et al.
Published: (2024)
SoftMimic: Learning Compliant Whole-body Control from Examples
by: Margolis, Gabriel B., et al.
Published: (2025)
by: Margolis, Gabriel B., et al.
Published: (2025)
PEAS: A Strategy for Crafting Transferable Adversarial Examples
by: Avraham, Bar, et al.
Published: (2024)
by: Avraham, Bar, et al.
Published: (2024)
Online Adversarial Knowledge Distillation for Graph Neural Networks
by: Wang, Can, et al.
Published: (2021)
by: Wang, Can, et al.
Published: (2021)
Provably Invincible Adversarial Attacks on Reinforcement Learning Systems: A Rate-Distortion Information-Theoretic Approach
by: Lu, Ziqing, et al.
Published: (2025)
by: Lu, Ziqing, et al.
Published: (2025)
On Membership Inference Attacks in Knowledge Distillation
by: Cui, Ziyao, et al.
Published: (2025)
by: Cui, Ziyao, et al.
Published: (2025)
MemLoss: Enhancing Adversarial Training with Recycling Adversarial Examples
by: Mahdi, Soroush, et al.
Published: (2025)
by: Mahdi, Soroush, et al.
Published: (2025)
Exploiting Edge Features for Transferable Adversarial Attacks in Distributed Machine Learning
by: Rossolini, Giulio, et al.
Published: (2025)
by: Rossolini, Giulio, et al.
Published: (2025)
Harmonizing Intra-coherence and Inter-divergence in Ensemble Attacks for Adversarial Transferability
by: Ma, Zhaoyang, et al.
Published: (2025)
by: Ma, Zhaoyang, et al.
Published: (2025)
Generative Diffusion Prior Distillation for Long-Context Knowledge Transfer
by: Udayangani, Nilushika, et al.
Published: (2026)
by: Udayangani, Nilushika, et al.
Published: (2026)
Learning to Reason: Temporal Saliency Distillation for Interpretable Knowledge Transfer
by: Dehigahawattage, Nilushika Udayangani Hewa, et al.
Published: (2026)
by: Dehigahawattage, Nilushika Udayangani Hewa, et al.
Published: (2026)
Understanding Adversarial Transfer: Why Representation-Space Attacks Fail Where Data-Space Attacks Succeed
by: Gupta, Isha, et al.
Published: (2025)
by: Gupta, Isha, et al.
Published: (2025)
Adversarial Examples Might be Avoidable: The Role of Data Concentration in Adversarial Robustness
by: Pal, Ambar, et al.
Published: (2023)
by: Pal, Ambar, et al.
Published: (2023)
Attacking the Spike: On the Transferability and Security of Spiking Neural Networks to Adversarial Examples
by: Xu, Nuo, et al.
Published: (2022)
by: Xu, Nuo, et al.
Published: (2022)
Untargeted Adversarial Attack on Knowledge Graph Embeddings
by: Zhao, Tianzhe, et al.
Published: (2024)
by: Zhao, Tianzhe, et al.
Published: (2024)
A New Type of Adversarial Examples
by: Nie, Xingyang, et al.
Published: (2025)
by: Nie, Xingyang, et al.
Published: (2025)
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
by: Oulkadda, Ilyas, et al.
Published: (2025)
by: Oulkadda, Ilyas, et al.
Published: (2025)
Analyzing the Impact of Adversarial Examples on Explainable Machine Learning
by: Devabhakthini, Prathyusha, et al.
Published: (2023)
by: Devabhakthini, Prathyusha, et al.
Published: (2023)
Catch-Only-One: Non-Transferable Examples for Model-Specific Authorization
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
Provably Efficient Action-Manipulation Attack Against Continuous Reinforcement Learning
by: Luo, Zhi, et al.
Published: (2024)
by: Luo, Zhi, et al.
Published: (2024)
Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation
by: Kim, Juno, et al.
Published: (2025)
by: Kim, Juno, et al.
Published: (2025)
Towards a Novel Perspective on Adversarial Examples Driven by Frequency
by: Zhang, Zhun, et al.
Published: (2024)
by: Zhang, Zhun, et al.
Published: (2024)
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
by: Shing, Makoto, et al.
Published: (2025)
by: Shing, Makoto, et al.
Published: (2025)
Provably Mitigating Overoptimization in RLHF: Your SFT Loss is Implicitly an Adversarial Regularizer
by: Liu, Zhihan, et al.
Published: (2024)
by: Liu, Zhihan, et al.
Published: (2024)
Is Value Functions Estimation with Classification Plug-and-play for Offline Reinforcement Learning?
by: Tarasov, Denis, et al.
Published: (2024)
by: Tarasov, Denis, et al.
Published: (2024)
Adversarial Attacks on Hyperbolic Networks
by: van Spengler, Max, et al.
Published: (2024)
by: van Spengler, Max, et al.
Published: (2024)
DP-TRAE: A Dual-Phase Merging Transferable Reversible Adversarial Example for Image Privacy Protection
by: Du, Xia, et al.
Published: (2025)
by: Du, Xia, et al.
Published: (2025)
Generating Realistic Adversarial Examples for Business Processes using Variational Autoencoders
by: Stevens, Alexander, et al.
Published: (2024)
by: Stevens, Alexander, et al.
Published: (2024)
Self-Play with Adversarial Critic: Provable and Scalable Offline Alignment for Language Models
by: Ji, Xiang, et al.
Published: (2024)
by: Ji, Xiang, et al.
Published: (2024)
FACL-Attack: Frequency-Aware Contrastive Learning for Transferable Adversarial Attacks
by: Yang, Hunmin, et al.
Published: (2024)
by: Yang, Hunmin, et al.
Published: (2024)
Similar Items
-
Contract And Conquer: How to Provably Compute Adversarial Examples for a Black-Box Model?
by: Chistyakova, Anna, et al.
Published: (2026) -
GLiRA: Black-Box Membership Inference Attack via Knowledge Distillation
by: Galichin, Andrey V., et al.
Published: (2024) -
Graph Neural Network for Crawling Target Nodes in Social Networks
by: Lukyanov, Kirill, et al.
Published: (2024) -
Improving the Transferability of Adversarial Examples by Inverse Knowledge Distillation
by: Wu, Wenyuan, et al.
Published: (2025) -
Framework GNN-AID: Graph Neural Network Analysis Interpretation and Defense
by: Lukyanov, Kirill, et al.
Published: (2025)