GLiRA: Black-Box Membership Inference Attack via Knowledge Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Galichin, Andrey V., Pautov, Mikhail, Zhavoronkin, Alexey, Rogov, Oleg Y., Oseledets, Ivan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spread them Apart: Towards Robust Watermarking of Generated Content
by: Pautov, Mikhail, et al.
Published: (2025)
by: Pautov, Mikhail, et al.
Published: (2025)
The Rogue Scalpel: Activation Steering Compromises LLM Safety
by: Korznikov, Anton, et al.
Published: (2025)
by: Korznikov, Anton, et al.
Published: (2025)
Probabilistically Robust Watermarking of Neural Networks
by: Pautov, Mikhail, et al.
Published: (2024)
by: Pautov, Mikhail, et al.
Published: (2024)
Certification of Speaker Recognition Models to Additive Perturbations
by: Korzh, Dmitrii, et al.
Published: (2024)
by: Korzh, Dmitrii, et al.
Published: (2024)
Contract And Conquer: How to Provably Compute Adversarial Examples for a Black-Box Model?
by: Chistyakova, Anna, et al.
Published: (2026)
by: Chistyakova, Anna, et al.
Published: (2026)
General Lipschitz: Certified Robustness Against Resolvable Semantic Transformations via Transformation-Dependent Randomized Smoothing
by: Korzh, Dmitrii, et al.
Published: (2023)
by: Korzh, Dmitrii, et al.
Published: (2023)
Probabilistic Verification of Voice Anti-Spoofing Models
by: Kushnir, Evgeny, et al.
Published: (2026)
by: Kushnir, Evgeny, et al.
Published: (2026)
Model Mimic Attack: Knowledge Distillation for Provably Transferable Adversarial Examples
by: Lukyanov, Kirill, et al.
Published: (2024)
by: Lukyanov, Kirill, et al.
Published: (2024)
Towards Robust Speech Deepfake Detection via Human-Inspired Reasoning
by: Dvirniak, Artem, et al.
Published: (2026)
by: Dvirniak, Artem, et al.
Published: (2026)
Black-Box Membership Inference Attack for LVLMs via Prior Knowledge-Calibrated Memory Probing
by: Yin, Jinhua, et al.
Published: (2025)
by: Yin, Jinhua, et al.
Published: (2025)
On Membership Inference Attacks in Knowledge Distillation
by: Cui, Ziyao, et al.
Published: (2025)
by: Cui, Ziyao, et al.
Published: (2025)
OrtSAE: Orthogonal Sparse Autoencoders Uncover Atomic Features
by: Korznikov, Anton, et al.
Published: (2025)
by: Korznikov, Anton, et al.
Published: (2025)
Sanity Checks for Sparse Autoencoders: Do SAEs Beat Random Baselines?
by: Korznikov, Anton, et al.
Published: (2026)
by: Korznikov, Anton, et al.
Published: (2026)
I Have Covered All the Bases Here: Interpreting Reasoning Features in Large Language Models via Sparse Autoencoders
by: Galichin, Andrey, et al.
Published: (2025)
by: Galichin, Andrey, et al.
Published: (2025)
GigaEvo: An Open Source Optimization Framework Powered By LLMs And Evolution Algorithms
by: Khrulkov, Valentin, et al.
Published: (2025)
by: Khrulkov, Valentin, et al.
Published: (2025)
Towards Black-Box Membership Inference Attack for Diffusion Models
by: Li, Jingwei, et al.
Published: (2024)
by: Li, Jingwei, et al.
Published: (2024)
E-MIA: Exam-Style Black-Box Membership Inference Attacks against RAG Systems
by: Guan, Zelin, et al.
Published: (2026)
by: Guan, Zelin, et al.
Published: (2026)
MrM: Black-Box Membership Inference Attacks against Multimodal RAG Systems
by: Yang, Peiru, et al.
Published: (2025)
by: Yang, Peiru, et al.
Published: (2025)
ActiveMark: on watermarking of visual foundation models via massive activations
by: Chistyakova, Anna, et al.
Published: (2025)
by: Chistyakova, Anna, et al.
Published: (2025)
Listener-Rewarded Thinking in VLMs for Image Preferences
by: Gambashidze, Alexander, et al.
Published: (2025)
by: Gambashidze, Alexander, et al.
Published: (2025)
RandMark: On Random Watermarking of Visual Foundation Models
by: Chistyakova, Anna, et al.
Published: (2026)
by: Chistyakova, Anna, et al.
Published: (2026)
Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment
by: Li, Jiaqing, et al.
Published: (2026)
by: Li, Jiaqing, et al.
Published: (2026)
LoRA-Leak: Membership Inference Attacks Against LoRA Fine-tuned Language Models
by: Ran, Delong, et al.
Published: (2025)
by: Ran, Delong, et al.
Published: (2025)
CLEAR: Character Unlearning in Textual and Visual Modalities
by: Dontsov, Alexey, et al.
Published: (2024)
by: Dontsov, Alexey, et al.
Published: (2024)
DistractMIA: Black-Box Membership Inference on Vision-Language Models via Semantic Distraction
by: Tang, Hongyi, et al.
Published: (2026)
by: Tang, Hongyi, et al.
Published: (2026)
Bayesian Inverse Problems Meet Flow Matching: Efficient and Flexible Inference via Transformers
by: Sherki, Daniil, et al.
Published: (2025)
by: Sherki, Daniil, et al.
Published: (2025)
Membership and Memorization in LLM Knowledge Distillation
by: Zhang, Ziqi, et al.
Published: (2025)
by: Zhang, Ziqi, et al.
Published: (2025)
AASIST3: KAN-Enhanced AASIST Speech Deepfake Detection using SSL Features and Additional Regularization for the ASVspoof 2024 Challenge
by: Borodin, Kirill, et al.
Published: (2024)
by: Borodin, Kirill, et al.
Published: (2024)
Ensemble Privacy Defense for Knowledge-Intensive LLMs against Membership Inference Attacks
by: Fu, Haowei, et al.
Published: (2025)
by: Fu, Haowei, et al.
Published: (2025)
Hyperparameters in Score-Based Membership Inference Attacks
by: Pradhan, Gauri, et al.
Published: (2025)
by: Pradhan, Gauri, et al.
Published: (2025)
Membership Inference Attack with Partial Features
by: Wang, Xurun, et al.
Published: (2025)
by: Wang, Xurun, et al.
Published: (2025)
Defending Membership Inference Attacks via Privacy-aware Sparsity Tuning
by: Hu, Qiang, et al.
Published: (2024)
by: Hu, Qiang, et al.
Published: (2024)
LeakBoost: Perceptual-Loss-Based Membership Inference Attack
by: Taub, Amit Kravchik, et al.
Published: (2026)
by: Taub, Amit Kravchik, et al.
Published: (2026)
CLMIA: Membership Inference Attacks via Unsupervised Contrastive Learning
by: Chen, Depeng, et al.
Published: (2024)
by: Chen, Depeng, et al.
Published: (2024)
Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models
by: Qi, Tao, et al.
Published: (2026)
by: Qi, Tao, et al.
Published: (2026)
Membership Inference Attacks on Discrete Diffusion Language Models
by: Kasivelrajan, Shailesh
Published: (2026)
by: Kasivelrajan, Shailesh
Published: (2026)
Generalization and Membership Inference Attack a Practical Perspective
by: Rahmani, Fateme, et al.
Published: (2026)
by: Rahmani, Fateme, et al.
Published: (2026)
Membership Inference Attacks Against Vision-Language Models
by: Hu, Yuke, et al.
Published: (2025)
by: Hu, Yuke, et al.
Published: (2025)
Membership Inference Attacks on Tokenizers of Large Language Models
by: Tong, Meng, et al.
Published: (2025)
by: Tong, Meng, et al.
Published: (2025)
Membership Inference Attacks Against Time-Series Models
by: Koren, Noam, et al.
Published: (2024)
by: Koren, Noam, et al.
Published: (2024)
Similar Items
-
Spread them Apart: Towards Robust Watermarking of Generated Content
by: Pautov, Mikhail, et al.
Published: (2025) -
The Rogue Scalpel: Activation Steering Compromises LLM Safety
by: Korznikov, Anton, et al.
Published: (2025) -
Probabilistically Robust Watermarking of Neural Networks
by: Pautov, Mikhail, et al.
Published: (2024) -
Certification of Speaker Recognition Models to Additive Perturbations
by: Korzh, Dmitrii, et al.
Published: (2024) -
Contract And Conquer: How to Provably Compute Adversarial Examples for a Black-Box Model?
by: Chistyakova, Anna, et al.
Published: (2026)