Hide in Plain Sight: Clean-Label Backdoor for Auditing Membership Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Depeng, Chen, Hao, Jin, Hulin, Cui, Jie, Zhong, Hong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLMIA: Membership Inference Attacks via Unsupervised Contrastive Learning
von: Chen, Depeng, et al.
Veröffentlicht: (2024)
von: Chen, Depeng, et al.
Veröffentlicht: (2024)
BAFFLE: Hiding Backdoors in Offline Reinforcement Learning Datasets
von: Gong, Chen, et al.
Veröffentlicht: (2022)
von: Gong, Chen, et al.
Veröffentlicht: (2022)
Strategic Sample Selection for Improved Clean-Label Backdoor Attacks in Text Classification
von: Kirci, Onur Alp, et al.
Veröffentlicht: (2025)
von: Kirci, Onur Alp, et al.
Veröffentlicht: (2025)
On Membership Inference Attacks in Knowledge Distillation
von: Cui, Ziyao, et al.
Veröffentlicht: (2025)
von: Cui, Ziyao, et al.
Veröffentlicht: (2025)
Hiding-in-Plain-Sight (HiPS) Attack on CLIP for Targetted Object Removal from Images
von: Daw, Arka, et al.
Veröffentlicht: (2024)
von: Daw, Arka, et al.
Veröffentlicht: (2024)
Privacy Auditing of Multi-domain Graph Pre-trained Model under Membership Inference Attacks
von: Luo, Jiayi, et al.
Veröffentlicht: (2025)
von: Luo, Jiayi, et al.
Veröffentlicht: (2025)
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
Double-Dip: Thwarting Label-Only Membership Inference Attacks with Transfer Learning and Randomization
von: Rajabi, Arezoo, et al.
Veröffentlicht: (2024)
von: Rajabi, Arezoo, et al.
Veröffentlicht: (2024)
Hiding in Plain Sight: Detectability-Aware Antidistillation of Reasoning Models
von: Hartman, Max, et al.
Veröffentlicht: (2026)
von: Hartman, Max, et al.
Veröffentlicht: (2026)
Hiding in Plain Sight: A Steganographic Approach to Stealthy LLM Jailbreaks
von: Geng, Jianing, et al.
Veröffentlicht: (2025)
von: Geng, Jianing, et al.
Veröffentlicht: (2025)
Membership Inference Attack with Partial Features
von: Wang, Xurun, et al.
Veröffentlicht: (2025)
von: Wang, Xurun, et al.
Veröffentlicht: (2025)
A Semantic and Clean-label Backdoor Attack against Graph Convolutional Networks
von: Dai, Jiazhu, et al.
Veröffentlicht: (2025)
von: Dai, Jiazhu, et al.
Veröffentlicht: (2025)
Purifying Generative LLMs from Backdoors without Prior Knowledge or Clean Reference
von: Li, Jianwei, et al.
Veröffentlicht: (2026)
von: Li, Jianwei, et al.
Veröffentlicht: (2026)
BadReward: Clean-Label Poisoning of Reward Models in Text-to-Image RLHF
von: Duan, Kaiwen, et al.
Veröffentlicht: (2025)
von: Duan, Kaiwen, et al.
Veröffentlicht: (2025)
Honeyfile Camouflage: Hiding Fake Files in Plain Sight
von: Timmer, Roelien C., et al.
Veröffentlicht: (2024)
von: Timmer, Roelien C., et al.
Veröffentlicht: (2024)
On the Evidentiary Limits of Membership Inference for Copyright Auditing
von: Ertan, Murat Bilgehan, et al.
Veröffentlicht: (2026)
von: Ertan, Murat Bilgehan, et al.
Veröffentlicht: (2026)
Similarity-based Label Inference Attack against Training and Inference of Split Learning
von: Liu, Junlin, et al.
Veröffentlicht: (2022)
von: Liu, Junlin, et al.
Veröffentlicht: (2022)
Center-Based Relaxed Learning Against Membership Inference Attacks
von: Fang, Xingli, et al.
Veröffentlicht: (2024)
von: Fang, Xingli, et al.
Veröffentlicht: (2024)
Do Parameters Reveal More than Loss for Membership Inference?
von: Suri, Anshuman, et al.
Veröffentlicht: (2024)
von: Suri, Anshuman, et al.
Veröffentlicht: (2024)
Learning-Based Difficulty Calibration for Enhanced Membership Inference Attacks
von: Shi, Haonan, et al.
Veröffentlicht: (2024)
von: Shi, Haonan, et al.
Veröffentlicht: (2024)
Improved Membership Inference Attacks Against Language Classification Models
von: Shachor, Shlomit, et al.
Veröffentlicht: (2023)
von: Shachor, Shlomit, et al.
Veröffentlicht: (2023)
Towards Sample-specific Backdoor Attack with Clean Labels via Attribute Trigger
von: Zhu, Mingyan, et al.
Veröffentlicht: (2023)
von: Zhu, Mingyan, et al.
Veröffentlicht: (2023)
Membership Inference over Diffusion-models-based Synthetic Tabular Data
von: Cheng, Peini, et al.
Veröffentlicht: (2025)
von: Cheng, Peini, et al.
Veröffentlicht: (2025)
SMI: Statistical Membership Inference for Reliable Unlearned Model Auditing
von: Sun, Jialong, et al.
Veröffentlicht: (2026)
von: Sun, Jialong, et al.
Veröffentlicht: (2026)
Recalling The Forgotten Class Memberships: Unlearned Models Can Be Noisy Labelers to Leak Privacy
von: Sui, Zhihao, et al.
Veröffentlicht: (2025)
von: Sui, Zhihao, et al.
Veröffentlicht: (2025)
Accuracy-Privacy Trade-off in the Mitigation of Membership Inference Attack in Federated Learning
von: Ahamed, Sayyed Farid, et al.
Veröffentlicht: (2024)
von: Ahamed, Sayyed Farid, et al.
Veröffentlicht: (2024)
Clean-Label Physical Backdoor Attacks with Data Distillation
von: Dao, Thinh, et al.
Veröffentlicht: (2024)
von: Dao, Thinh, et al.
Veröffentlicht: (2024)
How to Backdoor the Knowledge Distillation
von: Wu, Chen, et al.
Veröffentlicht: (2025)
von: Wu, Chen, et al.
Veröffentlicht: (2025)
DCMI: A Differential Calibration Membership Inference Attack Against Retrieval-Augmented Generation
von: Gao, Xinyu, et al.
Veröffentlicht: (2025)
von: Gao, Xinyu, et al.
Veröffentlicht: (2025)
Architectural Backdoors for Within-Batch Data Stealing and Model Inference Manipulation
von: Küchler, Nicolas, et al.
Veröffentlicht: (2025)
von: Küchler, Nicolas, et al.
Veröffentlicht: (2025)
Position: Retire the "Positive Backdoor" Label -- Secret Alignment Requires Strict and Systematic Evaluation
von: Li, Jianwei, et al.
Veröffentlicht: (2026)
von: Li, Jianwei, et al.
Veröffentlicht: (2026)
Heterogeneous Graph Backdoor Attack
von: Chen, Jiawei, et al.
Veröffentlicht: (2025)
von: Chen, Jiawei, et al.
Veröffentlicht: (2025)
Invisible Backdoor Attack Through Singular Value Decomposition
von: Chen, Wenmin, et al.
Veröffentlicht: (2024)
von: Chen, Wenmin, et al.
Veröffentlicht: (2024)
Injecting Universal Jailbreak Backdoors into LLMs in Minutes
von: Chen, Zhuowei, et al.
Veröffentlicht: (2025)
von: Chen, Zhuowei, et al.
Veröffentlicht: (2025)
Winning the MIDST Challenge: New Membership Inference Attacks on Diffusion Models for Tabular Data Synthesis
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
Res-MIA: A Training-Free Resolution-Based Membership Inference Attack on Federated Learning Models
von: Zare, Mohammad, et al.
Veröffentlicht: (2026)
von: Zare, Mohammad, et al.
Veröffentlicht: (2026)
The Sample Complexity of Membership Inference and Privacy Auditing
von: Haghifam, Mahdi, et al.
Veröffentlicht: (2025)
von: Haghifam, Mahdi, et al.
Veröffentlicht: (2025)
Membership Inference Attacks on LLM-based Recommender Systems
von: He, Jiajie, et al.
Veröffentlicht: (2025)
von: He, Jiajie, et al.
Veröffentlicht: (2025)
Structure-Aware Distributed Backdoor Attacks in Federated Learning
von: Jian, Wang, et al.
Veröffentlicht: (2026)
von: Jian, Wang, et al.
Veröffentlicht: (2026)
PSBD: Prediction Shift Uncertainty Unlocks Backdoor Detection
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CLMIA: Membership Inference Attacks via Unsupervised Contrastive Learning
von: Chen, Depeng, et al.
Veröffentlicht: (2024) -
BAFFLE: Hiding Backdoors in Offline Reinforcement Learning Datasets
von: Gong, Chen, et al.
Veröffentlicht: (2022) -
Strategic Sample Selection for Improved Clean-Label Backdoor Attacks in Text Classification
von: Kirci, Onur Alp, et al.
Veröffentlicht: (2025) -
On Membership Inference Attacks in Knowledge Distillation
von: Cui, Ziyao, et al.
Veröffentlicht: (2025) -
Hiding-in-Plain-Sight (HiPS) Attack on CLIP for Targetted Object Removal from Images
von: Daw, Arka, et al.
Veröffentlicht: (2024)