What Really is a Member? Discrediting Membership Inference via Poisoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mangaokar, Neal, Hooda, Ashish, Li, Zhuohang, Malin, Bradley A., Fawaz, Kassem, Jha, Somesh, Prakash, Atul, Chowdhury, Amrita Roy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PRP: Propagating Universal Perturbations to Attack Large Language Model Guard-Rails
von: Mangaokar, Neal, et al.
Veröffentlicht: (2024)
von: Mangaokar, Neal, et al.
Veröffentlicht: (2024)
PolicyLR: A Logic Representation For Privacy Policies
von: Hooda, Ashish, et al.
Veröffentlicht: (2024)
von: Hooda, Ashish, et al.
Veröffentlicht: (2024)
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
von: Feng, Ryan, et al.
Veröffentlicht: (2022)
von: Feng, Ryan, et al.
Veröffentlicht: (2022)
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
Robustness of Locally Differentially Private Graph Analysis Against Poisoning
von: Imola, Jacob, et al.
Veröffentlicht: (2022)
von: Imola, Jacob, et al.
Veröffentlicht: (2022)
Functional Homotopy: Smoothing Discrete Optimization via Continuous Parameters for LLM Jailbreak Attacks
von: Wang, Zi, et al.
Veröffentlicht: (2024)
von: Wang, Zi, et al.
Veröffentlicht: (2024)
Confused ChatGPT: Cross-App Context Poisoning via First-Party APIs
von: Wang, Chao, et al.
Veröffentlicht: (2026)
von: Wang, Chao, et al.
Veröffentlicht: (2026)
Bayes-Nash Generative Privacy Against Membership Inference Attacks
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
Pr$εε$mpt: Sanitizing Sensitive Prompts for LLMs
von: Chowdhury, Amrita Roy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Amrita Roy, et al.
Veröffentlicht: (2025)
Taming Data Challenges in ML-based Security Tasks Using Generative AI
von: Kanchi, Shravya, et al.
Veröffentlicht: (2025)
von: Kanchi, Shravya, et al.
Veröffentlicht: (2025)
Harmful Terms and Where to Find Them: Measuring and Modeling Unfavorable Financial Terms and Conditions in Shopping Websites at Scale
von: Tsai, Elisa, et al.
Veröffentlicht: (2025)
von: Tsai, Elisa, et al.
Veröffentlicht: (2025)
SEA: Shareable and Explainable Attribution for Query-based Black-box Attacks
von: Gao, Yue, et al.
Veröffentlicht: (2023)
von: Gao, Yue, et al.
Veröffentlicht: (2023)
Prediction with Expert Advice under Local Differential Privacy
von: Jacobsen, Ben, et al.
Veröffentlicht: (2025)
von: Jacobsen, Ben, et al.
Veröffentlicht: (2025)
Private Continual Counting of Unbounded Streams
von: Jacobsen, Ben, et al.
Veröffentlicht: (2025)
von: Jacobsen, Ben, et al.
Veröffentlicht: (2025)
Metric Differential Privacy at the User-Level Via the Earth Mover's Distance
von: Imola, Jacob, et al.
Veröffentlicht: (2024)
von: Imola, Jacob, et al.
Veröffentlicht: (2024)
Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025)
Analyzing Inference Privacy Risks Through Gradients in Machine Learning
von: Li, Zhuohang, et al.
Veröffentlicht: (2024)
von: Li, Zhuohang, et al.
Veröffentlicht: (2024)
Dependency-Aware Privacy for Multi-turn Agents
von: Anshumaan, Divyam, et al.
Veröffentlicht: (2026)
von: Anshumaan, Divyam, et al.
Veröffentlicht: (2026)
Systems Security Foundations for Agentic Computing
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2025)
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2025)
Personalizing Agent Privacy Decisions via Logical Entailment
von: Flemings, James, et al.
Veröffentlicht: (2025)
von: Flemings, James, et al.
Veröffentlicht: (2025)
Piquant$\varepsilon$: Private Quantile Estimation in the Two-Server Model
von: Keller, Hannah, et al.
Veröffentlicht: (2025)
von: Keller, Hannah, et al.
Veröffentlicht: (2025)
Text-Based Personas for Simulating User Privacy Decisions
von: Fawaz, Kassem, et al.
Veröffentlicht: (2026)
von: Fawaz, Kassem, et al.
Veröffentlicht: (2026)
When Reasoning Leaks Membership: Membership Inference Attack on Black-box Large Reasoning Models
von: Hu, Ruihan, et al.
Veröffentlicht: (2026)
von: Hu, Ruihan, et al.
Veröffentlicht: (2026)
What Hard Tokens Reveal: Exploiting Low-confidence Tokens for Membership Inference Attacks against Large Language Models
von: Jawad, Md Tasnim, et al.
Veröffentlicht: (2026)
von: Jawad, Md Tasnim, et al.
Veröffentlicht: (2026)
Membership Inference Attacks on Vision-Language-Action Models
von: Peng, Yuefeng, et al.
Veröffentlicht: (2026)
von: Peng, Yuefeng, et al.
Veröffentlicht: (2026)
Dynamic Probabilistic Noise Injection for Membership Inference Defense
von: Forough, Javad, et al.
Veröffentlicht: (2025)
von: Forough, Javad, et al.
Veröffentlicht: (2025)
Agent Security is a Systems Problem
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2026)
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2026)
WebSP-Eval: Evaluating Web Agents on Website Security and Privacy Tasks
von: Ramesh, Guruprasad Viswanathan, et al.
Veröffentlicht: (2026)
von: Ramesh, Guruprasad Viswanathan, et al.
Veröffentlicht: (2026)
Find a Scapegoat: Poisoning Membership Inference Attack and Defense to Federated Learning
von: Mo, Wenjin, et al.
Veröffentlicht: (2025)
von: Mo, Wenjin, et al.
Veröffentlicht: (2025)
Imitative Membership Inference Attack
von: Du, Yuntao, et al.
Veröffentlicht: (2025)
von: Du, Yuntao, et al.
Veröffentlicht: (2025)
Neighborhood Blending: A Lightweight Inference-Time Defense Against Membership Inference Attacks
von: Zafar, Osama, et al.
Veröffentlicht: (2026)
von: Zafar, Osama, et al.
Veröffentlicht: (2026)
White-box Membership Inference Attacks against Diffusion Models
von: Pang, Yan, et al.
Veröffentlicht: (2023)
von: Pang, Yan, et al.
Veröffentlicht: (2023)
Membership Inference Attacks and Defenses in Federated Learning: A Survey
von: Bai, Li, et al.
Veröffentlicht: (2024)
von: Bai, Li, et al.
Veröffentlicht: (2024)
Membership Inference Attacks Against Video Large Language Models
von: Song, Wei, et al.
Veröffentlicht: (2026)
von: Song, Wei, et al.
Veröffentlicht: (2026)
Privacy Leaks by Adversaries: Adversarial Iterations for Membership Inference Attack
von: Xue, Jing, et al.
Veröffentlicht: (2025)
von: Xue, Jing, et al.
Veröffentlicht: (2025)
Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface
von: Labunets, Andrey, et al.
Veröffentlicht: (2025)
von: Labunets, Andrey, et al.
Veröffentlicht: (2025)
ARMOR: Aligning Secure and Safe Large Language Models via Meticulous Reasoning
von: Zhao, Zhengyue, et al.
Veröffentlicht: (2025)
von: Zhao, Zhengyue, et al.
Veröffentlicht: (2025)
Evaluating the Defense Potential of Machine Unlearning against Membership Inference Attacks
von: Tsiolakis, Theodoros, et al.
Veröffentlicht: (2025)
von: Tsiolakis, Theodoros, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PRP: Propagating Universal Perturbations to Attack Large Language Model Guard-Rails
von: Mangaokar, Neal, et al.
Veröffentlicht: (2024) -
PolicyLR: A Logic Representation For Privacy Policies
von: Hooda, Ashish, et al.
Veröffentlicht: (2024) -
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
von: Feng, Ryan, et al.
Veröffentlicht: (2022) -
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026) -
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)