Position: Privacy Is Not Just Memorization!
Fuente:
arXiv
Salvato in:
| Autori principali: | Mireshghallah, Niloofar, Li, Tianshi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024)
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024)
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
di: Naseh, Ali, et al.
Pubblicazione: (2025)
di: Naseh, Ali, et al.
Pubblicazione: (2025)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
A False Sense of Privacy: Evaluating Textual Data Sanitization Beyond Surface-level Privacy Leakage
di: Xin, Rui, et al.
Pubblicazione: (2025)
di: Xin, Rui, et al.
Pubblicazione: (2025)
Unlocking Memorization in Large Language Models with Dynamic Soft Prompting
di: Wang, Zhepeng, et al.
Pubblicazione: (2024)
di: Wang, Zhepeng, et al.
Pubblicazione: (2024)
PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action
di: Shao, Yijia, et al.
Pubblicazione: (2024)
di: Shao, Yijia, et al.
Pubblicazione: (2024)
Trustworthy AI: Safety, Bias, and Privacy -- A Survey
di: Fang, Xingli, et al.
Pubblicazione: (2025)
di: Fang, Xingli, et al.
Pubblicazione: (2025)
Do Phone-Use Agents Respect Your Privacy?
di: Tang, Zhengyang, et al.
Pubblicazione: (2026)
di: Tang, Zhengyang, et al.
Pubblicazione: (2026)
Unveiling Privacy, Memorization, and Input Curvature Links
di: Ravikumar, Deepak, et al.
Pubblicazione: (2024)
di: Ravikumar, Deepak, et al.
Pubblicazione: (2024)
Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models
di: Ruzzetti, Elena Sofia, et al.
Pubblicazione: (2025)
di: Ruzzetti, Elena Sofia, et al.
Pubblicazione: (2025)
Fine-Tuning Language Models with Differential Privacy through Adaptive Noise Allocation
di: Li, Xianzhi, et al.
Pubblicazione: (2024)
di: Li, Xianzhi, et al.
Pubblicazione: (2024)
EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
di: Liao, Zeyi, et al.
Pubblicazione: (2024)
di: Liao, Zeyi, et al.
Pubblicazione: (2024)
Learnable Privacy Neurons Localization in Language Models
di: Chen, Ruizhe, et al.
Pubblicazione: (2024)
di: Chen, Ruizhe, et al.
Pubblicazione: (2024)
Generalization or Memorization: Data Contamination and Trustworthy Evaluation for Large Language Models
di: Dong, Yihong, et al.
Pubblicazione: (2024)
di: Dong, Yihong, et al.
Pubblicazione: (2024)
Beyond Gradient and Priors in Privacy Attacks: Leveraging Pooler Layer Inputs of Language Models in Federated Learning
di: Li, Jianwei, et al.
Pubblicazione: (2023)
di: Li, Jianwei, et al.
Pubblicazione: (2023)
How Private is Your Attention? Bridging Privacy with In-Context Learning
di: Bonnerjee, Soham, et al.
Pubblicazione: (2025)
di: Bonnerjee, Soham, et al.
Pubblicazione: (2025)
HARMONIC: Harnessing LLMs for Tabular Data Synthesis and Privacy Protection
di: Wang, Yuxin, et al.
Pubblicazione: (2024)
di: Wang, Yuxin, et al.
Pubblicazione: (2024)
Just Do It!? Computer-Use Agents Exhibit Blind Goal-Directedness
di: Shayegani, Erfan, et al.
Pubblicazione: (2025)
di: Shayegani, Erfan, et al.
Pubblicazione: (2025)
Preserving Privacy in Large Language Models: A Survey on Current Threats and Solutions
di: Miranda, Michele, et al.
Pubblicazione: (2024)
di: Miranda, Michele, et al.
Pubblicazione: (2024)
FedMentor: Domain-Aware Differential Privacy for Heterogeneous Federated LLMs in Mental Health
di: Sarwar, Nobin, et al.
Pubblicazione: (2025)
di: Sarwar, Nobin, et al.
Pubblicazione: (2025)
PBa-LLM: Privacy- and Bias-aware NLP using Named-Entity Recognition (NER)
di: Mancera, Gonzalo, et al.
Pubblicazione: (2025)
di: Mancera, Gonzalo, et al.
Pubblicazione: (2025)
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks
di: Chen, Xiaoyi, et al.
Pubblicazione: (2023)
di: Chen, Xiaoyi, et al.
Pubblicazione: (2023)
DP-MemArc: Differential Privacy Transfer Learning for Memory Efficient Language Models
di: Liu, Yanming, et al.
Pubblicazione: (2024)
di: Liu, Yanming, et al.
Pubblicazione: (2024)
Privacy-Preserving Data Deduplication for Enhancing Federated Learning of Language Models (Extended Version)
di: Abadi, Aydin, et al.
Pubblicazione: (2024)
di: Abadi, Aydin, et al.
Pubblicazione: (2024)
BinaryShield: Cross-Service Threat Intelligence in LLM Services using Privacy-Preserving Fingerprints
di: Gill, Waris, et al.
Pubblicazione: (2025)
di: Gill, Waris, et al.
Pubblicazione: (2025)
Reconstruct Your Previous Conversations! Comprehensively Investigating Privacy Leakage Risks in Conversations with GPT Models
di: Chu, Junjie, et al.
Pubblicazione: (2024)
di: Chu, Junjie, et al.
Pubblicazione: (2024)
MetaMoE: Diversity-Aware Proxy Selection for Privacy-Preserving Mixture-of-Experts Unification
di: Jiang, Weisen, et al.
Pubblicazione: (2026)
di: Jiang, Weisen, et al.
Pubblicazione: (2026)
IncogniText: Privacy-enhancing Conditional Text Anonymization via LLM-based Private Attribute Randomization
di: Frikha, Ahmed, et al.
Pubblicazione: (2024)
di: Frikha, Ahmed, et al.
Pubblicazione: (2024)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
Optimizing the Privacy-Utility Balance using Synthetic Data and Configurable Perturbation Pipelines
di: Sharma, Anantha, et al.
Pubblicazione: (2025)
di: Sharma, Anantha, et al.
Pubblicazione: (2025)
Clio: Privacy-Preserving Insights into Real-World AI Use
di: Tamkin, Alex, et al.
Pubblicazione: (2024)
di: Tamkin, Alex, et al.
Pubblicazione: (2024)
The Resurgence of GCG Adversarial Attacks on Large Language Models
di: Tan, Yuting, et al.
Pubblicazione: (2025)
di: Tan, Yuting, et al.
Pubblicazione: (2025)
Safety Alignment Can Be Not Superficial With Explicit Safety Signals
di: Li, Jianwei, et al.
Pubblicazione: (2025)
di: Li, Jianwei, et al.
Pubblicazione: (2025)
Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models
di: Li, Xiao, et al.
Pubblicazione: (2024)
di: Li, Xiao, et al.
Pubblicazione: (2024)
SentinelLMs: Encrypted Input Adaptation and Fine-tuning of Language Models for Private and Secure Inference
di: Mishra, Abhijit, et al.
Pubblicazione: (2023)
di: Mishra, Abhijit, et al.
Pubblicazione: (2023)
Revealing Weaknesses in Text Watermarking Through Self-Information Rewrite Attacks
di: Cheng, Yixin, et al.
Pubblicazione: (2025)
di: Cheng, Yixin, et al.
Pubblicazione: (2025)
Scalable Defense against In-the-wild Jailbreaking Attacks with Safety Context Retrieval
di: Chen, Taiye, et al.
Pubblicazione: (2025)
di: Chen, Taiye, et al.
Pubblicazione: (2025)
Federated In-Context LLM Agent Learning
di: Wu, Panlong, et al.
Pubblicazione: (2024)
di: Wu, Panlong, et al.
Pubblicazione: (2024)
On Evaluating The Performance of Watermarked Machine-Generated Texts Under Adversarial Attacks
di: Liu, Zesen, et al.
Pubblicazione: (2024)
di: Liu, Zesen, et al.
Pubblicazione: (2024)
Preference Tuning For Toxicity Mitigation Generalizes Across Languages
di: Li, Xiaochen, et al.
Pubblicazione: (2024)
di: Li, Xiaochen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024) -
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
di: Naseh, Ali, et al.
Pubblicazione: (2025) -
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023) -
A False Sense of Privacy: Evaluating Textual Data Sanitization Beyond Surface-level Privacy Leakage
di: Xin, Rui, et al.
Pubblicazione: (2025) -
Unlocking Memorization in Large Language Models with Dynamic Soft Prompting
di: Wang, Zhepeng, et al.
Pubblicazione: (2024)