Privacy-Preserving Instructions for Aligning Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Da, Kairouz, Peter, Oh, Sewoong, Xu, Zheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AirGapAgent: Protecting Privacy-Conscious Conversational Agents
by: Bagdasarian, Eugene, et al.
Published: (2024)
by: Bagdasarian, Eugene, et al.
Published: (2024)
Improved Communication-Privacy Trade-offs in $L_2$ Mean Estimation under Streaming Differential Privacy
by: Chen, Wei-Ning, et al.
Published: (2024)
by: Chen, Wei-Ning, et al.
Published: (2024)
User Inference Attacks on Large Language Models
by: Kandpal, Nikhil, et al.
Published: (2023)
by: Kandpal, Nikhil, et al.
Published: (2023)
Randomization Techniques to Mitigate the Risk of Copyright Infringement
by: Chen, Wei-Ning, et al.
Published: (2024)
by: Chen, Wei-Ning, et al.
Published: (2024)
Privacy-Preserving Parameter-Efficient Fine-Tuning for Large Language Model Services
by: Li, Yansong, et al.
Published: (2023)
by: Li, Yansong, et al.
Published: (2023)
Behavioral Canaries: Auditing Private Retrieved Context Usage in RL Fine-Tuning
by: Chen, Chaoran, et al.
Published: (2026)
by: Chen, Chaoran, et al.
Published: (2026)
PAPILLON: Privacy Preservation from Internet-based and Local Language Model Ensembles
by: Siyan, Li, et al.
Published: (2024)
by: Siyan, Li, et al.
Published: (2024)
Privacy Preserving In-Context-Learning Framework for Large Language Models
by: Bhusal, Bishnu, et al.
Published: (2025)
by: Bhusal, Bishnu, et al.
Published: (2025)
Token-Level Privacy in Large Language Models
by: Harel, Re'em, et al.
Published: (2025)
by: Harel, Re'em, et al.
Published: (2025)
Privacy in Large Language Models: Attacks, Defenses and Future Directions
by: Li, Haoran, et al.
Published: (2023)
by: Li, Haoran, et al.
Published: (2023)
One-shot Empirical Privacy Estimation for Federated Learning
by: Andrew, Galen, et al.
Published: (2023)
by: Andrew, Galen, et al.
Published: (2023)
A False Sense of Privacy: Evaluating Textual Data Sanitization Beyond Surface-level Privacy Leakage
by: Xin, Rui, et al.
Published: (2025)
by: Xin, Rui, et al.
Published: (2025)
Continual Pretraining on Encrypted Synthetic Data for Privacy-Preserving LLMs
by: Liu, Honghao, et al.
Published: (2026)
by: Liu, Honghao, et al.
Published: (2026)
Secure Stateful Aggregation: A Practical Protocol with Applications in Differentially-Private Federated Learning
by: Ball, Marshall, et al.
Published: (2024)
by: Ball, Marshall, et al.
Published: (2024)
MAPLE: Metadata Augmented Private Language Evolution
by: Chien, Eli, et al.
Published: (2026)
by: Chien, Eli, et al.
Published: (2026)
MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents
by: Chen, Yining, et al.
Published: (2026)
by: Chen, Yining, et al.
Published: (2026)
BitBypass: A New Direction in Jailbreaking Aligned Large Language Models with Bitstream Camouflage
by: Nakka, Kalyan, et al.
Published: (2025)
by: Nakka, Kalyan, et al.
Published: (2025)
GoldCoin: Grounding Large Language Models in Privacy Laws via Contextual Integrity Theory
by: Fan, Wei, et al.
Published: (2024)
by: Fan, Wei, et al.
Published: (2024)
Preserving Privacy in Large Language Models: A Survey on Current Threats and Solutions
by: Miranda, Michele, et al.
Published: (2024)
by: Miranda, Michele, et al.
Published: (2024)
Quantifying Association Capabilities of Large Language Models and Its Implications on Privacy Leakage
by: Shao, Hanyin, et al.
Published: (2023)
by: Shao, Hanyin, et al.
Published: (2023)
ForgeDAN: An Evolutionary Framework for Jailbreaking Aligned Large Language Models
by: Cheng, Siyang, et al.
Published: (2025)
by: Cheng, Siyang, et al.
Published: (2025)
SecFormer: Fast and Accurate Privacy-Preserving Inference for Transformer Models via SMPC
by: Luo, Jinglong, et al.
Published: (2024)
by: Luo, Jinglong, et al.
Published: (2024)
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
by: Yan, Jun, et al.
Published: (2023)
by: Yan, Jun, et al.
Published: (2023)
AgentBnB: A Browser-Based Cybersecurity Tabletop Exercise with Large Language Model Support and Retrieval-Aligned Scaffolding
by: Anwar, Arman, et al.
Published: (2025)
by: Anwar, Arman, et al.
Published: (2025)
FakeZero: Real-Time, Privacy-Preserving Misinformation Detection for Facebook and X
by: Essahli, Soufiane, et al.
Published: (2025)
by: Essahli, Soufiane, et al.
Published: (2025)
QueryAttack: Jailbreaking Aligned Large Language Models Using Structured Non-natural Query Language
by: Zou, Qingsong, et al.
Published: (2025)
by: Zou, Qingsong, et al.
Published: (2025)
Privacy-Preserving Retrieval-Augmented Generation with Differential Privacy
by: Koga, Tatsuki, et al.
Published: (2024)
by: Koga, Tatsuki, et al.
Published: (2024)
The Model's Language Matters: A Comparative Privacy Analysis of LLMs
by: Mishra, Abhishek K., et al.
Published: (2025)
by: Mishra, Abhishek K., et al.
Published: (2025)
GenBreak: Red Teaming Text-to-Image Generators Using Large Language Models
by: Wang, Zilong, et al.
Published: (2025)
by: Wang, Zilong, et al.
Published: (2025)
Instructional Fingerprinting of Large Language Models
by: Xu, Jiashu, et al.
Published: (2024)
by: Xu, Jiashu, et al.
Published: (2024)
Advancing Adversarial Suffix Transfer Learning on Aligned Large Language Models
by: Liu, Hongfu, et al.
Published: (2024)
by: Liu, Hongfu, et al.
Published: (2024)
CCJA: Context-Coherent Jailbreak Attack for Aligned Large Language Models
by: Zhou, Guanghao, et al.
Published: (2025)
by: Zhou, Guanghao, et al.
Published: (2025)
Safely Learning with Private Data: A Federated Learning Framework for Large Language Model
by: Zheng, JiaYing, et al.
Published: (2024)
by: Zheng, JiaYing, et al.
Published: (2024)
Security and Privacy Challenges of Large Language Models: A Survey
by: Das, Badhan Chandra, et al.
Published: (2024)
by: Das, Badhan Chandra, et al.
Published: (2024)
LLMAtKGE: Large Language Models as Explainable Attackers against Knowledge Graph Embeddings
by: Li, Ting, et al.
Published: (2025)
by: Li, Ting, et al.
Published: (2025)
Synthetic Query Generation for Privacy-Preserving Deep Retrieval Systems using Differentially Private Language Models
by: Carranza, Aldo Gael, et al.
Published: (2023)
by: Carranza, Aldo Gael, et al.
Published: (2023)
Majority Bit-Aware Watermarking For Large Language Models
by: Xu, Jiahao, et al.
Published: (2025)
by: Xu, Jiahao, et al.
Published: (2025)
Privacy-Preserving Transformers: SwiftKey's Differential Privacy Implementation
by: Abouelenin, Abdelrahman, et al.
Published: (2025)
by: Abouelenin, Abdelrahman, et al.
Published: (2025)
Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
by: Xu, Jiashu, et al.
Published: (2023)
by: Xu, Jiashu, et al.
Published: (2023)
Can Reinforcement Learning Unlock the Hidden Dangers in Aligned Large Language Models?
by: Karkevandi, Mohammad Bahrami, et al.
Published: (2024)
by: Karkevandi, Mohammad Bahrami, et al.
Published: (2024)
Similar Items
-
AirGapAgent: Protecting Privacy-Conscious Conversational Agents
by: Bagdasarian, Eugene, et al.
Published: (2024) -
Improved Communication-Privacy Trade-offs in $L_2$ Mean Estimation under Streaming Differential Privacy
by: Chen, Wei-Ning, et al.
Published: (2024) -
User Inference Attacks on Large Language Models
by: Kandpal, Nikhil, et al.
Published: (2023) -
Randomization Techniques to Mitigate the Risk of Copyright Infringement
by: Chen, Wei-Ning, et al.
Published: (2024) -
Privacy-Preserving Parameter-Efficient Fine-Tuning for Large Language Model Services
by: Li, Yansong, et al.
Published: (2023)