Improved Algorithms for Differentially Private Language Model Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Keyu, Tang, Hao, Liu, Qinglin, Xu, Yizhao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Differentially Private Preference Data Synthesis for Large Language Model Alignment
by: Gao, Fengyu, et al.
Published: (2026)
by: Gao, Fengyu, et al.
Published: (2026)
DR-Encoder: Encode Low-rank Gradients with Random Prior for Large Language Models Differentially Privately
by: Wu, Huiwen, et al.
Published: (2024)
by: Wu, Huiwen, et al.
Published: (2024)
Differentially Private Model Merging
by: Yin, Qichuan, et al.
Published: (2026)
by: Yin, Qichuan, et al.
Published: (2026)
Reconstruction of Differentially Private Text Sanitization via Large Language Models
by: Pang, Shuchao, et al.
Published: (2024)
by: Pang, Shuchao, et al.
Published: (2024)
Can Differentially Private Fine-tuning LLMs Protect Against Privacy Attacks?
by: Du, Hao, et al.
Published: (2025)
by: Du, Hao, et al.
Published: (2025)
Differentially Private Deep Model-Based Reinforcement Learning
by: Rio, Alexandre, et al.
Published: (2024)
by: Rio, Alexandre, et al.
Published: (2024)
Privacy-Preserving Federated Learning with Differentially Private Hyperdimensional Computing
by: Piran, Fardin Jalil, et al.
Published: (2024)
by: Piran, Fardin Jalil, et al.
Published: (2024)
Differentially Private Worst-group Risk Minimization
by: Zhou, Xinyu, et al.
Published: (2024)
by: Zhou, Xinyu, et al.
Published: (2024)
Too Good to be True? Turn Any Model Differentially Private With DP-Weights
by: Zagardo, David
Published: (2024)
by: Zagardo, David
Published: (2024)
Clustering and Median Aggregation Improve Differentially Private Inference
by: Amin, Kareem, et al.
Published: (2025)
by: Amin, Kareem, et al.
Published: (2025)
Differentially Private Iterative Screening Rules for Linear Regression
by: Khanna, Amol, et al.
Published: (2025)
by: Khanna, Amol, et al.
Published: (2025)
Differentially Private In-Context Learning with Nearest Neighbor Search
by: Koskela, Antti, et al.
Published: (2025)
by: Koskela, Antti, et al.
Published: (2025)
Evaluating Differentially Private Generation of Domain-Specific Text
by: Sun, Yidan, et al.
Published: (2025)
by: Sun, Yidan, et al.
Published: (2025)
Provable Differentially Private Computation of the Cross-Attention Mechanism
by: Ke, Yekun, et al.
Published: (2024)
by: Ke, Yekun, et al.
Published: (2024)
DC-SGD: Differentially Private SGD with Dynamic Clipping through Gradient Norm Distribution Estimation
by: Wei, Chengkun, et al.
Published: (2025)
by: Wei, Chengkun, et al.
Published: (2025)
dpmm: Differentially Private Marginal Models, a Library for Synthetic Tabular Data Generation
by: Mahiou, Sofiane, et al.
Published: (2025)
by: Mahiou, Sofiane, et al.
Published: (2025)
Differentially Private Synthetic Data Release for Topics API Outputs
by: Dick, Travis, et al.
Published: (2025)
by: Dick, Travis, et al.
Published: (2025)
Differentially Private Block-wise Gradient Shuffle for Deep Learning
by: Zagardo, David
Published: (2024)
by: Zagardo, David
Published: (2024)
Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
by: Gao, Fengyu, et al.
Published: (2024)
by: Gao, Fengyu, et al.
Published: (2024)
Differentially Private Synthetic Data Generation Using Context-Aware GANs
by: Kotal, Anantaa, et al.
Published: (2025)
by: Kotal, Anantaa, et al.
Published: (2025)
Smooth Sensitivity for Learning Differentially-Private yet Accurate Rule Lists
by: Ly, Timothée, et al.
Published: (2024)
by: Ly, Timothée, et al.
Published: (2024)
DP-TabICL: In-Context Learning with Differentially Private Tabular Data
by: Carey, Alycia N., et al.
Published: (2024)
by: Carey, Alycia N., et al.
Published: (2024)
Differentially Private Publication of Electricity Time Series Data in Smart Grids
by: Shaham, Sina, et al.
Published: (2024)
by: Shaham, Sina, et al.
Published: (2024)
Clients Collaborate: Flexible Differentially Private Federated Learning with Guaranteed Improvement of Utility-Privacy Trade-off
by: Li, Yuecheng, et al.
Published: (2024)
by: Li, Yuecheng, et al.
Published: (2024)
GraMFedDHAR: Graph Based Multimodal Differentially Private Federated HAR
by: Halder, Labani, et al.
Published: (2025)
by: Halder, Labani, et al.
Published: (2025)
PACZero: PAC-Private Fine-Tuning of Language Models via Sign Quantization
by: Ertan, Murat Bilgehan, et al.
Published: (2026)
by: Ertan, Murat Bilgehan, et al.
Published: (2026)
RQP-SGD: Differential Private Machine Learning through Noisy SGD and Randomized Quantization
by: Feng, Ce, et al.
Published: (2024)
by: Feng, Ce, et al.
Published: (2024)
Jailbreaking and Mitigation of Vulnerabilities in Large Language Models
by: Peng, Benji, et al.
Published: (2024)
by: Peng, Benji, et al.
Published: (2024)
Differentially Private Distributed Inference
by: Papachristou, Marios, et al.
Published: (2024)
by: Papachristou, Marios, et al.
Published: (2024)
On Differentially Private String Distances
by: Hu, Jerry Yao-Chieh, et al.
Published: (2024)
by: Hu, Jerry Yao-Chieh, et al.
Published: (2024)
Pharmacist: Safety Alignment Data Curation for Large Language Models against Harmful Fine-tuning
by: Liu, Guozhi, et al.
Published: (2025)
by: Liu, Guozhi, et al.
Published: (2025)
On Mitigating the Utility-Loss in Differentially Private Learning: A new Perspective by a Geometrically Inspired Kernel Approach
by: Kumar, Mohit, et al.
Published: (2023)
by: Kumar, Mohit, et al.
Published: (2023)
A Robust Framework for Secure Cardiovascular Risk Prediction: An Architectural Case Study of Differentially Private Federated Learning
by: Tertulino, Rodrigo, et al.
Published: (2026)
by: Tertulino, Rodrigo, et al.
Published: (2026)
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
by: Panda, Ashwinee, et al.
Published: (2022)
by: Panda, Ashwinee, et al.
Published: (2022)
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
by: Ngong, Ivoline C., et al.
Published: (2024)
by: Ngong, Ivoline C., et al.
Published: (2024)
Linearizing Models for Efficient yet Robust Private Inference
by: Sarkar, Sreetama, et al.
Published: (2024)
by: Sarkar, Sreetama, et al.
Published: (2024)
Safety Tax: Safety Alignment Makes Your Large Reasoning Models Less Reasonable
by: Huang, Tiansheng, et al.
Published: (2025)
by: Huang, Tiansheng, et al.
Published: (2025)
Differentially Private Reinforcement Learning with Self-Play
by: Qiao, Dan, et al.
Published: (2024)
by: Qiao, Dan, et al.
Published: (2024)
Improved Membership Inference Attacks Against Language Classification Models
by: Shachor, Shlomit, et al.
Published: (2023)
by: Shachor, Shlomit, et al.
Published: (2023)
Improved Large Language Model Jailbreak Detection via Pretrained Embeddings
by: Galinkin, Erick, et al.
Published: (2024)
by: Galinkin, Erick, et al.
Published: (2024)
Similar Items
-
Differentially Private Preference Data Synthesis for Large Language Model Alignment
by: Gao, Fengyu, et al.
Published: (2026) -
DR-Encoder: Encode Low-rank Gradients with Random Prior for Large Language Models Differentially Privately
by: Wu, Huiwen, et al.
Published: (2024) -
Differentially Private Model Merging
by: Yin, Qichuan, et al.
Published: (2026) -
Reconstruction of Differentially Private Text Sanitization via Large Language Models
by: Pang, Shuchao, et al.
Published: (2024) -
Can Differentially Private Fine-tuning LLMs Protect Against Privacy Attacks?
by: Du, Hao, et al.
Published: (2025)