Generated Data with Fake Privacy: Hidden Dangers of Fine-tuning Large Language Models on Generated Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Akkus, Atilla, Aghdam, Masoud Poorghaffar, Li, Mingjie, Chu, Junjie, Backes, Michael, Zhang, Yang, Sav, Sinem |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bridging Local and Federated Data Normalization in Federated Learning: A Privacy-Preserving Approach
von: Coşğun, Melih, et al.
Veröffentlicht: (2025)
von: Coşğun, Melih, et al.
Veröffentlicht: (2025)
ASAP-Bilkent/fake-privacy-finetuning-llms: v1
von: Masoud Poorghaffar Aghdam, et al.
Veröffentlicht: (2025)
von: Masoud Poorghaffar Aghdam, et al.
Veröffentlicht: (2025)
How to Privately Tune Hyperparameters in Federated Learning? Insights from a Benchmark Study
von: Mitic, Natalija, et al.
Veröffentlicht: (2024)
von: Mitic, Natalija, et al.
Veröffentlicht: (2024)
CURE: Privacy-Preserving Split Learning Done Right
von: Kanpak, Halil Ibrahim, et al.
Veröffentlicht: (2024)
von: Kanpak, Halil Ibrahim, et al.
Veröffentlicht: (2024)
Adjacent Words, Divergent Intents: Jailbreaking Large Language Models via Task Concurrency
von: Jiang, Yukun, et al.
Veröffentlicht: (2025)
von: Jiang, Yukun, et al.
Veröffentlicht: (2025)
SERSEM: Selective Entropy-Weighted Scoring for Membership Inference in Code Language Models
von: Dikici, Kıvanç Kuzey, et al.
Veröffentlicht: (2026)
von: Dikici, Kıvanç Kuzey, et al.
Veröffentlicht: (2026)
A Taxonomy of Attacks and Defenses in Split Learning
von: Shabbir, Aqsa, et al.
Veröffentlicht: (2025)
von: Shabbir, Aqsa, et al.
Veröffentlicht: (2025)
Reconstruct Your Previous Conversations! Comprehensively Investigating Privacy Leakage Risks in Conversations with GPT Models
von: Chu, Junjie, et al.
Veröffentlicht: (2024)
von: Chu, Junjie, et al.
Veröffentlicht: (2024)
Navigating the Designs of Privacy-Preserving Fine-tuning for Large Language Models
von: Shi, Haonan, et al.
Veröffentlicht: (2025)
von: Shi, Haonan, et al.
Veröffentlicht: (2025)
PriFFT: Privacy-preserving Federated Fine-tuning of Large Language Models via Hybrid Secret Sharing
von: You, Zhichao, et al.
Veröffentlicht: (2025)
von: You, Zhichao, et al.
Veröffentlicht: (2025)
Efficient Data-Free Model Stealing with Label Diversity
von: Liu, Yiyong, et al.
Veröffentlicht: (2024)
von: Liu, Yiyong, et al.
Veröffentlicht: (2024)
When Understanding Becomes a Risk: Authenticity and Safety Risks in the Emerging Image Generation Paradigm
von: Leng, Ye, et al.
Veröffentlicht: (2026)
von: Leng, Ye, et al.
Veröffentlicht: (2026)
Understanding Data Importance in Machine Learning Attacks: Does Valuable Data Pose Greater Harm?
von: Wen, Rui, et al.
Veröffentlicht: (2024)
von: Wen, Rui, et al.
Veröffentlicht: (2024)
The Hidden Dangers of Public Serverless Repositories: An Empirical Security Assessment
von: Marin, Eduard, et al.
Veröffentlicht: (2025)
von: Marin, Eduard, et al.
Veröffentlicht: (2025)
The Hidden Dangers of Outdated Software: A Cyber Security Perspective
von: Thiyagarajan, Gogulakrishnan, et al.
Veröffentlicht: (2025)
von: Thiyagarajan, Gogulakrishnan, et al.
Veröffentlicht: (2025)
The Hidden Dangers of Browsing AI Agents
von: Mudryi, Mykyta, et al.
Veröffentlicht: (2025)
von: Mudryi, Mykyta, et al.
Veröffentlicht: (2025)
Shake to Leak: Fine-tuning Diffusion Models Can Amplify the Generative Privacy Risk
von: Li, Zhangheng, et al.
Veröffentlicht: (2024)
von: Li, Zhangheng, et al.
Veröffentlicht: (2024)
Peering Behind the Shield: Guardrail Identification in Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
Hidden Data Privacy Breaches in Federated Learning
von: Gong, Xueluan, et al.
Veröffentlicht: (2024)
von: Gong, Xueluan, et al.
Veröffentlicht: (2024)
JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring
von: Chu, Junjie, et al.
Veröffentlicht: (2025)
von: Chu, Junjie, et al.
Veröffentlicht: (2025)
Fine-tuning Large Language Models for DGA and DNS Exfiltration Detection
von: Sayed, Md Abu, et al.
Veröffentlicht: (2024)
von: Sayed, Md Abu, et al.
Veröffentlicht: (2024)
Security and Privacy on Generative Data in AIGC: A Survey
von: Wang, Tao, et al.
Veröffentlicht: (2023)
von: Wang, Tao, et al.
Veröffentlicht: (2023)
Can Reinforcement Learning Unlock the Hidden Dangers in Aligned Large Language Models?
von: Karkevandi, Mohammad Bahrami, et al.
Veröffentlicht: (2024)
von: Karkevandi, Mohammad Bahrami, et al.
Veröffentlicht: (2024)
Fine-grained Manipulation Attacks to Local Differential Privacy Protocols for Data Streams
von: Li, Xinyu, et al.
Veröffentlicht: (2025)
von: Li, Xinyu, et al.
Veröffentlicht: (2025)
On Protecting the Data Privacy of Large Language Models (LLMs): A Survey
von: Yan, Biwei, et al.
Veröffentlicht: (2024)
von: Yan, Biwei, et al.
Veröffentlicht: (2024)
Watermarking LLM-Generated Datasets in Downstream Tasks
von: Liu, Yugeng, et al.
Veröffentlicht: (2025)
von: Liu, Yugeng, et al.
Veröffentlicht: (2025)
SafeSynthDP: Leveraging Large Language Models for Privacy-Preserving Synthetic Data Generation Using Differential Privacy
von: Nahid, Md Mahadi Hasan, et al.
Veröffentlicht: (2024)
von: Nahid, Md Mahadi Hasan, et al.
Veröffentlicht: (2024)
RewardDS: Privacy-Preserving Fine-Tuning for Large Language Models via Reward Driven Data Synthesis
von: Wang, Jianwei, et al.
Veröffentlicht: (2025)
von: Wang, Jianwei, et al.
Veröffentlicht: (2025)
Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
Excessive Reasoning Attack on Reasoning LLMs
von: Si, Wai Man, et al.
Veröffentlicht: (2025)
von: Si, Wai Man, et al.
Veröffentlicht: (2025)
A Review of Privacy Metrics for Privacy-Preserving Synthetic Data Generation
von: Trudslev, Frederik Marinus, et al.
Veröffentlicht: (2025)
von: Trudslev, Frederik Marinus, et al.
Veröffentlicht: (2025)
Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
How Much Do Code Language Models Remember? An Investigation on Data Extraction Attacks before and after Fine-tuning
von: Salerno, Fabio, et al.
Veröffentlicht: (2025)
von: Salerno, Fabio, et al.
Veröffentlicht: (2025)
On the Proactive Generation of Unsafe Images From Text-To-Image Models Using Benign Prompts
von: Wu, Yixin, et al.
Veröffentlicht: (2023)
von: Wu, Yixin, et al.
Veröffentlicht: (2023)
Defeating Cerberus: Concept-Guided Privacy-Leakage Mitigation in Multimodal Language Models
von: Zhang, Boyang, et al.
Veröffentlicht: (2025)
von: Zhang, Boyang, et al.
Veröffentlicht: (2025)
Fine-tuning is Not Fine: Mitigating Backdoor Attacks in GNNs with Limited Clean Data
von: Zhang, Jiale, et al.
Veröffentlicht: (2025)
von: Zhang, Jiale, et al.
Veröffentlicht: (2025)
Pharmacist: Safety Alignment Data Curation for Large Language Models against Harmful Fine-tuning
von: Liu, Guozhi, et al.
Veröffentlicht: (2025)
von: Liu, Guozhi, et al.
Veröffentlicht: (2025)
Robustness Over Time: Understanding Adversarial Examples' Effectiveness on Longitudinal Versions of Large Language Models
von: Liu, Yugeng, et al.
Veröffentlicht: (2023)
von: Liu, Yugeng, et al.
Veröffentlicht: (2023)
User Behavior Analysis in Privacy Protection with Large Language Models: A Study on Privacy Preferences with Limited Data
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
Practical Secure Inference Algorithm for Fine-tuned Large Language Model Based on Fully Homomorphic Encryption
von: Ruoyan, Zhang, et al.
Veröffentlicht: (2025)
von: Ruoyan, Zhang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bridging Local and Federated Data Normalization in Federated Learning: A Privacy-Preserving Approach
von: Coşğun, Melih, et al.
Veröffentlicht: (2025) -
ASAP-Bilkent/fake-privacy-finetuning-llms: v1
von: Masoud Poorghaffar Aghdam, et al.
Veröffentlicht: (2025) -
How to Privately Tune Hyperparameters in Federated Learning? Insights from a Benchmark Study
von: Mitic, Natalija, et al.
Veröffentlicht: (2024) -
CURE: Privacy-Preserving Split Learning Done Right
von: Kanpak, Halil Ibrahim, et al.
Veröffentlicht: (2024) -
Adjacent Words, Divergent Intents: Jailbreaking Large Language Models via Task Concurrency
von: Jiang, Yukun, et al.
Veröffentlicht: (2025)