PrivUn: Unveiling Latent Ripple Effects and Shallow Forgetting in Privacy Unlearning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xiaoyi, Wang, Haoyuan, Tang, Siyuan, Liu, Sijia, Su, Liya, Wang, XiaoFeng, Tang, Haixu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks
von: Chen, Xiaoyi, et al.
Veröffentlicht: (2023)
von: Chen, Xiaoyi, et al.
Veröffentlicht: (2023)
Selective Amnesia: On Efficient, High-Fidelity and Blind Suppression of Backdoor Effects in Trojaned Machine Learning Models
von: Zhu, Rui, et al.
Veröffentlicht: (2022)
von: Zhu, Rui, et al.
Veröffentlicht: (2022)
Forgetting to Forget: Attention Sink as A Gateway for Backdooring LLM Unlearning
von: Shang, Bingqi, et al.
Veröffentlicht: (2025)
von: Shang, Bingqi, et al.
Veröffentlicht: (2025)
Beyond Local vs. External: A Game-Theoretic Framework for Trustworthy Knowledge Acquisition
von: Yao, Rujing, et al.
Veröffentlicht: (2026)
von: Yao, Rujing, et al.
Veröffentlicht: (2026)
Challenging Forgets: Unveiling the Worst-Case Forget Sets in Machine Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
Efficiently Quantifying and Mitigating Ripple Effects in Model Editing
von: Wang, Jianchen, et al.
Veröffentlicht: (2024)
von: Wang, Jianchen, et al.
Veröffentlicht: (2024)
From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks
von: Zhang, Zhexin, et al.
Veröffentlicht: (2024)
von: Zhang, Zhexin, et al.
Veröffentlicht: (2024)
To Forget or Not? Towards Practical Knowledge Unlearning for Large Language Models
von: Tian, Bozhong, et al.
Veröffentlicht: (2024)
von: Tian, Bozhong, et al.
Veröffentlicht: (2024)
Reversing the Forget-Retain Objectives: An Efficient LLM Unlearning Framework from Logit Difference
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
LLM-Enhanced Software Patch Localization
von: Yu, Jinhong, et al.
Veröffentlicht: (2024)
von: Yu, Jinhong, et al.
Veröffentlicht: (2024)
PrivAct: Internalizing Contextual Privacy Preservation via Multi-Agent Preference Training
von: Cheng, Yuhan, et al.
Veröffentlicht: (2026)
von: Cheng, Yuhan, et al.
Veröffentlicht: (2026)
UIPE: Enhancing LLM Unlearning by Removing Knowledge Related to Forgetting Targets
von: Wang, Wenyu, et al.
Veröffentlicht: (2025)
von: Wang, Wenyu, et al.
Veröffentlicht: (2025)
PrivGemo: Privacy-Preserving Dual-Tower Graph Retrieval for Empowering LLM Reasoning with Memory Augmentation
von: Tan, Xingyu, et al.
Veröffentlicht: (2026)
von: Tan, Xingyu, et al.
Veröffentlicht: (2026)
Picachv: Formally Verified Data Use Policy Enforcement for Secure Data Analytics
von: Chen, Haobin Hiroki, et al.
Veröffentlicht: (2025)
von: Chen, Haobin Hiroki, et al.
Veröffentlicht: (2025)
Privacy Ripple Effects from Adding or Removing Personal Information in Language Model Training
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2025)
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2025)
Clues in Tweets: Twitter-Guided Discovery and Analysis of SMS Spam
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
Selective Forgetting: Advancing Machine Unlearning Techniques and Evaluation in Language Models
von: Wang, Lingzhi, et al.
Veröffentlicht: (2024)
von: Wang, Lingzhi, et al.
Veröffentlicht: (2024)
PrivLM-Bench: A Multi-level Privacy Evaluation Benchmark for Language Models
von: Li, Haoran, et al.
Veröffentlicht: (2023)
von: Li, Haoran, et al.
Veröffentlicht: (2023)
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current Benchmarks
von: Pal, Soumyadeep, et al.
Veröffentlicht: (2025)
von: Pal, Soumyadeep, et al.
Veröffentlicht: (2025)
DP-MGTD: Privacy-Preserving Machine-Generated Text Detection via Adaptive Differentially Private Entity Sanitization
von: Wang, Lionel Z., et al.
Veröffentlicht: (2026)
von: Wang, Lionel Z., et al.
Veröffentlicht: (2026)
Answer When Needed, Forget When Not: Language Models Pretend to Forget via In-Context Knowledge Unlearning
von: Takashiro, Shota, et al.
Veröffentlicht: (2024)
von: Takashiro, Shota, et al.
Veröffentlicht: (2024)
DPAdapter: Improving Differentially Private Deep Learning through Noise Tolerance Pre-training
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
LLM Unlearning via Loss Adjustment with Only Forget Data
von: Wang, Yaxuan, et al.
Veröffentlicht: (2024)
von: Wang, Yaxuan, et al.
Veröffentlicht: (2024)
Improve Decoding Factuality by Token-wise Cross Layer Entropy of Large Language Models
von: Wu, Jialiang, et al.
Veröffentlicht: (2025)
von: Wu, Jialiang, et al.
Veröffentlicht: (2025)
Textual Unlearning Gives a False Sense of Unlearning
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
Not Every Token Needs Forgetting: Selective Unlearning to Limit Change in Utility in Large Language Model Unlearning
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
FIT to Forget: Robust Continual Unlearning for Large Language Models
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2026)
MedPriv-Bench: Benchmarking the Privacy-Utility Trade-off of Large Language Models in Medical Open-End Question Answering
von: Guan, Shaowei, et al.
Veröffentlicht: (2026)
von: Guan, Shaowei, et al.
Veröffentlicht: (2026)
iCLP: Large Language Model Reasoning with Implicit Cognition Latent Planning
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis
von: Ma, Weitao, et al.
Veröffentlicht: (2024)
von: Ma, Weitao, et al.
Veröffentlicht: (2024)
Recover-to-Forget: Gradient Reconstruction from LoRA for Efficient LLM Unlearning
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
When Machine Unlearning Meets Retrieval-Augmented Generation (RAG): Keep Secret or Forget Knowledge?
von: Wang, Shang, et al.
Veröffentlicht: (2024)
von: Wang, Shang, et al.
Veröffentlicht: (2024)
On Effects of Steering Latent Representation for Large Language Model Unlearning
von: Huu-Tien, Dang, et al.
Veröffentlicht: (2024)
von: Huu-Tien, Dang, et al.
Veröffentlicht: (2024)
LLM Unlearning Under the Microscope: A Full-Stack View on Methods and Metrics
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
Beyond Forgetting: Machine Unlearning Elicits Controllable Side Behaviors and Capabilities
von: Dang, Tien, et al.
Veröffentlicht: (2026)
von: Dang, Tien, et al.
Veröffentlicht: (2026)
Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution
von: Sadhu, Saisab, et al.
Veröffentlicht: (2026)
von: Sadhu, Saisab, et al.
Veröffentlicht: (2026)
Stealthy Peers: Understanding Security Risks of WebRTC-Based Peer-Assisted Video Streaming
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
Learn and Don't Forget: Adding a New Language to ASR Foundation Models
von: Qian, Mengjie, et al.
Veröffentlicht: (2024)
von: Qian, Mengjie, et al.
Veröffentlicht: (2024)
Leaky Cauldron on the Dark Land: Understanding Memory Side-Channel Hazards in SGX
von: Wang, Wenhao, et al.
Veröffentlicht: (2017)
von: Wang, Wenhao, et al.
Veröffentlicht: (2017)
ACU: Analytic Continual Unlearning for Efficient and Exact Forgetting with Privacy Preservation
von: Tang, Jianheng, et al.
Veröffentlicht: (2025)
von: Tang, Jianheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks
von: Chen, Xiaoyi, et al.
Veröffentlicht: (2023) -
Selective Amnesia: On Efficient, High-Fidelity and Blind Suppression of Backdoor Effects in Trojaned Machine Learning Models
von: Zhu, Rui, et al.
Veröffentlicht: (2022) -
Forgetting to Forget: Attention Sink as A Gateway for Backdooring LLM Unlearning
von: Shang, Bingqi, et al.
Veröffentlicht: (2025) -
Beyond Local vs. External: A Game-Theoretic Framework for Trustworthy Knowledge Acquisition
von: Yao, Rujing, et al.
Veröffentlicht: (2026) -
Challenging Forgets: Unveiling the Worst-Case Forget Sets in Machine Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)