Unlearning or Obfuscating? Jogging the Memory of Unlearned LLMs via Benign Relearning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hu, Shengyuan, Fu, Yiwei, Wu, Zhiwei Steven, Smith, Virginia |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
BLUR: A Benchmark for LLM Unlearning Robust to Forget-Retain Overlap
par: Hu, Shengyuan, et autres
Publié: (2025)
par: Hu, Shengyuan, et autres
Publié: (2025)
Rethinking Benign Relearning: Syntax as the Hidden Driver of Unlearning Failures
par: Yoon, Sangyeon, et autres
Publié: (2026)
par: Yoon, Sangyeon, et autres
Publié: (2026)
Layered Unlearning for Adversarial Relearning
par: Qian, Timothy, et autres
Publié: (2025)
par: Qian, Timothy, et autres
Publié: (2025)
Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts
par: Gao, Hongcheng, et autres
Publié: (2024)
par: Gao, Hongcheng, et autres
Publié: (2024)
Unlearning's Blind Spots: Over-Unlearning and Prototypical Relearning Attack
par: Ha, SeungBum, et autres
Publié: (2025)
par: Ha, SeungBum, et autres
Publié: (2025)
Guardrail Baselines for Unlearning in LLMs
par: Thaker, Pratiksha, et autres
Publié: (2024)
par: Thaker, Pratiksha, et autres
Publié: (2024)
Efficient Unlearning through Maximizing Relearning Convergence Delay
par: Tran, Khoa, et autres
Publié: (2026)
par: Tran, Khoa, et autres
Publié: (2026)
Unlearning Isn't Invisible: Detecting Unlearning Traces in LLMs from Model Outputs
par: Chen, Yiwei, et autres
Publié: (2025)
par: Chen, Yiwei, et autres
Publié: (2025)
COLUR: Confidence-Oriented Learning, Unlearning and Relearning with Noisy-Label Data for Model Restoration and Refinement
par: Sui, Zhihao, et autres
Publié: (2025)
par: Sui, Zhihao, et autres
Publié: (2025)
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
par: Wu, Xiaoyu, et autres
Publié: (2025)
par: Wu, Xiaoyu, et autres
Publié: (2025)
FedCARE: Federated Unlearning with Conflict-Aware Projection and Relearning-Resistant Recovery
par: Li, Yue, et autres
Publié: (2026)
par: Li, Yue, et autres
Publié: (2026)
SAEs $\textit{Can}$ Improve Unlearning: Dynamic Sparse Autoencoder Guardrails for Precision Unlearning in LLMs
par: Muhamed, Aashiq, et autres
Publié: (2025)
par: Muhamed, Aashiq, et autres
Publié: (2025)
Unlearn to Relearn Backdoors: Deferred Backdoor Functionality Attacks on Deep Learning Models
par: Shin, Jeongjin, et autres
Publié: (2024)
par: Shin, Jeongjin, et autres
Publié: (2024)
Unlearning vs. Obfuscation: Are We Truly Removing Knowledge?
par: Sun, Guangzhi, et autres
Publié: (2025)
par: Sun, Guangzhi, et autres
Publié: (2025)
Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond
par: Fan, Chongyu, et autres
Publié: (2025)
par: Fan, Chongyu, et autres
Publié: (2025)
Exact Unlearning of Finetuning Data via Model Merging at Scale
par: Kuo, Kevin, et autres
Publié: (2025)
par: Kuo, Kevin, et autres
Publié: (2025)
Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs
par: Xu, Xiaoyu, et autres
Publié: (2025)
par: Xu, Xiaoyu, et autres
Publié: (2025)
Position: LLM Unlearning Benchmarks are Weak Measures of Progress
par: Thaker, Pratiksha, et autres
Publié: (2024)
par: Thaker, Pratiksha, et autres
Publié: (2024)
Enhancing One-run Privacy Auditing with Quantile Regression-Based Membership Inference
par: Liu, Terrance, et autres
Publié: (2025)
par: Liu, Terrance, et autres
Publié: (2025)
Reconstruction Attacks on Machine Unlearning: Simple Models are Vulnerable
par: Bertran, Martin, et autres
Publié: (2024)
par: Bertran, Martin, et autres
Publié: (2024)
Leak@$k$: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding
par: Reisizadeh, Hadi, et autres
Publié: (2025)
par: Reisizadeh, Hadi, et autres
Publié: (2025)
DP2Unlearning: An Efficient and Guaranteed Unlearning Framework for LLMs
par: Mahmud, Tamim Al, et autres
Publié: (2025)
par: Mahmud, Tamim Al, et autres
Publié: (2025)
FUNU: Boosting Machine Unlearning Efficiency by Filtering Unnecessary Unlearning
par: Li, Zitong, et autres
Publié: (2025)
par: Li, Zitong, et autres
Publié: (2025)
Unified Parameter-Efficient Unlearning for LLMs
par: Ding, Chenlu, et autres
Publié: (2024)
par: Ding, Chenlu, et autres
Publié: (2024)
Learning-Time Encoding Shapes Unlearning in LLMs
par: Wu, Ruihan, et autres
Publié: (2025)
par: Wu, Ruihan, et autres
Publié: (2025)
Leverage Unlearning to Sanitize LLMs
par: Boutet, Antoine, et autres
Publié: (2025)
par: Boutet, Antoine, et autres
Publié: (2025)
Online Learning and Unlearning
par: Hu, Yaxi, et autres
Publié: (2025)
par: Hu, Yaxi, et autres
Publié: (2025)
UCD: Unlearning in LLMs via Contrastive Decoding
par: Suriyakumar, Vinith M., et autres
Publié: (2025)
par: Suriyakumar, Vinith M., et autres
Publié: (2025)
Targeted Unlearning with Single Layer Unlearning Gradient
par: Cai, Zikui, et autres
Publié: (2024)
par: Cai, Zikui, et autres
Publié: (2024)
Probing Knowledge Holes in Unlearned LLMs
par: Ko, Myeongseob, et autres
Publié: (2025)
par: Ko, Myeongseob, et autres
Publié: (2025)
Split, Unlearn, Merge: Leveraging Data Attributes for More Effective Unlearning in LLMs
par: Kadhe, Swanand Ravindra, et autres
Publié: (2024)
par: Kadhe, Swanand Ravindra, et autres
Publié: (2024)
Unlearning through Knowledge Overwriting: Reversible Federated Unlearning via Selective Sparse Adapter
par: Zhong, Zhengyi, et autres
Publié: (2025)
par: Zhong, Zhengyi, et autres
Publié: (2025)
Mechanistic Unlearning: Robust Knowledge Unlearning and Editing via Mechanistic Localization
par: Guo, Phillip, et autres
Publié: (2024)
par: Guo, Phillip, et autres
Publié: (2024)
Federated Graph Unlearning
par: Ai, Yuming, et autres
Publié: (2025)
par: Ai, Yuming, et autres
Publié: (2025)
How Data Inter-connectivity Shapes LLMs Unlearning: A Structural Unlearning Perspective
par: Qiu, Xinchi, et autres
Publié: (2024)
par: Qiu, Xinchi, et autres
Publié: (2024)
Conformal Unlearning: A New Paradigm for Unlearning in Conformal Predictors
par: Alkhatib, Yahya, et autres
Publié: (2025)
par: Alkhatib, Yahya, et autres
Publié: (2025)
An Illusion of Unlearning? Assessing Machine Unlearning Through Internal Representations
par: Gao, Yichen, et autres
Publié: (2026)
par: Gao, Yichen, et autres
Publié: (2026)
Erase or Hide? Suppressing Spurious Unlearning Neurons for Robust Unlearning
par: Yang, Nakyeong, et autres
Publié: (2025)
par: Yang, Nakyeong, et autres
Publié: (2025)
Learning to Unlearn: Instance-wise Unlearning for Pre-trained Classifiers
par: Cha, Sungmin, et autres
Publié: (2023)
par: Cha, Sungmin, et autres
Publié: (2023)
Align-then-Unlearn: Embedding Alignment for LLM Unlearning
par: Spohn, Philipp, et autres
Publié: (2025)
par: Spohn, Philipp, et autres
Publié: (2025)
Documents similaires
-
BLUR: A Benchmark for LLM Unlearning Robust to Forget-Retain Overlap
par: Hu, Shengyuan, et autres
Publié: (2025) -
Rethinking Benign Relearning: Syntax as the Hidden Driver of Unlearning Failures
par: Yoon, Sangyeon, et autres
Publié: (2026) -
Layered Unlearning for Adversarial Relearning
par: Qian, Timothy, et autres
Publié: (2025) -
Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts
par: Gao, Hongcheng, et autres
Publié: (2024) -
Unlearning's Blind Spots: Over-Unlearning and Prototypical Relearning Attack
par: Ha, SeungBum, et autres
Publié: (2025)