Fast Exact Unlearning for In-Context Learning Data for LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Muresanu, Andrei I., Thudi, Anvith, Zhang, Michael R., Papernot, Nicolas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gradients Look Alike: Sensitivity is Often Overestimated in DP-SGD
von: Thudi, Anvith, et al.
Veröffentlicht: (2023)
von: Thudi, Anvith, et al.
Veröffentlicht: (2023)
Efficient Public Verification of Private ML via Regularization
von: Bell, Zoë Ruha, et al.
Veröffentlicht: (2025)
von: Bell, Zoë Ruha, et al.
Veröffentlicht: (2025)
UnUnlearning: Unlearning is not sufficient for content regulation in advanced generative AI
von: Shumailov, Ilia, et al.
Veröffentlicht: (2024)
von: Shumailov, Ilia, et al.
Veröffentlicht: (2024)
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
ACU: Analytic Continual Unlearning for Efficient and Exact Forgetting with Privacy Preservation
von: Tang, Jianheng, et al.
Veröffentlicht: (2025)
von: Tang, Jianheng, et al.
Veröffentlicht: (2025)
Have it your way: Individualized Privacy Assignment for DP-SGD
von: Boenisch, Franziska, et al.
Veröffentlicht: (2023)
von: Boenisch, Franziska, et al.
Veröffentlicht: (2023)
Data-Free Privacy-Preserving for LLMs via Model Inversion and Selective Unlearning
von: Zhou, Xinjie, et al.
Veröffentlicht: (2026)
von: Zhou, Xinjie, et al.
Veröffentlicht: (2026)
Backdoor Detection through Replicated Execution of Outsourced Training
von: Jia, Hengrui, et al.
Veröffentlicht: (2025)
von: Jia, Hengrui, et al.
Veröffentlicht: (2025)
In-Context Unlearning: Language Models as Few Shot Unlearners
von: Pawelczyk, Martin, et al.
Veröffentlicht: (2023)
von: Pawelczyk, Martin, et al.
Veröffentlicht: (2023)
Unlink to Unlearn: Simplifying Edge Unlearning in GNNs
von: Tan, Jiajun, et al.
Veröffentlicht: (2024)
von: Tan, Jiajun, et al.
Veröffentlicht: (2024)
Causal Unlearning in Collaborative Optimization: Exact and Approximate Influence Reversal under Adversarial Contributions
von: Mahdavi, Ali, et al.
Veröffentlicht: (2026)
von: Mahdavi, Ali, et al.
Veröffentlicht: (2026)
Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2025)
Hierarchical Dual-Strategy Unlearning for Biomedical and Healthcare Intelligence Using Imperfect and Privacy-Sensitive Medical Data
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
Machine Unlearning with Minimal Gradient Dependence for High Unlearning Ratios
von: Huang, Tao, et al.
Veröffentlicht: (2024)
von: Huang, Tao, et al.
Veröffentlicht: (2024)
Unlearning Inversion Attacks for Graph Neural Networks
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
GraphToxin: Reconstructing Full Unlearned Graphs from Graph Unlearning
von: Song, Ying, et al.
Veröffentlicht: (2025)
von: Song, Ying, et al.
Veröffentlicht: (2025)
Oblivionis: A Lightweight Learning and Unlearning Framework for Federated Large Language Models
von: Zhang, Fuyao, et al.
Veröffentlicht: (2025)
von: Zhang, Fuyao, et al.
Veröffentlicht: (2025)
Data Unlearning Beyond Uniform Forgetting via Diffusion Time and Frequency Selection
von: Park, Jinseong, et al.
Veröffentlicht: (2025)
von: Park, Jinseong, et al.
Veröffentlicht: (2025)
Towards Lifecycle Unlearning Commitment Management: Measuring Sample-level Unlearning Completeness
von: Wang, Cheng-Long, et al.
Veröffentlicht: (2025)
von: Wang, Cheng-Long, et al.
Veröffentlicht: (2025)
Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
von: Gao, Fengyu, et al.
Veröffentlicht: (2024)
von: Gao, Fengyu, et al.
Veröffentlicht: (2024)
SoK: Unlearnability and Unlearning for Model Dememorization
von: Zhang, Mengying, et al.
Veröffentlicht: (2026)
von: Zhang, Mengying, et al.
Veröffentlicht: (2026)
FRAMU: Attention-based Machine Unlearning using Federated Reinforcement Learning
von: Shaik, Thanveer, et al.
Veröffentlicht: (2023)
von: Shaik, Thanveer, et al.
Veröffentlicht: (2023)
Inexact Unlearning Needs More Careful Evaluations to Avoid a False Sense of Privacy
von: Hayes, Jamie, et al.
Veröffentlicht: (2024)
von: Hayes, Jamie, et al.
Veröffentlicht: (2024)
SAEs $\textit{Can}$ Improve Unlearning: Dynamic Sparse Autoencoder Guardrails for Precision Unlearning in LLMs
von: Muhamed, Aashiq, et al.
Veröffentlicht: (2025)
von: Muhamed, Aashiq, et al.
Veröffentlicht: (2025)
DP-TabICL: In-Context Learning with Differentially Private Tabular Data
von: Carey, Alycia N., et al.
Veröffentlicht: (2024)
von: Carey, Alycia N., et al.
Veröffentlicht: (2024)
Unlearn to Relearn Backdoors: Deferred Backdoor Functionality Attacks on Deep Learning Models
von: Shin, Jeongjin, et al.
Veröffentlicht: (2024)
von: Shin, Jeongjin, et al.
Veröffentlicht: (2024)
LAMD: Context-driven Android Malware Detection and Classification with LLMs
von: Qian, Xingzhi, et al.
Veröffentlicht: (2025)
von: Qian, Xingzhi, et al.
Veröffentlicht: (2025)
Confidential Guardian: Cryptographically Prohibiting the Abuse of Model Abstention
von: Rabanser, Stephan, et al.
Veröffentlicht: (2025)
von: Rabanser, Stephan, et al.
Veröffentlicht: (2025)
TracLLM: A Generic Framework for Attributing Long Context LLMs
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
Breaking the Trilemma of Privacy, Utility, Efficiency via Controllable Machine Unlearning
von: Liu, Zheyuan, et al.
Veröffentlicht: (2023)
von: Liu, Zheyuan, et al.
Veröffentlicht: (2023)
PRUNE: A Patching Based Repair Framework for Certifiable Unlearning of Neural Networks
von: Li, Xuran, et al.
Veröffentlicht: (2025)
von: Li, Xuran, et al.
Veröffentlicht: (2025)
Scalable Federated Unlearning via Isolated and Coded Sharding
von: Lin, Yijing, et al.
Veröffentlicht: (2024)
von: Lin, Yijing, et al.
Veröffentlicht: (2024)
WARP: Weight Teleportation for Attack-Resilient Unlearning Protocols
von: Maheri, Mohammad M, et al.
Veröffentlicht: (2025)
von: Maheri, Mohammad M, et al.
Veröffentlicht: (2025)
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024)
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024)
Scaling Trends for Data Poisoning in LLMs
von: Bowen, Dillon, et al.
Veröffentlicht: (2024)
von: Bowen, Dillon, et al.
Veröffentlicht: (2024)
MCP Safety Audit: LLMs with the Model Context Protocol Allow Major Security Exploits
von: Radosevich, Brandon, et al.
Veröffentlicht: (2025)
von: Radosevich, Brandon, et al.
Veröffentlicht: (2025)
Recalling The Forgotten Class Memberships: Unlearned Models Can Be Noisy Labelers to Leak Privacy
von: Sui, Zhihao, et al.
Veröffentlicht: (2025)
von: Sui, Zhihao, et al.
Veröffentlicht: (2025)
Concealing Backdoor Model Updates in Federated Learning by Trigger-Optimized Data Poisoning
von: Zhang, Yujie, et al.
Veröffentlicht: (2024)
von: Zhang, Yujie, et al.
Veröffentlicht: (2024)
LMEraser: Large Model Unlearning through Adaptive Prompt Tuning
von: Xu, Jie, et al.
Veröffentlicht: (2024)
von: Xu, Jie, et al.
Veröffentlicht: (2024)
Textual Unlearning Gives a False Sense of Unlearning
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Gradients Look Alike: Sensitivity is Often Overestimated in DP-SGD
von: Thudi, Anvith, et al.
Veröffentlicht: (2023) -
Efficient Public Verification of Private ML via Regularization
von: Bell, Zoë Ruha, et al.
Veröffentlicht: (2025) -
UnUnlearning: Unlearning is not sufficient for content regulation in advanced generative AI
von: Shumailov, Ilia, et al.
Veröffentlicht: (2024) -
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025) -
ACU: Analytic Continual Unlearning for Efficient and Exact Forgetting with Privacy Preservation
von: Tang, Jianheng, et al.
Veröffentlicht: (2025)