LMEraser: Large Model Unlearning through Adaptive Prompt Tuning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Xu, Jie, Wu, Zihan, Wang, Cong, Jia, Xiaohua |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PrivTune: Efficient and Privacy-Preserving Fine-Tuning of Large Language Models via Device-Cloud Collaboration
par: Liu, Yi, et autres
Publié: (2025)
par: Liu, Yi, et autres
Publié: (2025)
Reconstructing Training Data from Adapter-based Federated Large Language Models
par: Chen, Silong, et autres
Publié: (2026)
par: Chen, Silong, et autres
Publié: (2026)
Oblivionis: A Lightweight Learning and Unlearning Framework for Federated Large Language Models
par: Zhang, Fuyao, et autres
Publié: (2025)
par: Zhang, Fuyao, et autres
Publié: (2025)
Machine Unlearning with Minimal Gradient Dependence for High Unlearning Ratios
par: Huang, Tao, et autres
Publié: (2024)
par: Huang, Tao, et autres
Publié: (2024)
Arondight: Red Teaming Large Vision Language Models with Auto-generated Multi-modal Jailbreak Prompts
par: Liu, Yi, et autres
Publié: (2024)
par: Liu, Yi, et autres
Publié: (2024)
SoK: Unlearnability and Unlearning for Model Dememorization
par: Zhang, Mengying, et autres
Publié: (2026)
par: Zhang, Mengying, et autres
Publié: (2026)
P$^2$RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-$k$ Retrieval
par: Ming, Yulong, et autres
Publié: (2026)
par: Ming, Yulong, et autres
Publié: (2026)
Towards Lifecycle Unlearning Commitment Management: Measuring Sample-level Unlearning Completeness
par: Wang, Cheng-Long, et autres
Publié: (2025)
par: Wang, Cheng-Long, et autres
Publié: (2025)
Unlink to Unlearn: Simplifying Edge Unlearning in GNNs
par: Tan, Jiajun, et autres
Publié: (2024)
par: Tan, Jiajun, et autres
Publié: (2024)
OBLIVIATE: Robust and Practical Machine Unlearning for Large Language Models
par: Xu, Xiaoyu, et autres
Publié: (2025)
par: Xu, Xiaoyu, et autres
Publié: (2025)
Data-Free Privacy-Preserving for LLMs via Model Inversion and Selective Unlearning
par: Zhou, Xinjie, et autres
Publié: (2026)
par: Zhou, Xinjie, et autres
Publié: (2026)
Prompt Injection Attacks on Large Language Models in Oncology
par: Clusmann, Jan, et autres
Publié: (2024)
par: Clusmann, Jan, et autres
Publié: (2024)
Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
par: Gao, Fengyu, et autres
Publié: (2024)
par: Gao, Fengyu, et autres
Publié: (2024)
Textual Unlearning Gives a False Sense of Unlearning
par: Du, Jiacheng, et autres
Publié: (2024)
par: Du, Jiacheng, et autres
Publié: (2024)
Machine Unlearning of Pre-trained Large Language Models
par: Yao, Jin, et autres
Publié: (2024)
par: Yao, Jin, et autres
Publié: (2024)
FIT to Forget: Robust Continual Unlearning for Large Language Models
par: Xu, Xiaoyu, et autres
Publié: (2026)
par: Xu, Xiaoyu, et autres
Publié: (2026)
GraphToxin: Reconstructing Full Unlearned Graphs from Graph Unlearning
par: Song, Ying, et autres
Publié: (2025)
par: Song, Ying, et autres
Publié: (2025)
Unlearning Inversion Attacks for Graph Neural Networks
par: Zhang, Jiahao, et autres
Publié: (2025)
par: Zhang, Jiahao, et autres
Publié: (2025)
In-Context Unlearning: Language Models as Few Shot Unlearners
par: Pawelczyk, Martin, et autres
Publié: (2023)
par: Pawelczyk, Martin, et autres
Publié: (2023)
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
par: Rashid, Md Rafi Ur, et autres
Publié: (2024)
par: Rashid, Md Rafi Ur, et autres
Publié: (2024)
Prompt, Divide, and Conquer: Bypassing Large Language Model Safety Filters via Segmented and Distributed Prompt Processing
par: Wahréus, Johan, et autres
Publié: (2025)
par: Wahréus, Johan, et autres
Publié: (2025)
Harry Potter is Still Here! Probing Knowledge Leakage in Targeted Unlearned Large Language Models via Automated Adversarial Prompting
par: To, Bang Trinh Tran, et autres
Publié: (2025)
par: To, Bang Trinh Tran, et autres
Publié: (2025)
PLeak: Prompt Leaking Attacks against Large Language Model Applications
par: Hui, Bo, et autres
Publié: (2024)
par: Hui, Bo, et autres
Publié: (2024)
Fine-Tuning Language Models with Differential Privacy through Adaptive Noise Allocation
par: Li, Xianzhi, et autres
Publié: (2024)
par: Li, Xianzhi, et autres
Publié: (2024)
Towards Robust Knowledge Unlearning: An Adversarial Framework for Assessing and Improving Unlearning Robustness in Large Language Models
par: Yuan, Hongbang, et autres
Publié: (2024)
par: Yuan, Hongbang, et autres
Publié: (2024)
Adaptive PII Mitigation Framework for Large Language Models
par: Asthana, Shubhi, et autres
Publié: (2025)
par: Asthana, Shubhi, et autres
Publié: (2025)
Machine Unlearning: Solutions and Challenges
par: Xu, Jie, et autres
Publié: (2023)
par: Xu, Jie, et autres
Publié: (2023)
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
par: Wu, Xiaoyu, et autres
Publié: (2025)
par: Wu, Xiaoyu, et autres
Publié: (2025)
Unlearn to Relearn Backdoors: Deferred Backdoor Functionality Attacks on Deep Learning Models
par: Shin, Jeongjin, et autres
Publié: (2024)
par: Shin, Jeongjin, et autres
Publié: (2024)
Jailbroken Frontier Models Retain Their Capabilities
par: Zhu, Daniel, et autres
Publié: (2026)
par: Zhu, Daniel, et autres
Publié: (2026)
SafeRedir: Prompt Embedding Redirection for Robust Unlearning in Image Generation Models
par: Liu, Renyang, et autres
Publié: (2026)
par: Liu, Renyang, et autres
Publié: (2026)
Recalling The Forgotten Class Memberships: Unlearned Models Can Be Noisy Labelers to Leak Privacy
par: Sui, Zhihao, et autres
Publié: (2025)
par: Sui, Zhihao, et autres
Publié: (2025)
Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions
par: Li, Wenjuan, et autres
Publié: (2026)
par: Li, Wenjuan, et autres
Publié: (2026)
DP-Adam-AC: Privacy-preserving Fine-Tuning of Localizable Language Models Using Adam Optimization with Adaptive Clipping
par: Yang, Ruoxing
Publié: (2025)
par: Yang, Ruoxing
Publié: (2025)
PRUNE: A Patching Based Repair Framework for Certifiable Unlearning of Neural Networks
par: Li, Xuran, et autres
Publié: (2025)
par: Li, Xuran, et autres
Publié: (2025)
SALAD: Systematic Assessment of Machine Unlearning on LLM-Aided Hardware Design
par: Wang, Zeng, et autres
Publié: (2025)
par: Wang, Zeng, et autres
Publié: (2025)
Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
par: Xu, Jiashu, et autres
Publié: (2023)
par: Xu, Jiashu, et autres
Publié: (2023)
Fast Exact Unlearning for In-Context Learning Data for LLMs
par: Muresanu, Andrei I., et autres
Publié: (2024)
par: Muresanu, Andrei I., et autres
Publié: (2024)
Scalable Federated Unlearning via Isolated and Coded Sharding
par: Lin, Yijing, et autres
Publié: (2024)
par: Lin, Yijing, et autres
Publié: (2024)
WARP: Weight Teleportation for Attack-Resilient Unlearning Protocols
par: Maheri, Mohammad M, et autres
Publié: (2025)
par: Maheri, Mohammad M, et autres
Publié: (2025)
Documents similaires
-
PrivTune: Efficient and Privacy-Preserving Fine-Tuning of Large Language Models via Device-Cloud Collaboration
par: Liu, Yi, et autres
Publié: (2025) -
Reconstructing Training Data from Adapter-based Federated Large Language Models
par: Chen, Silong, et autres
Publié: (2026) -
Oblivionis: A Lightweight Learning and Unlearning Framework for Federated Large Language Models
par: Zhang, Fuyao, et autres
Publié: (2025) -
Machine Unlearning with Minimal Gradient Dependence for High Unlearning Ratios
par: Huang, Tao, et autres
Publié: (2024) -
Arondight: Red Teaming Large Vision Language Models with Auto-generated Multi-modal Jailbreak Prompts
par: Liu, Yi, et autres
Publié: (2024)