SoK: Machine Unlearning for Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Jie, Xing, Yue, Cui, Yingqian, Aggarwal, Charu C., Liu, Hui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SoK: Unlearnability and Unlearning for Model Dememorization
by: Zhang, Mengying, et al.
Published: (2026)
by: Zhang, Mengying, et al.
Published: (2026)
SoK: Data Minimization in Machine Learning
by: Staab, Robin, et al.
Published: (2025)
by: Staab, Robin, et al.
Published: (2025)
SoK: Dataset Copyright Auditing in Machine Learning Systems
by: Du, Linkang, et al.
Published: (2024)
by: Du, Linkang, et al.
Published: (2024)
SoK: Data Reconstruction Attacks Against Machine Learning Models: Definition, Metrics, and Benchmark
by: Wen, Rui, et al.
Published: (2025)
by: Wen, Rui, et al.
Published: (2025)
SoK: Unintended Interactions among Machine Learning Defenses and Risks
by: Duddu, Vasisht, et al.
Published: (2023)
by: Duddu, Vasisht, et al.
Published: (2023)
SoK: Reducing the Vulnerability of Fine-tuned Language Models to Membership Inference Attacks
by: Amit, Guy, et al.
Published: (2024)
by: Amit, Guy, et al.
Published: (2024)
SoK: What Makes Private Learning Unfair?
by: Yao, Kai, et al.
Published: (2025)
by: Yao, Kai, et al.
Published: (2025)
SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity
by: McFadden, Shae, et al.
Published: (2026)
by: McFadden, Shae, et al.
Published: (2026)
SoK: Privacy Preserving Machine Learning using Functional Encryption: Opportunities and Challenges
by: Panzade, Prajwal, et al.
Published: (2022)
by: Panzade, Prajwal, et al.
Published: (2022)
SoK: Can Trajectory Generation Combine Privacy and Utility?
by: Buchholz, Erik, et al.
Published: (2024)
by: Buchholz, Erik, et al.
Published: (2024)
SoK: Security and Privacy Risks of Healthcare AI
by: Chang, Yuanhaur, et al.
Published: (2024)
by: Chang, Yuanhaur, et al.
Published: (2024)
SoK: Verifiable Cross-Silo FL
by: Korneev, Aleksei, et al.
Published: (2024)
by: Korneev, Aleksei, et al.
Published: (2024)
SoK: Watermarking for AI-Generated Content
by: Zhao, Xuandong, et al.
Published: (2024)
by: Zhao, Xuandong, et al.
Published: (2024)
SoK: Analyzing Adversarial Examples: A Framework to Study Adversary Knowledge
by: Fenaux, Lucas, et al.
Published: (2024)
by: Fenaux, Lucas, et al.
Published: (2024)
SoK: Privacy-aware LLM in Healthcare: Threat Model, Privacy Techniques, Challenges and Recommendations
by: Tahera, Mohoshin Ara, et al.
Published: (2026)
by: Tahera, Mohoshin Ara, et al.
Published: (2026)
SoK: Benchmarking Poisoning Attacks and Defenses in Federated Learning
by: Zhang, Heyi, et al.
Published: (2025)
by: Zhang, Heyi, et al.
Published: (2025)
SoK: On the Offensive Potential of AI
by: Schröer, Saskia Laura, et al.
Published: (2024)
by: Schröer, Saskia Laura, et al.
Published: (2024)
SoK: Blockchain-Based Decentralized AI (DeAI)
by: Lui, Elizabeth, et al.
Published: (2024)
by: Lui, Elizabeth, et al.
Published: (2024)
SoK: Enhancing Cryptographic Collaborative Learning with Differential Privacy
by: Capano, Francesco, et al.
Published: (2026)
by: Capano, Francesco, et al.
Published: (2026)
SoK: Semantic Privacy in Large Language Models
by: Ma, Baihe, et al.
Published: (2025)
by: Ma, Baihe, et al.
Published: (2025)
SoK: A Systems Perspective on Compound AI Threats and Countermeasures
by: Banerjee, Sarbartha, et al.
Published: (2024)
by: Banerjee, Sarbartha, et al.
Published: (2024)
SoK: A Review of Differentially Private Linear Models For High-Dimensional Data
by: Khanna, Amol, et al.
Published: (2024)
by: Khanna, Amol, et al.
Published: (2024)
Sharpness-Aware Data Poisoning Attack
by: He, Pengfei, et al.
Published: (2023)
by: He, Pengfei, et al.
Published: (2023)
SoK: Understanding Vulnerabilities in the Large Language Model Supply Chain
by: Wang, Shenao, et al.
Published: (2025)
by: Wang, Shenao, et al.
Published: (2025)
SoK: Membership Inference Attacks on LLMs are Rushing Nowhere (and How to Fix It)
by: Meeus, Matthieu, et al.
Published: (2024)
by: Meeus, Matthieu, et al.
Published: (2024)
Machine Unlearning for Traditional Models and Large Language Models: A Short Survey
by: Xu, Yi
Published: (2024)
by: Xu, Yi
Published: (2024)
DiffusionShield: A Watermark for Copyright Protection against Generative Diffusion Models
by: Cui, Yingqian, et al.
Published: (2023)
by: Cui, Yingqian, et al.
Published: (2023)
SoK: Potentials and Challenges of Large Language Models for Reverse Engineering
by: Hu, Xinyu, et al.
Published: (2025)
by: Hu, Xinyu, et al.
Published: (2025)
On Large Language Model Continual Unlearning
by: Gao, Chongyang, et al.
Published: (2024)
by: Gao, Chongyang, et al.
Published: (2024)
SoK: Realistic Adversarial Attacks and Defenses for Intelligent Network Intrusion Detection
by: Vitorino, João, et al.
Published: (2023)
by: Vitorino, João, et al.
Published: (2023)
SoK: Systematization and Benchmarking of Deepfake Detectors in a Unified Framework
by: Le, Binh M., et al.
Published: (2024)
by: Le, Binh M., et al.
Published: (2024)
Towards Unveiling Vulnerabilities of Large Reasoning Models in Machine Unlearning
by: Chen, Aobo, et al.
Published: (2026)
by: Chen, Aobo, et al.
Published: (2026)
Machine Unlearning of Pre-trained Large Language Models
by: Yao, Jin, et al.
Published: (2024)
by: Yao, Jin, et al.
Published: (2024)
SoK: Evaluating Jailbreak Guardrails for Large Language Models
by: Wang, Xunguang, et al.
Published: (2025)
by: Wang, Xunguang, et al.
Published: (2025)
SoK: Pitfalls in Evaluating Black-Box Attacks
by: Suya, Fnu, et al.
Published: (2023)
by: Suya, Fnu, et al.
Published: (2023)
Adversarial Machine Unlearning
by: Di, Zonglin, et al.
Published: (2024)
by: Di, Zonglin, et al.
Published: (2024)
Rectifying Privacy and Efficacy Measurements in Machine Unlearning: A New Inference Attack Perspective
by: Naderloui, Nima, et al.
Published: (2025)
by: Naderloui, Nima, et al.
Published: (2025)
SoK: Runtime Integrity
by: Ammar, Mahmoud, et al.
Published: (2024)
by: Ammar, Mahmoud, et al.
Published: (2024)
SoK: Taxonomy and Evaluation of Prompt Security in Large Language Models
by: Hong, Hanbin, et al.
Published: (2025)
by: Hong, Hanbin, et al.
Published: (2025)
SoK: Robustness in Large Language Models against Jailbreak Attacks
by: Xu, Feiyue, et al.
Published: (2026)
by: Xu, Feiyue, et al.
Published: (2026)
Similar Items
-
SoK: Unlearnability and Unlearning for Model Dememorization
by: Zhang, Mengying, et al.
Published: (2026) -
SoK: Data Minimization in Machine Learning
by: Staab, Robin, et al.
Published: (2025) -
SoK: Dataset Copyright Auditing in Machine Learning Systems
by: Du, Linkang, et al.
Published: (2024) -
SoK: Data Reconstruction Attacks Against Machine Learning Models: Definition, Metrics, and Benchmark
by: Wen, Rui, et al.
Published: (2025) -
SoK: Unintended Interactions among Machine Learning Defenses and Risks
by: Duddu, Vasisht, et al.
Published: (2023)