Mirror Mirror on the Wall, Have I Forgotten it All? A New Framework for Evaluating Machine Unlearning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Brimhall, Brennon, Mathew, Philip, Fendley, Neil, Cao, Yinzhi, Green, Matthew |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Comment and Control: Hijacking Agentic Workflows via Context-Grounded Evolution
von: Fendley, Neil, et al.
Veröffentlicht: (2026)
von: Fendley, Neil, et al.
Veröffentlicht: (2026)
MirrorMark: Generalizable Mirrored Sampling for Multi-bit LLM Watermarking
von: Jiang, Ya, et al.
Veröffentlicht: (2026)
von: Jiang, Ya, et al.
Veröffentlicht: (2026)
MirrorShield: Towards Universal Defense Against Jailbreaks via Entropy-Guided Mirror Crafting
von: Pu, Rui, et al.
Veröffentlicht: (2025)
von: Pu, Rui, et al.
Veröffentlicht: (2025)
Towards Lifecycle Unlearning Commitment Management: Measuring Sample-level Unlearning Completeness
von: Wang, Cheng-Long, et al.
Veröffentlicht: (2025)
von: Wang, Cheng-Long, et al.
Veröffentlicht: (2025)
Recalling The Forgotten Class Memberships: Unlearned Models Can Be Noisy Labelers to Leak Privacy
von: Sui, Zhihao, et al.
Veröffentlicht: (2025)
von: Sui, Zhihao, et al.
Veröffentlicht: (2025)
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
PLeak: Prompt Leaking Attacks against Large Language Model Applications
von: Hui, Bo, et al.
Veröffentlicht: (2024)
von: Hui, Bo, et al.
Veröffentlicht: (2024)
The Mirror Design Pattern: Strict Data Geometry over Model Scale for Prompt Injection Detection
von: Corll, J Alex
Veröffentlicht: (2026)
von: Corll, J Alex
Veröffentlicht: (2026)
Data Lineage Inference: Uncovering Privacy Vulnerabilities of Dataset Pruning
von: Li, Qi, et al.
Veröffentlicht: (2024)
von: Li, Qi, et al.
Veröffentlicht: (2024)
A Robust Certified Machine Unlearning Method Under Distribution Shift
von: Guo, Jinduo, et al.
Veröffentlicht: (2026)
von: Guo, Jinduo, et al.
Veröffentlicht: (2026)
Machine Unlearning: A Comprehensive Survey
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
Threats, Attacks, and Defenses in Machine Unlearning: A Survey
von: Liu, Ziyao, et al.
Veröffentlicht: (2024)
von: Liu, Ziyao, et al.
Veröffentlicht: (2024)
Injection, Attack and Erasure: Revocable Backdoor Attacks via Machine Unlearning
von: Song, Baogang, et al.
Veröffentlicht: (2025)
von: Song, Baogang, et al.
Veröffentlicht: (2025)
ERASER: Machine Unlearning in MLaaS via an Inference Serving-Aware Approach
von: Hu, Yuke, et al.
Veröffentlicht: (2023)
von: Hu, Yuke, et al.
Veröffentlicht: (2023)
Releasing Malevolence from Benevolence: The Menace of Benign Data on Machine Unlearning
von: Ma, Binhao, et al.
Veröffentlicht: (2024)
von: Ma, Binhao, et al.
Veröffentlicht: (2024)
Machine Unlearning with Minimal Gradient Dependence for High Unlearning Ratios
von: Huang, Tao, et al.
Veröffentlicht: (2024)
von: Huang, Tao, et al.
Veröffentlicht: (2024)
Detecting Privilege Escalation in Polyglot Microservices via Agentic Program Analysis
von: Li, Penghui, et al.
Veröffentlicht: (2026)
von: Li, Penghui, et al.
Veröffentlicht: (2026)
Data Duplication: A Novel Multi-Purpose Attack Paradigm in Machine Unlearning
von: Ye, Dayong, et al.
Veröffentlicht: (2025)
von: Ye, Dayong, et al.
Veröffentlicht: (2025)
zkUnlearner: A Zero-Knowledge Framework for Verifiable Unlearning with Multi-Granularity and Forgery-Resistance
von: Wang, Nan, et al.
Veröffentlicht: (2025)
von: Wang, Nan, et al.
Veröffentlicht: (2025)
Corrective Machine Unlearning
von: Goel, Shashwat, et al.
Veröffentlicht: (2024)
von: Goel, Shashwat, et al.
Veröffentlicht: (2024)
Erased but Not Forgotten: How Backdoors Compromise Concept Erasure
von: Braun, Tobias, et al.
Veröffentlicht: (2025)
von: Braun, Tobias, et al.
Veröffentlicht: (2025)
Oblivionis: A Lightweight Learning and Unlearning Framework for Federated Large Language Models
von: Zhang, Fuyao, et al.
Veröffentlicht: (2025)
von: Zhang, Fuyao, et al.
Veröffentlicht: (2025)
QUEEN: Query Unlearning against Model Extraction
von: Chen, Huajie, et al.
Veröffentlicht: (2024)
von: Chen, Huajie, et al.
Veröffentlicht: (2024)
Model Inversion Attack against Federated Unlearning
von: Zhou, Lei, et al.
Veröffentlicht: (2025)
von: Zhou, Lei, et al.
Veröffentlicht: (2025)
Jailbreaking Safeguarded Text-to-Image Models via Large Language Models
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
Unlearning at Scale: Implementing the Right to be Forgotten in Large Language Models
von: X, Abdullah
Veröffentlicht: (2025)
von: X, Abdullah
Veröffentlicht: (2025)
Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2025)
Towards Robust Knowledge Unlearning: An Adversarial Framework for Assessing and Improving Unlearning Robustness in Large Language Models
von: Yuan, Hongbang, et al.
Veröffentlicht: (2024)
von: Yuan, Hongbang, et al.
Veröffentlicht: (2024)
FRAMU: Attention-based Machine Unlearning using Federated Reinforcement Learning
von: Shaik, Thanveer, et al.
Veröffentlicht: (2023)
von: Shaik, Thanveer, et al.
Veröffentlicht: (2023)
Breaking the Trilemma of Privacy, Utility, Efficiency via Controllable Machine Unlearning
von: Liu, Zheyuan, et al.
Veröffentlicht: (2023)
von: Liu, Zheyuan, et al.
Veröffentlicht: (2023)
SALAD: Systematic Assessment of Machine Unlearning on LLM-Aided Hardware Design
von: Wang, Zeng, et al.
Veröffentlicht: (2025)
von: Wang, Zeng, et al.
Veröffentlicht: (2025)
Revisiting Privacy Leakage in Machine Unlearning: Membership Inference Beyond the Forgotten Set
von: Fu, Jie, et al.
Veröffentlicht: (2026)
von: Fu, Jie, et al.
Veröffentlicht: (2026)
Unlink to Unlearn: Simplifying Edge Unlearning in GNNs
von: Tan, Jiajun, et al.
Veröffentlicht: (2024)
von: Tan, Jiajun, et al.
Veröffentlicht: (2024)
JPU: Bridging Jailbreak Defense and Unlearning via On-Policy Path Rectification
von: Wang, Xi, et al.
Veröffentlicht: (2026)
von: Wang, Xi, et al.
Veröffentlicht: (2026)
Mitigating Sensitive Information Leakage in LLMs4Code through Machine Unlearning
von: Gu, Shanzhi, et al.
Veröffentlicht: (2025)
von: Gu, Shanzhi, et al.
Veröffentlicht: (2025)
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
von: Rashid, Md Rafi Ur, et al.
Veröffentlicht: (2024)
von: Rashid, Md Rafi Ur, et al.
Veröffentlicht: (2024)
ForgetMark: Stealthy Fingerprint Embedding via Targeted Unlearning in Language Models
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
A Framework for Formalizing LLM Agent Security
von: Siu, Vincent, et al.
Veröffentlicht: (2026)
von: Siu, Vincent, et al.
Veröffentlicht: (2026)
Scrub It Out! Erasing Sensitive Memorization in Code Language Models via Machine Unlearning
von: Chu, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Chu, Zhaoyang, et al.
Veröffentlicht: (2025)
Unified Neural Backdoor Removal with Only Few Clean Samples through Unlearning and Relearning
von: Min, Nay Myat, et al.
Veröffentlicht: (2024)
von: Min, Nay Myat, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Comment and Control: Hijacking Agentic Workflows via Context-Grounded Evolution
von: Fendley, Neil, et al.
Veröffentlicht: (2026) -
MirrorMark: Generalizable Mirrored Sampling for Multi-bit LLM Watermarking
von: Jiang, Ya, et al.
Veröffentlicht: (2026) -
MirrorShield: Towards Universal Defense Against Jailbreaks via Entropy-Guided Mirror Crafting
von: Pu, Rui, et al.
Veröffentlicht: (2025) -
Towards Lifecycle Unlearning Commitment Management: Measuring Sample-level Unlearning Completeness
von: Wang, Cheng-Long, et al.
Veröffentlicht: (2025) -
Recalling The Forgotten Class Memberships: Unlearned Models Can Be Noisy Labelers to Leak Privacy
von: Sui, Zhihao, et al.
Veröffentlicht: (2025)