Unlearning's Blind Spots: Over-Unlearning and Prototypical Relearning Attack
Fuente:
arXiv
Saved in:
| Main Authors: | Ha, SeungBum, Park, Saerom, Yoon, Sung Whan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Approximate Domain Unlearning for Vision-Language Models
by: Kawamura, Kodai, et al.
Published: (2025)
by: Kawamura, Kodai, et al.
Published: (2025)
OFMU: Optimization-Driven Framework for Machine Unlearning
by: Asif, Sadia, et al.
Published: (2025)
by: Asif, Sadia, et al.
Published: (2025)
Towards Independence Criterion in Machine Unlearning of Features and Labels
by: Han, Ling, et al.
Published: (2024)
by: Han, Ling, et al.
Published: (2024)
Reliable Unlearning Harmful Information in LLMs with Metamorphosis Representation Projection
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
Unlearning at Scale: Implementing the Right to be Forgotten in Large Language Models
by: X, Abdullah
Published: (2025)
by: X, Abdullah
Published: (2025)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
Scalable and Robust LLM Unlearning by Correcting Responses with Retrieved Exclusions
by: Kim, Junbeom, et al.
Published: (2025)
by: Kim, Junbeom, et al.
Published: (2025)
Machine Unlearning for Masked Diffusion Language Models
by: Lee, Georu, et al.
Published: (2026)
by: Lee, Georu, et al.
Published: (2026)
TACO: Tackling Over-correction in Federated Learning with Tailored Adaptive Correction
by: Liu, Weijie, et al.
Published: (2025)
by: Liu, Weijie, et al.
Published: (2025)
Inference-Time Machine Unlearning via Gated Activation Redirection
by: Turani, Vinícius Conte, et al.
Published: (2026)
by: Turani, Vinícius Conte, et al.
Published: (2026)
Learning to Unlearn for Robust Machine Unlearning
by: Huang, Mark He, et al.
Published: (2024)
by: Huang, Mark He, et al.
Published: (2024)
Digital Forgetting in Large Language Models: A Survey of Unlearning Methods
by: Blanco-Justicia, Alberto, et al.
Published: (2024)
by: Blanco-Justicia, Alberto, et al.
Published: (2024)
Memory-efficient Continual Learning with Prototypical Exemplar Condensation
by: Nguyen, Minh-Duong, et al.
Published: (2026)
by: Nguyen, Minh-Duong, et al.
Published: (2026)
Prototype Transformer: Towards Language Model Architectures Interpretable by Design
by: Yordanov, Yordan, et al.
Published: (2026)
by: Yordanov, Yordan, et al.
Published: (2026)
Quantization-Robust LLM Unlearning via Low-Rank Adaptation
by: Abitante, João Vitor Boer, et al.
Published: (2026)
by: Abitante, João Vitor Boer, et al.
Published: (2026)
Fusion-Based Neural Generalization for Predicting Temperature Fields in Industrial PET Preform Heating
by: Alsheikh, Ahmad, et al.
Published: (2025)
by: Alsheikh, Ahmad, et al.
Published: (2025)
DataRater: Meta-Learned Dataset Curation
by: Calian, Dan A., et al.
Published: (2025)
by: Calian, Dan A., et al.
Published: (2025)
The Lattice Geometry of Neural Network Quantization -- A Short Equivalence Proof of GPTQ and Babai's Algorithm
by: Birnick, Johann
Published: (2025)
by: Birnick, Johann
Published: (2025)
Residual Reservoir Memory Networks
by: Pinna, Matteo, et al.
Published: (2025)
by: Pinna, Matteo, et al.
Published: (2025)
FreRA: A Frequency-Refined Augmentation for Contrastive Learning on Time Series Classification
by: Tian, Tian, et al.
Published: (2025)
by: Tian, Tian, et al.
Published: (2025)
Deep Residual Echo State Networks: exploring residual orthogonal connections in untrained Recurrent Neural Networks
by: Pinna, Matteo, et al.
Published: (2025)
by: Pinna, Matteo, et al.
Published: (2025)
1 bit is all we need: binary normalized neural networks
by: Cabral, Eduardo Lobo Lustoda, et al.
Published: (2025)
by: Cabral, Eduardo Lobo Lustoda, et al.
Published: (2025)
Multi-Task Reinforcement Learning with Language-Encoded Gated Policy Networks
by: Arora, Rushiv
Published: (2025)
by: Arora, Rushiv
Published: (2025)
Global-Order GFlowNets
by: Pastor-Pérez, Lluís, et al.
Published: (2025)
by: Pastor-Pérez, Lluís, et al.
Published: (2025)
Large Language Models as Attribution Regularizers for Efficient Model Training
by: Vukadin, Davor, et al.
Published: (2025)
by: Vukadin, Davor, et al.
Published: (2025)
A Geometric Perspective for High-Dimensional Multiplex Graphs
by: Abdous, Kamel, et al.
Published: (2025)
by: Abdous, Kamel, et al.
Published: (2025)
Evaluating Model Explanations without Ground Truth
by: Rawal, Kaivalya, et al.
Published: (2025)
by: Rawal, Kaivalya, et al.
Published: (2025)
Evidential Deep Active Learning for Semi-Supervised Classification
by: Zhao, Shenkai, et al.
Published: (2025)
by: Zhao, Shenkai, et al.
Published: (2025)
A Practical Approach to using Supervised Machine Learning Models to Classify Aviation Safety Occurrences
by: Siow, Bryan Y.
Published: (2025)
by: Siow, Bryan Y.
Published: (2025)
Scaling Offline RL via Efficient and Expressive Shortcut Models
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
One Router to Route Them All: Homogeneous Expert Routing for Heterogeneous Graph Transformers
by: Shakirov, Georgiy, et al.
Published: (2025)
by: Shakirov, Georgiy, et al.
Published: (2025)
A Simple Generalisation of the Implicit Dynamics of In-Context Learning
by: Innocenti, Francesco, et al.
Published: (2025)
by: Innocenti, Francesco, et al.
Published: (2025)
Load and Renewable Energy Forecasting Using Deep Learning for Grid Stability
by: Sarkar, Kamal
Published: (2025)
by: Sarkar, Kamal
Published: (2025)
Relevance-driven Input Dropout: an Explanation-guided Regularization Technique
by: Gururaj, Shreyas, et al.
Published: (2025)
by: Gururaj, Shreyas, et al.
Published: (2025)
Leveraging Personalized PageRank and Higher-Order Topological Structures for Heterophily Mitigation in Graph Neural Networks
by: Wang, Yumeng, et al.
Published: (2025)
by: Wang, Yumeng, et al.
Published: (2025)
Grokking Beyond the Euclidean Norm of Model Parameters
by: Notsawo, Pascal Jr Tikeng, et al.
Published: (2025)
by: Notsawo, Pascal Jr Tikeng, et al.
Published: (2025)
Manipulating 3D Molecules in a Fixed-Dimensional E(3)-Equivariant Latent Space
by: Chen, Zitao, et al.
Published: (2025)
by: Chen, Zitao, et al.
Published: (2025)
Integrating Causality with Neurochaos Learning: Proposed Approach and Research Agenda
by: Narendra, Nanjangud C., et al.
Published: (2025)
by: Narendra, Nanjangud C., et al.
Published: (2025)
On the Role of Pre-trained Embeddings in Binary Code Analysis
by: Maier, Alwin, et al.
Published: (2025)
by: Maier, Alwin, et al.
Published: (2025)
Expressive Value Learning for Scalable Offline Reinforcement Learning
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
Similar Items
-
Approximate Domain Unlearning for Vision-Language Models
by: Kawamura, Kodai, et al.
Published: (2025) -
OFMU: Optimization-Driven Framework for Machine Unlearning
by: Asif, Sadia, et al.
Published: (2025) -
Towards Independence Criterion in Machine Unlearning of Features and Labels
by: Han, Ling, et al.
Published: (2024) -
Reliable Unlearning Harmful Information in LLMs with Metamorphosis Representation Projection
by: Wu, Chengcan, et al.
Published: (2025) -
Unlearning at Scale: Implementing the Right to be Forgotten in Large Language Models
by: X, Abdullah
Published: (2025)