Erase or Hide? Suppressing Spurious Unlearning Neurons for Robust Unlearning
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Nakyeong, Kim, Dong-Kyum, Kwon, Jea, Kim, Minsung, Jung, Kyomin, Cha, Meeyoung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bilinear representation mitigates reversal curse and enables consistent model editing
by: Kim, Dong-Kyum, et al.
Published: (2025)
by: Kim, Dong-Kyum, et al.
Published: (2025)
Rethinking Post-Unlearning Behavior of Large Vision-Language Models
by: Kim, Minsung, et al.
Published: (2025)
by: Kim, Minsung, et al.
Published: (2025)
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models
by: Kim, Minsung, et al.
Published: (2025)
by: Kim, Minsung, et al.
Published: (2025)
Erase at the Core: Representation Unlearning for Machine Unlearning
by: Lee, Jaewon, et al.
Published: (2026)
by: Lee, Jaewon, et al.
Published: (2026)
Reliability-Aware Adaptive Self-Consistency for Efficient Sampling in LLM Reasoning
by: Kim, Junseok, et al.
Published: (2026)
by: Kim, Junseok, et al.
Published: (2026)
Mitigating Biases for Instruction-following Language Models via Bias Neurons Elimination
by: Yang, Nakyeong, et al.
Published: (2023)
by: Yang, Nakyeong, et al.
Published: (2023)
Reference-Specific Unlearning Metrics Can Hide the Truth: A Reality Check
by: Cho, Sungjun, et al.
Published: (2025)
by: Cho, Sungjun, et al.
Published: (2025)
Are We Truly Forgetting? A Critical Re-examination of Machine Unlearning Evaluation Protocols
by: Kim, Yongwoo, et al.
Published: (2025)
by: Kim, Yongwoo, et al.
Published: (2025)
FaithUn: Toward Faithful Forgetting in Language Models by Investigating the Interconnectedness of Knowledge
by: Yang, Nakyeong, et al.
Published: (2025)
by: Yang, Nakyeong, et al.
Published: (2025)
MVMR: A New Framework for Evaluating Faithfulness of Video Moment Retrieval against Multiple Distractors
by: Yang, Nakyeong, et al.
Published: (2023)
by: Yang, Nakyeong, et al.
Published: (2023)
Erase then Rectify: A Training-Free Parameter Editing Approach for Cost-Effective Graph Unlearning
by: Yang, Zhe-Rui, et al.
Published: (2024)
by: Yang, Zhe-Rui, et al.
Published: (2024)
TraceHiding: Scalable Machine Unlearning for Mobility Data
by: Faraji, Ali, et al.
Published: (2025)
by: Faraji, Ali, et al.
Published: (2025)
Learning to Unlearn: Instance-wise Unlearning for Pre-trained Classifiers
by: Cha, Sungmin, et al.
Published: (2023)
by: Cha, Sungmin, et al.
Published: (2023)
Towards Robust and Parameter-Efficient Knowledge Unlearning for LLMs
by: Cha, Sungmin, et al.
Published: (2024)
by: Cha, Sungmin, et al.
Published: (2024)
Knowledge Vector Weakening: Efficient Training-free Unlearning for Large Vision-Language Models
by: Kim, Yejin, et al.
Published: (2026)
by: Kim, Yejin, et al.
Published: (2026)
Easy to Learn, Yet Hard to Forget: Towards Robust Unlearning Under Bias
by: Kwon, JuneHyoung, et al.
Published: (2026)
by: Kwon, JuneHyoung, et al.
Published: (2026)
Persona Switch: Mixing Distinct Perspectives in Decoding Time
by: Kim, Junseok, et al.
Published: (2026)
by: Kim, Junseok, et al.
Published: (2026)
Persona is a Double-edged Sword: Mitigating the Negative Impact of Role-playing Prompts in Zero-shot Reasoning Tasks
by: Kim, Junseok, et al.
Published: (2024)
by: Kim, Junseok, et al.
Published: (2024)
Erase to Enhance: Data-Efficient Machine Unlearning in MRI Reconstruction
by: Xue, Yuyang, et al.
Published: (2024)
by: Xue, Yuyang, et al.
Published: (2024)
Machine Unlearning for Robust DNNs: Attribution-Guided Partitioning and Neuron Pruning in Noisy Environments
by: Jin, Deliang, et al.
Published: (2025)
by: Jin, Deliang, et al.
Published: (2025)
CLEAR: Unlearning Spurious Style-Content Associations with Contrastive LEarning with Anti-contrastive Regularization
by: Sun, Minghui, et al.
Published: (2025)
by: Sun, Minghui, et al.
Published: (2025)
Revisiting Machine Unlearning with Dimensional Alignment
by: Seo, Seonguk, et al.
Published: (2024)
by: Seo, Seonguk, et al.
Published: (2024)
Erasing Concepts from Text-to-Image Diffusion Models with Few-shot Unlearning
by: Fuchi, Masane, et al.
Published: (2024)
by: Fuchi, Masane, et al.
Published: (2024)
Disentangled Sparse Representations for Concept-Separated Diffusion Unlearning
by: Kim, Hyeonjin, et al.
Published: (2026)
by: Kim, Hyeonjin, et al.
Published: (2026)
Remaining-data-free Machine Unlearning by Suppressing Sample Contribution
by: Cheng, Xinwen, et al.
Published: (2024)
by: Cheng, Xinwen, et al.
Published: (2024)
Learning to Unlearn for Robust Machine Unlearning
by: Huang, Mark He, et al.
Published: (2024)
by: Huang, Mark He, et al.
Published: (2024)
Fairness and Robustness in Machine Unlearning
by: Tran, Khoa, et al.
Published: (2025)
by: Tran, Khoa, et al.
Published: (2025)
Robust LLM Unlearning with MUDMAN: Meta-Unlearning with Disruption Masking And Normalization
by: Sondej, Filip, et al.
Published: (2025)
by: Sondej, Filip, et al.
Published: (2025)
Forget and Explain: Transparent Verification of GNN Unlearning
by: Ahsan, Imran, et al.
Published: (2025)
by: Ahsan, Imran, et al.
Published: (2025)
Robust Optimization in Protein Fitness Landscapes Using Reinforcement Learning in Latent Space
by: Lee, Minji, et al.
Published: (2024)
by: Lee, Minji, et al.
Published: (2024)
How You Ask Matters! Adaptive RAG Robustness to Query Variations
by: Jang, Yunah, et al.
Published: (2026)
by: Jang, Yunah, et al.
Published: (2026)
Mechanistic Unlearning: Robust Knowledge Unlearning and Editing via Mechanistic Localization
by: Guo, Phillip, et al.
Published: (2024)
by: Guo, Phillip, et al.
Published: (2024)
Retain-Neutral Surrogates for Min-Max Unlearning
by: Cai, Junhao, et al.
Published: (2026)
by: Cai, Junhao, et al.
Published: (2026)
Verifying Robust Unlearning: Probing Residual Knowledge in Unlearned Models
by: Xuan, Hao, et al.
Published: (2025)
by: Xuan, Hao, et al.
Published: (2025)
Severing Spurious Correlations with Data Pruning
by: Mulchandani, Varun, et al.
Published: (2025)
by: Mulchandani, Varun, et al.
Published: (2025)
Targeted Unlearning with Single Layer Unlearning Gradient
by: Cai, Zikui, et al.
Published: (2024)
by: Cai, Zikui, et al.
Published: (2024)
Adversarial Mixup Unlearning
by: Peng, Zhuoyi, et al.
Published: (2025)
by: Peng, Zhuoyi, et al.
Published: (2025)
Improving Fisher Information Estimation and Efficiency for LoRA-based LLM Unlearning
by: Kim, Yejin, et al.
Published: (2025)
by: Kim, Yejin, et al.
Published: (2025)
Holistic Unlearning Benchmark: A Multi-Faceted Evaluation for Text-to-Image Diffusion Model Unlearning
by: Moon, Saemi, et al.
Published: (2024)
by: Moon, Saemi, et al.
Published: (2024)
Contrastive Unlearning: A Contrastive Approach to Machine Unlearning
by: Lee, Hong kyu, et al.
Published: (2024)
by: Lee, Hong kyu, et al.
Published: (2024)
Similar Items
-
Bilinear representation mitigates reversal curse and enables consistent model editing
by: Kim, Dong-Kyum, et al.
Published: (2025) -
Rethinking Post-Unlearning Behavior of Large Vision-Language Models
by: Kim, Minsung, et al.
Published: (2025) -
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models
by: Kim, Minsung, et al.
Published: (2025) -
Erase at the Core: Representation Unlearning for Machine Unlearning
by: Lee, Jaewon, et al.
Published: (2026) -
Reliability-Aware Adaptive Self-Consistency for Efficient Sampling in LLM Reasoning
by: Kim, Junseok, et al.
Published: (2026)