Which Retain Set Matters for LLM Unlearning? A Case Study on Entity Unlearning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Hwan, Lee, Hwanhee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents
von: Chang, Hwan, et al.
Veröffentlicht: (2025)
von: Chang, Hwan, et al.
Veröffentlicht: (2025)
Reversing the Forget-Retain Objectives: An Efficient LLM Unlearning Framework from Logit Difference
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
Hallucinate at the Last in Long Response Generation: A Case Study on Long Document Summarization
von: Yang, Joonho, et al.
Veröffentlicht: (2025)
von: Yang, Joonho, et al.
Veröffentlicht: (2025)
Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models
von: Jang, Haeun, et al.
Veröffentlicht: (2026)
von: Jang, Haeun, et al.
Veröffentlicht: (2026)
Align-then-Unlearn: Embedding Alignment for LLM Unlearning
von: Spohn, Philipp, et al.
Veröffentlicht: (2025)
von: Spohn, Philipp, et al.
Veröffentlicht: (2025)
Does Unlearning Truly Unlearn? A Black Box Evaluation of LLM Unlearning Methods
von: Doshi, Jai, et al.
Veröffentlicht: (2024)
von: Doshi, Jai, et al.
Veröffentlicht: (2024)
Unlearning What Matters: Token-Level Attribution for Precise Language Model Unlearning
von: Wu, Jiawei, et al.
Veröffentlicht: (2026)
von: Wu, Jiawei, et al.
Veröffentlicht: (2026)
Can Small Language Models Learn, Unlearn, and Retain Noise Patterns?
von: Scaria, Nicy, et al.
Veröffentlicht: (2024)
von: Scaria, Nicy, et al.
Veröffentlicht: (2024)
Keep Security! Benchmarking Security Policy Preservation in Large Language Model Contexts Against Indirect Attacks in Question Answering
von: Chang, Hwan, et al.
Veröffentlicht: (2025)
von: Chang, Hwan, et al.
Veröffentlicht: (2025)
OpenUnlearning: Accelerating LLM Unlearning via Unified Benchmarking of Methods and Metrics
von: Dorna, Vineeth, et al.
Veröffentlicht: (2025)
von: Dorna, Vineeth, et al.
Veröffentlicht: (2025)
LLM Unlearning with LLM Beliefs
von: Li, Kemou, et al.
Veröffentlicht: (2025)
von: Li, Kemou, et al.
Veröffentlicht: (2025)
LUME: LLM Unlearning with Multitask Evaluations
von: Ramakrishna, Anil, et al.
Veröffentlicht: (2025)
von: Ramakrishna, Anil, et al.
Veröffentlicht: (2025)
No Encore: Unlearning as Opt-Out in Music Generation
von: Kim, Jinju, et al.
Veröffentlicht: (2025)
von: Kim, Jinju, et al.
Veröffentlicht: (2025)
Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning
von: Zhai, Naixin, et al.
Veröffentlicht: (2026)
von: Zhai, Naixin, et al.
Veröffentlicht: (2026)
Robust LLM Unlearning Against Relearning Attacks: The Minor Components in Representations Matter
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
Opt-Out: Investigating Entity-Level Unlearning for Large Language Models via Optimal Transport
von: Choi, Minseok, et al.
Veröffentlicht: (2024)
von: Choi, Minseok, et al.
Veröffentlicht: (2024)
Probing-RAG: Self-Probing to Guide Language Models in Selective Document Retrieval
von: Baek, Ingeol, et al.
Veröffentlicht: (2024)
von: Baek, Ingeol, et al.
Veröffentlicht: (2024)
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis
von: Ma, Weitao, et al.
Veröffentlicht: (2024)
von: Ma, Weitao, et al.
Veröffentlicht: (2024)
Consistency-Aware Editing for Entity-level Unlearning in Language Models
von: Han, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Han, Xiaoqi, et al.
Veröffentlicht: (2025)
Robust LLM Unlearning with MUDMAN: Meta-Unlearning with Disruption Masking And Normalization
von: Sondej, Filip, et al.
Veröffentlicht: (2025)
von: Sondej, Filip, et al.
Veröffentlicht: (2025)
Rotation Control Unlearning: Quantifying and Controlling Continuous Unlearning for LLM with The Cognitive Rotation Space
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
Representation-Guided Parameter-Efficient LLM Unlearning
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
Not Every Token Needs Forgetting: Selective Unlearning to Limit Change in Utility in Large Language Model Unlearning
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
Exclusive Unlearning
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2026)
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2026)
Measuring the Depth of LLM Unlearning via Activation Patching
von: Lee, Jaeung, et al.
Veröffentlicht: (2026)
von: Lee, Jaeung, et al.
Veröffentlicht: (2026)
Improving LLM Unlearning Robustness via Random Perturbations
von: Huu-Tien, Dang, et al.
Veröffentlicht: (2025)
von: Huu-Tien, Dang, et al.
Veröffentlicht: (2025)
Position: LLM Unlearning Benchmarks are Weak Measures of Progress
von: Thaker, Pratiksha, et al.
Veröffentlicht: (2024)
von: Thaker, Pratiksha, et al.
Veröffentlicht: (2024)
Agents Are All You Need for LLM Unlearning
von: Sanyal, Debdeep, et al.
Veröffentlicht: (2025)
von: Sanyal, Debdeep, et al.
Veröffentlicht: (2025)
Catastrophic Failure of LLM Unlearning via Quantization
von: Zhang, Zhiwei, et al.
Veröffentlicht: (2024)
von: Zhang, Zhiwei, et al.
Veröffentlicht: (2024)
LLM Unlearning Should Be Form-Independent
von: Ye, Xiaotian, et al.
Veröffentlicht: (2025)
von: Ye, Xiaotian, et al.
Veröffentlicht: (2025)
Explainable LLM Unlearning Through Reasoning
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
CURaTE: Continual Unlearning in Real Time with Ensured Preservation of LLM Knowledge
von: Bae, Seyun, et al.
Veröffentlicht: (2026)
von: Bae, Seyun, et al.
Veröffentlicht: (2026)
Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
Does Localization Inform Unlearning? A Rigorous Examination of Local Parameter Attribution for Knowledge Unlearning in Language Models
von: Lee, Hwiyeong, et al.
Veröffentlicht: (2025)
von: Lee, Hwiyeong, et al.
Veröffentlicht: (2025)
Unlearning Climate Misinformation in Large Language Models
von: Fore, Michael, et al.
Veröffentlicht: (2024)
von: Fore, Michael, et al.
Veröffentlicht: (2024)
Label Smoothing Improves Gradient Ascent in LLM Unlearning
von: Pang, Zirui, et al.
Veröffentlicht: (2025)
von: Pang, Zirui, et al.
Veröffentlicht: (2025)
CLUE: Conflict-guided Localization for LLM Unlearning Framework
von: Chen, Hang, et al.
Veröffentlicht: (2025)
von: Chen, Hang, et al.
Veröffentlicht: (2025)
Dual-Space Smoothness for Robust and Balanced LLM Unlearning
von: Yan, Han, et al.
Veröffentlicht: (2025)
von: Yan, Han, et al.
Veröffentlicht: (2025)
On the Hidden Costs of Counterfactual Knowledge Training in LLM Unlearning
von: Ye, Xiaotian, et al.
Veröffentlicht: (2026)
von: Ye, Xiaotian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents
von: Chang, Hwan, et al.
Veröffentlicht: (2025) -
Reversing the Forget-Retain Objectives: An Efficient LLM Unlearning Framework from Logit Difference
von: Ji, Jiabao, et al.
Veröffentlicht: (2024) -
Hallucinate at the Last in Long Response Generation: A Case Study on Long Document Summarization
von: Yang, Joonho, et al.
Veröffentlicht: (2025) -
Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models
von: Jang, Haeun, et al.
Veröffentlicht: (2026) -
Align-then-Unlearn: Embedding Alignment for LLM Unlearning
von: Spohn, Philipp, et al.
Veröffentlicht: (2025)