LLM Unlearning on Noisy Forget Sets: A Study of Incomplete, Rewritten, and Watermarked Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Changsheng, Zhang, Yihua, Wei, Dennis, Jia, Jinghan, Chen, Pin-Yu, Liu, Sijia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Quantization-Robust LLM Unlearning via Low-Rank Adaptation
von: Abitante, João Vitor Boer, et al.
Veröffentlicht: (2026)
von: Abitante, João Vitor Boer, et al.
Veröffentlicht: (2026)
MMiC: Mitigating Modality Incompleteness in Clustered Federated Learning
von: Yang, Lishan, et al.
Veröffentlicht: (2025)
von: Yang, Lishan, et al.
Veröffentlicht: (2025)
Concept-based Rubrics Improve LLM Formative Assessment and Data Synthesis
von: Wei, Yuchen, et al.
Veröffentlicht: (2025)
von: Wei, Yuchen, et al.
Veröffentlicht: (2025)
Mr. Snuffleupagus at SemEval-2025 Task 4: Unlearning Factual Knowledge from LLMs Using Adaptive RMU
von: Dosajh, Arjun, et al.
Veröffentlicht: (2025)
von: Dosajh, Arjun, et al.
Veröffentlicht: (2025)
Spectral Clustering in Convex and Constrained Settings
von: Behera, Swarup Ranjan, et al.
Veröffentlicht: (2024)
von: Behera, Swarup Ranjan, et al.
Veröffentlicht: (2024)
Reconstructing Syllable Sequences in Abugida Scripts with Incomplete Inputs
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2025)
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2025)
Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders
von: Patel, Het, et al.
Veröffentlicht: (2026)
von: Patel, Het, et al.
Veröffentlicht: (2026)
Scaling Laws for Forgetting When Fine-Tuning Large Language Models
von: Kalajdzievski, Damjan
Veröffentlicht: (2024)
von: Kalajdzievski, Damjan
Veröffentlicht: (2024)
Digital Forgetting in Large Language Models: A Survey of Unlearning Methods
von: Blanco-Justicia, Alberto, et al.
Veröffentlicht: (2024)
von: Blanco-Justicia, Alberto, et al.
Veröffentlicht: (2024)
OFMU: Optimization-Driven Framework for Machine Unlearning
von: Asif, Sadia, et al.
Veröffentlicht: (2025)
von: Asif, Sadia, et al.
Veröffentlicht: (2025)
Generalizable LLM Learning of Graph Synthetic Data with Post-training Alignment
von: Zhang, Yizhuo, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhuo, et al.
Veröffentlicht: (2025)
VIGOR+: Iterative Confounder Generation and Validation via LLM-CEVAE Feedback Loop
von: Zhu, JiaWei, et al.
Veröffentlicht: (2025)
von: Zhu, JiaWei, et al.
Veröffentlicht: (2025)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
Large Language Model (LLM) Bias Index -- LLMBI
von: Oketunji, Abiodun Finbarrs, et al.
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs, et al.
Veröffentlicht: (2023)
WSM: Decay-Free Learning Rate Schedule via Checkpoint Merging for LLM Pre-training
von: Tian, Changxin, et al.
Veröffentlicht: (2025)
von: Tian, Changxin, et al.
Veröffentlicht: (2025)
FlexQuant: A Flexible and Efficient Dynamic Precision Switching Framework for LLM Quantization
von: Liu, Fangxin, et al.
Veröffentlicht: (2025)
von: Liu, Fangxin, et al.
Veröffentlicht: (2025)
End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning
von: Chen, Guanzhong, et al.
Veröffentlicht: (2025)
von: Chen, Guanzhong, et al.
Veröffentlicht: (2025)
KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Balancing Efficiency and Effectiveness: An LLM-Infused Approach for Optimized CTR Prediction
von: Zhang, Guoxiao, et al.
Veröffentlicht: (2024)
von: Zhang, Guoxiao, et al.
Veröffentlicht: (2024)
Merge-Bench: Resolve Merge Conflicts with Large Language Models
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
Synergy over Discrepancy: A Partition-Based Approach to Multi-Domain LLM Fine-Tuning
von: Ye, Hua, et al.
Veröffentlicht: (2025)
von: Ye, Hua, et al.
Veröffentlicht: (2025)
Representation-Aware Unlearning via Activation Signatures: From Suppression to Entity-Signature Erasure
von: Mahmood, Syed Naveed, et al.
Veröffentlicht: (2026)
von: Mahmood, Syed Naveed, et al.
Veröffentlicht: (2026)
Random Heterogeneous Neurochaos Learning Architecture for Data Classification
von: S, Remya Ajai A, et al.
Veröffentlicht: (2024)
von: S, Remya Ajai A, et al.
Veröffentlicht: (2024)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
von: Pan, Leyi, et al.
Veröffentlicht: (2024)
von: Pan, Leyi, et al.
Veröffentlicht: (2024)
QUAD: Quantization and Parameter-Efficient Tuning of LLM with Activation Decomposition
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
FastForward Pruning: Efficient LLM Pruning via Single-Step Reinforcement Learning
von: Yuan, Xin, et al.
Veröffentlicht: (2025)
von: Yuan, Xin, et al.
Veröffentlicht: (2025)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Dual-Phase Federated Deep Unlearning via Weight-Aware Rollback and Reconstruction
von: Zhou, Changjun, et al.
Veröffentlicht: (2025)
von: Zhou, Changjun, et al.
Veröffentlicht: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
Forget Attention: Importance-Aware Attention Is All You Need
von: Shin, Soohyeong, et al.
Veröffentlicht: (2026)
von: Shin, Soohyeong, et al.
Veröffentlicht: (2026)
Watermarking Degrades Alignment in Language Models: Analysis and Mitigation
von: Verma, Apurv, et al.
Veröffentlicht: (2025)
von: Verma, Apurv, et al.
Veröffentlicht: (2025)
Does Localization Inform Unlearning? A Rigorous Examination of Local Parameter Attribution for Knowledge Unlearning in Language Models
von: Lee, Hwiyeong, et al.
Veröffentlicht: (2025)
von: Lee, Hwiyeong, et al.
Veröffentlicht: (2025)
AMELI: Enhancing Multimodal Entity Linking with Fine-Grained Attributes
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2023)
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2023)
PersonalLLM: Tailoring LLMs to Individual Preferences
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
HYPEROFA: Expanding LLM Vocabulary to New Languages via Hypernetwork-Based Embedding Initialization
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
Linguistically-Informed Multilingual Instruction Tuning: Is There an Optimal Set of Languages to Tune?
von: Soykan, Gürkan, et al.
Veröffentlicht: (2024)
von: Soykan, Gürkan, et al.
Veröffentlicht: (2024)
Solving the Granularity Mismatch: Hierarchical Preference Learning for Long-Horizon LLM Agents
von: Gao, Heyang, et al.
Veröffentlicht: (2025)
von: Gao, Heyang, et al.
Veröffentlicht: (2025)
Every Sample Matters: Leveraging Mixture-of-Experts and High-Quality Data for Efficient and Accurate Code LLM
von: Codefuse, et al.
Veröffentlicht: (2025)
von: Codefuse, et al.
Veröffentlicht: (2025)
PairCFR: Enhancing Model Training on Paired Counterfactually Augmented Data through Contrastive Learning
von: Qiu, Xiaoqi, et al.
Veröffentlicht: (2024)
von: Qiu, Xiaoqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025) -
Quantization-Robust LLM Unlearning via Low-Rank Adaptation
von: Abitante, João Vitor Boer, et al.
Veröffentlicht: (2026) -
MMiC: Mitigating Modality Incompleteness in Clustered Federated Learning
von: Yang, Lishan, et al.
Veröffentlicht: (2025) -
Concept-based Rubrics Improve LLM Formative Assessment and Data Synthesis
von: Wei, Yuchen, et al.
Veröffentlicht: (2025) -
Mr. Snuffleupagus at SemEval-2025 Task 4: Unlearning Factual Knowledge from LLMs Using Adaptive RMU
von: Dosajh, Arjun, et al.
Veröffentlicht: (2025)