Gespeichert in:
| Hauptverfasser: | Liu, Zhihao, Lou, Jian, Hu, Yuke, Li, Xiaochen, Chen, Yitian, Chen, Tailun, Qin, Zhizhen, Ren, Kui, Qin, Zhan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2508.20443 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Evaluation for Real-World LLM Unlearning
von: Miao, Ke, et al.
Veröffentlicht: (2025)
von: Miao, Ke, et al.
Veröffentlicht: (2025)
Module-Aware Parameter-Efficient Machine Unlearning on Transformers
von: Bao, Wenjie, et al.
Veröffentlicht: (2025)
von: Bao, Wenjie, et al.
Veröffentlicht: (2025)
ERASER: Machine Unlearning in MLaaS via an Inference Serving-Aware Approach
von: Hu, Yuke, et al.
Veröffentlicht: (2023)
von: Hu, Yuke, et al.
Veröffentlicht: (2023)
Certified Minimax Unlearning with Generalization Rates and Deletion Capacity
von: Liu, Jiaqi, et al.
Veröffentlicht: (2023)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2023)
MIRAGE: Misleading Retrieval-Augmented Generation via Black-box and Query-agnostic Poisoning Attacks
von: Chen, Tailun, et al.
Veröffentlicht: (2025)
von: Chen, Tailun, et al.
Veröffentlicht: (2025)
Membership Inference Attacks Against Vision-Language Models
von: Hu, Yuke, et al.
Veröffentlicht: (2025)
von: Hu, Yuke, et al.
Veröffentlicht: (2025)
JANUS: A Lightweight Framework for Jailbreaking Text-to-Image Models via Distribution Optimization
von: Zheng, Haolun, et al.
Veröffentlicht: (2026)
von: Zheng, Haolun, et al.
Veröffentlicht: (2026)
FINER: Enhancing State-of-the-art Classifiers with Feature Attribution to Facilitate Security Analysis
von: He, Yiling, et al.
Veröffentlicht: (2023)
von: He, Yiling, et al.
Veröffentlicht: (2023)
Shadow in the Cache: Unveiling and Mitigating Privacy Risks of KV-cache in LLM Inference
von: Luo, Zhifan, et al.
Veröffentlicht: (2025)
von: Luo, Zhifan, et al.
Veröffentlicht: (2025)
Prompting Forgetting: Unlearning in GANs via Textual Guidance
von: Nagasubramaniam, Piyush, et al.
Veröffentlicht: (2025)
von: Nagasubramaniam, Piyush, et al.
Veröffentlicht: (2025)
SWAT: A System-Wide Approach to Tunable Leakage Mitigation in Encrypted Data Stores
von: Zheng, Leqian, et al.
Veröffentlicht: (2023)
von: Zheng, Leqian, et al.
Veröffentlicht: (2023)
Aligned but Fragile: Enhancing LLM Safety Robustness via Zeroth-Order Optimization
von: Liu, Zhihao, et al.
Veröffentlicht: (2026)
von: Liu, Zhihao, et al.
Veröffentlicht: (2026)
UIPE: Enhancing LLM Unlearning by Removing Knowledge Related to Forgetting Targets
von: Wang, Wenyu, et al.
Veröffentlicht: (2025)
von: Wang, Wenyu, et al.
Veröffentlicht: (2025)
Machine Unlearning under Retain-Forget Entanglement
von: Cheng, Jingpu, et al.
Veröffentlicht: (2026)
von: Cheng, Jingpu, et al.
Veröffentlicht: (2026)
Releasing Malevolence from Benevolence: The Menace of Benign Data on Machine Unlearning
von: Ma, Binhao, et al.
Veröffentlicht: (2024)
von: Ma, Binhao, et al.
Veröffentlicht: (2024)
Forgetting to Forget: Attention Sink as A Gateway for Backdooring LLM Unlearning
von: Shang, Bingqi, et al.
Veröffentlicht: (2025)
von: Shang, Bingqi, et al.
Veröffentlicht: (2025)
Mitigating Privacy Risk via Forget Set-Free Unlearning
von: Newatia, Aviraj, et al.
Veröffentlicht: (2026)
von: Newatia, Aviraj, et al.
Veröffentlicht: (2026)
Eguard: Defending LLM Embeddings Against Inversion Attacks via Text Mutual Information Optimization
von: Liu, Tiantian, et al.
Veröffentlicht: (2024)
von: Liu, Tiantian, et al.
Veröffentlicht: (2024)
To Forget or Not? Towards Practical Knowledge Unlearning for Large Language Models
von: Tian, Bozhong, et al.
Veröffentlicht: (2024)
von: Tian, Bozhong, et al.
Veröffentlicht: (2024)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
Quantifying and Defending against Privacy Threats on Federated Knowledge Graph Embedding
von: Hu, Yuke, et al.
Veröffentlicht: (2023)
von: Hu, Yuke, et al.
Veröffentlicht: (2023)
Combating Concept Drift with Explanatory Detection and Adaptation for Android Malware Classification
von: He, Yiling, et al.
Veröffentlicht: (2024)
von: He, Yiling, et al.
Veröffentlicht: (2024)
LLM-Guided Multi-View Hypergraph Learning for Human-Centric Explainable Recommendation
von: Chu, Zhixuan, et al.
Veröffentlicht: (2024)
von: Chu, Zhixuan, et al.
Veröffentlicht: (2024)
LLM Unlearning via Loss Adjustment with Only Forget Data
von: Wang, Yaxuan, et al.
Veröffentlicht: (2024)
von: Wang, Yaxuan, et al.
Veröffentlicht: (2024)
Towards Aligned Data Forgetting via Twin Machine Unlearning
von: Niu, Zhenxing, et al.
Veröffentlicht: (2025)
von: Niu, Zhenxing, et al.
Veröffentlicht: (2025)
Machine Unlearning for Streaming Forgetting
von: Shen, Shaofei, et al.
Veröffentlicht: (2025)
von: Shen, Shaofei, et al.
Veröffentlicht: (2025)
Forget-It-All: Multi-Concept Machine Unlearning via Concept-Aware Neuron Masking
von: Deng, Kaiyuan, et al.
Veröffentlicht: (2026)
von: Deng, Kaiyuan, et al.
Veröffentlicht: (2026)
Prompt-Consistency Image Generation (PCIG): A Unified Framework Integrating LLMs, Knowledge Graphs, and Controllable Diffusion Models
von: Sun, Yichen, et al.
Veröffentlicht: (2024)
von: Sun, Yichen, et al.
Veröffentlicht: (2024)
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
Don't Say No: Jailbreaking LLM by Suppressing Refusal
von: Zhou, Yukai, et al.
Veröffentlicht: (2024)
von: Zhou, Yukai, et al.
Veröffentlicht: (2024)
Towards Reliable Forgetting: A Survey on Machine Unlearning Verification
von: Xue, Lulu, et al.
Veröffentlicht: (2025)
von: Xue, Lulu, et al.
Veröffentlicht: (2025)
Forgetting-MarI: LLM Unlearning via Marginal Information Regularization
von: Xu, Shizhou, et al.
Veröffentlicht: (2025)
von: Xu, Shizhou, et al.
Veröffentlicht: (2025)
BLUR: A Benchmark for LLM Unlearning Robust to Forget-Retain Overlap
von: Hu, Shengyuan, et al.
Veröffentlicht: (2025)
von: Hu, Shengyuan, et al.
Veröffentlicht: (2025)
Duplex-GS: Proxy-Guided Weighted Blending for Real-Time Order-Independent Gaussian Splatting
von: Liu, Weihang, et al.
Veröffentlicht: (2025)
von: Liu, Weihang, et al.
Veröffentlicht: (2025)
Machine Unlearning in Speech Emotion Recognition via Forget Set Alone
von: Ren, Zhao, et al.
Veröffentlicht: (2025)
von: Ren, Zhao, et al.
Veröffentlicht: (2025)
Recover-to-Forget: Gradient Reconstruction from LoRA for Efficient LLM Unlearning
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
HarmMetric Eval: Benchmarking Metrics and Judges for LLM Harmfulness Assessment
von: Yang, Langqi, et al.
Veröffentlicht: (2025)
von: Yang, Langqi, et al.
Veröffentlicht: (2025)
Textual Unlearning Gives a False Sense of Unlearning
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
DualMind: Towards Understanding Cognitive-Affective Cascades in Public Opinion Dissemination via Multi-Agent Simulation
von: Huang, Enhao, et al.
Veröffentlicht: (2026)
von: Huang, Enhao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Evaluation for Real-World LLM Unlearning
von: Miao, Ke, et al.
Veröffentlicht: (2025) -
Module-Aware Parameter-Efficient Machine Unlearning on Transformers
von: Bao, Wenjie, et al.
Veröffentlicht: (2025) -
ERASER: Machine Unlearning in MLaaS via an Inference Serving-Aware Approach
von: Hu, Yuke, et al.
Veröffentlicht: (2023) -
Certified Minimax Unlearning with Generalization Rates and Deletion Capacity
von: Liu, Jiaqi, et al.
Veröffentlicht: (2023) -
MIRAGE: Misleading Retrieval-Augmented Generation via Black-box and Query-agnostic Poisoning Attacks
von: Chen, Tailun, et al.
Veröffentlicht: (2025)