An Information Theoretic Evaluation Metric For Strong Unlearning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jeon, Dongjae, Jeung, Wonje, Kim, Taeheon, No, Albert, Choi, Jonghyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-Level Knowledge Distillation and Dynamic Self-Supervised Learning for Continual Learning
von: Kim, Taeheon, et al.
Veröffentlicht: (2025)
von: Kim, Taeheon, et al.
Veröffentlicht: (2025)
TTA-DAME: Test-Time Adaptation with Domain Augmentation and Model Ensemble for Dynamic Driving Conditions
von: Jeon, Dongjae, et al.
Veröffentlicht: (2025)
von: Jeon, Dongjae, et al.
Veröffentlicht: (2025)
Rethinking Benign Relearning: Syntax as the Hidden Driver of Unlearning Failures
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2026)
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2026)
Rainbow Padding: Mitigating Early Termination in Instruction-Tuned Diffusion LLMs
von: Kim, Bumjun, et al.
Veröffentlicht: (2025)
von: Kim, Bumjun, et al.
Veröffentlicht: (2025)
Few-Shot Truly Benign DPO Attack for Jailbreaking LLMs
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2026)
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2026)
Understanding and Mitigating Memorization in Generative Models via Sharpness of Probability Landscapes
von: Jeon, Dongjae, et al.
Veröffentlicht: (2024)
von: Jeon, Dongjae, et al.
Veröffentlicht: (2024)
VLMs Trace Without Tracking: Diagnosing Failures in Visual Path Following
von: Hong, Hyesoo, et al.
Veröffentlicht: (2026)
von: Hong, Hyesoo, et al.
Veröffentlicht: (2026)
Preserve-Then-Quantize: Balancing Rank Budgets for Quantization Error Reconstruction in LLMs
von: Cho, Yoonjun, et al.
Veröffentlicht: (2026)
von: Cho, Yoonjun, et al.
Veröffentlicht: (2026)
Large Language Models Still Exhibit Bias in Long Text
von: Jeung, Wonje, et al.
Veröffentlicht: (2024)
von: Jeung, Wonje, et al.
Veröffentlicht: (2024)
Co-LoRA: Collaborative Model Personalization on Heterogeneous Multi-Modal Clients
von: Seo, Minhyuk, et al.
Veröffentlicht: (2025)
von: Seo, Minhyuk, et al.
Veröffentlicht: (2025)
Learning Equi-angular Representations for Online Continual Learning
von: Seo, Minhyuk, et al.
Veröffentlicht: (2024)
von: Seo, Minhyuk, et al.
Veröffentlicht: (2024)
A2D: Any-Order, Any-Step Safety Alignment for Diffusion Language Models
von: Jeung, Wonje, et al.
Veröffentlicht: (2025)
von: Jeung, Wonje, et al.
Veröffentlicht: (2025)
ReALFRED: An Embodied Instruction Following Benchmark in Photo-Realistic Environments
von: Kim, Taewoong, et al.
Veröffentlicht: (2024)
von: Kim, Taewoong, et al.
Veröffentlicht: (2024)
Information-Theoretic Discrete Diffusion
von: Jeon, Moongyu, et al.
Veröffentlicht: (2025)
von: Jeon, Moongyu, et al.
Veröffentlicht: (2025)
Assigning Distinct Roles to Quantized and Low-Rank Matrices Toward Optimal Weight Decomposition
von: Cho, Yoonjun, et al.
Veröffentlicht: (2025)
von: Cho, Yoonjun, et al.
Veröffentlicht: (2025)
Adversarial Sample-Based Approach for Tighter Privacy Auditing in Final Model-Only Scenarios
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2024)
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2024)
An Information Theoretic Approach to Machine Unlearning
von: Foster, Jack, et al.
Veröffentlicht: (2024)
von: Foster, Jack, et al.
Veröffentlicht: (2024)
Online Continual Learning For Interactive Instruction Following Agents
von: Kim, Byeonghwi, et al.
Veröffentlicht: (2024)
von: Kim, Byeonghwi, et al.
Veröffentlicht: (2024)
Incremental Learning of Retrievable Skills For Efficient Continual Task Adaptation
von: Lee, Daehee, et al.
Veröffentlicht: (2024)
von: Lee, Daehee, et al.
Veröffentlicht: (2024)
Representation Bending for Large Language Model Safety
von: Yousefpour, Ashkan, et al.
Veröffentlicht: (2025)
von: Yousefpour, Ashkan, et al.
Veröffentlicht: (2025)
Information-Theoretic Foundations for Machine Learning
von: Jeon, Hong Jun, et al.
Veröffentlicht: (2024)
von: Jeon, Hong Jun, et al.
Veröffentlicht: (2024)
SAFEPATH: Preventing Harmful Reasoning in Chain-of-Thought via Early Alignment
von: Jeung, Wonje, et al.
Veröffentlicht: (2025)
von: Jeung, Wonje, et al.
Veröffentlicht: (2025)
Weak-to-Strong Generalization under Distribution Shifts
von: Jeon, Myeongho, et al.
Veröffentlicht: (2025)
von: Jeon, Myeongho, et al.
Veröffentlicht: (2025)
Becoming Experienced Judges: Selective Test-Time Learning for Evaluators
von: Jwa, Seungyeon, et al.
Veröffentlicht: (2025)
von: Jwa, Seungyeon, et al.
Veröffentlicht: (2025)
Information-Theoretic Foundations for Neural Scaling Laws
von: Jeon, Hong Jun, et al.
Veröffentlicht: (2024)
von: Jeon, Hong Jun, et al.
Veröffentlicht: (2024)
Machine Unlearning via Information Theoretic Regularization
von: Xu, Shizhou, et al.
Veröffentlicht: (2025)
von: Xu, Shizhou, et al.
Veröffentlicht: (2025)
ROKA: Robust Knowledge Unlearning against Adversaries
von: Shin, Jinmyeong, et al.
Veröffentlicht: (2026)
von: Shin, Jinmyeong, et al.
Veröffentlicht: (2026)
MOGAM: A Multimodal Object-oriented Graph Attention Model for Depression Detection
von: Cha, Junyeop, et al.
Veröffentlicht: (2024)
von: Cha, Junyeop, et al.
Veröffentlicht: (2024)
DAPD: Dependency-Aware Parallel Decoding via Attention for Diffusion LLMs
von: Kim, Bumjun, et al.
Veröffentlicht: (2026)
von: Kim, Bumjun, et al.
Veröffentlicht: (2026)
Unlearning Information Bottleneck: Machine Unlearning of Systematic Patterns and Biases
von: Han, Ling, et al.
Veröffentlicht: (2024)
von: Han, Ling, et al.
Veröffentlicht: (2024)
R-TOFU: Unlearning in Large Reasoning Models
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2025)
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2025)
Forget and Explain: Transparent Verification of GNN Unlearning
von: Ahsan, Imran, et al.
Veröffentlicht: (2025)
von: Ahsan, Imran, et al.
Veröffentlicht: (2025)
SCALE: Self-uncertainty Conditioned Adaptive Looking and Execution for Vision-Language-Action Models
von: Choi, Hyeonbeom, et al.
Veröffentlicht: (2026)
von: Choi, Hyeonbeom, et al.
Veröffentlicht: (2026)
SEPS: A Separability Measure for Robust Unlearning in LLMs
von: Jeung, Wonje, et al.
Veröffentlicht: (2025)
von: Jeung, Wonje, et al.
Veröffentlicht: (2025)
Budgeted Online Continual Learning by Adaptive Layer Freezing and Frequency-based Sampling
von: Seo, Minhyuk, et al.
Veröffentlicht: (2024)
von: Seo, Minhyuk, et al.
Veröffentlicht: (2024)
Forget What Matters, Keep the Rest: Selective Unlearning of Informative Tokens
von: Koh, Seunghee, et al.
Veröffentlicht: (2026)
von: Koh, Seunghee, et al.
Veröffentlicht: (2026)
Hospitality-VQA: Decision-Oriented Informativeness Evaluation for Vision-Language Models
von: Lee, Jeongwoo, et al.
Veröffentlicht: (2026)
von: Lee, Jeongwoo, et al.
Veröffentlicht: (2026)
Forgetting Any Data at Any Time: A Theoretically Certified Unlearning Framework for Vertical Federated Learning
von: Wang, Linian, et al.
Veröffentlicht: (2025)
von: Wang, Linian, et al.
Veröffentlicht: (2025)
Efficient Knowledge Graph Unlearning with Zeroth-order Information
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Auditing Language Model Unlearning via Information Decomposition
von: Goel, Anmol, et al.
Veröffentlicht: (2026)
von: Goel, Anmol, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Multi-Level Knowledge Distillation and Dynamic Self-Supervised Learning for Continual Learning
von: Kim, Taeheon, et al.
Veröffentlicht: (2025) -
TTA-DAME: Test-Time Adaptation with Domain Augmentation and Model Ensemble for Dynamic Driving Conditions
von: Jeon, Dongjae, et al.
Veröffentlicht: (2025) -
Rethinking Benign Relearning: Syntax as the Hidden Driver of Unlearning Failures
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2026) -
Rainbow Padding: Mitigating Early Termination in Instruction-Tuned Diffusion LLMs
von: Kim, Bumjun, et al.
Veröffentlicht: (2025) -
Few-Shot Truly Benign DPO Attack for Jailbreaking LLMs
von: Yoon, Sangyeon, et al.
Veröffentlicht: (2026)