Anatomy of Unlearning: The Dual Impact of Fact Salience and Model Fine-Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Borisiuk, Anna, Savchenko, Andrey, Panchenko, Alexander, Tutubalina, Elena |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SynthDetoxM: Modern LLMs are Few-Shot Parallel Detoxification Data Annotators
von: Moskovskiy, Daniil, et al.
Veröffentlicht: (2025)
von: Moskovskiy, Daniil, et al.
Veröffentlicht: (2025)
The Chronicles of RiDiC: Generating Datasets with Controlled Popularity Distribution for Long-form Factuality Evaluation
von: Braslavski, Pavel, et al.
Veröffentlicht: (2026)
von: Braslavski, Pavel, et al.
Veröffentlicht: (2026)
Leveraging LLM Parametric Knowledge for Fact Checking without Retrieval
von: Vazhentsev, Artem, et al.
Veröffentlicht: (2026)
von: Vazhentsev, Artem, et al.
Veröffentlicht: (2026)
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
von: Sakhovskiy, Andrey, et al.
Veröffentlicht: (2025)
von: Sakhovskiy, Andrey, et al.
Veröffentlicht: (2025)
Evolutionary Search for Automated Design of Uncertainty Quantification Methods
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2026)
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2026)
The benefits of query-based KGQA systems for complex and temporal questions in LLM era
von: Alekseev, Artem, et al.
Veröffentlicht: (2025)
von: Alekseev, Artem, et al.
Veröffentlicht: (2025)
When Punctuation Matters: A Large-Scale Comparison of Prompt Robustness Methods for LLMs
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2025)
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2025)
Team Anotheroption at SemEval-2025 Task 8: Bridging the Gap Between Open-Source and Proprietary LLMs in Table QA
von: Evkarpidi, Nikolas, et al.
Veröffentlicht: (2025)
von: Evkarpidi, Nikolas, et al.
Veröffentlicht: (2025)
Refining Salience-Aware Sparse Fine-Tuning Strategies for Language Models
von: Liu, Xinxin, et al.
Veröffentlicht: (2024)
von: Liu, Xinxin, et al.
Veröffentlicht: (2024)
Dissecting Fine-Tuning Unlearning in Large Language Models
von: Hong, Yihuai, et al.
Veröffentlicht: (2024)
von: Hong, Yihuai, et al.
Veröffentlicht: (2024)
Confidence Estimation for Error Detection in Text-to-SQL Systems
von: Somov, Oleg, et al.
Veröffentlicht: (2025)
von: Somov, Oleg, et al.
Veröffentlicht: (2025)
SONAR-LLM: Autoregressive Transformer that Thinks in Sentence Embeddings and Speaks in Tokens
von: Dragunov, Nikita, et al.
Veröffentlicht: (2025)
von: Dragunov, Nikita, et al.
Veröffentlicht: (2025)
SparseGrad: A Selective Method for Efficient Fine-tuning of MLP Layers
von: Chekalina, Viktoriia, et al.
Veröffentlicht: (2024)
von: Chekalina, Viktoriia, et al.
Veröffentlicht: (2024)
One Task Vector is not Enough: A Large-Scale Study for In-Context Learning
von: Tikhonov, Pavel, et al.
Veröffentlicht: (2025)
von: Tikhonov, Pavel, et al.
Veröffentlicht: (2025)
CLEAR: Character Unlearning in Textual and Visual Modalities
von: Dontsov, Alexey, et al.
Veröffentlicht: (2024)
von: Dontsov, Alexey, et al.
Veröffentlicht: (2024)
Geopolitical biases in LLMs: what are the "good" and the "bad" countries according to contemporary language models
von: Salnikov, Mikhail, et al.
Veröffentlicht: (2025)
von: Salnikov, Mikhail, et al.
Veröffentlicht: (2025)
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
I Have Covered All the Bases Here: Interpreting Reasoning Features in Large Language Models via Sparse Autoencoders
von: Galichin, Andrey, et al.
Veröffentlicht: (2025)
von: Galichin, Andrey, et al.
Veröffentlicht: (2025)
Latent Knowledge as a Predictor of Fact Acquisition in Fine-Tuned Large Language Models
von: Hier, Daniel B., et al.
Veröffentlicht: (2026)
von: Hier, Daniel B., et al.
Veröffentlicht: (2026)
Surprising Efficacy of Fine-Tuned Transformers for Fact-Checking over Larger Language Models
von: Setty, Vinay
Veröffentlicht: (2024)
von: Setty, Vinay
Veröffentlicht: (2024)
FineDialFact: A benchmark for Fine-grained Dialogue Fact Verification
von: Chen, Xiangyan, et al.
Veröffentlicht: (2025)
von: Chen, Xiangyan, et al.
Veröffentlicht: (2025)
Confidence Is All You Need: Few-Shot RL Fine-Tuning of Language Models
von: Li, Pengyi, et al.
Veröffentlicht: (2025)
von: Li, Pengyi, et al.
Veröffentlicht: (2025)
Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs
von: Afonin, Nikita, et al.
Veröffentlicht: (2025)
von: Afonin, Nikita, et al.
Veröffentlicht: (2025)
Beyond Detection: Rethinking Education in the Age of AI-writing
von: Marina, Maria, et al.
Veröffentlicht: (2026)
von: Marina, Maria, et al.
Veröffentlicht: (2026)
Beyond QA Pairs: Assessing Parameter-Efficient Fine-Tuning for Fact Embedding in LLMs
von: Ratnakar, Shivam, et al.
Veröffentlicht: (2025)
von: Ratnakar, Shivam, et al.
Veröffentlicht: (2025)
FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing
von: Wang, Peng, et al.
Veröffentlicht: (2026)
von: Wang, Peng, et al.
Veröffentlicht: (2026)
On the Impact of Fine-Tuning on Chain-of-Thought Reasoning
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
MEOW: MEMOry Supervised LLM Unlearning Via Inverted Facts
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
LUNE: Efficient LLM Unlearning via LoRA Fine-Tuning with Negative Examples
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
Cross-Lingual Learning vs. Low-Resource Fine-Tuning: A Case Study with Fact-Checking in Turkish
von: Cekinel, Recep Firat, et al.
Veröffentlicht: (2024)
von: Cekinel, Recep Firat, et al.
Veröffentlicht: (2024)
Reviving Your MNEME: Predicting The Side Effects of LLM Unlearning and Fine-Tuning via Sparse Model Diffing
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
SPARTA: Evaluating Reasoning Segmentation Robustness through Black-Box Adversarial Paraphrasing in Text Autoencoder Latent Space
von: Zinkovich, Viktoriia, et al.
Veröffentlicht: (2025)
von: Zinkovich, Viktoriia, et al.
Veröffentlicht: (2025)
Impact of Fine-Tuning Methods on Memorization in Large Language Models
von: Hou, Jie, et al.
Veröffentlicht: (2025)
von: Hou, Jie, et al.
Veröffentlicht: (2025)
MultiParaDetox: Extending Text Detoxification with Parallel Data to New Languages
von: Dementieva, Daryna, et al.
Veröffentlicht: (2024)
von: Dementieva, Daryna, et al.
Veröffentlicht: (2024)
<think> So let's replace this phrase with insult... </think> Lessons learned from generation of toxic texts with LLMs
von: Pletenev, Sergey, et al.
Veröffentlicht: (2025)
von: Pletenev, Sergey, et al.
Veröffentlicht: (2025)
S3: A Simple Strong Sample-effective Multimodal Dialog System
von: Rykov, Elisei, et al.
Veröffentlicht: (2024)
von: Rykov, Elisei, et al.
Veröffentlicht: (2024)
Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2024)
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2024)
FactLens: Benchmarking Fine-Grained Fact Verification
von: Mitra, Kushan, et al.
Veröffentlicht: (2024)
von: Mitra, Kushan, et al.
Veröffentlicht: (2024)
TaxoLLaMA: WordNet-based Model for Solving Multiple Lexical Semantic Tasks
von: Moskvoretskii, Viktor, et al.
Veröffentlicht: (2024)
von: Moskvoretskii, Viktor, et al.
Veröffentlicht: (2024)
Team Trifecta at Factify5WQA: Setting the Standard in Fact Verification with Fine-Tuning
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SynthDetoxM: Modern LLMs are Few-Shot Parallel Detoxification Data Annotators
von: Moskovskiy, Daniil, et al.
Veröffentlicht: (2025) -
The Chronicles of RiDiC: Generating Datasets with Controlled Popularity Distribution for Long-form Factuality Evaluation
von: Braslavski, Pavel, et al.
Veröffentlicht: (2026) -
Leveraging LLM Parametric Knowledge for Fact Checking without Retrieval
von: Vazhentsev, Artem, et al.
Veröffentlicht: (2026) -
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
von: Sakhovskiy, Andrey, et al.
Veröffentlicht: (2025) -
Evolutionary Search for Automated Design of Uncertainty Quantification Methods
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2026)