Understanding Finetuning for Factual Knowledge Extraction
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ghosal, Gaurav, Hashimoto, Tatsunori, Raghunathan, Aditi |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Language Models with Conformal Factuality Guarantees
par: Mohri, Christopher, et autres
Publié: (2024)
par: Mohri, Christopher, et autres
Publié: (2024)
Context-Parametric Inversion: Why Instruction Finetuning Can Worsen Context Reliance
par: Goyal, Sachin, et autres
Publié: (2024)
par: Goyal, Sachin, et autres
Publié: (2024)
Understanding Catastrophic Forgetting in Language Models via Implicit Inference
par: Kotha, Suhas, et autres
Publié: (2023)
par: Kotha, Suhas, et autres
Publié: (2023)
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
par: Zhong, Ziqian, et autres
Publié: (2025)
par: Zhong, Ziqian, et autres
Publié: (2025)
Improving Pretraining Data Using Perplexity Correlations
par: Thrush, Tristan, et autres
Publié: (2024)
par: Thrush, Tristan, et autres
Publié: (2024)
ImpossibleBench: Measuring LLMs' Propensity of Exploiting Test Cases
par: Zhong, Ziqian, et autres
Publié: (2025)
par: Zhong, Ziqian, et autres
Publié: (2025)
Memorization Sinks: Isolating Memorization during LLM Training
par: Ghosal, Gaurav R., et autres
Publié: (2025)
par: Ghosal, Gaurav R., et autres
Publié: (2025)
Self-Trained Verification for Training- and Test-Time Self-Improvement
par: Wu, Chen Henry, et autres
Publié: (2026)
par: Wu, Chen Henry, et autres
Publié: (2026)
Towards Reliable Latent Knowledge Estimation in LLMs: Zero-Prompt Many-Shot Based Factual Knowledge Extraction
par: Wu, Qinyuan, et autres
Publié: (2024)
par: Wu, Qinyuan, et autres
Publié: (2024)
Mitigating Bias in RAG: Controlling the Embedder
par: Kim, Taeyoun, et autres
Publié: (2025)
par: Kim, Taeyoun, et autres
Publié: (2025)
Testing the Limits of Jailbreaking Defenses with the Purple Problem
par: Kim, Taeyoun, et autres
Publié: (2024)
par: Kim, Taeyoun, et autres
Publié: (2024)
Observational Scaling Laws and the Predictability of Language Model Performance
par: Ruan, Yangjun, et autres
Publié: (2024)
par: Ruan, Yangjun, et autres
Publié: (2024)
Mode-Conditioning Unlocks Superior Test-Time Scaling
par: Wu, Chen Henry, et autres
Publié: (2025)
par: Wu, Chen Henry, et autres
Publié: (2025)
Understanding Contextual Recall in Transformers: How Finetuning Enables In-Context Reasoning over Pretraining Knowledge
par: Vasudeva, Bhavya, et autres
Publié: (2026)
par: Vasudeva, Bhavya, et autres
Publié: (2026)
Putting It All into Context: Simplifying Agents with LCLMs
par: Jiang, Mingjian, et autres
Publié: (2025)
par: Jiang, Mingjian, et autres
Publié: (2025)
Persuasion Tokens for Editing Factual Knowledge in LLMs
par: Youssef, Paul, et autres
Publié: (2026)
par: Youssef, Paul, et autres
Publié: (2026)
Linguistic Calibration of Long-Form Generations
par: Band, Neil, et autres
Publié: (2024)
par: Band, Neil, et autres
Publié: (2024)
Robust Distortion-free Watermarks for Language Models
par: Kuditipudi, Rohith, et autres
Publié: (2023)
par: Kuditipudi, Rohith, et autres
Publié: (2023)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
par: Si, Chenglei, et autres
Publié: (2024)
par: Si, Chenglei, et autres
Publié: (2024)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
par: Si, Chenglei, et autres
Publié: (2025)
par: Si, Chenglei, et autres
Publié: (2025)
Repetition Improves Language Model Embeddings
par: Springer, Jacob Mitchell, et autres
Publié: (2024)
par: Springer, Jacob Mitchell, et autres
Publié: (2024)
Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting
par: Watts, Ishaan, et autres
Publié: (2026)
par: Watts, Ishaan, et autres
Publié: (2026)
Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
par: Dubois, Yann, et autres
Publié: (2024)
par: Dubois, Yann, et autres
Publié: (2024)
On the Learnability of Watermarks for Language Models
par: Gu, Chenchen, et autres
Publié: (2023)
par: Gu, Chenchen, et autres
Publié: (2023)
Reasoning to Learn from Latent Thoughts
par: Ruan, Yangjun, et autres
Publié: (2025)
par: Ruan, Yangjun, et autres
Publié: (2025)
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction
par: Nagarajan, Vaishnavh, et autres
Publié: (2025)
par: Nagarajan, Vaishnavh, et autres
Publié: (2025)
KnowLA: Enhancing Parameter-efficient Finetuning with Knowledgeable Adaptation
par: Luo, Xindi, et autres
Publié: (2024)
par: Luo, Xindi, et autres
Publié: (2024)
From Style to Facts: Mapping the Boundaries of Knowledge Injection with Finetuning
par: Zhao, Eric, et autres
Publié: (2025)
par: Zhao, Eric, et autres
Publié: (2025)
Synthetic continued pretraining
par: Yang, Zitong, et autres
Publié: (2024)
par: Yang, Zitong, et autres
Publié: (2024)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
par: Jiang, Mingjian, et autres
Publié: (2024)
par: Jiang, Mingjian, et autres
Publié: (2024)
Agentic Adversarial QA for Improving Domain-Specific LLMs
par: Grari, Vincent, et autres
Publié: (2026)
par: Grari, Vincent, et autres
Publié: (2026)
Alternate Preference Optimization for Unlearning Factual Knowledge in Large Language Models
par: Mekala, Anmol, et autres
Publié: (2024)
par: Mekala, Anmol, et autres
Publié: (2024)
AutoBencher: Towards Declarative Benchmark Construction
par: Li, Xiang Lisa, et autres
Publié: (2024)
par: Li, Xiang Lisa, et autres
Publié: (2024)
Polynomial Regression as a Task for Understanding In-context Learning Through Finetuning and Alignment
par: Wilcoxson, Max, et autres
Publié: (2024)
par: Wilcoxson, Max, et autres
Publié: (2024)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
par: Wang, Qianli, et autres
Publié: (2025)
par: Wang, Qianli, et autres
Publié: (2025)
Failure Modes of LLMs for Causal Reasoning on Narratives
par: Yamin, Khurram, et autres
Publié: (2024)
par: Yamin, Khurram, et autres
Publié: (2024)
Auditing Prompt Caching in Language Model APIs
par: Gu, Chenchen, et autres
Publié: (2025)
par: Gu, Chenchen, et autres
Publié: (2025)
Understanding Factual Recall in Transformers via Associative Memories
par: Nichani, Eshaan, et autres
Publié: (2024)
par: Nichani, Eshaan, et autres
Publié: (2024)
PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective
par: Huang, Yangyi, et autres
Publié: (2026)
par: Huang, Yangyi, et autres
Publié: (2026)
Factual Knowledge in Language Models: Robustness and Anomalies under Simple Temporal Context Variations
par: Khodja, Hichem Ammar, et autres
Publié: (2025)
par: Khodja, Hichem Ammar, et autres
Publié: (2025)
Documents similaires
-
Language Models with Conformal Factuality Guarantees
par: Mohri, Christopher, et autres
Publié: (2024) -
Context-Parametric Inversion: Why Instruction Finetuning Can Worsen Context Reliance
par: Goyal, Sachin, et autres
Publié: (2024) -
Understanding Catastrophic Forgetting in Language Models via Implicit Inference
par: Kotha, Suhas, et autres
Publié: (2023) -
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
par: Zhong, Ziqian, et autres
Publié: (2025) -
Improving Pretraining Data Using Perplexity Correlations
par: Thrush, Tristan, et autres
Publié: (2024)