Fact Recall, Heuristics or Pure Guesswork? Precise Interpretations of Language Models for Fact Completion
Fuente:
arXiv
Saved in:
| Main Authors: | Saynova, Denitsa, Hagström, Lovisa, Johansson, Moa, Johansson, Richard, Kuhlmann, Marco |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Identifying Non-Replicable Social Science Studies with Language Models
by: Saynova, Denitsa, et al.
Published: (2025)
by: Saynova, Denitsa, et al.
Published: (2025)
Language Model Re-rankers are Fooled by Lexical Similarities
by: Hagström, Lovisa, et al.
Published: (2025)
by: Hagström, Lovisa, et al.
Published: (2025)
Setting the AI Agenda -- Evidence from Sweden in the ChatGPT Era
by: Bruinsma, Bastiaan, et al.
Published: (2024)
by: Bruinsma, Bastiaan, et al.
Published: (2024)
Fact or Guesswork? Evaluating Large Language Models' Medical Knowledge with Structured One-Hop Judgments
by: Li, Jiaxi, et al.
Published: (2025)
by: Li, Jiaxi, et al.
Published: (2025)
CUB: Benchmarking Context Utilisation Techniques for Language Models
by: Hagström, Lovisa, et al.
Published: (2025)
by: Hagström, Lovisa, et al.
Published: (2025)
To Copy or Not to Copy: Copying Is Easier to Induce Than Recall
by: Farahani, Mehrdad, et al.
Published: (2026)
by: Farahani, Mehrdad, et al.
Published: (2026)
Benchmarking Debiasing Methods for LLM-based Parameter Estimates
by: de Pieuchon, Nicolas Audinet, et al.
Published: (2025)
by: de Pieuchon, Nicolas Audinet, et al.
Published: (2025)
What Matters in Memorizing and Recalling Facts? Multifaceted Benchmarks for Knowledge Probing in Language Models
by: Zhao, Xin, et al.
Published: (2024)
by: Zhao, Xin, et al.
Published: (2024)
Time Awareness in Large Language Models: Benchmarking Fact Recall Across Time
by: Herel, David, et al.
Published: (2024)
by: Herel, David, et al.
Published: (2024)
Can Large Language Models (or Humans) Disentangle Text?
by: de Pieuchon, Nicolas Audinet, et al.
Published: (2024)
by: de Pieuchon, Nicolas Audinet, et al.
Published: (2024)
How Well Do Large Language Models Disambiguate Swedish Words?
by: Johansson, Richard
Published: (2024)
by: Johansson, Richard
Published: (2024)
Reasoning in Transformers -- Mitigating Spurious Correlations and Reasoning Shortcuts
by: Enström, Daniel, et al.
Published: (2024)
by: Enström, Daniel, et al.
Published: (2024)
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
by: Pradeep, Ronak, et al.
Published: (2025)
by: Pradeep, Ronak, et al.
Published: (2025)
Do Factual Recall Mechanisms Carry over from Text to Speech in Multimodal Language Models?
by: Modica, Luca, et al.
Published: (2026)
by: Modica, Luca, et al.
Published: (2026)
Summing Up the Facts: Additive Mechanisms Behind Factual Recall in LLMs
by: Chughtai, Bilal, et al.
Published: (2024)
by: Chughtai, Bilal, et al.
Published: (2024)
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
by: Zhang, Ying, et al.
Published: (2025)
by: Zhang, Ying, et al.
Published: (2025)
Deciphering the Interplay of Parametric and Non-parametric Memory in Retrieval-augmented Language Models
by: Farahani, Mehrdad, et al.
Published: (2024)
by: Farahani, Mehrdad, et al.
Published: (2024)
Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking
by: Russo, Daniel, et al.
Published: (2024)
by: Russo, Daniel, et al.
Published: (2024)
PACE: Procedural Abstractions for Communicating Efficiently
by: Thomas, Jonathan D., et al.
Published: (2024)
by: Thomas, Jonathan D., et al.
Published: (2024)
Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling
by: Yang, Linyao, et al.
Published: (2023)
by: Yang, Linyao, et al.
Published: (2023)
FactSim: Fact-Checking for Opinion Summarization
by: Anghinoni, Leandro, et al.
Published: (2026)
by: Anghinoni, Leandro, et al.
Published: (2026)
Learning Efficient Recursive Numeral Systems via Reinforcement Learning
by: Silvi, Andrea, et al.
Published: (2024)
by: Silvi, Andrea, et al.
Published: (2024)
Extractive Fact Decomposition for Interpretable Natural Language Inference in one Forward Pass
by: Popovič, Nicholas, et al.
Published: (2025)
by: Popovič, Nicholas, et al.
Published: (2025)
VeriFact: Enhancing Long-Form Factuality Evaluation with Refined Fact Extraction and Reference Facts
by: Liu, Xin, et al.
Published: (2025)
by: Liu, Xin, et al.
Published: (2025)
ChronoFact: Timeline-based Temporal Fact Verification
by: Barik, Anab Maulana, et al.
Published: (2024)
by: Barik, Anab Maulana, et al.
Published: (2024)
MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts
by: He, Jiayi, et al.
Published: (2025)
by: He, Jiayi, et al.
Published: (2025)
What Happens to a Dataset Transformed by a Projection-based Concept Removal Method?
by: Johansson, Richard
Published: (2024)
by: Johansson, Richard
Published: (2024)
TrendFact: A Benchmark for Explainable Hotspot Perception in Fact-Checking with Natural Language Explanation
by: Zhang, Xiaocheng, et al.
Published: (2024)
by: Zhang, Xiaocheng, et al.
Published: (2024)
How Do Multilingual Language Models Remember Facts?
by: Fierro, Constanza, et al.
Published: (2024)
by: Fierro, Constanza, et al.
Published: (2024)
Logical Consistency of Large Language Models in Fact-checking
by: Ghosh, Bishwamittra, et al.
Published: (2024)
by: Ghosh, Bishwamittra, et al.
Published: (2024)
Scaling Laws for Fact Memorization of Large Language Models
by: Lu, Xingyu, et al.
Published: (2024)
by: Lu, Xingyu, et al.
Published: (2024)
Language Models Hallucinate, but May Excel at Fact Verification
by: Guan, Jian, et al.
Published: (2023)
by: Guan, Jian, et al.
Published: (2023)
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Do All Autoregressive Transformers Remember Facts the Same Way? A Cross-Architecture Analysis of Recall Mechanisms
by: Choe, Minyeong, et al.
Published: (2025)
by: Choe, Minyeong, et al.
Published: (2025)
Molecular Facts: Desiderata for Decontextualization in LLM Fact Verification
by: Gunjal, Anisha, et al.
Published: (2024)
by: Gunjal, Anisha, et al.
Published: (2024)
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models
by: Satriani, Dario, et al.
Published: (2025)
by: Satriani, Dario, et al.
Published: (2025)
MuLan: A Study of Fact Mutability in Language Models
by: Fierro, Constanza, et al.
Published: (2024)
by: Fierro, Constanza, et al.
Published: (2024)
Scalable Influence and Fact Tracing for Large Language Model Pretraining
by: Chang, Tyler A., et al.
Published: (2024)
by: Chang, Tyler A., et al.
Published: (2024)
Kongzi: A Historical Large Language Model with Fact Enhancement
by: Yang, Jiashu, et al.
Published: (2025)
by: Yang, Jiashu, et al.
Published: (2025)
The Perils & Promises of Fact-checking with Large Language Models
by: Quelle, Dorian, et al.
Published: (2023)
by: Quelle, Dorian, et al.
Published: (2023)
Similar Items
-
Identifying Non-Replicable Social Science Studies with Language Models
by: Saynova, Denitsa, et al.
Published: (2025) -
Language Model Re-rankers are Fooled by Lexical Similarities
by: Hagström, Lovisa, et al.
Published: (2025) -
Setting the AI Agenda -- Evidence from Sweden in the ChatGPT Era
by: Bruinsma, Bastiaan, et al.
Published: (2024) -
Fact or Guesswork? Evaluating Large Language Models' Medical Knowledge with Structured One-Hop Judgments
by: Li, Jiaxi, et al.
Published: (2025) -
CUB: Benchmarking Context Utilisation Techniques for Language Models
by: Hagström, Lovisa, et al.
Published: (2025)