Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency
Fuente:
arXiv
Saved in:
| Main Authors: | Smith, Matthew L., Shock, Jonathan P., Segun, Samuel T., Olatunji, Iyiola E., Bissyandé, Tegawendé F. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluation Drift in LLM Personality Induction: Are We Moving the Goalpost?
by: Rajput, Prateek, et al.
Published: (2026)
by: Rajput, Prateek, et al.
Published: (2026)
How Secure is Secure Code Generation? Adversarial Prompts Put LLM Defenses to the Test
by: Tessa, Melissa, et al.
Published: (2026)
by: Tessa, Melissa, et al.
Published: (2026)
Reinforcement Learning-Guided Chain-of-Draft for Token-Efficient Code Generation
by: Tang, Xunzhu, et al.
Published: (2025)
by: Tang, Xunzhu, et al.
Published: (2025)
Geometric Factual Recall in Transformers
by: Ravfogel, Shauli, et al.
Published: (2026)
by: Ravfogel, Shauli, et al.
Published: (2026)
Dynamic Stability of LLM-Generated Code
by: Rajput, Prateek, et al.
Published: (2025)
by: Rajput, Prateek, et al.
Published: (2025)
Beyond Real Faces: Synthetic Datasets Can Achieve Reliable Recognition Performance without Privacy Compromise
by: Borsukiewicz, Paweł, et al.
Published: (2025)
by: Borsukiewicz, Paweł, et al.
Published: (2025)
Towards a Holistic Evaluation of LLMs on Factual Knowledge Recall
by: Yuan, Jiaqing, et al.
Published: (2024)
by: Yuan, Jiaqing, et al.
Published: (2024)
Why Low-Resource NLP Needs More Than Cross-Lingual Transfer: Lessons Learned from Luxembourgish
by: Philippy, Fred, et al.
Published: (2026)
by: Philippy, Fred, et al.
Published: (2026)
LuxEmbedder: A Cross-Lingual Approach to Enhanced Luxembourgish Sentence Embeddings
by: Philippy, Fred, et al.
Published: (2024)
by: Philippy, Fred, et al.
Published: (2024)
Critical Confabulation: Can LLMs Hallucinate for Social Good?
by: Sui, Peiqi, et al.
Published: (2025)
by: Sui, Peiqi, et al.
Published: (2025)
Summing Up the Facts: Additive Mechanisms Behind Factual Recall in LLMs
by: Chughtai, Bilal, et al.
Published: (2024)
by: Chughtai, Bilal, et al.
Published: (2024)
From Rookie to Expert: Manipulating LLMs for Automated Vulnerability Exploitation in Enterprise Software
by: Diouf, Moustapha Awwalou, et al.
Published: (2025)
by: Diouf, Moustapha Awwalou, et al.
Published: (2025)
Revisiting Code Similarity Evaluation with Abstract Syntax Tree Edit Distance
by: Song, Yewei, et al.
Published: (2024)
by: Song, Yewei, et al.
Published: (2024)
Enhancing Small Language Models for Cross-Lingual Generalized Zero-Shot Classification with Soft Prompt Tuning
by: Philippy, Fred, et al.
Published: (2025)
by: Philippy, Fred, et al.
Published: (2025)
Correctness isnt Efficiency: Runtime Memory Divergence in LLM-Generated Code
by: Rajput, Prateek, et al.
Published: (2026)
by: Rajput, Prateek, et al.
Published: (2026)
Injecting Falsehoods: Adversarial Man-in-the-Middle Attacks Undermining Factual Recall in LLMs
by: Fastowski, Alina, et al.
Published: (2025)
by: Fastowski, Alina, et al.
Published: (2025)
Can LLMs Detect Their Confabulations? Estimating Reliability in Uncertainty-Aware Language Models
by: Zhou, Tianyi, et al.
Published: (2025)
by: Zhou, Tianyi, et al.
Published: (2025)
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
by: Liu, Yihong, et al.
Published: (2026)
by: Liu, Yihong, et al.
Published: (2026)
Can VLMs Recall Factual Associations From Visual References?
by: Ashok, Dhananjay, et al.
Published: (2025)
by: Ashok, Dhananjay, et al.
Published: (2025)
LuxInstruct: A Cross-Lingual Instruction Tuning Dataset For Luxembourgish
by: Philippy, Fred, et al.
Published: (2025)
by: Philippy, Fred, et al.
Published: (2025)
Assessing Spear-Phishing Website Generation in Large Language Model Coding Agents
by: Malloy, Tailia, et al.
Published: (2026)
by: Malloy, Tailia, et al.
Published: (2026)
Anchored Confabulation: Partial Evidence Non-Monotonically Amplifies Confident Hallucination in LLMs
by: Lathkar, Ashish Balkishan
Published: (2026)
by: Lathkar, Ashish Balkishan
Published: (2026)
Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline
by: Lu, Meng, et al.
Published: (2025)
by: Lu, Meng, et al.
Published: (2025)
Characterizing Build Compromises Through Vulnerability Disclosure Analysis
by: Diao, Maimouna Tamah, et al.
Published: (2025)
by: Diao, Maimouna Tamah, et al.
Published: (2025)
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models
by: Lv, Ang, et al.
Published: (2024)
by: Lv, Ang, et al.
Published: (2024)
Adversarial Attacks and Defenses on Graph-aware Large Language Models (LLMs)
by: Olatunji, Iyiola E., et al.
Published: (2025)
by: Olatunji, Iyiola E., et al.
Published: (2025)
Understanding Factual Recall in Transformers via Associative Memories
by: Nichani, Eshaan, et al.
Published: (2024)
by: Nichani, Eshaan, et al.
Published: (2024)
Unveiling Factual Recall Behaviors of Large Language Models through Knowledge Neurons
by: Wang, Yifei, et al.
Published: (2024)
by: Wang, Yifei, et al.
Published: (2024)
Confabulation: The Surprising Value of Large Language Model Hallucinations
by: Sui, Peiqi, et al.
Published: (2024)
by: Sui, Peiqi, et al.
Published: (2024)
ACE: Attribution-Controlled Knowledge Editing for Multi-hop Factual Recall
by: Yang, Jiayu, et al.
Published: (2025)
by: Yang, Jiayu, et al.
Published: (2025)
Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality
by: Calderon, Nitay, et al.
Published: (2026)
by: Calderon, Nitay, et al.
Published: (2026)
Layerwise Recall and the Geometry of Interwoven Knowledge in LLMs
by: Lei, Ge, et al.
Published: (2025)
by: Lei, Ge, et al.
Published: (2025)
Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation
by: Dejl, Adam, et al.
Published: (2025)
by: Dejl, Adam, et al.
Published: (2025)
Boosting Open-Source LLMs for Program Repair via Reasoning Transfer and LLM-Guided Reinforcement Learning
by: Tang, Xunzhu, et al.
Published: (2025)
by: Tang, Xunzhu, et al.
Published: (2025)
Measuring LLM Code Generation Stability via Structural Entropy
by: Song, Yewei, et al.
Published: (2025)
by: Song, Yewei, et al.
Published: (2025)
LaFiCMIL: Rethinking Large File Classification from the Perspective of Correlated Multiple Instance Learning
by: Sun, Tiezhu, et al.
Published: (2023)
by: Sun, Tiezhu, et al.
Published: (2023)
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
by: Djiré, Albérick Euraste, et al.
Published: (2025)
by: Djiré, Albérick Euraste, et al.
Published: (2025)
Soft Prompt Tuning for Cross-Lingual Transfer: When Less is More
by: Philippy, Fred, et al.
Published: (2024)
by: Philippy, Fred, et al.
Published: (2024)
Do Factual Recall Mechanisms Carry over from Text to Speech in Multimodal Language Models?
by: Modica, Luca, et al.
Published: (2026)
by: Modica, Luca, et al.
Published: (2026)
Relation Also Knows: Rethinking the Recall and Editing of Factual Associations in Auto-Regressive Transformer Language Models
by: Liu, Xiyu, et al.
Published: (2024)
by: Liu, Xiyu, et al.
Published: (2024)
Similar Items
-
Evaluation Drift in LLM Personality Induction: Are We Moving the Goalpost?
by: Rajput, Prateek, et al.
Published: (2026) -
How Secure is Secure Code Generation? Adversarial Prompts Put LLM Defenses to the Test
by: Tessa, Melissa, et al.
Published: (2026) -
Reinforcement Learning-Guided Chain-of-Draft for Token-Efficient Code Generation
by: Tang, Xunzhu, et al.
Published: (2025) -
Geometric Factual Recall in Transformers
by: Ravfogel, Shauli, et al.
Published: (2026) -
Dynamic Stability of LLM-Generated Code
by: Rajput, Prateek, et al.
Published: (2025)