Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality
Fuente:
arXiv
Salvato in:
| Autori principali: | Calderon, Nitay, Ben-David, Eyal, Gekhman, Zorik, Ofek, Eran, Yona, Gal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Confidence Improves Self-Consistency in LLMs
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
di: Gekhman, Zorik, et al.
Pubblicazione: (2026)
di: Gekhman, Zorik, et al.
Pubblicazione: (2026)
Inside-Out: Hidden Factual Knowledge in LLMs
di: Gekhman, Zorik, et al.
Pubblicazione: (2025)
di: Gekhman, Zorik, et al.
Pubblicazione: (2025)
NL-Eye: Abductive NLI for Images
di: Ventura, Mor, et al.
Pubblicazione: (2024)
di: Ventura, Mor, et al.
Pubblicazione: (2024)
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
di: Gekhman, Zorik, et al.
Pubblicazione: (2024)
di: Gekhman, Zorik, et al.
Pubblicazione: (2024)
Measuring the Robustness of NLP Models to Domain Shifts
di: Calderon, Nitay, et al.
Pubblicazione: (2023)
di: Calderon, Nitay, et al.
Pubblicazione: (2023)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2024)
di: Calderon, Nitay, et al.
Pubblicazione: (2024)
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models
di: Lv, Ang, et al.
Pubblicazione: (2024)
di: Lv, Ang, et al.
Pubblicazione: (2024)
Keep Guessing? When Considering Inference Scaling, Mind the Baselines
di: Yona, Gal, et al.
Pubblicazione: (2024)
di: Yona, Gal, et al.
Pubblicazione: (2024)
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
di: Lissak, Shir, et al.
Pubblicazione: (2024)
di: Lissak, Shir, et al.
Pubblicazione: (2024)
Unveiling Factual Recall Behaviors of Large Language Models through Knowledge Neurons
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
PrefixNLI: Detecting Factual Inconsistencies as Soon as They Arise
di: Harary, Sapir, et al.
Pubblicazione: (2025)
di: Harary, Sapir, et al.
Pubblicazione: (2025)
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
di: Toker, Gilat, et al.
Pubblicazione: (2026)
di: Toker, Gilat, et al.
Pubblicazione: (2026)
Towards a Holistic Evaluation of LLMs on Factual Knowledge Recall
di: Yuan, Jiaqing, et al.
Pubblicazione: (2024)
di: Yuan, Jiaqing, et al.
Pubblicazione: (2024)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
Why Fine-Tuning Encourages Hallucinations and How to Fix It
di: Kaplan, Guy, et al.
Pubblicazione: (2026)
di: Kaplan, Guy, et al.
Pubblicazione: (2026)
Can VLMs Recall Factual Associations From Visual References?
di: Ashok, Dhananjay, et al.
Pubblicazione: (2025)
di: Ashok, Dhananjay, et al.
Pubblicazione: (2025)
Locate-then-edit for Multi-hop Factual Recall under Knowledge Editing
di: Zhang, Zhuoran, et al.
Pubblicazione: (2024)
di: Zhang, Zhuoran, et al.
Pubblicazione: (2024)
StratMem-Bench: Evaluating Strategic Memory Use in Virtual Character Conversation Beyond Factual Recall
di: Wu, Yerong, et al.
Pubblicazione: (2026)
di: Wu, Yerong, et al.
Pubblicazione: (2026)
Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency
di: Smith, Matthew L., et al.
Pubblicazione: (2026)
di: Smith, Matthew L., et al.
Pubblicazione: (2026)
Injecting Falsehoods: Adversarial Man-in-the-Middle Attacks Undermining Factual Recall in LLMs
di: Fastowski, Alina, et al.
Pubblicazione: (2025)
di: Fastowski, Alina, et al.
Pubblicazione: (2025)
SimpleQA Verified: A Reliable Factuality Benchmark to Measure Parametric Knowledge
di: Haas, Lukas, et al.
Pubblicazione: (2025)
di: Haas, Lukas, et al.
Pubblicazione: (2025)
On Early Detection of Hallucinations in Factual Question Answering
di: Snyder, Ben, et al.
Pubblicazione: (2023)
di: Snyder, Ben, et al.
Pubblicazione: (2023)
The FACTS Leaderboard: A Comprehensive Benchmark for Large Language Model Factuality
di: Cheng, Aileen, et al.
Pubblicazione: (2025)
di: Cheng, Aileen, et al.
Pubblicazione: (2025)
Reversing Large Language Models for Efficient Training and Fine-Tuning
di: Gal, Eshed, et al.
Pubblicazione: (2025)
di: Gal, Eshed, et al.
Pubblicazione: (2025)
Factuality on Demand: Controlling the Factuality-Informativeness Trade-off in Text Generation
di: Gong, Ziwei, et al.
Pubblicazione: (2026)
di: Gong, Ziwei, et al.
Pubblicazione: (2026)
Do Automatic Factuality Metrics Measure Factuality? A Critical Evaluation
di: Ramprasad, Sanjana, et al.
Pubblicazione: (2024)
di: Ramprasad, Sanjana, et al.
Pubblicazione: (2024)
Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
di: Wagner, Eitan, et al.
Pubblicazione: (2024)
di: Wagner, Eitan, et al.
Pubblicazione: (2024)
AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
di: Nakash, Itay, et al.
Pubblicazione: (2025)
di: Nakash, Itay, et al.
Pubblicazione: (2025)
Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models
di: Yona, Itay, et al.
Pubblicazione: (2026)
di: Yona, Itay, et al.
Pubblicazione: (2026)
Time Awareness in Large Language Models: Benchmarking Fact Recall Across Time
di: Herel, David, et al.
Pubblicazione: (2024)
di: Herel, David, et al.
Pubblicazione: (2024)
Is Factuality Enhancement a Free Lunch For LLMs? Better Factuality Can Lead to Worse Context-Faithfulness
di: Bi, Baolong, et al.
Pubblicazione: (2024)
di: Bi, Baolong, et al.
Pubblicazione: (2024)
Navigating Cultural Chasms: Exploring and Unlocking the Cultural POV of Text-To-Image Models
di: Ventura, Mor, et al.
Pubblicazione: (2023)
di: Ventura, Mor, et al.
Pubblicazione: (2023)
Can LLMs Learn Macroeconomic Narratives from Social Media?
di: Gueta, Almog, et al.
Pubblicazione: (2024)
di: Gueta, Almog, et al.
Pubblicazione: (2024)
Dual Debiasing: Remove Stereotypes and Keep Factual Gender for Fair Language Modeling and Translation
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2025)
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2025)
Fine-Grained Detection of Context-Grounded Hallucinations Using LLMs
di: Peisakhovsky, Yehonatan, et al.
Pubblicazione: (2025)
di: Peisakhovsky, Yehonatan, et al.
Pubblicazione: (2025)
Generating Benchmarks for Factuality Evaluation of Language Models
di: Muhlgay, Dor, et al.
Pubblicazione: (2023)
di: Muhlgay, Dor, et al.
Pubblicazione: (2023)
Language Models' Factuality Depends on the Language of Inquiry
di: Aggarwal, Tushar, et al.
Pubblicazione: (2025)
di: Aggarwal, Tushar, et al.
Pubblicazione: (2025)
From Confidence to Collapse in LLM Factual Robustness
di: Fastowski, Alina, et al.
Pubblicazione: (2025)
di: Fastowski, Alina, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Confidence Improves Self-Consistency in LLMs
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025) -
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
di: Gekhman, Zorik, et al.
Pubblicazione: (2026) -
Inside-Out: Hidden Factual Knowledge in LLMs
di: Gekhman, Zorik, et al.
Pubblicazione: (2025) -
NL-Eye: Abductive NLI for Images
di: Ventura, Mor, et al.
Pubblicazione: (2024) -
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
di: Gekhman, Zorik, et al.
Pubblicazione: (2024)