LLMs' Reading Comprehension Is Affected by Parametric Knowledge and Struggles with Hypothetical Statements
Fuente:
arXiv
Saved in:
| Main Authors: | Basmov, Victoria, Goldberg, Yoav, Tsarfaty, Reut |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Simple Linguistic Inferences of Large Language Models (LLMs): Blind Spots and Blinds
by: Basmov, Victoria, et al.
Published: (2023)
by: Basmov, Victoria, et al.
Published: (2023)
HeQ: a Large and Diverse Hebrew Reading Comprehension Benchmark
by: Cohen, Amir DN, et al.
Published: (2025)
by: Cohen, Amir DN, et al.
Published: (2025)
NoviCode: Generating Programs from Natural Language Utterances by Novices
by: Mordechai, Asaf Achi, et al.
Published: (2024)
by: Mordechai, Asaf Achi, et al.
Published: (2024)
A Novel Computational and Modeling Foundation for Automatic Coherence Assessment
by: Maimon, Aviya, et al.
Published: (2023)
by: Maimon, Aviya, et al.
Published: (2023)
Beyond English: The Impact of Prompt Translation Strategies across Languages and Tasks in Multilingual LLMs
by: Mondshine, Itai, et al.
Published: (2025)
by: Mondshine, Itai, et al.
Published: (2025)
Mitigating Hallucinations in Large Vision-Language Models (LVLMs) via Language-Contrastive Decoding (LCD)
by: Manevich, Avshalom, et al.
Published: (2024)
by: Manevich, Avshalom, et al.
Published: (2024)
Beyond Word Boundaries: A Hebrew Coreference Benchmark and an Evaluation Protocol for Morphologically Complex Text
by: Greenfeld, Refael Shaked, et al.
Published: (2026)
by: Greenfeld, Refael Shaked, et al.
Published: (2026)
A Truly Joint Neural Architecture for Segmentation and Parsing
by: Levi, Danit Yshaayahu, et al.
Published: (2024)
by: Levi, Danit Yshaayahu, et al.
Published: (2024)
IQ Test for LLMs: An Evaluation Framework for Uncovering Core Skills in LLMs
by: Maimon, Aviya, et al.
Published: (2025)
by: Maimon, Aviya, et al.
Published: (2025)
MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents
by: Wolfson, Tomer, et al.
Published: (2025)
by: Wolfson, Tomer, et al.
Published: (2025)
Beyond N-Grams: Rethinking Evaluation Metrics and Strategies for Multilingual Abstractive Summarization
by: Mondshine, Itai, et al.
Published: (2025)
by: Mondshine, Itai, et al.
Published: (2025)
MRL Parsing Without Tears: The Case of Hebrew
by: Shmidman, Shaltiel, et al.
Published: (2024)
by: Shmidman, Shaltiel, et al.
Published: (2024)
Superlatives in Context: Modeling the Implicit Semantics of Superlatives
by: Pyatkin, Valentina, et al.
Published: (2024)
by: Pyatkin, Valentina, et al.
Published: (2024)
Do Pretrained Contextual Language Models Distinguish between Hebrew Homograph Analyses?
by: Shmidman, Avi, et al.
Published: (2024)
by: Shmidman, Avi, et al.
Published: (2024)
HeSum: a Novel Dataset for Abstractive Text Summarization in Hebrew
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
Into the Unknown: Generating Geospatial Descriptions for New Environments
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
Where Do We Go from Here? Multi-scale Allocentric Relational Inference from Natural Spatial Descriptions
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
Effective QA-driven Annotation of Predicate-Argument Relations Across Languages
by: Davidov, Jonathan, et al.
Published: (2026)
by: Davidov, Jonathan, et al.
Published: (2026)
Knowledge Navigator: LLM-guided Browsing Framework for Exploratory Search in Scientific Literature
by: Katz, Uri, et al.
Published: (2024)
by: Katz, Uri, et al.
Published: (2024)
Large Reasoning Models Struggle to Transfer Parametric Knowledge Across Scripts
by: Bandarkar, Lucas, et al.
Published: (2026)
by: Bandarkar, Lucas, et al.
Published: (2026)
Is It Really Long Context if All You Need Is Retrieval? Towards Genuinely Difficult Long Context NLP
by: Goldman, Omer, et al.
Published: (2024)
by: Goldman, Omer, et al.
Published: (2024)
Fine-Grained Prediction of Reading Comprehension from Eye Movements
by: Shubi, Omer, et al.
Published: (2024)
by: Shubi, Omer, et al.
Published: (2024)
Multilingual Instruction Tuning With Just a Pinch of Multilinguality
by: Shaham, Uri, et al.
Published: (2024)
by: Shaham, Uri, et al.
Published: (2024)
Unpacking Tokenization: Evaluating Text Compression and its Correlation with Model Performance
by: Goldman, Omer, et al.
Published: (2024)
by: Goldman, Omer, et al.
Published: (2024)
Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs
by: Mor-Lan, Guy, et al.
Published: (2026)
by: Mor-Lan, Guy, et al.
Published: (2026)
LLMs Struggle with Abstract Meaning Comprehension More Than Expected
by: Alhazmi, Hamoud, et al.
Published: (2026)
by: Alhazmi, Hamoud, et al.
Published: (2026)
Leveraging Collection-Wide Similarities for Unsupervised Document Structure Extraction
by: Lior, Gili, et al.
Published: (2024)
by: Lior, Gili, et al.
Published: (2024)
Encyclo-K: Evaluating LLMs with Dynamically Composed Knowledge Statements
by: Liang, Yiming, et al.
Published: (2025)
by: Liang, Yiming, et al.
Published: (2025)
Open Problems and a Hypothetical Path Forward in LLM Knowledge Paradigms
by: Ye, Xiaotian, et al.
Published: (2025)
by: Ye, Xiaotian, et al.
Published: (2025)
Same Task, More Tokens: the Impact of Input Length on the Reasoning Performance of Large Language Models
by: Levy, Mosh, et al.
Published: (2024)
by: Levy, Mosh, et al.
Published: (2024)
PRISM: PRIor from corpus Statistics for topic Modeling
by: Ishon, Tal, et al.
Published: (2026)
by: Ishon, Tal, et al.
Published: (2026)
Log-linear Guardedness and its Implications
by: Ravfogel, Shauli, et al.
Published: (2022)
by: Ravfogel, Shauli, et al.
Published: (2022)
Humans Perceive Wrong Narratives from AI Reasoning Texts
by: Levy, Mosh, et al.
Published: (2025)
by: Levy, Mosh, et al.
Published: (2025)
HypoTermQA: Hypothetical Terms Dataset for Benchmarking Hallucination Tendency of LLMs
by: Uluoglakci, Cem, et al.
Published: (2024)
by: Uluoglakci, Cem, et al.
Published: (2024)
Leveraging LLMs for Hypothetical Deduction in Logical Inference: A Neuro-Symbolic Approach
by: Li, Qingchuan, et al.
Published: (2024)
by: Li, Qingchuan, et al.
Published: (2024)
Breaking the Language Barrier: Can Direct Inference Outperform Pre-Translation in Multilingual LLM Applications?
by: Intrator, Yotam, et al.
Published: (2024)
by: Intrator, Yotam, et al.
Published: (2024)
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
by: Gekhman, Zorik, et al.
Published: (2026)
by: Gekhman, Zorik, et al.
Published: (2026)
Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs
by: Yang, Wanli, et al.
Published: (2026)
by: Yang, Wanli, et al.
Published: (2026)
The Effect of Surprisal on Reading Times in Information Seeking and Repeated Reading
by: Klein, Keren Gruteke, et al.
Published: (2024)
by: Klein, Keren Gruteke, et al.
Published: (2024)
BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models
by: Ben-Zaken, Elad, et al.
Published: (2021)
by: Ben-Zaken, Elad, et al.
Published: (2021)
Similar Items
-
Simple Linguistic Inferences of Large Language Models (LLMs): Blind Spots and Blinds
by: Basmov, Victoria, et al.
Published: (2023) -
HeQ: a Large and Diverse Hebrew Reading Comprehension Benchmark
by: Cohen, Amir DN, et al.
Published: (2025) -
NoviCode: Generating Programs from Natural Language Utterances by Novices
by: Mordechai, Asaf Achi, et al.
Published: (2024) -
A Novel Computational and Modeling Foundation for Automatic Coherence Assessment
by: Maimon, Aviya, et al.
Published: (2023) -
Beyond English: The Impact of Prompt Translation Strategies across Languages and Tasks in Multilingual LLMs
by: Mondshine, Itai, et al.
Published: (2025)