What Matters in Memorizing and Recalling Facts? Multifaceted Benchmarks for Knowledge Probing in Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Xin, Yoshinaga, Naoki, Oba, Daisuke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tracing the Roots of Facts in Multilingual Language Models: Independent, Shared, and Transferred Knowledge
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
Tracing Multilingual Knowledge Acquisition Dynamics in Domain Adaptation: A Case Study of English-Japanese Biomedical Adaptation
von: Zhao, Xin, et al.
Veröffentlicht: (2025)
von: Zhao, Xin, et al.
Veröffentlicht: (2025)
Time Awareness in Large Language Models: Benchmarking Fact Recall Across Time
von: Herel, David, et al.
Veröffentlicht: (2024)
von: Herel, David, et al.
Veröffentlicht: (2024)
Scaling Laws for Fact Memorization of Large Language Models
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
Fact Recall, Heuristics or Pure Guesswork? Precise Interpretations of Language Models for Fact Completion
von: Saynova, Denitsa, et al.
Veröffentlicht: (2024)
von: Saynova, Denitsa, et al.
Veröffentlicht: (2024)
MultifacetEval: Multifaceted Evaluation to Probe LLMs in Mastering Medical Knowledge
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2024)
Neuron Empirical Gradient: Discovering and Quantifying Neurons Global Linear Controllability
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
Facts Fade Fast: Evaluating Memorization of Outdated Medical Knowledge in Large Language Models
von: Vladika, Juraj, et al.
Veröffentlicht: (2025)
von: Vladika, Juraj, et al.
Veröffentlicht: (2025)
OWL: Probing Cross-Lingual Recall of Memorized Texts via World Literature
von: Srivastava, Alisha, et al.
Veröffentlicht: (2025)
von: Srivastava, Alisha, et al.
Veröffentlicht: (2025)
In-Contextual Gender Bias Suppression for Large Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2023)
von: Oba, Daisuke, et al.
Veröffentlicht: (2023)
Commentary Generation from Data Records of Multiplayer Strategy Esports Game
von: Wang, Zihan, et al.
Veröffentlicht: (2022)
von: Wang, Zihan, et al.
Veröffentlicht: (2022)
Drifting Objectives for Refining Discrete Diffusion Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
Memorization Dynamics in Knowledge Distillation for Language Models
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026)
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026)
Diffusion-State Policy Optimization for Masked Diffusion Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
Guess or Recall? Training CNNs to Classify and Localize Memorization in LLMs
von: Dentan, Jérémie, et al.
Veröffentlicht: (2025)
von: Dentan, Jérémie, et al.
Veröffentlicht: (2025)
Tracing Relational Knowledge Recall in Large Language Models
von: Popovič, Nicholas, et al.
Veröffentlicht: (2026)
von: Popovič, Nicholas, et al.
Veröffentlicht: (2026)
Functional Abstraction of Knowledge Recall in Large Language Models
von: Wang, Zijian, et al.
Veröffentlicht: (2025)
von: Wang, Zijian, et al.
Veröffentlicht: (2025)
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
von: Pradeep, Ronak, et al.
Veröffentlicht: (2025)
von: Pradeep, Ronak, et al.
Veröffentlicht: (2025)
Recite, Reconstruct, Recollect: Memorization in LMs as a Multifaceted Phenomenon
von: Prashanth, USVSN Sai, et al.
Veröffentlicht: (2024)
von: Prashanth, USVSN Sai, et al.
Veröffentlicht: (2024)
CxMP: A Linguistic Minimal-Pair Benchmark for Evaluating Constructional Understanding in Language Models
von: Oba, Miyu, et al.
Veröffentlicht: (2026)
von: Oba, Miyu, et al.
Veröffentlicht: (2026)
Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling
von: Yang, Linyao, et al.
Veröffentlicht: (2023)
von: Yang, Linyao, et al.
Veröffentlicht: (2023)
Bidirectional LMs are Better Knowledge Memorizers? A Benchmark for Real-world Knowledge Injection
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
M-QALM: A Benchmark to Assess Clinical Reading Comprehension and Knowledge Recall in Large Language Models via Question Answering
von: Subramanian, Anand, et al.
Veröffentlicht: (2024)
von: Subramanian, Anand, et al.
Veröffentlicht: (2024)
Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms
von: Nishida, Yuto, et al.
Veröffentlicht: (2026)
von: Nishida, Yuto, et al.
Veröffentlicht: (2026)
Retentive or Forgetful? Diving into the Knowledge Memorizing Mechanism of Language Models
von: Cao, Boxi, et al.
Veröffentlicht: (2023)
von: Cao, Boxi, et al.
Veröffentlicht: (2023)
Has this Fact been Edited? Detecting Knowledge Edits in Language Models
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
Beyond Facts: Benchmarking Distributional Reading Comprehension in Large Language Models
von: Guo, Pei-Fu, et al.
Veröffentlicht: (2026)
von: Guo, Pei-Fu, et al.
Veröffentlicht: (2026)
Summing Up the Facts: Additive Mechanisms Behind Factual Recall in LLMs
von: Chughtai, Bilal, et al.
Veröffentlicht: (2024)
von: Chughtai, Bilal, et al.
Veröffentlicht: (2024)
Rethinking Response Evaluation from Interlocutor's Eye for Open-Domain Dialogue Systems
von: Tsuta, Yuma, et al.
Veröffentlicht: (2024)
von: Tsuta, Yuma, et al.
Veröffentlicht: (2024)
Benchmarking and Rethinking Knowledge Editing for Large Language Models
von: He, Guoxiu, et al.
Veröffentlicht: (2025)
von: He, Guoxiu, et al.
Veröffentlicht: (2025)
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
Unveiling Factual Recall Behaviors of Large Language Models through Knowledge Neurons
von: Wang, Yifei, et al.
Veröffentlicht: (2024)
von: Wang, Yifei, et al.
Veröffentlicht: (2024)
To Words and Beyond: Probing Large Language Models for Sentence-Level Psycholinguistic Norms of Memorability and Reading Times
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
Continual Memorization of Factoids in Language Models
von: Chen, Howard, et al.
Veröffentlicht: (2024)
von: Chen, Howard, et al.
Veröffentlicht: (2024)
Probing Language Models on Their Knowledge Source
von: Tighidet, Zineddine, et al.
Veröffentlicht: (2024)
von: Tighidet, Zineddine, et al.
Veröffentlicht: (2024)
MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts
von: He, Jiayi, et al.
Veröffentlicht: (2025)
von: He, Jiayi, et al.
Veröffentlicht: (2025)
When Facts Change: Probing LLMs on Evolving Knowledge with evolveQA
von: Nakshatri, Nishanth Sridhar, et al.
Veröffentlicht: (2025)
von: Nakshatri, Nishanth Sridhar, et al.
Veröffentlicht: (2025)
What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning
von: Weng, Zhaotian, et al.
Veröffentlicht: (2025)
von: Weng, Zhaotian, et al.
Veröffentlicht: (2025)
TrendFact: A Benchmark for Explainable Hotspot Perception in Fact-Checking with Natural Language Explanation
von: Zhang, Xiaocheng, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaocheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Tracing the Roots of Facts in Multilingual Language Models: Independent, Shared, and Transferred Knowledge
von: Zhao, Xin, et al.
Veröffentlicht: (2024) -
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
von: Zhang, Ying, et al.
Veröffentlicht: (2025) -
Tracing Multilingual Knowledge Acquisition Dynamics in Domain Adaptation: A Case Study of English-Japanese Biomedical Adaptation
von: Zhao, Xin, et al.
Veröffentlicht: (2025) -
Time Awareness in Large Language Models: Benchmarking Fact Recall Across Time
von: Herel, David, et al.
Veröffentlicht: (2024) -
Scaling Laws for Fact Memorization of Large Language Models
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)