Are LLM Belief Updates Consistent with Bayes' Theorem?
Fuente:
arXiv
Saved in:
| Main Authors: | Imran, Sohaib, Kendiukhov, Ihor, Broerman, Matthew, Thomas, Aditya, Campanella, Riccardo, Lamb, Rob, Atkinson, Peter M. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Out-of-Context Abduction: LLMs Make Inferences About Procedural Data Leveraging Declarative Facts in Earlier Training Data
by: Imran, Sohaib, et al.
Published: (2025)
by: Imran, Sohaib, et al.
Published: (2025)
Consistency Training while Mitigating Obfuscation via Rate Matching
by: Imran, Sohaib, et al.
Published: (2026)
by: Imran, Sohaib, et al.
Published: (2026)
Systematic Evaluation of Single-Cell Foundation Model Interpretability Reveals Attention Captures Co-Expression Rather Than Unique Regulatory Signal
by: Kendiukhov, Ihor
Published: (2026)
by: Kendiukhov, Ihor
Published: (2026)
Multi-Dimensional Spectral Geometry of Biological Knowledge in Single-Cell Transformer Representations
by: Kendiukhov, Ihor
Published: (2026)
by: Kendiukhov, Ihor
Published: (2026)
A Review of Developmental Interpretability in Large Language Models
by: Kendiukhov, Ihor
Published: (2025)
by: Kendiukhov, Ihor
Published: (2025)
PABU: Progress-Aware Belief Update for Efficient LLM Agents
by: Jiang, Haitao, et al.
Published: (2026)
by: Jiang, Haitao, et al.
Published: (2026)
VAL-Bench: Belief Consistency as a measure for Value Alignment in Language Models
by: Gupta, Aman, et al.
Published: (2025)
by: Gupta, Aman, et al.
Published: (2025)
Steamroller Problems: An Evaluation of LLM Reasoning Capability with Automated Theorem Prover Strategies
by: McGinness, Lachlan, et al.
Published: (2024)
by: McGinness, Lachlan, et al.
Published: (2024)
The Belief State Transformer
by: Hu, Edward S., et al.
Published: (2024)
by: Hu, Edward S., et al.
Published: (2024)
DeepTheorem: Advancing LLM Reasoning for Theorem Proving Through Natural Language and Reinforcement Learning
by: Zhang, Ziyin, et al.
Published: (2025)
by: Zhang, Ziyin, et al.
Published: (2025)
Fundamental Problems With Model Editing: How Should Rational Belief Revision Work in LLMs?
by: Hase, Peter, et al.
Published: (2024)
by: Hase, Peter, et al.
Published: (2024)
Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content
by: Stepanov, Ihor, et al.
Published: (2026)
by: Stepanov, Ihor, et al.
Published: (2026)
Ergodicity Library: A Python Toolkit for Stochastic-Process Simulation, Time-Average Diagnostics, and Agent-Based Experiments
by: Kendiukhov, Ihor
Published: (2026)
by: Kendiukhov, Ihor
Published: (2026)
FairBelief -- Assessing Harmful Beliefs in Language Models
by: Setzu, Mattia, et al.
Published: (2024)
by: Setzu, Mattia, et al.
Published: (2024)
Preference-Aware Memory Update for Long-Term LLM Agents
by: Sun, Haoran, et al.
Published: (2025)
by: Sun, Haoran, et al.
Published: (2025)
Mining the Mind: What 100M Beliefs Reveal About Frontier LLM Knowledge
by: Ghosh, Shrestha, et al.
Published: (2025)
by: Ghosh, Shrestha, et al.
Published: (2025)
Faithful and Robust LLM-Driven Theorem Proving for NLI Explanations
by: Quan, Xin, et al.
Published: (2025)
by: Quan, Xin, et al.
Published: (2025)
Accumulating Context Changes the Beliefs of Language Models
by: Geng, Jiayi, et al.
Published: (2025)
by: Geng, Jiayi, et al.
Published: (2025)
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
by: Borah, Angana, et al.
Published: (2026)
by: Borah, Angana, et al.
Published: (2026)
Harnessing Consistency for Robust Test-Time LLM Ensemble
by: Zeng, Zhichen, et al.
Published: (2025)
by: Zeng, Zhichen, et al.
Published: (2025)
AXCEL: Automated eXplainable Consistency Evaluation using LLMs
by: Sreekar, P Aditya, et al.
Published: (2024)
by: Sreekar, P Aditya, et al.
Published: (2024)
Trustworthy LLM-Mediated Communication: Evaluating Information Fidelity in LLM as a Communicator (LAAC) Framework in Multiple Application Domains
by: Rafi, Mohammed Musthafa, et al.
Published: (2025)
by: Rafi, Mohammed Musthafa, et al.
Published: (2025)
Resource-Efficient Fine-Tuning of LLaMA-3.2-3B for Medical Chain-of-Thought Reasoning
by: Mansha, Imran
Published: (2025)
by: Mansha, Imran
Published: (2025)
ConsistencyChecker: Tree-based Evaluation of LLM Generalization Capabilities
by: Hong, Zhaochen, et al.
Published: (2025)
by: Hong, Zhaochen, et al.
Published: (2025)
TheoremExplainAgent: Towards Video-based Multimodal Explanations for LLM Theorem Understanding
by: Ku, Max, et al.
Published: (2025)
by: Ku, Max, et al.
Published: (2025)
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
by: Clark, Peter, et al.
Published: (2023)
by: Clark, Peter, et al.
Published: (2023)
Simulating LLM-to-LLM Tutoring for Multilingual Math Feedback
by: Tonga, Junior Cedric, et al.
Published: (2025)
by: Tonga, Junior Cedric, et al.
Published: (2025)
CSCE: Boosting LLM Reasoning by Simultaneous Enhancing of Causal Significance and Consistency
by: Wang, Kangsheng, et al.
Published: (2024)
by: Wang, Kangsheng, et al.
Published: (2024)
TriAlign: Towards Universal Truth Consistency in Personalized LLM Alignment
by: Nguyen, Thi-Nhung, et al.
Published: (2026)
by: Nguyen, Thi-Nhung, et al.
Published: (2026)
Curved Inference: Concern-Sensitive Geometry in Large Language Model Residual Streams
by: Manson, Rob
Published: (2025)
by: Manson, Rob
Published: (2025)
GraphMind: Theorem Selection and Conclusion Generation Framework with Dynamic GNN for LLM Reasoning
by: Li, Yutong, et al.
Published: (2025)
by: Li, Yutong, et al.
Published: (2025)
Narrative Theory-Driven LLM Methods for Automatic Story Generation and Understanding: A Survey
by: Liu, David Y., et al.
Published: (2026)
by: Liu, David Y., et al.
Published: (2026)
Vulnerability of LLMs' Stated Beliefs? LLMs Belief Resistance Check Through Strategic Persuasive Conversation Interventions
by: Huang, Fan, et al.
Published: (2026)
by: Huang, Fan, et al.
Published: (2026)
ABBEL: LLM Agents Acting through Belief Bottlenecks Expressed in Language
by: Lidayan, Aly, et al.
Published: (2025)
by: Lidayan, Aly, et al.
Published: (2025)
Mind the Gap: How Elicitation Protocols Shape the Stated-Revealed Preference Gap in Language Models
by: Mahajan, Pranav, et al.
Published: (2026)
by: Mahajan, Pranav, et al.
Published: (2026)
GLiNER multi-task: Generalist Lightweight Model for Various Information Extraction Tasks
by: Stepanov, Ihor, et al.
Published: (2024)
by: Stepanov, Ihor, et al.
Published: (2024)
Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries
by: Kale, Sahil, et al.
Published: (2025)
by: Kale, Sahil, et al.
Published: (2025)
Evaluating Consistencies in LLM responses through a Semantic Clustering of Question Answering
by: Lee, Yanggyu, et al.
Published: (2024)
by: Lee, Yanggyu, et al.
Published: (2024)
Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling
by: Wan, Guangya, et al.
Published: (2024)
by: Wan, Guangya, et al.
Published: (2024)
DecMetrics: Structured Claim Decomposition Scoring for Factually Consistent LLM Outputs
by: Huang, Minghui
Published: (2025)
by: Huang, Minghui
Published: (2025)
Similar Items
-
Out-of-Context Abduction: LLMs Make Inferences About Procedural Data Leveraging Declarative Facts in Earlier Training Data
by: Imran, Sohaib, et al.
Published: (2025) -
Consistency Training while Mitigating Obfuscation via Rate Matching
by: Imran, Sohaib, et al.
Published: (2026) -
Systematic Evaluation of Single-Cell Foundation Model Interpretability Reveals Attention Captures Co-Expression Rather Than Unique Regulatory Signal
by: Kendiukhov, Ihor
Published: (2026) -
Multi-Dimensional Spectral Geometry of Biological Knowledge in Single-Cell Transformer Representations
by: Kendiukhov, Ihor
Published: (2026) -
A Review of Developmental Interpretability in Large Language Models
by: Kendiukhov, Ihor
Published: (2025)