Reasoning Capabilities of Large Language Models. Lessons Learned from General Game Playing
Fuente:
arXiv
Saved in:
| Main Authors: | Świechowski, Maciej, Żychowski, Adam, Mańdziuk, Jacek |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generics and Default Reasoning in Large Language Models
by: Kirkpatrick, James Ravi, et al.
Published: (2025)
by: Kirkpatrick, James Ravi, et al.
Published: (2025)
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark
by: Li, Fangjun, et al.
Published: (2024)
by: Li, Fangjun, et al.
Published: (2024)
Satisfiability Solving with LLMs: A Matched-Pair Evaluation of Reasoning Capability
by: Zhang, Leizhen, et al.
Published: (2026)
by: Zhang, Leizhen, et al.
Published: (2026)
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models
by: Wan, Yuxuan, et al.
Published: (2024)
by: Wan, Yuxuan, et al.
Published: (2024)
Can Large Language Models Understand DL-Lite Ontologies? An Empirical Study
by: Wang, Keyu, et al.
Published: (2024)
by: Wang, Keyu, et al.
Published: (2024)
Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning
by: Bourigault, Pauline, et al.
Published: (2026)
by: Bourigault, Pauline, et al.
Published: (2026)
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
by: Liu, Yinhong, et al.
Published: (2024)
by: Liu, Yinhong, et al.
Published: (2024)
JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
by: Chen, Michael K., et al.
Published: (2025)
by: Chen, Michael K., et al.
Published: (2025)
EXPLORER: Exploration-guided Reasoning for Textual Reinforcement Learning
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Recursive Decomposition of Logical Thoughts: Framework for Superior Reasoning and Knowledge Propagation in Large Language Models
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
Ontology for Policing: Conceptual Knowledge Learning for Semantic Understanding and Reasoning in Law Enforcement Reports
by: Srbinovska, Anita, et al.
Published: (2026)
by: Srbinovska, Anita, et al.
Published: (2026)
Are Language Models Efficient Reasoners? A Perspective from Logic Programming
by: Opedal, Andreas, et al.
Published: (2025)
by: Opedal, Andreas, et al.
Published: (2025)
Llemma: An Open Language Model For Mathematics
by: Azerbayev, Zhangir, et al.
Published: (2023)
by: Azerbayev, Zhangir, et al.
Published: (2023)
Towards Logically Sound Natural Language Reasoning with Logic-Enhanced Language Model Agents
by: Mensfelt, Agnieszka, et al.
Published: (2024)
by: Mensfelt, Agnieszka, et al.
Published: (2024)
DECIDER: A Dual-System Rule-Controllable Decoding Framework for Language Generation
by: Xu, Chen, et al.
Published: (2024)
by: Xu, Chen, et al.
Published: (2024)
Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning
by: Su, DiJia, et al.
Published: (2025)
by: Su, DiJia, et al.
Published: (2025)
LLM-ARC: Enhancing LLMs with an Automated Reasoning Critic
by: Kalyanpur, Aditya, et al.
Published: (2024)
by: Kalyanpur, Aditya, et al.
Published: (2024)
Sound and Complete Neurosymbolic Reasoning with LLM-Grounded Interpretations
by: Allen, Bradley P., et al.
Published: (2025)
by: Allen, Bradley P., et al.
Published: (2025)
Bridging LLMs and Symbolic Reasoning in Educational QA Systems: Insights from the XAI Challenge at IJCNN 2025
by: Nguyen, Long S. T., et al.
Published: (2025)
by: Nguyen, Long S. T., et al.
Published: (2025)
Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning
by: Zhang, Xinglang, et al.
Published: (2026)
by: Zhang, Xinglang, et al.
Published: (2026)
Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study
by: Zhou, Yujun, et al.
Published: (2025)
by: Zhou, Yujun, et al.
Published: (2025)
Logic-Parametric Neuro-Symbolic NLI: Controlling Logical Formalisms for Verifiable LLM Reasoning
by: Farjami, Ali, et al.
Published: (2026)
by: Farjami, Ali, et al.
Published: (2026)
Towards Advanced Mathematical Reasoning for LLMs via First-Order Logic Theorem Proving
by: Cao, Chuxue, et al.
Published: (2025)
by: Cao, Chuxue, et al.
Published: (2025)
Large Language Models Imitate Logical Reasoning, but at what Cost?
by: McGinness, Lachlan, et al.
Published: (2025)
by: McGinness, Lachlan, et al.
Published: (2025)
A Reliable Common-Sense Reasoning Socialbot Built Using LLMs and Goal-Directed ASP
by: Zeng, Yankai, et al.
Published: (2024)
by: Zeng, Yankai, et al.
Published: (2024)
A Neurosymbolic Approach to Loop Invariant Generation via Weakest Precondition Reasoning
by: King, Daragh, et al.
Published: (2025)
by: King, Daragh, et al.
Published: (2025)
Elenchus: Generating Knowledge Bases from Prover-Skeptic Dialogues
by: Allen, Bradley P.
Published: (2026)
by: Allen, Bradley P.
Published: (2026)
Defining implication relation for classical logic
by: Fu, Li
Published: (2013)
by: Fu, Li
Published: (2013)
VeriThoughts: Enabling Automated Verilog Code Generation using Reasoning and Formal Verification
by: Yubeaton, Patrick, et al.
Published: (2025)
by: Yubeaton, Patrick, et al.
Published: (2025)
Question Answering with LLMs and Learning from Answer Sets
by: Borroto, Manuel, et al.
Published: (2025)
by: Borroto, Manuel, et al.
Published: (2025)
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast
by: Fan, Wen, et al.
Published: (2024)
by: Fan, Wen, et al.
Published: (2024)
Reasoning Inconsistencies and How to Mitigate Them in Deep Learning
by: Arakelyan, Erik
Published: (2025)
by: Arakelyan, Erik
Published: (2025)
ASP-Bench: From Natural Language to Logic Programs
by: Szeider, Stefan
Published: (2026)
by: Szeider, Stefan
Published: (2026)
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Reliable Reasoning with Large Language Models via Preference-Based Maximum Satisfiability
by: Orvalho, Pedro, et al.
Published: (2026)
by: Orvalho, Pedro, et al.
Published: (2026)
A Syllogistic Probe: Tracing the Evolution of Logic Reasoning in Large Language Models
by: Zang, Zhengqing, et al.
Published: (2026)
by: Zang, Zhengqing, et al.
Published: (2026)
Why Cannot Large Language Models Ever Make True Correct Reasoning?
by: Cheng, Jingde
Published: (2025)
by: Cheng, Jingde
Published: (2025)
Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors
by: Bao, Qiming, et al.
Published: (2025)
by: Bao, Qiming, et al.
Published: (2025)
Toward Guarantees for Clinical Reasoning in Vision Language Models via Formal Verification
by: Singh, Vikash, et al.
Published: (2026)
by: Singh, Vikash, et al.
Published: (2026)
Multi-Step Deductive Reasoning Over Natural Language: An Empirical Study on Out-of-Distribution Generalisation
by: Bao, Qiming, et al.
Published: (2022)
by: Bao, Qiming, et al.
Published: (2022)
Similar Items
-
Generics and Default Reasoning in Large Language Models
by: Kirkpatrick, James Ravi, et al.
Published: (2025) -
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark
by: Li, Fangjun, et al.
Published: (2024) -
Satisfiability Solving with LLMs: A Matched-Pair Evaluation of Reasoning Capability
by: Zhang, Leizhen, et al.
Published: (2026) -
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models
by: Wan, Yuxuan, et al.
Published: (2024) -
Can Large Language Models Understand DL-Lite Ontologies? An Empirical Study
by: Wang, Keyu, et al.
Published: (2024)