Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Li, Lixing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Chain-Oriented Objective Logic with Neural Network Feedback Control and Cascade Filtering for Dynamic Multi-DSL Regulation
von: Han, Jipeng
Veröffentlicht: (2024)
von: Han, Jipeng
Veröffentlicht: (2024)
Theoretical Foundations of Latent Posterior Factors: Formal Guarantees for Multi-Evidence Reasoning
von: Alege, Aliyu Agboola
Veröffentlicht: (2026)
von: Alege, Aliyu Agboola
Veröffentlicht: (2026)
COBRA-PPM: A Causal Bayesian Reasoning Architecture Using Probabilistic Programming for Robot Manipulation Under Uncertainty
von: Cannizzaro, Ricardo, et al.
Veröffentlicht: (2024)
von: Cannizzaro, Ricardo, et al.
Veröffentlicht: (2024)
Enhancing Large Language Models through Neuro-Symbolic Integration and Ontological Reasoning
von: Vsevolodovna, Ruslan Idelfonso Magana, et al.
Veröffentlicht: (2025)
von: Vsevolodovna, Ruslan Idelfonso Magana, et al.
Veröffentlicht: (2025)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
Trustworthiness Preservation by Copies of Machine Learning Systems
von: Ceragioli, Leonardo, et al.
Veröffentlicht: (2025)
von: Ceragioli, Leonardo, et al.
Veröffentlicht: (2025)
Normative Conditional Reasoning as a Fragment of HOL
von: Parent, Xavier, et al.
Veröffentlicht: (2023)
von: Parent, Xavier, et al.
Veröffentlicht: (2023)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
LLM-Assisted Formalization Enables Deterministic Detection of Statutory Inconsistency in the Internal Revenue Code
von: Yadamsuren, Borchuluun, et al.
Veröffentlicht: (2025)
von: Yadamsuren, Borchuluun, et al.
Veröffentlicht: (2025)
LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics
von: Peyronnet, Antoine, et al.
Veröffentlicht: (2026)
von: Peyronnet, Antoine, et al.
Veröffentlicht: (2026)
A Prompt Learning Framework for Source Code Summarization
von: Xu, Tingting, et al.
Veröffentlicht: (2023)
von: Xu, Tingting, et al.
Veröffentlicht: (2023)
Many Logics, One Methodology: A Plea for Logical Pluralism in Formalised Reasoning (preprint)
von: Benzmüller, Christoph, et al.
Veröffentlicht: (2026)
von: Benzmüller, Christoph, et al.
Veröffentlicht: (2026)
I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning
von: Alege, Aliyu Agboola
Veröffentlicht: (2026)
von: Alege, Aliyu Agboola
Veröffentlicht: (2026)
Chase Anonymisation: Privacy-Preserving Knowledge Graphs with Logical Reasoning
von: Bellomarini, Luigi, et al.
Veröffentlicht: (2024)
von: Bellomarini, Luigi, et al.
Veröffentlicht: (2024)
Imandra CodeLogician: Neuro-Symbolic Reasoning for Precise Analysis of Software Logic
von: Lin, Hongyu, et al.
Veröffentlicht: (2026)
von: Lin, Hongyu, et al.
Veröffentlicht: (2026)
Conformal Path Reasoning: Trustworthy Knowledge Graph Question Answering via Path-Level Calibration
von: Lin, Shuhang, et al.
Veröffentlicht: (2026)
von: Lin, Shuhang, et al.
Veröffentlicht: (2026)
An Encoding of Abstract Dialectical Frameworks into Higher-Order Logic
von: Martina, Antoine, et al.
Veröffentlicht: (2023)
von: Martina, Antoine, et al.
Veröffentlicht: (2023)
Anomaly Detection and Inlet Pressure Prediction in Water Distribution Systems Using Machine Learning
von: Khoa, Tran Dang
Veröffentlicht: (2024)
von: Khoa, Tran Dang
Veröffentlicht: (2024)
Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining
von: Cao, Deyu, et al.
Veröffentlicht: (2025)
von: Cao, Deyu, et al.
Veröffentlicht: (2025)
LTL Verification of Memoryful Neural Agents
von: Hosseini, Mehran, et al.
Veröffentlicht: (2025)
von: Hosseini, Mehran, et al.
Veröffentlicht: (2025)
Reasoning and Planning with Dynamically Changing Norms
von: Olson, Taylor, et al.
Veröffentlicht: (2026)
von: Olson, Taylor, et al.
Veröffentlicht: (2026)
From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction
von: Petrov, Alex, et al.
Veröffentlicht: (2026)
von: Petrov, Alex, et al.
Veröffentlicht: (2026)
TPTP World Infrastructure for Non-classical Logics
von: Steen, Alexander, et al.
Veröffentlicht: (2025)
von: Steen, Alexander, et al.
Veröffentlicht: (2025)
Intersymbolic AI: Interlinking Symbolic AI and Subsymbolic AI
von: Platzer, André
Veröffentlicht: (2024)
von: Platzer, André
Veröffentlicht: (2024)
Structure Transfer: an Inference-Based Calculus for the Transformation of Representations
von: Raggi, Daniel, et al.
Veröffentlicht: (2025)
von: Raggi, Daniel, et al.
Veröffentlicht: (2025)
Beyond Prediction -- Structuring Epistemic Integrity in Artificial Reasoning Systems
von: Wright, Craig Steven
Veröffentlicht: (2025)
von: Wright, Craig Steven
Veröffentlicht: (2025)
Faithful Logic Embeddings in HOL -- Deep and Shallow
von: Benzmüller, Christoph
Veröffentlicht: (2025)
von: Benzmüller, Christoph
Veröffentlicht: (2025)
A Proof System with Causal Labels (Part II): checking Counterfactual Fairness
von: Ceragioli, Leonardo, et al.
Veröffentlicht: (2025)
von: Ceragioli, Leonardo, et al.
Veröffentlicht: (2025)
A Proof System with Causal Labels (Part I): checking Individual Fairness and Intersectionality
von: Ceragioli, Leonardo, et al.
Veröffentlicht: (2025)
von: Ceragioli, Leonardo, et al.
Veröffentlicht: (2025)
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
Sliced-Wasserstein Distribution Alignment Loss Improves the Ultra-Low-Bit Quantization of Large Language Models
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
Graph2Tac: Online Representation Learning of Formal Math Concepts
von: Blaauwbroek, Lasse, et al.
Veröffentlicht: (2024)
von: Blaauwbroek, Lasse, et al.
Veröffentlicht: (2024)
BRIDGE: Building Representations In Domain Guided Program Synthesis
von: George, Robert Joseph, et al.
Veröffentlicht: (2025)
von: George, Robert Joseph, et al.
Veröffentlicht: (2025)
Prediction-space knowledge markets for communication-efficient federated learning on multimedia tasks
von: Du, Wenzhang
Veröffentlicht: (2025)
von: Du, Wenzhang
Veröffentlicht: (2025)
Definite Descriptions and Hybrid Tense Logic
von: Indrzejczak, Andrzej, et al.
Veröffentlicht: (2024)
von: Indrzejczak, Andrzej, et al.
Veröffentlicht: (2024)
SPARQL in N3: SPARQL CONSTRUCT as a rule language for the Semantic Web (Extended Version)
von: Arndt, Dörthe, et al.
Veröffentlicht: (2025)
von: Arndt, Dörthe, et al.
Veröffentlicht: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode
von: Vernie, Julius, et al.
Veröffentlicht: (2026)
von: Vernie, Julius, et al.
Veröffentlicht: (2026)
Fitting Ontologies and Constraints to Relational Structures
von: Hosemann, Simon, et al.
Veröffentlicht: (2025)
von: Hosemann, Simon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Chain-Oriented Objective Logic with Neural Network Feedback Control and Cascade Filtering for Dynamic Multi-DSL Regulation
von: Han, Jipeng
Veröffentlicht: (2024) -
Theoretical Foundations of Latent Posterior Factors: Formal Guarantees for Multi-Evidence Reasoning
von: Alege, Aliyu Agboola
Veröffentlicht: (2026) -
COBRA-PPM: A Causal Bayesian Reasoning Architecture Using Probabilistic Programming for Robot Manipulation Under Uncertainty
von: Cannizzaro, Ricardo, et al.
Veröffentlicht: (2024) -
Enhancing Large Language Models through Neuro-Symbolic Integration and Ontological Reasoning
von: Vsevolodovna, Ruslan Idelfonso Magana, et al.
Veröffentlicht: (2025) -
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)