Satisfiability Solving with LLMs: A Matched-Pair Evaluation of Reasoning Capability
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Leizhen, Chen, Shuhan, Chen, Sheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study
von: Zhou, Yujun, et al.
Veröffentlicht: (2025)
von: Zhou, Yujun, et al.
Veröffentlicht: (2025)
Solving Satisfiability Modulo Counting Exactly with Probabilistic Circuits
von: Li, Jinzhao, et al.
Veröffentlicht: (2025)
von: Li, Jinzhao, et al.
Veröffentlicht: (2025)
Reasoning Capabilities of Large Language Models. Lessons Learned from General Game Playing
von: Świechowski, Maciej, et al.
Veröffentlicht: (2026)
von: Świechowski, Maciej, et al.
Veröffentlicht: (2026)
LLM-ARC: Enhancing LLMs with an Automated Reasoning Critic
von: Kalyanpur, Aditya, et al.
Veröffentlicht: (2024)
von: Kalyanpur, Aditya, et al.
Veröffentlicht: (2024)
Towards Advanced Mathematical Reasoning for LLMs via First-Order Logic Theorem Proving
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
GaloisSAT: Differentiable Boolean Satisfiability Solving via Finite Field Algebra
von: Kim, Curie, et al.
Veröffentlicht: (2026)
von: Kim, Curie, et al.
Veröffentlicht: (2026)
A Reliable Common-Sense Reasoning Socialbot Built Using LLMs and Goal-Directed ASP
von: Zeng, Yankai, et al.
Veröffentlicht: (2024)
von: Zeng, Yankai, et al.
Veröffentlicht: (2024)
Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
Bridging LLMs and Symbolic Reasoning in Educational QA Systems: Insights from the XAI Challenge at IJCNN 2025
von: Nguyen, Long S. T., et al.
Veröffentlicht: (2025)
von: Nguyen, Long S. T., et al.
Veröffentlicht: (2025)
Solving Satisfiability Modulo Counting for Symbolic and Statistical AI Integration With Provable Guarantees
von: Li, Jinzhao, et al.
Veröffentlicht: (2023)
von: Li, Jinzhao, et al.
Veröffentlicht: (2023)
TimelineKGQA: A Comprehensive Question-Answer Pair Generator for Temporal Knowledge Graphs
von: Sun, Qiang, et al.
Veröffentlicht: (2025)
von: Sun, Qiang, et al.
Veröffentlicht: (2025)
Reliable Reasoning with Large Language Models via Preference-Based Maximum Satisfiability
von: Orvalho, Pedro, et al.
Veröffentlicht: (2026)
von: Orvalho, Pedro, et al.
Veröffentlicht: (2026)
Beyond Theorem Proving: Formulation, Framework and Benchmark for Formal Problem-Solving
von: Liu, Qi, et al.
Veröffentlicht: (2025)
von: Liu, Qi, et al.
Veröffentlicht: (2025)
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks
von: Ganguly, Debargha, et al.
Veröffentlicht: (2025)
von: Ganguly, Debargha, et al.
Veröffentlicht: (2025)
Question Answering with LLMs and Learning from Answer Sets
von: Borroto, Manuel, et al.
Veröffentlicht: (2025)
von: Borroto, Manuel, et al.
Veröffentlicht: (2025)
Generics and Default Reasoning in Large Language Models
von: Kirkpatrick, James Ravi, et al.
Veröffentlicht: (2025)
von: Kirkpatrick, James Ravi, et al.
Veröffentlicht: (2025)
EXPLORER: Exploration-guided Reasoning for Textual Reinforcement Learning
von: Basu, Kinjal, et al.
Veröffentlicht: (2024)
von: Basu, Kinjal, et al.
Veröffentlicht: (2024)
Sound and Complete Neurosymbolic Reasoning with LLM-Grounded Interpretations
von: Allen, Bradley P., et al.
Veröffentlicht: (2025)
von: Allen, Bradley P., et al.
Veröffentlicht: (2025)
Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning
von: Bourigault, Pauline, et al.
Veröffentlicht: (2026)
von: Bourigault, Pauline, et al.
Veröffentlicht: (2026)
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark
von: Li, Fangjun, et al.
Veröffentlicht: (2024)
von: Li, Fangjun, et al.
Veröffentlicht: (2024)
Transformer Encoder Satisfiability: Complexity and Impact on Formal Reasoning
von: Sälzer, Marco, et al.
Veröffentlicht: (2024)
von: Sälzer, Marco, et al.
Veröffentlicht: (2024)
From Blind Solvers to Logical Thinkers: Benchmarking LLMs' Logical Integrity on Faulty Mathematical Problems
von: Rahman, A M Muntasir, et al.
Veröffentlicht: (2024)
von: Rahman, A M Muntasir, et al.
Veröffentlicht: (2024)
Entailment vs. Verification for Partial-assignment Satisfiability and Enumeration
von: Sebastiani, Roberto
Veröffentlicht: (2025)
von: Sebastiani, Roberto
Veröffentlicht: (2025)
Logic-Parametric Neuro-Symbolic NLI: Controlling Logical Formalisms for Verifiable LLM Reasoning
von: Farjami, Ali, et al.
Veröffentlicht: (2026)
von: Farjami, Ali, et al.
Veröffentlicht: (2026)
Ontology for Policing: Conceptual Knowledge Learning for Semantic Understanding and Reasoning in Law Enforcement Reports
von: Srbinovska, Anita, et al.
Veröffentlicht: (2026)
von: Srbinovska, Anita, et al.
Veröffentlicht: (2026)
Probabilistic and Causal Satisfiability: Constraining the Model
von: Bläser, Markus, et al.
Veröffentlicht: (2025)
von: Bläser, Markus, et al.
Veröffentlicht: (2025)
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Do LLMs Really Struggle at NL-FOL Translation? Revealing their Strengths via a Novel Benchmarking Strategy
von: Brunello, Andrea, et al.
Veröffentlicht: (2025)
von: Brunello, Andrea, et al.
Veröffentlicht: (2025)
Defining implication relation for classical logic
von: Fu, Li
Veröffentlicht: (2013)
von: Fu, Li
Veröffentlicht: (2013)
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models
von: Wan, Yuxuan, et al.
Veröffentlicht: (2024)
von: Wan, Yuxuan, et al.
Veröffentlicht: (2024)
Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning
von: Luo, Ziyan, et al.
Veröffentlicht: (2023)
von: Luo, Ziyan, et al.
Veröffentlicht: (2023)
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
Quantified Linear and Polynomial Arithmetic Satisfiability via Template-based Skolemization
von: Chatterjee, Krishnendu, et al.
Veröffentlicht: (2024)
von: Chatterjee, Krishnendu, et al.
Veröffentlicht: (2024)
Constrained and Robust Policy Synthesis with Satisfiability-Modulo-Probabilistic-Model-Checking
von: Heck, Linus, et al.
Veröffentlicht: (2025)
von: Heck, Linus, et al.
Veröffentlicht: (2025)
Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
Enhancing Reasoning Capabilities of LLMs via Principled Synthetic Logic Corpus
von: Morishita, Terufumi, et al.
Veröffentlicht: (2024)
von: Morishita, Terufumi, et al.
Veröffentlicht: (2024)
DECIDER: A Dual-System Rule-Controllable Decoding Framework for Language Generation
von: Xu, Chen, et al.
Veröffentlicht: (2024)
von: Xu, Chen, et al.
Veröffentlicht: (2024)
Satisfying Rationality Postulates of Structured Argumentation Through Deductive Support -- Technical Report
von: Cramer, Marcos, et al.
Veröffentlicht: (2026)
von: Cramer, Marcos, et al.
Veröffentlicht: (2026)
Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving
von: Liang, Zhenwen, et al.
Veröffentlicht: (2025)
von: Liang, Zhenwen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study
von: Zhou, Yujun, et al.
Veröffentlicht: (2025) -
Solving Satisfiability Modulo Counting Exactly with Probabilistic Circuits
von: Li, Jinzhao, et al.
Veröffentlicht: (2025) -
Reasoning Capabilities of Large Language Models. Lessons Learned from General Game Playing
von: Świechowski, Maciej, et al.
Veröffentlicht: (2026) -
LLM-ARC: Enhancing LLMs with an Automated Reasoning Critic
von: Kalyanpur, Aditya, et al.
Veröffentlicht: (2024) -
Towards Advanced Mathematical Reasoning for LLMs via First-Order Logic Theorem Proving
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)