Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Yujun, Ye, Jiayi, Ling, Zipeng, Han, Yufei, Huang, Yue, Zhuang, Haomin, Liang, Zhenwen, Guo, Kehan, Guo, Taicheng, Wang, Xiangqi, Zhang, Xiangliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dual Optimal: Make Your LLM Peer-like with Dignity
von: Wang, Xiangqi, et al.
Veröffentlicht: (2026)
von: Wang, Xiangqi, et al.
Veröffentlicht: (2026)
Social Science Meets LLMs: How Reliable Are Large Language Models in Social Simulations?
von: Huang, Yue, et al.
Veröffentlicht: (2024)
von: Huang, Yue, et al.
Veröffentlicht: (2024)
Exploring Multi-Temperature Strategies for Token- and Rollout-Level Control in RLVR
von: Zhuang, Haomin, et al.
Veröffentlicht: (2025)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2025)
Reliable Control-Point Selection for Steering Reasoning in Large Language Models
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
Defending Jailbreak Prompts via In-Context Adversarial Game
von: Zhou, Yujun, et al.
Veröffentlicht: (2024)
von: Zhou, Yujun, et al.
Veröffentlicht: (2024)
Modal Logic for Reasoning About Uncertainty and Confusion
von: Bílková, Marta, et al.
Veröffentlicht: (2025)
von: Bílková, Marta, et al.
Veröffentlicht: (2025)
SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
Towards Advanced Mathematical Reasoning for LLMs via First-Order Logic Theorem Proving
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
SATQuest: A Verifier for Logical Reasoning Evaluation and Reinforcement Fine-Tuning of LLMs
von: Zhao, Yanxiao, et al.
Veröffentlicht: (2025)
von: Zhao, Yanxiao, et al.
Veröffentlicht: (2025)
Causally-Enhanced Reinforcement Policy Optimization
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
Learn from Failure: Fine-Tuning LLMs with Trial-and-Error Data for Intuitionistic Propositional Logic Proving
von: An, Chenyang, et al.
Veröffentlicht: (2024)
von: An, Chenyang, et al.
Veröffentlicht: (2024)
Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
Field Knowledge as a Dual to Distributed Knowledge: A Characterization by Weighted Modal Logic
von: Liang, Xiaolong, et al.
Veröffentlicht: (2024)
von: Liang, Xiaolong, et al.
Veröffentlicht: (2024)
Towards Solving More Challenging IMO Problems via Decoupled Reasoning and Proving
von: Liang, Zhenwen, et al.
Veröffentlicht: (2025)
von: Liang, Zhenwen, et al.
Veröffentlicht: (2025)
Agent-Knowledge Logic for Alternative Epistemic Logic
von: Nishimura, Yuki
Veröffentlicht: (2024)
von: Nishimura, Yuki
Veröffentlicht: (2024)
Total Outcome Logic: Unified Reasoning for a Taxonomy of Program Logics
von: Li, James, et al.
Veröffentlicht: (2024)
von: Li, James, et al.
Veröffentlicht: (2024)
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
Nested Sequents for Intermediate Logics: The Case of Gödel-Dummett Logics
von: Lyon, Tim S.
Veröffentlicht: (2023)
von: Lyon, Tim S.
Veröffentlicht: (2023)
Justification Logic for Intuitionistic Modal Logic (Extended Technical Report)
von: Marin, Sonia, et al.
Veröffentlicht: (2025)
von: Marin, Sonia, et al.
Veröffentlicht: (2025)
Proof Complexity of Linear Logics
von: Tabatabai, Amirhossein Akbar, et al.
Veröffentlicht: (2026)
von: Tabatabai, Amirhossein Akbar, et al.
Veröffentlicht: (2026)
Skolemization In Intermediate Logics
von: Baaz, Matthias, et al.
Veröffentlicht: (2025)
von: Baaz, Matthias, et al.
Veröffentlicht: (2025)
Constructive Quantum Logics
von: Aguilera, Juan P., et al.
Veröffentlicht: (2025)
von: Aguilera, Juan P., et al.
Veröffentlicht: (2025)
A Logic of Inability
von: Wang, Shanxia
Veröffentlicht: (2026)
von: Wang, Shanxia
Veröffentlicht: (2026)
From Blind Solvers to Logical Thinkers: Benchmarking LLMs' Logical Integrity on Faulty Mathematical Problems
von: Rahman, A M Muntasir, et al.
Veröffentlicht: (2024)
von: Rahman, A M Muntasir, et al.
Veröffentlicht: (2024)
FALCON: Scalable Reasoning over Inconsistent ALC Ontologies
von: Hinnerichs, Tilman, et al.
Veröffentlicht: (2022)
von: Hinnerichs, Tilman, et al.
Veröffentlicht: (2022)
Dependence Logics in Temporal Settings
von: Baltag, Alexandru, et al.
Veröffentlicht: (2022)
von: Baltag, Alexandru, et al.
Veröffentlicht: (2022)
Decidability of Quantum Modal Logic
von: Tokuo, Kenji
Veröffentlicht: (2026)
von: Tokuo, Kenji
Veröffentlicht: (2026)
Dynamic Cantor Derivative Logic
von: Fernández-Duque, David, et al.
Veröffentlicht: (2021)
von: Fernández-Duque, David, et al.
Veröffentlicht: (2021)
Reasoning about Medical Triage Optimization with Logic Programming
von: Patil, Jaikrishna Manojkumar, et al.
Veröffentlicht: (2025)
von: Patil, Jaikrishna Manojkumar, et al.
Veröffentlicht: (2025)
Sequent Calculi for Data-Aware Modal Logics
von: Areces, Carlos, et al.
Veröffentlicht: (2025)
von: Areces, Carlos, et al.
Veröffentlicht: (2025)
Dynamic Probability Logic: Decidability & Computability
von: Chopoghloo, Somayeh, et al.
Veröffentlicht: (2024)
von: Chopoghloo, Somayeh, et al.
Veröffentlicht: (2024)
Decidability of Quasi-Dense Modal Logics
von: Ostropolski-Nalewaja, Piotr, et al.
Veröffentlicht: (2024)
von: Ostropolski-Nalewaja, Piotr, et al.
Veröffentlicht: (2024)
A Study on Actions for Atomic Logics
von: Espejo-Boix, Raül
Veröffentlicht: (2024)
von: Espejo-Boix, Raül
Veröffentlicht: (2024)
Extending Action Logic with Omega Iteration
von: Pshenitsyn, Tikhon
Veröffentlicht: (2025)
von: Pshenitsyn, Tikhon
Veröffentlicht: (2025)
A Logic of Secrecy on Simplicial Models
von: Wang, Shanxia
Veröffentlicht: (2026)
von: Wang, Shanxia
Veröffentlicht: (2026)
Distribution-Free Normal Modal Logics
von: Hartonas, Chrysafis
Veröffentlicht: (2024)
von: Hartonas, Chrysafis
Veröffentlicht: (2024)
Base-extension Semantics for Modal Logic
von: Eckhardt, Timo, et al.
Veröffentlicht: (2024)
von: Eckhardt, Timo, et al.
Veröffentlicht: (2024)
Encoding and Reasoning about Arrays in Constraint Logic Programming with Sets
von: Cristiá, Maximiliano, et al.
Veröffentlicht: (2025)
von: Cristiá, Maximiliano, et al.
Veröffentlicht: (2025)
Reasoning About Probabilities, Actions, and Knowledge in Fuzzy Modal Logic
von: Kozhemiachenko, Daniil, et al.
Veröffentlicht: (2026)
von: Kozhemiachenko, Daniil, et al.
Veröffentlicht: (2026)
Partially Finite Model Reasoning in Description Logics Extended Version
von: Gogacz, Tomasz, et al.
Veröffentlicht: (2026)
von: Gogacz, Tomasz, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Dual Optimal: Make Your LLM Peer-like with Dignity
von: Wang, Xiangqi, et al.
Veröffentlicht: (2026) -
Social Science Meets LLMs: How Reliable Are Large Language Models in Social Simulations?
von: Huang, Yue, et al.
Veröffentlicht: (2024) -
Exploring Multi-Temperature Strategies for Token- and Rollout-Level Control in RLVR
von: Zhuang, Haomin, et al.
Veröffentlicht: (2025) -
Reliable Control-Point Selection for Steering Reasoning in Large Language Models
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026) -
Defending Jailbreak Prompts via In-Context Adversarial Game
von: Zhou, Yujun, et al.
Veröffentlicht: (2024)