Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yinhong, Guo, Zhijiang, Liang, Tianya, Shareghi, Ehsan, Vulić, Ivan, Collier, Nigel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
von: Zhou, Han, et al.
Veröffentlicht: (2024)
von: Zhou, Han, et al.
Veröffentlicht: (2024)
Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models
von: Wan, Yuxuan, et al.
Veröffentlicht: (2024)
von: Wan, Yuxuan, et al.
Veröffentlicht: (2024)
All Roads Lead to Rome: Graph-Based Confidence Estimation for Large Language Model Reasoning
von: Zhang, Caiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Caiqi, et al.
Veröffentlicht: (2025)
Logic Programming with Extensible Types
von: Perez, Ivan, et al.
Veröffentlicht: (2026)
von: Perez, Ivan, et al.
Veröffentlicht: (2026)
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
Improving Symbolic Translation of Language Models for Logical Reasoning
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2026)
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2026)
MedLogic-AQA: Enhancing Medical Question Answering with Abstractive Models Focusing on Logical Structures
von: Zafar, Aizan, et al.
Veröffentlicht: (2024)
von: Zafar, Aizan, et al.
Veröffentlicht: (2024)
ReasonGraph: Visualisation of Reasoning Paths
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
Prompt Compression for Large Language Models: A Survey
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study
von: Zhou, Yujun, et al.
Veröffentlicht: (2025)
von: Zhou, Yujun, et al.
Veröffentlicht: (2025)
Compositional Consistency-Guided Decoding for Three-Way Logical Question Answering
von: Huang, Tianyi, et al.
Veröffentlicht: (2026)
von: Huang, Tianyi, et al.
Veröffentlicht: (2026)
Gradual Exact Logic: Unifying Hoare Logic and Incorrectness Logic via Gradual Verification
von: Zimmerman, Conrad, et al.
Veröffentlicht: (2024)
von: Zimmerman, Conrad, et al.
Veröffentlicht: (2024)
Towards Logically Sound Natural Language Reasoning with Logic-Enhanced Language Model Agents
von: Mensfelt, Agnieszka, et al.
Veröffentlicht: (2024)
von: Mensfelt, Agnieszka, et al.
Veröffentlicht: (2024)
Hybrid Models for Natural Language Reasoning: The Case of Syllogistic Logic
von: Guzmán, Manuel Vargas, et al.
Veröffentlicht: (2025)
von: Guzmán, Manuel Vargas, et al.
Veröffentlicht: (2025)
Unrealizability Logic
von: Kim, Jinwoo, et al.
Veröffentlicht: (2022)
von: Kim, Jinwoo, et al.
Veröffentlicht: (2022)
ASP-Bench: From Natural Language to Logic Programs
von: Szeider, Stefan
Veröffentlicht: (2026)
von: Szeider, Stefan
Veröffentlicht: (2026)
Partial Incorrectness Logic
von: Verscht, Lena, et al.
Veröffentlicht: (2025)
von: Verscht, Lena, et al.
Veröffentlicht: (2025)
Outcome Logic: A Unified Approach to the Metatheory of Program Logics with Branching Effects
von: Zilberstein, Noam
Veröffentlicht: (2024)
von: Zilberstein, Noam
Veröffentlicht: (2024)
From Blind Solvers to Logical Thinkers: Benchmarking LLMs' Logical Integrity on Faulty Mathematical Problems
von: Rahman, A M Muntasir, et al.
Veröffentlicht: (2024)
von: Rahman, A M Muntasir, et al.
Veröffentlicht: (2024)
Quantifying Logical Consistency in Transformers via Query-Key Alignment
von: Tulchinskii, Eduard, et al.
Veröffentlicht: (2025)
von: Tulchinskii, Eduard, et al.
Veröffentlicht: (2025)
KOS-TL (Knowledge Operation System Type Logic)
von: Chen, Peng
Veröffentlicht: (2026)
von: Chen, Peng
Veröffentlicht: (2026)
Logic-Parametric Neuro-Symbolic NLI: Controlling Logical Formalisms for Verifiable LLM Reasoning
von: Farjami, Ali, et al.
Veröffentlicht: (2026)
von: Farjami, Ali, et al.
Veröffentlicht: (2026)
Positive First-order Logic on Words and Graphs
von: Kuperberg, Denis
Veröffentlicht: (2022)
von: Kuperberg, Denis
Veröffentlicht: (2022)
Recursive Mutexes in Separation Logic
von: Du, Ke, et al.
Veröffentlicht: (2026)
von: Du, Ke, et al.
Veröffentlicht: (2026)
Finite-Choice Logic Programming
von: Martens, Chris, et al.
Veröffentlicht: (2024)
von: Martens, Chris, et al.
Veröffentlicht: (2024)
Recursive Decomposition of Logical Thoughts: Framework for Superior Reasoning and Knowledge Propagation in Large Language Models
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
HistMSO: A Logic for Reasoning about Consistency Models with MONA
von: Coget, Isabelle, et al.
Veröffentlicht: (2026)
von: Coget, Isabelle, et al.
Veröffentlicht: (2026)
Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
von: Lalwani, Abhinav, et al.
Veröffentlicht: (2024)
von: Lalwani, Abhinav, et al.
Veröffentlicht: (2024)
Logic and Languages of Higher-Dimensional Automata
von: Amrane, Amazigh, et al.
Veröffentlicht: (2024)
von: Amrane, Amazigh, et al.
Veröffentlicht: (2024)
Logical forms complement probability in understanding language model (and human) performance
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
The Alternation Hierarchy of First-Order Logic on Words is Decidable
von: Barloy, Corentin, et al.
Veröffentlicht: (2025)
von: Barloy, Corentin, et al.
Veröffentlicht: (2025)
Ordered Adjoint Logic (Extended Version)
von: Roshal, Sophia, et al.
Veröffentlicht: (2026)
von: Roshal, Sophia, et al.
Veröffentlicht: (2026)
Towards Concurrent Quantitative Separation Logic
von: Fesefeldt, Ira, et al.
Veröffentlicht: (2022)
von: Fesefeldt, Ira, et al.
Veröffentlicht: (2022)
Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors
von: Bao, Qiming, et al.
Veröffentlicht: (2025)
von: Bao, Qiming, et al.
Veröffentlicht: (2025)
Proceedings Twentieth International Workshop on Logical Frameworks and Meta-Languages: Theory and Practice
von: Chaudhuri, Kaustuv, et al.
Veröffentlicht: (2025)
von: Chaudhuri, Kaustuv, et al.
Veröffentlicht: (2025)
Are Language Models Efficient Reasoners? A Perspective from Logic Programming
von: Opedal, Andreas, et al.
Veröffentlicht: (2025)
von: Opedal, Andreas, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
von: Liu, Yinhong, et al.
Veröffentlicht: (2024) -
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
von: Zhou, Han, et al.
Veröffentlicht: (2024) -
Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence
von: Liu, Yinhong, et al.
Veröffentlicht: (2024) -
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models
von: Wan, Yuxuan, et al.
Veröffentlicht: (2024) -
All Roads Lead to Rome: Graph-Based Confidence Estimation for Large Language Model Reasoning
von: Zhang, Caiqi, et al.
Veröffentlicht: (2025)