RULEBREAKERS: Challenging LLMs at the Crossroads between Formal Logic and Human-like Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Chan, Jason, Gaizauskas, Robert, Zhao, Zhixue |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs
por: Chan, Jason, et al.
Publicado: (2026)
por: Chan, Jason, et al.
Publicado: (2026)
Position: On the Methodological Pitfalls of Evaluating Base LLMs for Reasoning
por: Chan, Jason, et al.
Publicado: (2025)
por: Chan, Jason, et al.
Publicado: (2025)
Explanation Generation for Contradiction Reconciliation with LLMs
por: Chan, Jason, et al.
Publicado: (2026)
por: Chan, Jason, et al.
Publicado: (2026)
Are LLMs Stable Formal Logic Translators in Logical Reasoning Across Linguistically Diversified Texts?
por: Li, Qingchuan, et al.
Publicado: (2025)
por: Li, Qingchuan, et al.
Publicado: (2025)
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
por: Zhao, Zhixue, et al.
Publicado: (2024)
por: Zhao, Zhixue, et al.
Publicado: (2024)
It's All About In-Context Learning! Teaching Extremely Low-Resource Languages to LLMs
por: Li, Yue, et al.
Publicado: (2025)
por: Li, Yue, et al.
Publicado: (2025)
Tracing and Reversing Edits in LLMs
por: Youssef, Paul, et al.
Publicado: (2025)
por: Youssef, Paul, et al.
Publicado: (2025)
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
por: Nie, Shangrui, et al.
Publicado: (2025)
por: Nie, Shangrui, et al.
Publicado: (2025)
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits
por: Youssef, Paul, et al.
Publicado: (2024)
por: Youssef, Paul, et al.
Publicado: (2024)
Do LLMs Game Formalization? Evaluating Faithfulness in Logical Reasoning
por: Kim, Kyuhee, et al.
Publicado: (2026)
por: Kim, Kyuhee, et al.
Publicado: (2026)
Can Confidence Estimates Decide When Chain-of-Thought Is Necessary for LLMs?
por: Lewis-Lim, Samuel, et al.
Publicado: (2025)
por: Lewis-Lim, Samuel, et al.
Publicado: (2025)
Boosting Logical Fallacy Reasoning in LLMs via Logical Structure Tree
por: Lei, Yuanyuan, et al.
Publicado: (2024)
por: Lei, Yuanyuan, et al.
Publicado: (2024)
Empowering LLMs with Logical Reasoning: A Comprehensive Survey
por: Cheng, Fengxiang, et al.
Publicado: (2025)
por: Cheng, Fengxiang, et al.
Publicado: (2025)
From Early Encoding to Late Suppression: Interpreting LLMs on Character Counting Tasks
por: Datta, Ayan, et al.
Publicado: (2026)
por: Datta, Ayan, et al.
Publicado: (2026)
Beyond Memorization: Distinguishing between Reductive and Epistemic Reasoning in LLMs using Classic Logic Puzzles
por: Gabay, Adi, et al.
Publicado: (2026)
por: Gabay, Adi, et al.
Publicado: (2026)
Bridging the Gap: In-Context Learning for Modeling Human Disagreement
por: Muscato, Benedetta, et al.
Publicado: (2025)
por: Muscato, Benedetta, et al.
Publicado: (2025)
Logic-Parametric Neuro-Symbolic NLI: Controlling Logical Formalisms for Verifiable LLM Reasoning
por: Farjami, Ali, et al.
Publicado: (2026)
por: Farjami, Ali, et al.
Publicado: (2026)
Where Reasoning Breaks: Logic-Aware Path Selection by Controlling Logical Connectives in LLMs Reasoning Chains
por: Park, Seunghyun, et al.
Publicado: (2026)
por: Park, Seunghyun, et al.
Publicado: (2026)
Last Layer Logits to Logic: Empowering LLMs with Logic-Consistent Structured Knowledge Reasoning
por: Li, Songze, et al.
Publicado: (2025)
por: Li, Songze, et al.
Publicado: (2025)
HCR-Reasoner: Synergizing Large Language Models and Theory for Human-like Causal Reasoning
por: Zhang, Yanxi, et al.
Publicado: (2025)
por: Zhang, Yanxi, et al.
Publicado: (2025)
Can LLMs Reason with Rules? Logic Scaffolding for Stress-Testing and Improving LLMs
por: Wang, Siyuan, et al.
Publicado: (2024)
por: Wang, Siyuan, et al.
Publicado: (2024)
ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
por: Lin, Bill Yuchen, et al.
Publicado: (2025)
por: Lin, Bill Yuchen, et al.
Publicado: (2025)
HAL: Inducing Human-likeness in LLMs with Alignment
por: Hasan, Masum, et al.
Publicado: (2026)
por: Hasan, Masum, et al.
Publicado: (2026)
NumeroLogic: Number Encoding for Enhanced LLMs' Numerical Reasoning
por: Schwartz, Eli, et al.
Publicado: (2024)
por: Schwartz, Eli, et al.
Publicado: (2024)
Label Set Optimization via Activation Distribution Kurtosis for Zero-shot Classification with Generative Models
por: Li, Yue, et al.
Publicado: (2024)
por: Li, Yue, et al.
Publicado: (2024)
LLMEval-Logic: A Solver-Verified Chinese Benchmark for Logical Reasoning of LLMs with Adversarial Hardening
por: Zhang, Ming, et al.
Publicado: (2026)
por: Zhang, Ming, et al.
Publicado: (2026)
Do Large Language Models Excel in Complex Logical Reasoning with Formal Language?
por: Jiang, Jin, et al.
Publicado: (2025)
por: Jiang, Jin, et al.
Publicado: (2025)
LogicSkills: A Structured Benchmark for Formal Reasoning in Large Language Models
por: Rabern, Brian, et al.
Publicado: (2026)
por: Rabern, Brian, et al.
Publicado: (2026)
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models
por: Zhao, Zhixue, et al.
Publicado: (2024)
por: Zhao, Zhixue, et al.
Publicado: (2024)
Incorporating Attribution Importance for Improving Faithfulness Metrics
por: Zhao, Zhixue, et al.
Publicado: (2023)
por: Zhao, Zhixue, et al.
Publicado: (2023)
Mitigating Content Effects on Reasoning in Language Models through Fine-Grained Activation Steering
por: Valentino, Marco, et al.
Publicado: (2025)
por: Valentino, Marco, et al.
Publicado: (2025)
Toward Robust Legal Text Formalization into Defeasible Deontic Logic using LLMs
por: Horner, Elias, et al.
Publicado: (2025)
por: Horner, Elias, et al.
Publicado: (2025)
Logic Haystacks: Probing LLMs Long-Context Logical Reasoning (Without Easily Identifiable Unrelated Padding)
por: Sileo, Damien
Publicado: (2025)
por: Sileo, Damien
Publicado: (2025)
A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters
por: Lam, Long Hei Matthew, et al.
Publicado: (2024)
por: Lam, Long Hei Matthew, et al.
Publicado: (2024)
Bridging Legal Interpretation and Formal Logic: Faithfulness, Assumption, and the Future of AI Legal Reasoning
por: Wang, Olivia Peiyu, et al.
Publicado: (2026)
por: Wang, Olivia Peiyu, et al.
Publicado: (2026)
Logic-Regularized Verifier Elicits Reasoning from LLMs
por: Wang, Xinyu, et al.
Publicado: (2026)
por: Wang, Xinyu, et al.
Publicado: (2026)
How Deep is Love in LLMs' Hearts? Exploring Semantic Size in Human-like Cognition
por: Yao, Yao, et al.
Publicado: (2025)
por: Yao, Yao, et al.
Publicado: (2025)
P-FOLIO: Evaluating and Improving Logical Reasoning with Abundant Human-Written Reasoning Chains
por: Han, Simeng, et al.
Publicado: (2024)
por: Han, Simeng, et al.
Publicado: (2024)
Logic Jailbreak: Efficiently Unlocking LLM Safety Restrictions Through Formal Logical Expression
por: Peng, Jingyu, et al.
Publicado: (2025)
por: Peng, Jingyu, et al.
Publicado: (2025)
Exploring Vision Language Models for Multimodal and Multilingual Stance Detection
por: Vasilakes, Jake, et al.
Publicado: (2025)
por: Vasilakes, Jake, et al.
Publicado: (2025)
Ejemplares similares
-
Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs
por: Chan, Jason, et al.
Publicado: (2026) -
Position: On the Methodological Pitfalls of Evaluating Base LLMs for Reasoning
por: Chan, Jason, et al.
Publicado: (2025) -
Explanation Generation for Contradiction Reconciliation with LLMs
por: Chan, Jason, et al.
Publicado: (2026) -
Are LLMs Stable Formal Logic Translators in Logical Reasoning Across Linguistically Diversified Texts?
por: Li, Qingchuan, et al.
Publicado: (2025) -
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
por: Zhao, Zhixue, et al.
Publicado: (2024)