Presupposition and Reasoning in Conditionals: A Theory-Based Study of Humans and LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Azin, Tara, Yu, Yongan, Singh, Raj, Jouravlev, Olessia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Let's CONFER: A Dataset for Evaluating Natural Language Inference Models on CONditional InFERence and Presupposition
von: Azin, Tara, et al.
Veröffentlicht: (2025)
von: Azin, Tara, et al.
Veröffentlicht: (2025)
Do Language Models Know Theo Has a Wife? Investigating the Proviso Problem
von: Azin, Tara, et al.
Veröffentlicht: (2026)
von: Azin, Tara, et al.
Veröffentlicht: (2026)
Evaluating Reasoning Models for Queries with Presuppositions
von: Sathyanathan, Rose, et al.
Veröffentlicht: (2026)
von: Sathyanathan, Rose, et al.
Veröffentlicht: (2026)
LLMs Struggle to Reject False Presuppositions when Misinformation Stakes are High
von: Sieker, Judith, et al.
Veröffentlicht: (2025)
von: Sieker, Judith, et al.
Veröffentlicht: (2025)
The Presupposition Problem in Representation Genesis
von: Wu, Yiling
Veröffentlicht: (2026)
von: Wu, Yiling
Veröffentlicht: (2026)
Safer in Translation? Presupposition Robustness in Indic Languages
von: Palnitkar, Aadi, et al.
Veröffentlicht: (2025)
von: Palnitkar, Aadi, et al.
Veröffentlicht: (2025)
Persian Abstract Meaning Representation: Annotation Guidelines and Gold Standard Dataset
von: Takhshid, Reza, et al.
Veröffentlicht: (2022)
von: Takhshid, Reza, et al.
Veröffentlicht: (2022)
Cancer-Myth: Evaluating Large Language Models on Patient Questions with False Presuppositions
von: Zhu, Wang Bill, et al.
Veröffentlicht: (2025)
von: Zhu, Wang Bill, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for Health-related Queries with Presuppositions
von: Kaur, Navreet, et al.
Veröffentlicht: (2023)
von: Kaur, Navreet, et al.
Veröffentlicht: (2023)
If We May De-Presuppose: Robustly Verifying Claims through Presupposition-Free Question Decomposition
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
RiddleBench: A New Generative Reasoning Benchmark for LLMs
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
Do LLMs Exhibit Human-Like Reasoning? Evaluating Theory of Mind in LLMs for Open-Ended Responses
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
An Empirical Study of In-context Learning in LLMs for Machine Translation
von: Chitale, Pranjal A., et al.
Veröffentlicht: (2024)
von: Chitale, Pranjal A., et al.
Veröffentlicht: (2024)
LLMs as Span Annotators: A Comparative Study of LLMs and Humans
von: Kasner, Zdeněk, et al.
Veröffentlicht: (2025)
von: Kasner, Zdeněk, et al.
Veröffentlicht: (2025)
Reasoning or a Semblance of it? A Diagnostic Study of Transitive Reasoning in LLMs
von: Mehrafarin, Houman, et al.
Veröffentlicht: (2024)
von: Mehrafarin, Houman, et al.
Veröffentlicht: (2024)
LLMs and the Human Condition
von: Wallis, Peter
Veröffentlicht: (2024)
von: Wallis, Peter
Veröffentlicht: (2024)
Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
EnigmaToM: Improve LLMs' Theory-of-Mind Reasoning Capabilities with Neural Knowledge Base of Entity States
von: Xu, Hainiu, et al.
Veröffentlicht: (2025)
von: Xu, Hainiu, et al.
Veröffentlicht: (2025)
Can LLMs Capture Human Preferences?
von: Goli, Ali, et al.
Veröffentlicht: (2023)
von: Goli, Ali, et al.
Veröffentlicht: (2023)
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP
von: Hasanaath, Ahmed, et al.
Veröffentlicht: (2025)
von: Hasanaath, Ahmed, et al.
Veröffentlicht: (2025)
HCR-Reasoner: Synergizing Large Language Models and Theory for Human-like Causal Reasoning
von: Zhang, Yanxi, et al.
Veröffentlicht: (2025)
von: Zhang, Yanxi, et al.
Veröffentlicht: (2025)
Humans or LLMs as the Judge? A Study on Judgement Biases
von: Chen, Guiming Hardy, et al.
Veröffentlicht: (2024)
von: Chen, Guiming Hardy, et al.
Veröffentlicht: (2024)
HugAgent: Benchmarking LLMs for Simulation of Individualized Human Reasoning
von: Li, Chance Jiajie, et al.
Veröffentlicht: (2025)
von: Li, Chance Jiajie, et al.
Veröffentlicht: (2025)
What's Not Said Still Hurts: A Description-Based Evaluation Framework for Measuring Social Bias in LLMs
von: Pan, Jinhao, et al.
Veröffentlicht: (2025)
von: Pan, Jinhao, et al.
Veröffentlicht: (2025)
Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs
von: Yu, Xingyang, et al.
Veröffentlicht: (2026)
von: Yu, Xingyang, et al.
Veröffentlicht: (2026)
THiNK: Can Large Language Models Think-aloud?
von: Yu, Yongan, et al.
Veröffentlicht: (2025)
von: Yu, Yongan, et al.
Veröffentlicht: (2025)
Graph-Based Alternatives to LLMs for Human Simulation
von: Suh, Joseph, et al.
Veröffentlicht: (2025)
von: Suh, Joseph, et al.
Veröffentlicht: (2025)
Exploring the Potential of Offline RL for Reasoning in LLMs: A Preliminary Study
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
Can Vision Language Models Be Adaptive in Mathematics Education? A Learner Model-based Rubric Study
von: Gao, Jie, et al.
Veröffentlicht: (2026)
von: Gao, Jie, et al.
Veröffentlicht: (2026)
Training and Evaluation of Guideline-Based Medical Reasoning in LLMs
von: Staniek, Michael, et al.
Veröffentlicht: (2025)
von: Staniek, Michael, et al.
Veröffentlicht: (2025)
Position: On the Methodological Pitfalls of Evaluating Base LLMs for Reasoning
von: Chan, Jason, et al.
Veröffentlicht: (2025)
von: Chan, Jason, et al.
Veröffentlicht: (2025)
RULEBREAKERS: Challenging LLMs at the Crossroads between Formal Logic and Human-like Reasoning
von: Chan, Jason, et al.
Veröffentlicht: (2024)
von: Chan, Jason, et al.
Veröffentlicht: (2024)
Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
Mapping the Minds of LLMs: A Graph-Based Analysis of Reasoning LLM
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
SVeritas: Benchmark for Robust Speaker Verification under Diverse Conditions
von: Baali, Massa, et al.
Veröffentlicht: (2025)
von: Baali, Massa, et al.
Veröffentlicht: (2025)
On Code-Induced Reasoning in LLMs
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
Reasoning about concepts with LLMs: Inconsistencies abound
von: Uceda-Sosa, Rosario, et al.
Veröffentlicht: (2024)
von: Uceda-Sosa, Rosario, et al.
Veröffentlicht: (2024)
Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning
von: Zhu, Jiayuan, et al.
Veröffentlicht: (2025)
von: Zhu, Jiayuan, et al.
Veröffentlicht: (2025)
Hero Projects: The Russian Empire and Big Technology from Lenin to Putin by Paul R.Josephson. Oxford: Oxford University Press, 2024. 344 pp. $45.95. ISBN 978‐0‐19‐769839‐6
von: Olessia Kirtchik
Veröffentlicht: (2024)
von: Olessia Kirtchik
Veröffentlicht: (2024)
Rethinking Machine Ethics -- Can LLMs Perform Moral Reasoning through the Lens of Moral Theories?
von: Zhou, Jingyan, et al.
Veröffentlicht: (2023)
von: Zhou, Jingyan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Let's CONFER: A Dataset for Evaluating Natural Language Inference Models on CONditional InFERence and Presupposition
von: Azin, Tara, et al.
Veröffentlicht: (2025) -
Do Language Models Know Theo Has a Wife? Investigating the Proviso Problem
von: Azin, Tara, et al.
Veröffentlicht: (2026) -
Evaluating Reasoning Models for Queries with Presuppositions
von: Sathyanathan, Rose, et al.
Veröffentlicht: (2026) -
LLMs Struggle to Reject False Presuppositions when Misinformation Stakes are High
von: Sieker, Judith, et al.
Veröffentlicht: (2025) -
The Presupposition Problem in Representation Genesis
von: Wu, Yiling
Veröffentlicht: (2026)