Doing Good or Doing Right? Exploring the Weakness of Commonsense Causal Reasoning Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Mingyue, Wang, Yinglin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not
von: Karakaş, Sercan
Veröffentlicht: (2026)
von: Karakaş, Sercan
Veröffentlicht: (2026)
Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models
von: Xiong, Kai, et al.
Veröffentlicht: (2025)
von: Xiong, Kai, et al.
Veröffentlicht: (2025)
Right for Right Reasons: Large Language Models for Verifiable Commonsense Knowledge Graph Question Answering
von: Toroghi, Armin, et al.
Veröffentlicht: (2024)
von: Toroghi, Armin, et al.
Veröffentlicht: (2024)
The Odyssey of Commonsense Causality: From Foundational Benchmarks to Cutting-Edge Reasoning
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
Thinking Out Loud: Do Reasoning Models Know When They're Right?
von: Zeng, Qingcheng, et al.
Veröffentlicht: (2025)
von: Zeng, Qingcheng, et al.
Veröffentlicht: (2025)
Outcome Rewards Do Not Guarantee Verifiable or Causally Important Reasoning
von: Yu, Qinan, et al.
Veröffentlicht: (2026)
von: Yu, Qinan, et al.
Veröffentlicht: (2026)
CommonWhy: A Dataset for Evaluating Entity-Based Causal Commonsense Reasoning in Large Language Models
von: Toroghi, Armin, et al.
Veröffentlicht: (2026)
von: Toroghi, Armin, et al.
Veröffentlicht: (2026)
Commonsense Reasoning in Arab Culture
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2025)
When Do Language Models Endorse Limitations on Human Rights Principles?
von: Samway, Keenan, et al.
Veröffentlicht: (2026)
von: Samway, Keenan, et al.
Veröffentlicht: (2026)
Synthetic Data Generation for Training Diversified Commonsense Reasoning Models
von: Zhang, Tianhui, et al.
Veröffentlicht: (2026)
von: Zhang, Tianhui, et al.
Veröffentlicht: (2026)
IndoCulture: Exploring Geographically-Influenced Cultural Commonsense Reasoning Across Eleven Indonesian Provinces
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
Do LLMs Have the Generalization Ability in Conducting Causal Inference?
von: Wang, Chen, et al.
Veröffentlicht: (2024)
von: Wang, Chen, et al.
Veröffentlicht: (2024)
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2023)
von: Wang, Yuqing, et al.
Veröffentlicht: (2023)
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
Leveraging Explicit Reasoning for Inference Integration in Commonsense-Augmented Dialogue Models
von: Finch, Sarah E., et al.
Veröffentlicht: (2024)
von: Finch, Sarah E., et al.
Veröffentlicht: (2024)
ConKE: Conceptualization-Augmented Knowledge Editing in Large Language Models for Commonsense Reasoning
von: Zhang, Liyu, et al.
Veröffentlicht: (2024)
von: Zhang, Liyu, et al.
Veröffentlicht: (2024)
Exploring Defeasibility in Causal Reasoning
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
Do Large Language Models have Shared Weaknesses in Medical Question Answering?
von: Bean, Andrew M., et al.
Veröffentlicht: (2023)
von: Bean, Andrew M., et al.
Veröffentlicht: (2023)
Good Agentic Friends Do Not Just Give Verbal Advice: They Can Update Your Weights
von: Bao, Wenrui, et al.
Veröffentlicht: (2026)
von: Bao, Wenrui, et al.
Veröffentlicht: (2026)
CANDLE: Iterative Conceptualization and Instantiation Distillation from Large Language Models for Commonsense Reasoning
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
When More Is Less: A Systematic Analysis of Spatial and Commonsense Information for Visual Spatial Reasoning
von: Akasaka, Muku, et al.
Veröffentlicht: (2026)
von: Akasaka, Muku, et al.
Veröffentlicht: (2026)
How Do Latent Reasoning Methods Perform Under Weak and Strong Supervision?
von: Cui, Yingqian, et al.
Veröffentlicht: (2026)
von: Cui, Yingqian, et al.
Veröffentlicht: (2026)
"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer?
von: Reichman, Benjamin, et al.
Veröffentlicht: (2025)
von: Reichman, Benjamin, et al.
Veröffentlicht: (2025)
Benchmarking Chinese Commonsense Reasoning with a Multi-hop Reasoning Perspective
von: You, Wangjie, et al.
Veröffentlicht: (2025)
von: You, Wangjie, et al.
Veröffentlicht: (2025)
An Interdisciplinary Review of Commonsense Reasoning and Intent Detection
von: Sakib, Md Nazmus
Veröffentlicht: (2025)
von: Sakib, Md Nazmus
Veröffentlicht: (2025)
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
Do Language Models Reason Across Languages?
von: Meng, Yan, et al.
Veröffentlicht: (2026)
von: Meng, Yan, et al.
Veröffentlicht: (2026)
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
von: Elozeiri, Kareem, et al.
Veröffentlicht: (2025)
von: Elozeiri, Kareem, et al.
Veröffentlicht: (2025)
MM-THEBench: Do Reasoning MLLMs Think Reasonably?
von: Huang, Zhidian, et al.
Veröffentlicht: (2026)
von: Huang, Zhidian, et al.
Veröffentlicht: (2026)
SCoRE: Benchmarking Long-Chain Reasoning in Commonsense Scenarios
von: Zhan, Weidong, et al.
Veröffentlicht: (2025)
von: Zhan, Weidong, et al.
Veröffentlicht: (2025)
RESPONSE: Benchmarking the Ability of Language Models to Undertake Commonsense Reasoning in Crisis Situation
von: Diallo, Aissatou, et al.
Veröffentlicht: (2025)
von: Diallo, Aissatou, et al.
Veröffentlicht: (2025)
Do as We Do, Not as You Think: the Conformity of Large Language Models
von: Weng, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Weng, Zhiyuan, et al.
Veröffentlicht: (2025)
Where Do Reasoning Models Refuse?
von: Yamaguchi, Kureha, et al.
Veröffentlicht: (2025)
von: Yamaguchi, Kureha, et al.
Veröffentlicht: (2025)
CKBP v2: Better Annotation and Reasoning for Commonsense Knowledge Base Population
von: Fang, Tianqing, et al.
Veröffentlicht: (2023)
von: Fang, Tianqing, et al.
Veröffentlicht: (2023)
Detecting Emotional Incongruity of Sarcasm by Commonsense Reasoning
von: Qiu, Ziqi, et al.
Veröffentlicht: (2024)
von: Qiu, Ziqi, et al.
Veröffentlicht: (2024)
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
Thai Winograd Schemas: A Benchmark for Thai Commonsense Reasoning
von: Artkaew, Phakphum
Veröffentlicht: (2024)
von: Artkaew, Phakphum
Veröffentlicht: (2024)
The Mystery of Compositional Generalization in Graph-based Generative Commonsense Reasoning
von: Fu, Xiyan, et al.
Veröffentlicht: (2024)
von: Fu, Xiyan, et al.
Veröffentlicht: (2024)
Do LLMs Signal When They're Right? Evidence from Neuron Agreement
von: Chen, Kang, et al.
Veröffentlicht: (2025)
von: Chen, Kang, et al.
Veröffentlicht: (2025)
Optimizing Language Model's Reasoning Abilities with Weak Supervision
von: Tong, Yongqi, et al.
Veröffentlicht: (2024)
von: Tong, Yongqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not
von: Karakaş, Sercan
Veröffentlicht: (2026) -
Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models
von: Xiong, Kai, et al.
Veröffentlicht: (2025) -
Right for Right Reasons: Large Language Models for Verifiable Commonsense Knowledge Graph Question Answering
von: Toroghi, Armin, et al.
Veröffentlicht: (2024) -
The Odyssey of Commonsense Causality: From Foundational Benchmarks to Cutting-Edge Reasoning
von: Cui, Shaobo, et al.
Veröffentlicht: (2024) -
Thinking Out Loud: Do Reasoning Models Know When They're Right?
von: Zeng, Qingcheng, et al.
Veröffentlicht: (2025)