Gespeichert in:
| Hauptverfasser: | Toroghi, Armin, Kalarde, Faeze Moradi, Sanner, Scott |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.12918 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoLoTa: A Dataset for Entity-based Commonsense Reasoning over Long-Tail Knowledge
von: Toroghi, Armin, et al.
Veröffentlicht: (2025)
von: Toroghi, Armin, et al.
Veröffentlicht: (2025)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
von: Guo, Willis, et al.
Veröffentlicht: (2024)
von: Guo, Willis, et al.
Veröffentlicht: (2024)
Right for Right Reasons: Large Language Models for Verifiable Commonsense Knowledge Graph Question Answering
von: Toroghi, Armin, et al.
Veröffentlicht: (2024)
von: Toroghi, Armin, et al.
Veröffentlicht: (2024)
Goal-Oriented Reasoning for RAG-based Memory in Conversational Agentic LLM Systems
von: Liang, Jiazhou, et al.
Veröffentlicht: (2026)
von: Liang, Jiazhou, et al.
Veröffentlicht: (2026)
Bayesian Optimization with LLM-Based Acquisition Functions for Natural Language Preference Elicitation
von: Austin, David Eric, et al.
Veröffentlicht: (2024)
von: Austin, David Eric, et al.
Veröffentlicht: (2024)
Semantic XPath: Structured Agentic Memory Access for Conversational AI
von: Liu, Yifan Simon, et al.
Veröffentlicht: (2026)
von: Liu, Yifan Simon, et al.
Veröffentlicht: (2026)
Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models
von: Xiong, Kai, et al.
Veröffentlicht: (2025)
von: Xiong, Kai, et al.
Veröffentlicht: (2025)
Zero-Shot Commonsense Validation and Reasoning with Large Language Models: An Evaluation on SemEval-2020 Task 4 Dataset
von: Alfugaha, Rawand, et al.
Veröffentlicht: (2025)
von: Alfugaha, Rawand, et al.
Veröffentlicht: (2025)
ConstraintChecker: A Plugin for Large Language Models to Reason on Commonsense Knowledge Bases
von: Do, Quyet V., et al.
Veröffentlicht: (2024)
von: Do, Quyet V., et al.
Veröffentlicht: (2024)
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2023)
von: Wang, Yuqing, et al.
Veröffentlicht: (2023)
Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not
von: Karakaş, Sercan
Veröffentlicht: (2026)
von: Karakaş, Sercan
Veröffentlicht: (2026)
BrainBench: Exposing the Commonsense Reasoning Gap in Large Language Models
von: Tang, Yuzhe
Veröffentlicht: (2026)
von: Tang, Yuzhe
Veröffentlicht: (2026)
Commonsense Generation and Evaluation for Dialogue Systems using Large Language Models
von: Estecha-Garitagoitia, Marcos, et al.
Veröffentlicht: (2025)
von: Estecha-Garitagoitia, Marcos, et al.
Veröffentlicht: (2025)
Doing Good or Doing Right? Exploring the Weakness of Commonsense Causal Reasoning Models
von: Han, Mingyue, et al.
Veröffentlicht: (2021)
von: Han, Mingyue, et al.
Veröffentlicht: (2021)
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes
von: Diallo, Aissatou, et al.
Veröffentlicht: (2024)
von: Diallo, Aissatou, et al.
Veröffentlicht: (2024)
Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures
von: Chang, Tyler A., et al.
Veröffentlicht: (2025)
von: Chang, Tyler A., et al.
Veröffentlicht: (2025)
CANDLE: Iterative Conceptualization and Instantiation Distillation from Large Language Models for Commonsense Reasoning
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
ConKE: Conceptualization-Augmented Knowledge Editing in Large Language Models for Commonsense Reasoning
von: Zhang, Liyu, et al.
Veröffentlicht: (2024)
von: Zhang, Liyu, et al.
Veröffentlicht: (2024)
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
The Odyssey of Commonsense Causality: From Foundational Benchmarks to Cutting-Edge Reasoning
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
From Data to Commonsense Reasoning: The Use of Large Language Models for Explainable AI
von: Krause, Stefanie, et al.
Veröffentlicht: (2024)
von: Krause, Stefanie, et al.
Veröffentlicht: (2024)
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning
von: Tang, Zhisheng, et al.
Veröffentlicht: (2024)
von: Tang, Zhisheng, et al.
Veröffentlicht: (2024)
Modelling Commonsense Commonalities with Multi-Facet Concept Embeddings
von: Kteich, Hanane, et al.
Veröffentlicht: (2024)
von: Kteich, Hanane, et al.
Veröffentlicht: (2024)
Q-STRUM Debate: Query-Driven Contrastive Summarization for Recommendation Comparison
von: Saad, George-Kirollos, et al.
Veröffentlicht: (2025)
von: Saad, George-Kirollos, et al.
Veröffentlicht: (2025)
ReTraceQA: Evaluating Reasoning Traces of Small Language Models in Commonsense Question Answering
von: Molfese, Francesco Maria, et al.
Veröffentlicht: (2025)
von: Molfese, Francesco Maria, et al.
Veröffentlicht: (2025)
ViCor: Bridging Visual Understanding and Commonsense Reasoning with Large Language Models
von: Zhou, Kaiwen, et al.
Veröffentlicht: (2023)
von: Zhou, Kaiwen, et al.
Veröffentlicht: (2023)
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
Ko-PIQA: A Korean Physical Commonsense Reasoning Dataset with Cultural Context
von: Choi, Dasol, et al.
Veröffentlicht: (2025)
von: Choi, Dasol, et al.
Veröffentlicht: (2025)
mCSQA: Multilingual Commonsense Reasoning Dataset with Unified Creation Strategy by Language Models and Humans
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
Understanding the Capabilities and Limitations of Large Language Models for Cultural Commonsense
von: Shen, Siqi, et al.
Veröffentlicht: (2024)
von: Shen, Siqi, et al.
Veröffentlicht: (2024)
Commonsense Reasoning in Arab Culture
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2025)
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
HellaSwag-Pro: A Large-Scale Bilingual Benchmark for Evaluating the Robustness of LLMs in Commonsense Reasoning
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2025)
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2025)
Archer: A Human-Labeled Text-to-SQL Dataset with Arithmetic, Commonsense and Hypothetical Reasoning
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
Reasoning Up the Instruction Ladder for Controllable Language Models
von: Zheng, Zishuo, et al.
Veröffentlicht: (2025)
von: Zheng, Zishuo, et al.
Veröffentlicht: (2025)
BiasCause: Evaluate Socially Biased Causal Reasoning of Large Language Models
von: Xie, Tian, et al.
Veröffentlicht: (2025)
von: Xie, Tian, et al.
Veröffentlicht: (2025)
RESPONSE: Benchmarking the Ability of Language Models to Undertake Commonsense Reasoning in Crisis Situation
von: Diallo, Aissatou, et al.
Veröffentlicht: (2025)
von: Diallo, Aissatou, et al.
Veröffentlicht: (2025)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
Common to Whom? Regional Cultural Commonsense and LLM Bias in India
von: Madhusudan, Sangmitra, et al.
Veröffentlicht: (2026)
von: Madhusudan, Sangmitra, et al.
Veröffentlicht: (2026)
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models
von: Tu, Ruibo, et al.
Veröffentlicht: (2024)
von: Tu, Ruibo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CoLoTa: A Dataset for Entity-based Commonsense Reasoning over Long-Tail Knowledge
von: Toroghi, Armin, et al.
Veröffentlicht: (2025) -
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
von: Guo, Willis, et al.
Veröffentlicht: (2024) -
Right for Right Reasons: Large Language Models for Verifiable Commonsense Knowledge Graph Question Answering
von: Toroghi, Armin, et al.
Veröffentlicht: (2024) -
Goal-Oriented Reasoning for RAG-based Memory in Conversational Agentic LLM Systems
von: Liang, Jiazhou, et al.
Veröffentlicht: (2026) -
Bayesian Optimization with LLM-Based Acquisition Functions for Natural Language Preference Elicitation
von: Austin, David Eric, et al.
Veröffentlicht: (2024)