Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiong, Kai, Ding, Xiao, Cao, Yixin, Yan, Yuxiong, Du, Li, Zhang, Yufei, Gao, Jinglong, Liu, Jiaqian, Qin, Bing, Liu, Ting |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Complex Causality Extraction via Improved Subtask Interaction and Knowledge Fusion
von: Gao, Jinglong, et al.
Veröffentlicht: (2024)
von: Gao, Jinglong, et al.
Veröffentlicht: (2024)
Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
von: Xiong, Kai, et al.
Veröffentlicht: (2023)
von: Xiong, Kai, et al.
Veröffentlicht: (2023)
Meaningful Learning: Enhancing Abstract Reasoning in Large Language Models via Generic Fact Guidance
von: Xiong, Kai, et al.
Veröffentlicht: (2024)
von: Xiong, Kai, et al.
Veröffentlicht: (2024)
Towards Generalizable and Faithful Logic Reasoning over Natural Language via Resolution Refutation
von: Sun, Zhouhao, et al.
Veröffentlicht: (2024)
von: Sun, Zhouhao, et al.
Veröffentlicht: (2024)
CrossICL: Cross-Task In-Context Learning via Unsupervised Demonstration Transfer
von: Gao, Jinglong, et al.
Veröffentlicht: (2025)
von: Gao, Jinglong, et al.
Veröffentlicht: (2025)
Diagnosing and Remedying Knowledge Deficiencies in LLMs via Label-free Curricular Meaningful Learning
von: Xiong, Kai, et al.
Veröffentlicht: (2024)
von: Xiong, Kai, et al.
Veröffentlicht: (2024)
Causal-Guided Active Learning for Debiasing Large Language Models
von: Du, Li, et al.
Veröffentlicht: (2024)
von: Du, Li, et al.
Veröffentlicht: (2024)
Causal Debiasing for Visual Commonsense Reasoning
von: Zou, Jiayi, et al.
Veröffentlicht: (2025)
von: Zou, Jiayi, et al.
Veröffentlicht: (2025)
ReCo: Reliable Causal Chain Reasoning via Structural Causal Recurrent Neural Networks
von: Xiong, Kai, et al.
Veröffentlicht: (2022)
von: Xiong, Kai, et al.
Veröffentlicht: (2022)
Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models
von: Sun, Zhouhao, et al.
Veröffentlicht: (2025)
von: Sun, Zhouhao, et al.
Veröffentlicht: (2025)
SCoRE: Benchmarking Long-Chain Reasoning in Commonsense Scenarios
von: Zhan, Weidong, et al.
Veröffentlicht: (2025)
von: Zhan, Weidong, et al.
Veröffentlicht: (2025)
ExpeTrans: LLMs Are Experiential Transfer Learners
von: Gao, Jinglong, et al.
Veröffentlicht: (2025)
von: Gao, Jinglong, et al.
Veröffentlicht: (2025)
Self-Route: Automatic Mode Switching via Capability Estimation for Efficient Reasoning
von: He, Yang, et al.
Veröffentlicht: (2025)
von: He, Yang, et al.
Veröffentlicht: (2025)
DeepTool: Scaling Interleaved Deliberation in Tool-Integrated Reasoning via Process-Supervised Reinforcement Learning
von: He, Yang, et al.
Veröffentlicht: (2026)
von: He, Yang, et al.
Veröffentlicht: (2026)
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner
von: Gao, Jinglong, et al.
Veröffentlicht: (2024)
von: Gao, Jinglong, et al.
Veröffentlicht: (2024)
Supervised Fine-Tuning Achieve Rapid Task Adaption Via Alternating Attention Head Activation Patterns
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
Consolidation or Adaptation? PRISM: Disentangling SFT and RL Data via Gradient Concentration
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
GraCoRe: Benchmarking Graph Comprehension and Complex Reasoning in Large Language Models
von: Yuan, Zike, et al.
Veröffentlicht: (2024)
von: Yuan, Zike, et al.
Veröffentlicht: (2024)
Deciphering the Impact of Pretraining Data on Large Language Models through Machine Unlearning
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
The Odyssey of Commonsense Causality: From Foundational Benchmarks to Cutting-Edge Reasoning
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
GR-Ben: A General Reasoning Benchmark for Evaluating Process Reward Models
von: Sun, Zhouhao, et al.
Veröffentlicht: (2026)
von: Sun, Zhouhao, et al.
Veröffentlicht: (2026)
Doing Good or Doing Right? Exploring the Weakness of Commonsense Causal Reasoning Models
von: Han, Mingyue, et al.
Veröffentlicht: (2021)
von: Han, Mingyue, et al.
Veröffentlicht: (2021)
Beyond Similarity: A Gradient-based Graph Method for Instruction Tuning Data Selection
von: Zhao, Yang, et al.
Veröffentlicht: (2025)
von: Zhao, Yang, et al.
Veröffentlicht: (2025)
XFinBench: Benchmarking LLMs in Complex Financial Problem Solving and Reasoning
von: Zhang, Zhihan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihan, et al.
Veröffentlicht: (2025)
Good Learners Think Their Thinking: Generative PRM Makes Large Reasoning Model More Efficient Math Learner
von: He, Tao, et al.
Veröffentlicht: (2025)
von: He, Tao, et al.
Veröffentlicht: (2025)
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement
von: Zhao, Weixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Weixiang, et al.
Veröffentlicht: (2025)
Experimental Design For Causal Inference Through An Optimization Lens
von: Zhao, Jinglong
Veröffentlicht: (2024)
von: Zhao, Jinglong
Veröffentlicht: (2024)
CauTraj: A Causal-Knowledge-Guided Framework for Lane-Changing Trajectory Planning of Autonomous Vehicles
von: Lei, Cailin, et al.
Veröffentlicht: (2025)
von: Lei, Cailin, et al.
Veröffentlicht: (2025)
Text Difficulty Study: Do machines behave the same as humans regarding text difficulty?
von: Chen, Bowen, et al.
Veröffentlicht: (2022)
von: Chen, Bowen, et al.
Veröffentlicht: (2022)
LogicPro: Improving Complex Logical Reasoning via Program-Guided Learning
von: Jiang, Jin, et al.
Veröffentlicht: (2024)
von: Jiang, Jin, et al.
Veröffentlicht: (2024)
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
HellaSwag-Pro: A Large-Scale Bilingual Benchmark for Evaluating the Robustness of LLMs in Commonsense Reasoning
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2025)
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2025)
LOGICAL-COMMONSENSEQA: A Benchmark for Logical Commonsense Reasoning
von: Junias, Obed, et al.
Veröffentlicht: (2026)
von: Junias, Obed, et al.
Veröffentlicht: (2026)
Benchmarking Chinese Commonsense Reasoning with a Multi-hop Reasoning Perspective
von: You, Wangjie, et al.
Veröffentlicht: (2025)
von: You, Wangjie, et al.
Veröffentlicht: (2025)
Are LLMs Capable of Data-based Statistical and Causal Reasoning? Benchmarking Advanced Quantitative Reasoning with Data
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
Long or short CoT? Investigating Instance-level Switch of Large Reasoning Models
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2025)
Meta-RTL: Reinforcement-Based Meta-Transfer Learning for Low-Resource Commonsense Reasoning
von: Fu, Yu, et al.
Veröffentlicht: (2024)
von: Fu, Yu, et al.
Veröffentlicht: (2024)
CommonWhy: A Dataset for Evaluating Entity-Based Causal Commonsense Reasoning in Large Language Models
von: Toroghi, Armin, et al.
Veröffentlicht: (2026)
von: Toroghi, Armin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Enhancing Complex Causality Extraction via Improved Subtask Interaction and Knowledge Fusion
von: Gao, Jinglong, et al.
Veröffentlicht: (2024) -
Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
von: Xiong, Kai, et al.
Veröffentlicht: (2023) -
Meaningful Learning: Enhancing Abstract Reasoning in Large Language Models via Generic Fact Guidance
von: Xiong, Kai, et al.
Veröffentlicht: (2024) -
Towards Generalizable and Faithful Logic Reasoning over Natural Language via Resolution Refutation
von: Sun, Zhouhao, et al.
Veröffentlicht: (2024) -
CrossICL: Cross-Task In-Context Learning via Unsupervised Demonstration Transfer
von: Gao, Jinglong, et al.
Veröffentlicht: (2025)