Executable Counterfactuals: Improving LLMs' Causal Reasoning Through Code
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vashishtha, Aniket, Dai, Qirun, Mei, Hongyuan, Sharma, Amit, Tan, Chenhao, Peng, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Realizing LLMs' Causal Potential Requires Science-Grounded, Novel Benchmarks
von: Srivastava, Ashutosh, et al.
Veröffentlicht: (2025)
von: Srivastava, Ashutosh, et al.
Veröffentlicht: (2025)
Teaching Transformers Causal Reasoning through Axiomatic Training
von: Vashishtha, Aniket, et al.
Veröffentlicht: (2024)
von: Vashishtha, Aniket, et al.
Veröffentlicht: (2024)
Improving Clinical Diagnosis with Counterfactual Multi-Agent Reasoning
von: You, Zhiwen, et al.
Veröffentlicht: (2026)
von: You, Zhiwen, et al.
Veröffentlicht: (2026)
Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
von: Kıcıman, Emre, et al.
Veröffentlicht: (2023)
von: Kıcıman, Emre, et al.
Veröffentlicht: (2023)
The Best Instruction-Tuning Data are Those That Fit
von: Zhang, Dylan, et al.
Veröffentlicht: (2025)
von: Zhang, Dylan, et al.
Veröffentlicht: (2025)
Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities
von: Dai, Qirun, et al.
Veröffentlicht: (2025)
von: Dai, Qirun, et al.
Veröffentlicht: (2025)
Better Think Thrice: Learning to Reason Causally with Double Counterfactual Consistency
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists
von: Yang, Junlin, et al.
Veröffentlicht: (2026)
von: Yang, Junlin, et al.
Veröffentlicht: (2026)
Generalization of RLVR Using Causal Reasoning as a Testbed
von: Lu, Brian, et al.
Veröffentlicht: (2025)
von: Lu, Brian, et al.
Veröffentlicht: (2025)
CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation
von: Khatibi, Elahe, et al.
Veröffentlicht: (2026)
von: Khatibi, Elahe, et al.
Veröffentlicht: (2026)
Barriers to Counterfactual Credit Attribution for Autoregressive Models
von: Cohen, Aloni, et al.
Veröffentlicht: (2026)
von: Cohen, Aloni, et al.
Veröffentlicht: (2026)
Hypothesis Generation with Large Language Models
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
Towards Generalizable Reasoning: Group Causal Counterfactual Policy Optimization for LLM Reasoning
von: Wang, Jingyao, et al.
Veröffentlicht: (2026)
von: Wang, Jingyao, et al.
Veröffentlicht: (2026)
Improving Generative Methods for Causal Evaluation via Simulation-Based Inference
von: Amaranath, Pracheta, et al.
Veröffentlicht: (2025)
von: Amaranath, Pracheta, et al.
Veröffentlicht: (2025)
CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions
von: Heddaya, Mourad, et al.
Veröffentlicht: (2024)
von: Heddaya, Mourad, et al.
Veröffentlicht: (2024)
Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds
von: Biswas, Prateek, et al.
Veröffentlicht: (2026)
von: Biswas, Prateek, et al.
Veröffentlicht: (2026)
Reasoning Through Execution: Unifying Process and Outcome Rewards for Code Generation
von: Yu, Zhuohao, et al.
Veröffentlicht: (2024)
von: Yu, Zhuohao, et al.
Veröffentlicht: (2024)
Can LLMs Reconcile Knowledge Conflicts in Counterfactual Reasoning
von: Yamin, Khurram, et al.
Veröffentlicht: (2025)
von: Yamin, Khurram, et al.
Veröffentlicht: (2025)
Learning Counterfactually Fair Models via Improved Generation with Neural Causal Models
von: Kher, Krishn Vishwas, et al.
Veröffentlicht: (2025)
von: Kher, Krishn Vishwas, et al.
Veröffentlicht: (2025)
Self-Execution Simulation Improves Coding Models
von: Maimon, Gallil, et al.
Veröffentlicht: (2026)
von: Maimon, Gallil, et al.
Veröffentlicht: (2026)
Decision Focused Causal Learning for Direct Counterfactual Marketing Optimization
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
Leveraging priors on distribution functions for multi-arm bandits
von: Vashishtha, Sumit, et al.
Veröffentlicht: (2025)
von: Vashishtha, Sumit, et al.
Veröffentlicht: (2025)
Federated Causal Representation Learning in State-Space Systems for Decentralized Counterfactual Reasoning
von: Mohamed, Nazal, et al.
Veröffentlicht: (2026)
von: Mohamed, Nazal, et al.
Veröffentlicht: (2026)
DiffusionCounterfactuals: Inferring High-dimensional Counterfactuals with Guidance of Causal Representations
von: Zhu, Jiageng, et al.
Veröffentlicht: (2024)
von: Zhu, Jiageng, et al.
Veröffentlicht: (2024)
$\nabla$-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space
von: Wang, Peihao, et al.
Veröffentlicht: (2026)
von: Wang, Peihao, et al.
Veröffentlicht: (2026)
Out-of-Distribution Adaptation in Offline RL: Counterfactual Reasoning via Causal Normalizing Flows
von: Cho, Minjae, et al.
Veröffentlicht: (2024)
von: Cho, Minjae, et al.
Veröffentlicht: (2024)
Where and How to Attack? A Causality-Inspired Recipe for Generating Counterfactual Adversarial Examples
von: Cai, Ruichu, et al.
Veröffentlicht: (2023)
von: Cai, Ruichu, et al.
Veröffentlicht: (2023)
CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution
von: Gu, Alex, et al.
Veröffentlicht: (2024)
von: Gu, Alex, et al.
Veröffentlicht: (2024)
Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs
von: Haque, Mirazul, et al.
Veröffentlicht: (2025)
von: Haque, Mirazul, et al.
Veröffentlicht: (2025)
CARE: Turning LLMs Into Causal Reasoning Expert
von: Dong, Juncheng, et al.
Veröffentlicht: (2025)
von: Dong, Juncheng, et al.
Veröffentlicht: (2025)
Failure Modes of LLMs for Causal Reasoning on Narratives
von: Yamin, Khurram, et al.
Veröffentlicht: (2024)
von: Yamin, Khurram, et al.
Veröffentlicht: (2024)
Causal and Counterfactual Views of Missing Data Models
von: Nabi, Razieh, et al.
Veröffentlicht: (2022)
von: Nabi, Razieh, et al.
Veröffentlicht: (2022)
Understanding Lookahead Dynamics Through Laplace Transform
von: Sanyal, Aniket, et al.
Veröffentlicht: (2025)
von: Sanyal, Aniket, et al.
Veröffentlicht: (2025)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
CausalBench: A Comprehensive Benchmark for Causal Learning Capability of LLMs
von: Zhou, Yu, et al.
Veröffentlicht: (2024)
von: Zhou, Yu, et al.
Veröffentlicht: (2024)
The Story is Not the Science: Execution-Grounded Evaluation of Mechanistic Interpretability Research
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2026)
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2026)
The Double Life of Code World Models: Provably Unmasking Malicious Behavior Through Execution Traces
von: Sahoo, Subramanyam
Veröffentlicht: (2025)
von: Sahoo, Subramanyam
Veröffentlicht: (2025)
Towards Verified Code Reasoning by LLMs
von: Sistla, Meghana, et al.
Veröffentlicht: (2025)
von: Sistla, Meghana, et al.
Veröffentlicht: (2025)
Causal Micro-Narratives
von: Heddaya, Mourad, et al.
Veröffentlicht: (2024)
von: Heddaya, Mourad, et al.
Veröffentlicht: (2024)
Certifying Counterfactual Bias in LLMs
von: Chaudhary, Isha, et al.
Veröffentlicht: (2024)
von: Chaudhary, Isha, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Realizing LLMs' Causal Potential Requires Science-Grounded, Novel Benchmarks
von: Srivastava, Ashutosh, et al.
Veröffentlicht: (2025) -
Teaching Transformers Causal Reasoning through Axiomatic Training
von: Vashishtha, Aniket, et al.
Veröffentlicht: (2024) -
Improving Clinical Diagnosis with Counterfactual Multi-Agent Reasoning
von: You, Zhiwen, et al.
Veröffentlicht: (2026) -
Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
von: Kıcıman, Emre, et al.
Veröffentlicht: (2023) -
The Best Instruction-Tuning Data are Those That Fit
von: Zhang, Dylan, et al.
Veröffentlicht: (2025)