Towards a Benchmark for Causal Business Process Reasoning with LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Fournier, Fabiana, Limonad, Lior, Skarbovsky, Inna |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The WHY in Business Processes: Unification of Causal Process Models
by: David, Yuval, et al.
Published: (2025)
by: David, Yuval, et al.
Published: (2025)
The WHY in Business Processes: Discovery of Causal Execution Dependencies
by: Fournier, Fabiana, et al.
Published: (2023)
by: Fournier, Fabiana, et al.
Published: (2023)
How well can a large language model explain business processes as perceived by users?
by: Fahland, Dirk, et al.
Published: (2024)
by: Fahland, Dirk, et al.
Published: (2024)
Agentic AI Process Observability: Discovering Behavioral Variability
by: Fournier, Fabiana, et al.
Published: (2025)
by: Fournier, Fabiana, et al.
Published: (2025)
Monetizing Currency Pair Sentiments through LLM Explainability
by: Limonad, Lior, et al.
Published: (2024)
by: Limonad, Lior, et al.
Published: (2024)
XABPs: Towards eXplainable Autonomous Business Processes
by: Fettke, Peter, et al.
Published: (2025)
by: Fettke, Peter, et al.
Published: (2025)
Selecting the Right LLM for eGov Explanations
by: Limonad, Lior, et al.
Published: (2025)
by: Limonad, Lior, et al.
Published: (2025)
Agent Mentor: Framing Agent Knowledge through Semantic Trajectory Analysis
by: Ben-Gigi, Roi, et al.
Published: (2026)
by: Ben-Gigi, Roi, et al.
Published: (2026)
Agentic Business Process Management: A Research Manifesto
by: Calvanese, Diego, et al.
Published: (2026)
by: Calvanese, Diego, et al.
Published: (2026)
Beyond Black-Box Benchmarking: Observability, Analytics, and Optimization of Agentic Systems
by: Moshkovich, Dany, et al.
Published: (2025)
by: Moshkovich, Dany, et al.
Published: (2025)
Expecting the Unexpected: Developing Autonomous-System Design Principles for Reacting to Unpredicted Events and Conditions
by: Marron, Assaf, et al.
Published: (2020)
by: Marron, Assaf, et al.
Published: (2020)
Towards a Benchmark for Large Language Models for Business Process Management Tasks
by: Busch, Kiran, et al.
Published: (2024)
by: Busch, Kiran, et al.
Published: (2024)
Unmasking Reasoning Processes: A Process-aware Benchmark for Evaluating Structural Mathematical Reasoning in LLMs
by: Zheng, Xiang, et al.
Published: (2026)
by: Zheng, Xiang, et al.
Published: (2026)
Are LLMs Capable of Data-based Statistical and Causal Reasoning? Benchmarking Advanced Quantitative Reasoning with Data
by: Liu, Xiao, et al.
Published: (2024)
by: Liu, Xiao, et al.
Published: (2024)
CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation
by: Sawarni, Ayush, et al.
Published: (2026)
by: Sawarni, Ayush, et al.
Published: (2026)
Reasoning about Affordances: Causal and Compositional Reasoning in LLMs
by: Gjerde, Magnus F., et al.
Published: (2025)
by: Gjerde, Magnus F., et al.
Published: (2025)
InterveneBench: Benchmarking LLMs for Intervention Reasoning and Causal Study Design in Real Social Systems
by: Shi, Shaojie, et al.
Published: (2026)
by: Shi, Shaojie, et al.
Published: (2026)
Towards CausalGPT: A Multi-Agent Approach for Faithful Knowledge Reasoning via Promoting Causal Consistency in LLMs
by: Tang, Ziyi, et al.
Published: (2023)
by: Tang, Ziyi, et al.
Published: (2023)
CARE: Turning LLMs Into Causal Reasoning Expert
by: Dong, Juncheng, et al.
Published: (2025)
by: Dong, Juncheng, et al.
Published: (2025)
ACCESS : A Benchmark for Abstract Causal Event Discovery and Reasoning
by: Vo, Vy, et al.
Published: (2025)
by: Vo, Vy, et al.
Published: (2025)
CausalAbstain: Enhancing Multilingual LLMs with Causal Reasoning for Trustworthy Abstention
by: Sun, Yuxi, et al.
Published: (2025)
by: Sun, Yuxi, et al.
Published: (2025)
CausalEval: Towards Better Causal Reasoning in Language Models
by: Yu, Longxuan, et al.
Published: (2024)
by: Yu, Longxuan, et al.
Published: (2024)
Causal Temporal Reasoning for Markov Decision Processes
by: Kazemi, Milad, et al.
Published: (2022)
by: Kazemi, Milad, et al.
Published: (2022)
oMeBench: Towards Robust Benchmarking of LLMs in Organic Mechanism Elucidation and Reasoning
by: Xu, Ruiling, et al.
Published: (2025)
by: Xu, Ruiling, et al.
Published: (2025)
NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise
by: Xu, Zhi, et al.
Published: (2026)
by: Xu, Zhi, et al.
Published: (2026)
Overclocking LLM Reasoning: Monitoring and Controlling Thinking Path Lengths in LLMs
by: Eisenstadt, Roy, et al.
Published: (2025)
by: Eisenstadt, Roy, et al.
Published: (2025)
Correlation or Causation: Analyzing the Causal Structures of LLM and LRM Reasoning Process
by: FU, Zhizhang, et al.
Published: (2025)
by: FU, Zhizhang, et al.
Published: (2025)
USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents
by: Lai, Siqi, et al.
Published: (2025)
by: Lai, Siqi, et al.
Published: (2025)
FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning
by: Wang, Zeyu, et al.
Published: (2026)
by: Wang, Zeyu, et al.
Published: (2026)
CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos
by: Li, Xuchen, et al.
Published: (2025)
by: Li, Xuchen, et al.
Published: (2025)
GeoGramBench: Benchmarking the Geometric Program Reasoning in Modern LLMs
by: Luo, Shixian, et al.
Published: (2025)
by: Luo, Shixian, et al.
Published: (2025)
MSQA: Benchmarking LLMs on Graduate-Level Materials Science Reasoning and Knowledge
by: Cheung, Jerry Junyang, et al.
Published: (2025)
by: Cheung, Jerry Junyang, et al.
Published: (2025)
CARV: A Diagnostic Benchmark for Compositional Analogical Reasoning in Multimodal LLMs
by: Du, Yongkang, et al.
Published: (2026)
by: Du, Yongkang, et al.
Published: (2026)
Benchmarking LLMs for Pairwise Causal Discovery in Biomedical and Multi-Domain Contexts
by: Anuyah, Sydney, et al.
Published: (2026)
by: Anuyah, Sydney, et al.
Published: (2026)
Reasoning in a Combinatorial and Constrained World: Benchmarking LLMs on Natural-Language Combinatorial Optimization
by: Jiang, Xia, et al.
Published: (2026)
by: Jiang, Xia, et al.
Published: (2026)
CausalVLBench: Benchmarking Visual Causal Reasoning in Large Vision-Language Models
by: Komanduri, Aneesh, et al.
Published: (2025)
by: Komanduri, Aneesh, et al.
Published: (2025)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Can Post-Training Transform LLMs into Causal Reasoners?
by: Chen, Junqi, et al.
Published: (2026)
by: Chen, Junqi, et al.
Published: (2026)
TopoBench: Benchmarking LLMs on Hard Topological Reasoning
by: Maniparambil, Mayug, et al.
Published: (2026)
by: Maniparambil, Mayug, et al.
Published: (2026)
Can LLMs Leverage Observational Data? Towards Data-Driven Causal Discovery with LLMs
by: Susanti, Yuni, et al.
Published: (2025)
by: Susanti, Yuni, et al.
Published: (2025)
Similar Items
-
The WHY in Business Processes: Unification of Causal Process Models
by: David, Yuval, et al.
Published: (2025) -
The WHY in Business Processes: Discovery of Causal Execution Dependencies
by: Fournier, Fabiana, et al.
Published: (2023) -
How well can a large language model explain business processes as perceived by users?
by: Fahland, Dirk, et al.
Published: (2024) -
Agentic AI Process Observability: Discovering Behavioral Variability
by: Fournier, Fabiana, et al.
Published: (2025) -
Monetizing Currency Pair Sentiments through LLM Explainability
by: Limonad, Lior, et al.
Published: (2024)