OCDB: Revisiting Causal Discovery with a Comprehensive Benchmark and Evaluation Framework
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Wei, Huang, Hong, Zhang, Guowen, Shi, Ruize, Yin, Kehan, Lin, Yuanyuan, Liu, Bang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scalable Heterogeneous Graph Learning via Heterogeneous-aware Orthogonal Prototype Experts
von: Zhou, Wei, et al.
Veröffentlicht: (2026)
von: Zhou, Wei, et al.
Veröffentlicht: (2026)
From General to Specific: Tailoring Large Language Models for Personalized Healthcare
von: Shi, Ruize, et al.
Veröffentlicht: (2024)
von: Shi, Ruize, et al.
Veröffentlicht: (2024)
Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values
von: Dong, Haonan, et al.
Veröffentlicht: (2026)
von: Dong, Haonan, et al.
Veröffentlicht: (2026)
Comprehensive Review and Empirical Evaluation of Causal Discovery Algorithms for Numerical Data
von: Niu, Wenjin, et al.
Veröffentlicht: (2024)
von: Niu, Wenjin, et al.
Veröffentlicht: (2024)
Revisiting Few-Shot Learning from a Causal Perspective
von: Lin, Guoliang, et al.
Veröffentlicht: (2022)
von: Lin, Guoliang, et al.
Veröffentlicht: (2022)
Causal Discovery as Dialectical Aggregation: A Quantitative Argumentation Framework
von: Wei, Sheng, et al.
Veröffentlicht: (2026)
von: Wei, Sheng, et al.
Veröffentlicht: (2026)
JAILJUDGE: A Comprehensive Jailbreak Judge Benchmark with Multi-Agent Enhanced Explanation Evaluation Framework
von: Liu, Fan, et al.
Veröffentlicht: (2024)
von: Liu, Fan, et al.
Veröffentlicht: (2024)
Causally-Enhanced Reinforcement Policy Optimization
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
Hybrid Local Causal Discovery
von: Ling, Zhaolong, et al.
Veröffentlicht: (2024)
von: Ling, Zhaolong, et al.
Veröffentlicht: (2024)
DMCD: Semantic-Statistical Framework for Causal Discovery
von: KaPatel, Samarth, et al.
Veröffentlicht: (2026)
von: KaPatel, Samarth, et al.
Veröffentlicht: (2026)
Is Your VLM for Autonomous Driving Safety-Ready? A Comprehensive Benchmark for Evaluating External and In-Cabin Risks
von: Meng, Xianhui, et al.
Veröffentlicht: (2025)
von: Meng, Xianhui, et al.
Veröffentlicht: (2025)
ACCESS : A Benchmark for Abstract Causal Event Discovery and Reasoning
von: Vo, Vy, et al.
Veröffentlicht: (2025)
von: Vo, Vy, et al.
Veröffentlicht: (2025)
Evaluating Progress in Graph Foundation Models: A Comprehensive Benchmark and New Insights
von: Yu, Xingtong, et al.
Veröffentlicht: (2026)
von: Yu, Xingtong, et al.
Veröffentlicht: (2026)
SkillFlow:Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents
von: Zhang, Ziao, et al.
Veröffentlicht: (2026)
von: Zhang, Ziao, et al.
Veröffentlicht: (2026)
Revolutionizing Database Q&A with Large Language Models: Comprehensive Benchmark and Evaluation
von: Zheng, Yihang, et al.
Veröffentlicht: (2024)
von: Zheng, Yihang, et al.
Veröffentlicht: (2024)
AMSbench: A Comprehensive Benchmark for Evaluating MLLM Capabilities in AMS Circuits
von: Shi, Yichen, et al.
Veröffentlicht: (2025)
von: Shi, Yichen, et al.
Veröffentlicht: (2025)
TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents
von: Chen, Weiyi, et al.
Veröffentlicht: (2026)
von: Chen, Weiyi, et al.
Veröffentlicht: (2026)
CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation
von: Sawarni, Ayush, et al.
Veröffentlicht: (2026)
von: Sawarni, Ayush, et al.
Veröffentlicht: (2026)
WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics
von: Liu, Chenxu, et al.
Veröffentlicht: (2026)
von: Liu, Chenxu, et al.
Veröffentlicht: (2026)
Challenges and Considerations in the Evaluation of Bayesian Causal Discovery
von: Mamaghan, Amir Mohammad Karimi, et al.
Veröffentlicht: (2024)
von: Mamaghan, Amir Mohammad Karimi, et al.
Veröffentlicht: (2024)
SPA-Bench: A Comprehensive Benchmark for SmartPhone Agent Evaluation
von: Chen, Jingxuan, et al.
Veröffentlicht: (2024)
von: Chen, Jingxuan, et al.
Veröffentlicht: (2024)
Temporal Latent Variable Structural Causal Model for Causal Discovery under External Interferences
von: Cai, Ruichu, et al.
Veröffentlicht: (2025)
von: Cai, Ruichu, et al.
Veröffentlicht: (2025)
Federated Causal Discovery from Heterogeneous Data
von: Li, Loka, et al.
Veröffentlicht: (2024)
von: Li, Loka, et al.
Veröffentlicht: (2024)
SoK: a Comprehensive Causality Analysis Framework for Large Language Model Security
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
Dependency-based Anomaly Detection: a General Framework and Comprehensive Evaluation
von: Lu, Sha, et al.
Veröffentlicht: (2020)
von: Lu, Sha, et al.
Veröffentlicht: (2020)
Neural Information Causality
von: Bang, Jeongho, et al.
Veröffentlicht: (2026)
von: Bang, Jeongho, et al.
Veröffentlicht: (2026)
Argumentative Causal Discovery
von: Russo, Fabrizio, et al.
Veröffentlicht: (2024)
von: Russo, Fabrizio, et al.
Veröffentlicht: (2024)
UltraEval-Audio: A Unified Framework for Comprehensive Evaluation of Audio Foundation Models
von: Shi, Qundong, et al.
Veröffentlicht: (2026)
von: Shi, Qundong, et al.
Veröffentlicht: (2026)
ELABORATION: A Comprehensive Benchmark on Human-LLM Competitive Programming
von: Yang, Xinwei, et al.
Veröffentlicht: (2025)
von: Yang, Xinwei, et al.
Veröffentlicht: (2025)
Differentiable Constraint-Based Causal Discovery
von: Zhou, Jincheng, et al.
Veröffentlicht: (2025)
von: Zhou, Jincheng, et al.
Veröffentlicht: (2025)
What Would Happen Next? Predicting Consequences from An Event Causality Graph
von: Zhan, Chuanhong, et al.
Veröffentlicht: (2024)
von: Zhan, Chuanhong, et al.
Veröffentlicht: (2024)
MFE-ETP: A Comprehensive Evaluation Benchmark for Multi-modal Foundation Models on Embodied Task Planning
von: Zhang, Min, et al.
Veröffentlicht: (2024)
von: Zhang, Min, et al.
Veröffentlicht: (2024)
InsightVision: A Comprehensive, Multi-Level Chinese-based Benchmark for Evaluating Implicit Visual Semantics in Large Vision Language Models
von: Yin, Xiaofei, et al.
Veröffentlicht: (2025)
von: Yin, Xiaofei, et al.
Veröffentlicht: (2025)
Revisiting, Benchmarking and Understanding Unsupervised Graph Domain Adaptation
von: Liu, Meihan, et al.
Veröffentlicht: (2024)
von: Liu, Meihan, et al.
Veröffentlicht: (2024)
MMCircuitEval: A Comprehensive Multimodal Circuit-Focused Benchmark for Evaluating LLMs
von: Zhao, Chenchen, et al.
Veröffentlicht: (2025)
von: Zhao, Chenchen, et al.
Veröffentlicht: (2025)
Benchmarking LLMs for Pairwise Causal Discovery in Biomedical and Multi-Domain Contexts
von: Anuyah, Sydney, et al.
Veröffentlicht: (2026)
von: Anuyah, Sydney, et al.
Veröffentlicht: (2026)
Navigating the Dual Facets: A Comprehensive Evaluation of Sequential Memory Editing in Large Language Models
von: Lin, Zihao, et al.
Veröffentlicht: (2024)
von: Lin, Zihao, et al.
Veröffentlicht: (2024)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework
von: Tu, Ruibo, et al.
Veröffentlicht: (2024)
von: Tu, Ruibo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Scalable Heterogeneous Graph Learning via Heterogeneous-aware Orthogonal Prototype Experts
von: Zhou, Wei, et al.
Veröffentlicht: (2026) -
From General to Specific: Tailoring Large Language Models for Personalized Healthcare
von: Shi, Ruize, et al.
Veröffentlicht: (2024) -
Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values
von: Dong, Haonan, et al.
Veröffentlicht: (2026) -
Comprehensive Review and Empirical Evaluation of Causal Discovery Algorithms for Numerical Data
von: Niu, Wenjin, et al.
Veröffentlicht: (2024) -
Revisiting Few-Shot Learning from a Causal Perspective
von: Lin, Guoliang, et al.
Veröffentlicht: (2022)