DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Ahmed, Ahmed G. A. H, Sakar, C. Okan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
KramaBench: A Benchmark for AI Systems on Data-to-Insight Pipelines over Data Lakes
by: Lai, Eugenie, et al.
Published: (2025)
by: Lai, Eugenie, et al.
Published: (2025)
GraphAide: Advanced Graph-Assisted Query and Reasoning System
by: Purohit, Sumit, et al.
Published: (2024)
by: Purohit, Sumit, et al.
Published: (2024)
ELT-Bench-Verified: Benchmark Quality Issues Underestimate AI Agent Capabilities
by: Zanoli, Christopher, et al.
Published: (2026)
by: Zanoli, Christopher, et al.
Published: (2026)
ELT-Bench: An End-to-End Benchmark for Evaluating AI Agents on ELT Pipelines
by: Jin, Tengjun, et al.
Published: (2025)
by: Jin, Tengjun, et al.
Published: (2025)
ClinDet-Bench: Beyond Abstention, Evaluating Judgment Determinability of LLMs in Clinical Decision-Making
by: Watanabe, Yusuke, et al.
Published: (2026)
by: Watanabe, Yusuke, et al.
Published: (2026)
Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning
by: Li, Fangjun, et al.
Published: (2024)
by: Li, Fangjun, et al.
Published: (2024)
AmbiGraph-Eval: Can LLMs Effectively Handle Ambiguous Graph Queries?
by: Tian, Yuchen, et al.
Published: (2025)
by: Tian, Yuchen, et al.
Published: (2025)
Managing FAIR Knowledge Graphs as Polyglot Data End Points: A Benchmark based on the rdf2pg Framework and Plant Biology Data
by: Brandizi, Marco, et al.
Published: (2025)
by: Brandizi, Marco, et al.
Published: (2025)
PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?
by: Xu, Jingzhe, et al.
Published: (2026)
by: Xu, Jingzhe, et al.
Published: (2026)
RelBench: A Benchmark for Deep Learning on Relational Databases
by: Robinson, Joshua, et al.
Published: (2024)
by: Robinson, Joshua, et al.
Published: (2024)
OMNIA: Closing the Loop by Leveraging LLMs for Knowledge Graph Completion
by: Ieng, Frédéric, et al.
Published: (2026)
by: Ieng, Frédéric, et al.
Published: (2026)
Leveraging Knowledge Graphs and LLMs to Support and Monitor Legislative Systems
by: Colombo, Andrea
Published: (2024)
by: Colombo, Andrea
Published: (2024)
Adaptive Data Quality Scoring Operations Framework using Drift-Aware Mechanism for Industrial Applications
by: Bayram, Firas, et al.
Published: (2024)
by: Bayram, Firas, et al.
Published: (2024)
Knowledge Graph Construction for Stock Markets with LLM-Based Explainable Reasoning
by: Lee, Cheonsol, et al.
Published: (2025)
by: Lee, Cheonsol, et al.
Published: (2025)
A System and Benchmark for LLM-based Q&A on Heterogeneous Data
by: Fokoue, Achille, et al.
Published: (2024)
by: Fokoue, Achille, et al.
Published: (2024)
Conflict Detection for Temporal Knowledge Graphs:A Fast Constraint Mining Algorithm and New Benchmarks
by: Chen, Jianhao, et al.
Published: (2023)
by: Chen, Jianhao, et al.
Published: (2023)
Grid-Based Projection of Spatial Data into Knowledge Graphs
by: Anjomshoaa, Amin, et al.
Published: (2024)
by: Anjomshoaa, Amin, et al.
Published: (2024)
LLM-KG-Bench 3.0: A Compass for SemanticTechnology Capabilities in the Ocean of LLMs
by: Meyer, Lars-Peter, et al.
Published: (2025)
by: Meyer, Lars-Peter, et al.
Published: (2025)
Comprehending Semantic Types in JSON Data with Graph Neural Networks
by: Wei, Shuang, et al.
Published: (2023)
by: Wei, Shuang, et al.
Published: (2023)
TKG-Thinker: Towards Dynamic Reasoning over Temporal Knowledge Graphs via Agentic Reinforcement Learning
by: Jiang, Zihao, et al.
Published: (2026)
by: Jiang, Zihao, et al.
Published: (2026)
HCT-QA: A Benchmark for Question Answering on Human-Centric Tables
by: Ahmad, Mohammad S., et al.
Published: (2025)
by: Ahmad, Mohammad S., et al.
Published: (2025)
CANDY: A Benchmark for Continuous Approximate Nearest Neighbor Search with Dynamic Data Ingestion
by: Zeng, Xianzhi, et al.
Published: (2024)
by: Zeng, Xianzhi, et al.
Published: (2024)
CMDBench: A Benchmark for Coarse-to-fine Multimodal Data Discovery in Compound AI Systems
by: Feng, Yanlin, et al.
Published: (2024)
by: Feng, Yanlin, et al.
Published: (2024)
A Multi-Agent System for Semantic Mapping of Relational Data to Knowledge Graphs
by: Trajanoska, Milena, et al.
Published: (2025)
by: Trajanoska, Milena, et al.
Published: (2025)
AvalancheBench: Evaluating Enterprise Data Agents Through Latent World Recovery
by: Kleczek, Darek, et al.
Published: (2026)
by: Kleczek, Darek, et al.
Published: (2026)
EpiCastBench: Datasets and Benchmarks for Multivariate Epidemic Forecasting
by: Panja, Madhurima, et al.
Published: (2026)
by: Panja, Madhurima, et al.
Published: (2026)
CypherBench: Towards Precise Retrieval over Full-scale Modern Knowledge Graphs in the LLM Era
by: Feng, Yanlin, et al.
Published: (2024)
by: Feng, Yanlin, et al.
Published: (2024)
OpenGLT: A Comprehensive Benchmark of Graph Neural Networks for Graph-Level Tasks
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
The Phish, The Spam, and The Valid: Generating Feature-Rich Emails for Benchmarking LLMs
by: Toth, Rebeka, et al.
Published: (2025)
by: Toth, Rebeka, et al.
Published: (2025)
NormTab: Improving Symbolic Reasoning in LLMs Through Tabular Data Normalization
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
From Data Quality for AI to AI for Data Quality: A Systematic Review of Tools for AI-Augmented Data Quality Management in Data Warehouses
by: Tamm, Heidi Carolina, et al.
Published: (2024)
by: Tamm, Heidi Carolina, et al.
Published: (2024)
Using off-the-shelf LLMs to query enterprise data by progressively revealing ontologies
by: Civili, C., et al.
Published: (2024)
by: Civili, C., et al.
Published: (2024)
Mind the Data Gap: Bridging LLMs to Enterprise Data Integration
by: Kayali, Moe, et al.
Published: (2024)
by: Kayali, Moe, et al.
Published: (2024)
Global Benchmark Database
by: Iser, Ashlin, et al.
Published: (2024)
by: Iser, Ashlin, et al.
Published: (2024)
Making LLMs Work for Enterprise Data Tasks
by: Demiralp, Çağatay, et al.
Published: (2024)
by: Demiralp, Çağatay, et al.
Published: (2024)
TopoBench: Benchmarking LLMs on Hard Topological Reasoning
by: Maniparambil, Mayug, et al.
Published: (2026)
by: Maniparambil, Mayug, et al.
Published: (2026)
Trustworthy and Efficient LLMs Meet Databases
by: Kim, Kyoungmin, et al.
Published: (2024)
by: Kim, Kyoungmin, et al.
Published: (2024)
LLMs for Knowledge Graph Construction and Reasoning: Recent Capabilities and Future Opportunities
by: Zhu, Yuqi, et al.
Published: (2023)
by: Zhu, Yuqi, et al.
Published: (2023)
Graph Learning in the Era of LLMs: A Survey from the Perspective of Data, Models, and Tasks
by: Li, Xunkai, et al.
Published: (2024)
by: Li, Xunkai, et al.
Published: (2024)
Odin: Multi-Signal Graph Intelligence for Autonomous Discovery in Knowledge Graphs
by: Kizito, Muyukani, et al.
Published: (2026)
by: Kizito, Muyukani, et al.
Published: (2026)
Similar Items
-
KramaBench: A Benchmark for AI Systems on Data-to-Insight Pipelines over Data Lakes
by: Lai, Eugenie, et al.
Published: (2025) -
GraphAide: Advanced Graph-Assisted Query and Reasoning System
by: Purohit, Sumit, et al.
Published: (2024) -
ELT-Bench-Verified: Benchmark Quality Issues Underestimate AI Agent Capabilities
by: Zanoli, Christopher, et al.
Published: (2026) -
ELT-Bench: An End-to-End Benchmark for Evaluating AI Agents on ELT Pipelines
by: Jin, Tengjun, et al.
Published: (2025) -
ClinDet-Bench: Beyond Abstention, Evaluating Judgment Determinability of LLMs in Clinical Decision-Making
by: Watanabe, Yusuke, et al.
Published: (2026)