Can AI Agents Answer Your Data Questions? A Benchmark for Data Agents
Fuente:
arXiv
Guardado en:
| Autores principales: | Ma, Ruiying, Shankar, Shreya, Chen, Ruiqi, Lin, Yiming, Zeighami, Sepanta, Ghosh, Rajoshi, Gupta, Abhinav, Gupta, Anushrut, Gopal, Tanmai, Parameswaran, Aditya G. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Task Cascades for Efficient Unstructured Data Processing
por: Shankar, Shreya, et al.
Publicado: (2026)
por: Shankar, Shreya, et al.
Publicado: (2026)
Cut Costs, Not Accuracy: LLM-Powered Data Processing with Guarantees
por: Zeighami, Sepanta, et al.
Publicado: (2025)
por: Zeighami, Sepanta, et al.
Publicado: (2025)
LLM-Powered Proactive Data Systems
por: Zeighami, Sepanta, et al.
Publicado: (2025)
por: Zeighami, Sepanta, et al.
Publicado: (2025)
Featurized-Decomposition Join: Low-Cost Semantic Joins with Guarantees
por: Zeighami, Sepanta, et al.
Publicado: (2025)
por: Zeighami, Sepanta, et al.
Publicado: (2025)
Semantic Data Processing with Holistic Data Understanding
por: Sun, Youran, et al.
Publicado: (2026)
por: Sun, Youran, et al.
Publicado: (2026)
Multi-Objective Agentic Rewrites for Unstructured Data Processing
por: Wei, Lindsey Linxi, et al.
Publicado: (2025)
por: Wei, Lindsey Linxi, et al.
Publicado: (2025)
Arming Data Agents with Tribal Knowledge
por: Agarwal, Shubham, et al.
Publicado: (2026)
por: Agarwal, Shubham, et al.
Publicado: (2026)
Towards Accurate and Efficient Document Analytics with Large Language Models
por: Lin, Yiming, et al.
Publicado: (2024)
por: Lin, Yiming, et al.
Publicado: (2024)
Supporting Our AI Overlords: Redesigning Data Systems to be Agent-First
por: Liu, Shu, et al.
Publicado: (2025)
por: Liu, Shu, et al.
Publicado: (2025)
BiasBuster: a Neural Approach for Accurate Estimation of Population Statistics using Biased Location Data
por: Zeighami, Sepanta, et al.
Publicado: (2024)
por: Zeighami, Sepanta, et al.
Publicado: (2024)
Theoretical Analysis of Learned Database Operations under Distribution Shift through Distribution Learnability
por: Zeighami, Sepanta, et al.
Publicado: (2024)
por: Zeighami, Sepanta, et al.
Publicado: (2024)
Towards Establishing Guaranteed Error for Learned Database Operations
por: Zeighami, Sepanta, et al.
Publicado: (2024)
por: Zeighami, Sepanta, et al.
Publicado: (2024)
RAG Without the Lag: Interactive Debugging for Retrieval-Augmented Generation Pipelines
por: Lauro, Quentin Romero, et al.
Publicado: (2025)
por: Lauro, Quentin Romero, et al.
Publicado: (2025)
NUDGE: Lightweight Non-Parametric Fine-Tuning of Embeddings for Retrieval
por: Zeighami, Sepanta, et al.
Publicado: (2024)
por: Zeighami, Sepanta, et al.
Publicado: (2024)
DocETL: Agentic Query Rewriting and Evaluation for Complex Document Processing
por: Shankar, Shreya, et al.
Publicado: (2024)
por: Shankar, Shreya, et al.
Publicado: (2024)
SPADE: Synthesizing Data Quality Assertions for Large Language Model Pipelines
por: Shankar, Shreya, et al.
Publicado: (2024)
por: Shankar, Shreya, et al.
Publicado: (2024)
Steering Semantic Data Processing With DocWrangler
por: Shankar, Shreya, et al.
Publicado: (2025)
por: Shankar, Shreya, et al.
Publicado: (2025)
TWIX: Automatically Reconstructing Structured Data from Templatized Documents
por: Lin, Yiming, et al.
Publicado: (2025)
por: Lin, Yiming, et al.
Publicado: (2025)
FDABench: A Benchmark for Data Agents on Analytical Queries over Heterogeneous Data
por: Wang, Ziting, et al.
Publicado: (2025)
por: Wang, Ziting, et al.
Publicado: (2025)
QUIS: Question-guided Insights Generation for Automated Exploratory Data Analysis
por: Manatkar, Abhijit, et al.
Publicado: (2024)
por: Manatkar, Abhijit, et al.
Publicado: (2024)
Is Agent Memory a Database? Rethinking Data Foundations for Long-Term AI Agent Memory
por: Orogat, Abdelghny, et al.
Publicado: (2026)
por: Orogat, Abdelghny, et al.
Publicado: (2026)
DataClaw: An Autonomous Data Agent with Instant Messaging Integration
por: Li, Huahang, et al.
Publicado: (2026)
por: Li, Huahang, et al.
Publicado: (2026)
TARGET: Benchmarking Table Retrieval for Generative Tasks
por: Ji, Xingyu, et al.
Publicado: (2025)
por: Ji, Xingyu, et al.
Publicado: (2025)
UniDataBench: Evaluating Data Analytics Agents Across Structured and Unstructured Data
por: Weng, Han, et al.
Publicado: (2025)
por: Weng, Han, et al.
Publicado: (2025)
Data Agents: Levels, State of the Art, and Open Problems
por: Luo, Yuyu, et al.
Publicado: (2026)
por: Luo, Yuyu, et al.
Publicado: (2026)
Data Agent: A Holistic Architecture for Orchestrating Data+AI Ecosystems
por: Sun, Zhaoyan, et al.
Publicado: (2025)
por: Sun, Zhaoyan, et al.
Publicado: (2025)
Reliable Collaborative Conversational Agent System Based on LLMs and Answer Set Programming
por: Zeng, Yankai, et al.
Publicado: (2025)
por: Zeng, Yankai, et al.
Publicado: (2025)
Pneuma-Seeker: A Relational Reification Mechanism to Align AI Agents with Human Work over Relational Data
por: Balaka, Muhammad Imam Luthfi, et al.
Publicado: (2026)
por: Balaka, Muhammad Imam Luthfi, et al.
Publicado: (2026)
LLM/Agent-as-Data-Analyst: A Survey
por: Tang, Zirui, et al.
Publicado: (2025)
por: Tang, Zirui, et al.
Publicado: (2025)
DeepEye: A Steerable Self-driving Data Agent System
por: Li, Boyan, et al.
Publicado: (2026)
por: Li, Boyan, et al.
Publicado: (2026)
ELT-Bench-Verified: Benchmark Quality Issues Underestimate AI Agent Capabilities
por: Zanoli, Christopher, et al.
Publicado: (2026)
por: Zanoli, Christopher, et al.
Publicado: (2026)
ELT-Bench: An End-to-End Benchmark for Evaluating AI Agents on ELT Pipelines
por: Jin, Tengjun, et al.
Publicado: (2025)
por: Jin, Tengjun, et al.
Publicado: (2025)
Snowpark: Performant, Secure, User-Friendly Data Engineering and AI/ML Next To Your Data
por: Baker, Brandon, et al.
Publicado: (2025)
por: Baker, Brandon, et al.
Publicado: (2025)
DAgent: A Relational Database-Driven Data Analysis Report Generation Agent
por: Xu, Wenyi, et al.
Publicado: (2025)
por: Xu, Wenyi, et al.
Publicado: (2025)
Interactive Data Harmonization with LLM Agents: Opportunities and Challenges
por: Santos, Aécio, et al.
Publicado: (2025)
por: Santos, Aécio, et al.
Publicado: (2025)
Create Benchmarks for Data Lakes
por: Lyu, Yi, et al.
Publicado: (2026)
por: Lyu, Yi, et al.
Publicado: (2026)
From Questions to Queries: An AI-powered Multi-Agent Framework for Spatial Text-to-SQL
por: Kazazi, Ali Khosravi, et al.
Publicado: (2025)
por: Kazazi, Ali Khosravi, et al.
Publicado: (2025)
Blue Data Intelligence Layer: Streaming Data and Agents for Multi-source Multi-modal Data-Centric Applications
por: Aminnaseri, Moin, et al.
Publicado: (2026)
por: Aminnaseri, Moin, et al.
Publicado: (2026)
SiloFuse: Cross-silo Synthetic Data Generation with Latent Tabular Diffusion Models
por: Shankar, Aditya, et al.
Publicado: (2024)
por: Shankar, Aditya, et al.
Publicado: (2024)
Development of Data Evaluation Benchmark for Data Wrangling Recommendation System
por: Wang, Yuqing, et al.
Publicado: (2024)
por: Wang, Yuqing, et al.
Publicado: (2024)
Ejemplares similares
-
Task Cascades for Efficient Unstructured Data Processing
por: Shankar, Shreya, et al.
Publicado: (2026) -
Cut Costs, Not Accuracy: LLM-Powered Data Processing with Guarantees
por: Zeighami, Sepanta, et al.
Publicado: (2025) -
LLM-Powered Proactive Data Systems
por: Zeighami, Sepanta, et al.
Publicado: (2025) -
Featurized-Decomposition Join: Low-Cost Semantic Joins with Guarantees
por: Zeighami, Sepanta, et al.
Publicado: (2025) -
Semantic Data Processing with Holistic Data Understanding
por: Sun, Youran, et al.
Publicado: (2026)