Multi-Objective Agentic Rewrites for Unstructured Data Processing
Fuente:
arXiv
Salvato in:
| Autori principali: | Wei, Lindsey Linxi, Shankar, Shreya, Zeighami, Sepanta, Chung, Yeounoh, Ozcan, Fatma, Parameswaran, Aditya G. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Task Cascades for Efficient Unstructured Data Processing
di: Shankar, Shreya, et al.
Pubblicazione: (2026)
di: Shankar, Shreya, et al.
Pubblicazione: (2026)
Cut Costs, Not Accuracy: LLM-Powered Data Processing with Guarantees
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025)
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025)
Featurized-Decomposition Join: Low-Cost Semantic Joins with Guarantees
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025)
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025)
Semantic Data Processing with Holistic Data Understanding
di: Sun, Youran, et al.
Pubblicazione: (2026)
di: Sun, Youran, et al.
Pubblicazione: (2026)
LLM-Powered Proactive Data Systems
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025)
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025)
DocETL: Agentic Query Rewriting and Evaluation for Complex Document Processing
di: Shankar, Shreya, et al.
Pubblicazione: (2024)
di: Shankar, Shreya, et al.
Pubblicazione: (2024)
Can AI Agents Answer Your Data Questions? A Benchmark for Data Agents
di: Ma, Ruiying, et al.
Pubblicazione: (2026)
di: Ma, Ruiying, et al.
Pubblicazione: (2026)
Arming Data Agents with Tribal Knowledge
di: Agarwal, Shubham, et al.
Pubblicazione: (2026)
di: Agarwal, Shubham, et al.
Pubblicazione: (2026)
BiasBuster: a Neural Approach for Accurate Estimation of Population Statistics using Biased Location Data
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
High-Fidelity And Complex Test Data Generation For Google SQL Code Generation Services
di: Kannan, Shivasankari, et al.
Pubblicazione: (2025)
di: Kannan, Shivasankari, et al.
Pubblicazione: (2025)
Theoretical Analysis of Learned Database Operations under Distribution Shift through Distribution Learnability
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
Towards Establishing Guaranteed Error for Learned Database Operations
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
Towards Accurate and Efficient Document Analytics with Large Language Models
di: Lin, Yiming, et al.
Pubblicazione: (2024)
di: Lin, Yiming, et al.
Pubblicazione: (2024)
RAG Without the Lag: Interactive Debugging for Retrieval-Augmented Generation Pipelines
di: Lauro, Quentin Romero, et al.
Pubblicazione: (2025)
di: Lauro, Quentin Romero, et al.
Pubblicazione: (2025)
Is Long Context All You Need? Leveraging LLM's Extended Context for NL2SQL
di: Chung, Yeounoh, et al.
Pubblicazione: (2025)
di: Chung, Yeounoh, et al.
Pubblicazione: (2025)
Fine-Grained Table Retrieval Through the Lens of Complex Queries
di: Kosiuk, Wojciech, et al.
Pubblicazione: (2026)
di: Kosiuk, Wojciech, et al.
Pubblicazione: (2026)
Supporting Our AI Overlords: Redesigning Data Systems to be Agent-First
di: Liu, Shu, et al.
Pubblicazione: (2025)
di: Liu, Shu, et al.
Pubblicazione: (2025)
NUDGE: Lightweight Non-Parametric Fine-Tuning of Embeddings for Retrieval
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
di: Zeighami, Sepanta, et al.
Pubblicazione: (2024)
Steering Semantic Data Processing With DocWrangler
di: Shankar, Shreya, et al.
Pubblicazione: (2025)
di: Shankar, Shreya, et al.
Pubblicazione: (2025)
An Agentic Approach to Metadata Reasoning
di: Zhang, Jiani, et al.
Pubblicazione: (2026)
di: Zhang, Jiani, et al.
Pubblicazione: (2026)
CHASE-SQL: Multi-Path Reasoning and Preference Optimized Candidate Selection in Text-to-SQL
di: Pourreza, Mohammadreza, et al.
Pubblicazione: (2024)
di: Pourreza, Mohammadreza, et al.
Pubblicazione: (2024)
RACOON: An LLM-based Framework for Retrieval-Augmented Column Type Annotation with a Knowledge Graph
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2024)
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2024)
SPADE: Synthesizing Data Quality Assertions for Large Language Model Pipelines
di: Shankar, Shreya, et al.
Pubblicazione: (2024)
di: Shankar, Shreya, et al.
Pubblicazione: (2024)
100x Cost & Latency Reduction: Performance Analysis of AI Query Approximation using Lightweight Proxy Models
di: Chung, Yeounoh, et al.
Pubblicazione: (2026)
di: Chung, Yeounoh, et al.
Pubblicazione: (2026)
Analytical Queries for Unstructured Data
di: Kang, Daniel
Pubblicazione: (2025)
di: Kang, Daniel
Pubblicazione: (2025)
Demonstration of MaskSearch: Efficiently Querying Image Masks for Machine Learning Workflows
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2024)
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2024)
Process Mining for Unstructured Data: Challenges and Research Directions
di: Koschmider, Agnes, et al.
Pubblicazione: (2023)
di: Koschmider, Agnes, et al.
Pubblicazione: (2023)
TWIX: Automatically Reconstructing Structured Data from Templatized Documents
di: Lin, Yiming, et al.
Pubblicazione: (2025)
di: Lin, Yiming, et al.
Pubblicazione: (2025)
LEAP: LLM-powered End-to-end Automatic Library for Processing Social Science Queries on Unstructured Data
di: Hu, Chuxuan, et al.
Pubblicazione: (2025)
di: Hu, Chuxuan, et al.
Pubblicazione: (2025)
Continuous Prompts: LLM-Augmented Pipeline Processing over Unstructured Streams
di: Chen, Shu, et al.
Pubblicazione: (2025)
di: Chen, Shu, et al.
Pubblicazione: (2025)
A Case for Computing on Unstructured Data
di: Sadia, Mushtari, et al.
Pubblicazione: (2025)
di: Sadia, Mushtari, et al.
Pubblicazione: (2025)
GenRewrite: Query Rewriting via Large Language Models
di: Liu, Jie, et al.
Pubblicazione: (2024)
di: Liu, Jie, et al.
Pubblicazione: (2024)
UniDataBench: Evaluating Data Analytics Agents Across Structured and Unstructured Data
di: Weng, Han, et al.
Pubblicazione: (2025)
di: Weng, Han, et al.
Pubblicazione: (2025)
Unstructured Data Analysis using LLMs: A Comprehensive Benchmark
di: Deng, Qiyan, et al.
Pubblicazione: (2025)
di: Deng, Qiyan, et al.
Pubblicazione: (2025)
Multi-Objective Genetic Algorithm for Materialized View Optimization in Data Warehouses
di: Manavi, Mahdi
Pubblicazione: (2024)
di: Manavi, Mahdi
Pubblicazione: (2024)
CHASE: A Native Relational Database for Hybrid Queries on Structured and Unstructured Data
di: Ma, Rui, et al.
Pubblicazione: (2025)
di: Ma, Rui, et al.
Pubblicazione: (2025)
Query Rewriting via LLMs
di: Dharwada, Sriram, et al.
Pubblicazione: (2025)
di: Dharwada, Sriram, et al.
Pubblicazione: (2025)
TARGET: Benchmarking Table Retrieval for Generative Tasks
di: Ji, Xingyu, et al.
Pubblicazione: (2025)
di: Ji, Xingyu, et al.
Pubblicazione: (2025)
QUEST: Query Optimization in Unstructured Document Analysis
di: Sun, Zhaoze, et al.
Pubblicazione: (2025)
di: Sun, Zhaoze, et al.
Pubblicazione: (2025)
AgenticData: An Agentic Data Analytics System for Heterogeneous Data
di: Sun, Ji, et al.
Pubblicazione: (2025)
di: Sun, Ji, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Task Cascades for Efficient Unstructured Data Processing
di: Shankar, Shreya, et al.
Pubblicazione: (2026) -
Cut Costs, Not Accuracy: LLM-Powered Data Processing with Guarantees
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025) -
Featurized-Decomposition Join: Low-Cost Semantic Joins with Guarantees
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025) -
Semantic Data Processing with Holistic Data Understanding
di: Sun, Youran, et al.
Pubblicazione: (2026) -
LLM-Powered Proactive Data Systems
di: Zeighami, Sepanta, et al.
Pubblicazione: (2025)