Guardado en:
| Autores principales: | Shankar, Shreya, Li, Haotian, Asawa, Parth, Hulsebos, Madelon, Lin, Yiming, Zamfirescu-Pereira, J. D., Chase, Harrison, Fu-Hinthorn, Will, Parameswaran, Aditya G., Wu, Eugene |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2401.03038 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines
por: Vir, Reya, et al.
Publicado: (2025)
por: Vir, Reya, et al.
Publicado: (2025)
Towards Accurate and Efficient Document Analytics with Large Language Models
por: Lin, Yiming, et al.
Publicado: (2024)
por: Lin, Yiming, et al.
Publicado: (2024)
TARGET: Benchmarking Table Retrieval for Generative Tasks
por: Ji, Xingyu, et al.
Publicado: (2025)
por: Ji, Xingyu, et al.
Publicado: (2025)
LLM-Powered Proactive Data Systems
por: Zeighami, Sepanta, et al.
Publicado: (2025)
por: Zeighami, Sepanta, et al.
Publicado: (2025)
Task Cascades for Efficient Unstructured Data Processing
por: Shankar, Shreya, et al.
Publicado: (2026)
por: Shankar, Shreya, et al.
Publicado: (2026)
Featurized-Decomposition Join: Low-Cost Semantic Joins with Guarantees
por: Zeighami, Sepanta, et al.
Publicado: (2025)
por: Zeighami, Sepanta, et al.
Publicado: (2025)
Cut Costs, Not Accuracy: LLM-Powered Data Processing with Guarantees
por: Zeighami, Sepanta, et al.
Publicado: (2025)
por: Zeighami, Sepanta, et al.
Publicado: (2025)
Towards Contextual Sensitive Data Detection
por: Telkamp, Liang, et al.
Publicado: (2025)
por: Telkamp, Liang, et al.
Publicado: (2025)
DocETL: Agentic Query Rewriting and Evaluation for Complex Document Processing
por: Shankar, Shreya, et al.
Publicado: (2024)
por: Shankar, Shreya, et al.
Publicado: (2024)
Are We Asking the Right Questions? On Ambiguity in Natural Language Queries for Tabular Data Analysis
por: Gomm, Daniel, et al.
Publicado: (2025)
por: Gomm, Daniel, et al.
Publicado: (2025)
Rethinking Dataset Discovery with DataScout
por: Lin, Rachel, et al.
Publicado: (2025)
por: Lin, Rachel, et al.
Publicado: (2025)
Flow with FlorDB: Incremental Context Maintenance for the Machine Learning Lifecycle
por: Garcia, Rolando, et al.
Publicado: (2024)
por: Garcia, Rolando, et al.
Publicado: (2024)
Testing Database Systems with Large Language Model Synthesized Fragments
por: Zhong, Suyang, et al.
Publicado: (2025)
por: Zhong, Suyang, et al.
Publicado: (2025)
Semantic Data Processing with Holistic Data Understanding
por: Sun, Youran, et al.
Publicado: (2026)
por: Sun, Youran, et al.
Publicado: (2026)
Observatory: Characterizing Embeddings of Relational Tables
por: Cong, Tianji, et al.
Publicado: (2023)
por: Cong, Tianji, et al.
Publicado: (2023)
Fine-Grained Table Retrieval Through the Lens of Complex Queries
por: Kosiuk, Wojciech, et al.
Publicado: (2026)
por: Kosiuk, Wojciech, et al.
Publicado: (2026)
DIRT: Database-Integrated Random Testing
por: Keles, Alperen, et al.
Publicado: (2026)
por: Keles, Alperen, et al.
Publicado: (2026)
Multi-Objective Agentic Rewrites for Unstructured Data Processing
por: Wei, Lindsey Linxi, et al.
Publicado: (2025)
por: Wei, Lindsey Linxi, et al.
Publicado: (2025)
Prompt Migration: Stabilizing GenAI Applications with Evolving Large Language Models
por: Tripathi, Shivani, et al.
Publicado: (2025)
por: Tripathi, Shivani, et al.
Publicado: (2025)
CUBES: A Parallel Synthesizer for SQL Using Examples
por: Brancas, Ricardo, et al.
Publicado: (2022)
por: Brancas, Ricardo, et al.
Publicado: (2022)
Synthesizing Document Database Queries using Collection Abstractions
por: Liu, Qikang, et al.
Publicado: (2024)
por: Liu, Qikang, et al.
Publicado: (2024)
Steering Semantic Data Processing With DocWrangler
por: Shankar, Shreya, et al.
Publicado: (2025)
por: Shankar, Shreya, et al.
Publicado: (2025)
AI-Assisted SQL Authoring at Industry Scale
por: Maddila, Chandra, et al.
Publicado: (2024)
por: Maddila, Chandra, et al.
Publicado: (2024)
Quality Assessment of Tabular Data using Large Language Models and Code Generation
por: Akella, Ashlesha, et al.
Publicado: (2025)
por: Akella, Ashlesha, et al.
Publicado: (2025)
Can AI Agents Answer Your Data Questions? A Benchmark for Data Agents
por: Ma, Ruiying, et al.
Publicado: (2026)
por: Ma, Ruiying, et al.
Publicado: (2026)
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
por: Liu, Shu, et al.
Publicado: (2026)
por: Liu, Shu, et al.
Publicado: (2026)
Process Modeling With Large Language Models
por: Kourani, Humam, et al.
Publicado: (2024)
por: Kourani, Humam, et al.
Publicado: (2024)
Object Graph Programming
por: Thimmaiah, Aditya, et al.
Publicado: (2024)
por: Thimmaiah, Aditya, et al.
Publicado: (2024)
Enhancing LLM Fine-tuning for Text-to-SQLs by SQL Quality Measurement
por: Sarker, Shouvon, et al.
Publicado: (2024)
por: Sarker, Shouvon, et al.
Publicado: (2024)
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
por: Shankar, Shreya, et al.
Publicado: (2024)
por: Shankar, Shreya, et al.
Publicado: (2024)
Adaptive Data Quality Scoring Operations Framework using Drift-Aware Mechanism for Industrial Applications
por: Bayram, Firas, et al.
Publicado: (2024)
por: Bayram, Firas, et al.
Publicado: (2024)
Liberal Entity Matching as a Compound AI Toolchain
por: Fu, Silvery D., et al.
Publicado: (2024)
por: Fu, Silvery D., et al.
Publicado: (2024)
Automated Tensor-Relational Decomposition for Large-Scale Sparse Tensor Computation
por: Tang, Yuxin, et al.
Publicado: (2026)
por: Tang, Yuxin, et al.
Publicado: (2026)
How well do LLMs reason over tabular data, really?
por: Wolff, Cornelius, et al.
Publicado: (2025)
por: Wolff, Cornelius, et al.
Publicado: (2025)
GEE-OPs: An Operator Knowledge Base for Geospatial Code Generation on the Google Earth Engine Platform Powered by Large Language Models
por: Hou, Shuyang, et al.
Publicado: (2024)
por: Hou, Shuyang, et al.
Publicado: (2024)
AssertionBench: A Benchmark to Evaluate Large-Language Models for Assertion Generation
por: Pulavarthi, Vaishnavi, et al.
Publicado: (2024)
por: Pulavarthi, Vaishnavi, et al.
Publicado: (2024)
Mining Constraints from Reference Process Models for Detecting Best-Practice Violations in Event Logs
por: Rebmann, Adrian, et al.
Publicado: (2024)
por: Rebmann, Adrian, et al.
Publicado: (2024)
Dinkel: State-Aware and Granular Framework for Validating Graph Databases
por: Wüst, Celine, et al.
Publicado: (2024)
por: Wüst, Celine, et al.
Publicado: (2024)
Vextra: A Unified Middleware Abstraction for Heterogeneous Vector Database Systems
por: Suri, Chandan, et al.
Publicado: (2026)
por: Suri, Chandan, et al.
Publicado: (2026)
A Practical Framework for Flaky Failure Triage in Distributed Database Continuous Integration
por: Zhu, Jun-Peng, et al.
Publicado: (2026)
por: Zhu, Jun-Peng, et al.
Publicado: (2026)
Ejemplares similares
-
PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines
por: Vir, Reya, et al.
Publicado: (2025) -
Towards Accurate and Efficient Document Analytics with Large Language Models
por: Lin, Yiming, et al.
Publicado: (2024) -
TARGET: Benchmarking Table Retrieval for Generative Tasks
por: Ji, Xingyu, et al.
Publicado: (2025) -
LLM-Powered Proactive Data Systems
por: Zeighami, Sepanta, et al.
Publicado: (2025) -
Task Cascades for Efficient Unstructured Data Processing
por: Shankar, Shreya, et al.
Publicado: (2026)