Quality Assessment of Tabular Data using Large Language Models and Code Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Akella, Ashlesha, Kaul, Akshar, Narayanam, Krishnasuri, Mehta, Sameep |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data Wrangling Task Automation Using Code-Generating Language Models
by: Akella, Ashlesha, et al.
Published: (2025)
by: Akella, Ashlesha, et al.
Published: (2025)
QUIS: Question-guided Insights Generation for Automated Exploratory Data Analysis
by: Manatkar, Abhijit, et al.
Published: (2024)
by: Manatkar, Abhijit, et al.
Published: (2024)
GEE-OPs: An Operator Knowledge Base for Geospatial Code Generation on the Google Earth Engine Platform Powered by Large Language Models
by: Hou, Shuyang, et al.
Published: (2024)
by: Hou, Shuyang, et al.
Published: (2024)
Geo-FuB: A Method for Constructing an Operator-Function Knowledge Base for Geospatial Code Generation Tasks Using Large Language Models
by: Hou, Shuyang, et al.
Published: (2024)
by: Hou, Shuyang, et al.
Published: (2024)
SimClone: Detecting Tabular Data Clones using Value Similarity
by: Yang, Xu, et al.
Published: (2024)
by: Yang, Xu, et al.
Published: (2024)
Adaptive Data Quality Scoring Operations Framework using Drift-Aware Mechanism for Industrial Applications
by: Bayram, Firas, et al.
Published: (2024)
by: Bayram, Firas, et al.
Published: (2024)
Prompt Migration: Stabilizing GenAI Applications with Evolving Large Language Models
by: Tripathi, Shivani, et al.
Published: (2025)
by: Tripathi, Shivani, et al.
Published: (2025)
LeGo-Code: Can Modular Curriculum Learning Advance Complex Code Generation? Insights from Text-to-SQL
by: Chafik, Salmane, et al.
Published: (2026)
by: Chafik, Salmane, et al.
Published: (2026)
Enhancing LLM Fine-tuning for Text-to-SQLs by SQL Quality Measurement
by: Sarker, Shouvon, et al.
Published: (2024)
by: Sarker, Shouvon, et al.
Published: (2024)
AutoGEEval: A Multimodal and Automated Framework for Geospatial Code Generation on GEE with Large Language Models
by: Hou, Shuyang, et al.
Published: (2025)
by: Hou, Shuyang, et al.
Published: (2025)
Generating Reliable Adverse event Profiles for Health through Automated Integrated Data (GRAPH-AID): A Semi-Automated Ontology Building Approach
by: Gadusu, Srikar Reddy, et al.
Published: (2025)
by: Gadusu, Srikar Reddy, et al.
Published: (2025)
Natural Language Query Engine for Relational Databases using Generative AI
by: Fotso, Steve Tueno
Published: (2024)
by: Fotso, Steve Tueno
Published: (2024)
Declarative Techniques for NL Queries over Heterogeneous Data
by: Khabiri, Elham, et al.
Published: (2025)
by: Khabiri, Elham, et al.
Published: (2025)
A Pythonic Functional Approach for Semantic Data Harmonisation in the ILIAD Project
by: Nystad, Erik Johan, et al.
Published: (2026)
by: Nystad, Erik Johan, et al.
Published: (2026)
SPADE: Synthesizing Data Quality Assertions for Large Language Model Pipelines
by: Shankar, Shreya, et al.
Published: (2024)
by: Shankar, Shreya, et al.
Published: (2024)
From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability
by: Sisodia, Twinkll
Published: (2026)
by: Sisodia, Twinkll
Published: (2026)
Control-flow Reconstruction Attacks on Business Process Models
by: Kirchmann, Henrik, et al.
Published: (2024)
by: Kirchmann, Henrik, et al.
Published: (2024)
Towards Automated Data Sciences with Natural Language and SageCopilot: Practices and Lessons Learned
by: Liao, Yuan, et al.
Published: (2024)
by: Liao, Yuan, et al.
Published: (2024)
Enhancing High-Quality Code Generation in Large Language Models with Comparative Prefix-Tuning
by: Jiang, Yuan, et al.
Published: (2025)
by: Jiang, Yuan, et al.
Published: (2025)
Vibe Coding on Trial: Operating Characteristics of Unanimous LLM Juries
by: Ullah, Muhammad Aziz, et al.
Published: (2026)
by: Ullah, Muhammad Aziz, et al.
Published: (2026)
Tabularis Formatus: Predictive Formatting for Tables
by: Singh, Mukul, et al.
Published: (2025)
by: Singh, Mukul, et al.
Published: (2025)
KRONE: Scalable LLM-Augmented Log Anomaly Detection via Hierarchical Abstraction
by: Ma, Lei, et al.
Published: (2026)
by: Ma, Lei, et al.
Published: (2026)
Detect, Localize, and Explain: Interactive Hierarchical Log Anomaly Analytics with LLM Augmentation
by: Ma, Lei, et al.
Published: (2026)
by: Ma, Lei, et al.
Published: (2026)
Unraveling the Never-Ending Story of Lifecycles and Vitalizing Processes
by: Fahrenkrog-Petersen, Stephan A., et al.
Published: (2024)
by: Fahrenkrog-Petersen, Stephan A., et al.
Published: (2024)
Liberal Entity Matching as a Compound AI Toolchain
by: Fu, Silvery D., et al.
Published: (2024)
by: Fu, Silvery D., et al.
Published: (2024)
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
by: Agarwal, Prerna, et al.
Published: (2025)
by: Agarwal, Prerna, et al.
Published: (2025)
Augmenting Large Language Models with Static Code Analysis for Automated Code Quality Improvements
by: Abtahi, Seyed Moein, et al.
Published: (2025)
by: Abtahi, Seyed Moein, et al.
Published: (2025)
A Study on the Improvement of Code Generation Quality Using Large Language Models Leveraging Product Documentation
by: Morimoto, Takuro, et al.
Published: (2025)
by: Morimoto, Takuro, et al.
Published: (2025)
An Empirical Study on Self-correcting Large Language Models for Data Science Code Generation
by: Quoc, Thai Tang, et al.
Published: (2024)
by: Quoc, Thai Tang, et al.
Published: (2024)
Task Abstention for Large Language Models in Code Generation
by: Zhou, Yanke, et al.
Published: (2026)
by: Zhou, Yanke, et al.
Published: (2026)
Enhancing SQL Query Generation with Neurosymbolic Reasoning
by: Princis, Henrijs, et al.
Published: (2024)
by: Princis, Henrijs, et al.
Published: (2024)
Polygon: Symbolic Reasoning for SQL using Conflict-Driven Under-Approximation Search
by: Zhao, Pinhan, et al.
Published: (2025)
by: Zhao, Pinhan, et al.
Published: (2025)
Large Language Model Guided Self-Debugging Code Generation
by: Adnan, Muntasir, et al.
Published: (2025)
by: Adnan, Muntasir, et al.
Published: (2025)
Personality-Guided Code Generation Using Large Language Models
by: Guo, Yaoqi, et al.
Published: (2024)
by: Guo, Yaoqi, et al.
Published: (2024)
Bugs in Large Language Models Generated Code: An Empirical Study
by: Tambon, Florian, et al.
Published: (2024)
by: Tambon, Florian, et al.
Published: (2024)
Agentic DAG-Orchestrated Planner Framework for Multi-Modal, Multi-Hop Question Answering in Hybrid Data Lakes
by: B, Kirushikesh D, et al.
Published: (2026)
by: B, Kirushikesh D, et al.
Published: (2026)
The Performance of the LSTM-based Code Generated by Large Language Models (LLMs) in Forecasting Time Series Data
by: Gopali, Saroj, et al.
Published: (2024)
by: Gopali, Saroj, et al.
Published: (2024)
GeoCode-GPT: A Large Language Model for Geospatial Code Generation Tasks
by: Hou, Shuyang, et al.
Published: (2024)
by: Hou, Shuyang, et al.
Published: (2024)
Generating High-Quality Datasets for Code Editing via Open-Source Language Models
by: Zhang, Zekai, et al.
Published: (2025)
by: Zhang, Zekai, et al.
Published: (2025)
Generating Software Architecture Description from Source Code using Reverse Engineering and Large Language Model
by: Hatahet, Ahmad, et al.
Published: (2025)
by: Hatahet, Ahmad, et al.
Published: (2025)
Similar Items
-
Data Wrangling Task Automation Using Code-Generating Language Models
by: Akella, Ashlesha, et al.
Published: (2025) -
QUIS: Question-guided Insights Generation for Automated Exploratory Data Analysis
by: Manatkar, Abhijit, et al.
Published: (2024) -
GEE-OPs: An Operator Knowledge Base for Geospatial Code Generation on the Google Earth Engine Platform Powered by Large Language Models
by: Hou, Shuyang, et al.
Published: (2024) -
Geo-FuB: A Method for Constructing an Operator-Function Knowledge Base for Geospatial Code Generation Tasks Using Large Language Models
by: Hou, Shuyang, et al.
Published: (2024) -
SimClone: Detecting Tabular Data Clones using Value Similarity
by: Yang, Xu, et al.
Published: (2024)