Tabular Data Understanding with LLMs: A Survey of Recent Advances and Challenges
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Xiaofeng, Ritter, Alan, Xu, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs
by: Zhou, Wei, et al.
Published: (2026)
by: Zhou, Wei, et al.
Published: (2026)
Graph Learning in the Era of LLMs: A Survey from the Perspective of Data, Models, and Tasks
by: Li, Xunkai, et al.
Published: (2024)
by: Li, Xunkai, et al.
Published: (2024)
Meta-Tuning LLMs to Leverage Lexical Knowledge for Generalizable Language Style Understanding
by: Guo, Ruohao, et al.
Published: (2023)
by: Guo, Ruohao, et al.
Published: (2023)
CuTS: Customizable Tabular Synthetic Data Generation
by: Vero, Mark, et al.
Published: (2023)
by: Vero, Mark, et al.
Published: (2023)
Enhancing Temporal Understanding in LLMs for Semi-structured Tables
by: Deng, Irwin, et al.
Published: (2024)
by: Deng, Irwin, et al.
Published: (2024)
Tackling prediction tasks in relational databases with LLMs
by: Wydmuch, Marek, et al.
Published: (2024)
by: Wydmuch, Marek, et al.
Published: (2024)
LLMs for Knowledge Graph Construction and Reasoning: Recent Capabilities and Future Opportunities
by: Zhu, Yuqi, et al.
Published: (2023)
by: Zhu, Yuqi, et al.
Published: (2023)
ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs
by: Qi, Yanlin, et al.
Published: (2026)
by: Qi, Yanlin, et al.
Published: (2026)
Probabilistic Reasoning with LLMs for k-anonymity Estimation
by: Zheng, Jonathan, et al.
Published: (2025)
by: Zheng, Jonathan, et al.
Published: (2025)
When Large Language Models Meet Vector Databases: A Survey
by: Jing, Zhi, et al.
Published: (2024)
by: Jing, Zhi, et al.
Published: (2024)
KcMF: A Knowledge-compliant Framework for Schema and Entity Matching with Fine-tuning-free LLMs
by: Xu, Yongqin, et al.
Published: (2024)
by: Xu, Yongqin, et al.
Published: (2024)
PARROT: A Benchmark for Evaluating LLMs in Cross-System SQL Translation
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
GNN: Graph Neural Network and Large Language Model for Data Discovery
by: Hoang, Thomas
Published: (2024)
by: Hoang, Thomas
Published: (2024)
Exploring the Heterogeneity of Tabular Data: A Diversity-aware Data Generator via LLMs
by: Tang, Yafeng, et al.
Published: (2025)
by: Tang, Yafeng, et al.
Published: (2025)
Q-NL Verifier: Leveraging Synthetic Data for Robust Knowledge Graph Question Answering
by: Schwabe, Tim, et al.
Published: (2025)
by: Schwabe, Tim, et al.
Published: (2025)
MMTU: A Massive Multi-Task Table Understanding and Reasoning Benchmark
by: Xing, Junjie, et al.
Published: (2025)
by: Xing, Junjie, et al.
Published: (2025)
A Survey of LLM $\times$ DATA
by: Zhou, Xuanhe, et al.
Published: (2025)
by: Zhou, Xuanhe, et al.
Published: (2025)
Pruning Minimal Reasoning Graphs for Efficient Retrieval-Augmented Generation
by: Wang, Ning, et al.
Published: (2026)
by: Wang, Ning, et al.
Published: (2026)
Hunt Instead of Wait: Evaluating Deep Data Research on Large Language Models
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Data Agent: A Holistic Architecture for Orchestrating Data+AI Ecosystems
by: Sun, Zhaoyan, et al.
Published: (2025)
by: Sun, Zhaoyan, et al.
Published: (2025)
TabSTAR: A Tabular Foundation Model for Tabular Data with Text Fields
by: Arazi, Alan, et al.
Published: (2025)
by: Arazi, Alan, et al.
Published: (2025)
Proof-Carrying Numbers (PCN): A Protocol for Trustworthy Numeric Answers from LLMs via Claim Verification
by: Solatorio, Aivin V.
Published: (2025)
by: Solatorio, Aivin V.
Published: (2025)
Cross-table Synthetic Tabular Data Detection
by: Kindji, G. Charbel N., et al.
Published: (2024)
by: Kindji, G. Charbel N., et al.
Published: (2024)
Jellyfish: A Large Language Model for Data Preprocessing
by: Zhang, Haochen, et al.
Published: (2023)
by: Zhang, Haochen, et al.
Published: (2023)
PromptMind Team at EHRSQL-2024: Improving Reliability of SQL Generation using Ensemble LLMs
by: Gundabathula, Satya K, et al.
Published: (2024)
by: Gundabathula, Satya K, et al.
Published: (2024)
Hierarchical Conditional Tabular GAN for Multi-Tabular Synthetic Data Generation
by: Ågren, Wilhelm, et al.
Published: (2024)
by: Ågren, Wilhelm, et al.
Published: (2024)
GAN-based Tabular Data Generator for Constructing Synopsis in Approximate Query Processing: Challenges and Solutions
by: Fallahian, Mohammadali, et al.
Published: (2022)
by: Fallahian, Mohammadali, et al.
Published: (2022)
RADAR: Benchmarking Language Models on Imperfect Tabular Data
by: Gu, Ken, et al.
Published: (2025)
by: Gu, Ken, et al.
Published: (2025)
Investigating and Alleviating Harm Amplification in LLM Interactions
by: Guo, Ruohao, et al.
Published: (2026)
by: Guo, Ruohao, et al.
Published: (2026)
LLM-Enhanced Data Management
by: Zhou, Xuanhe, et al.
Published: (2024)
by: Zhou, Xuanhe, et al.
Published: (2024)
ReCellTy: Domain-Specific Knowledge Graph Retrieval-Augmented LLMs Reasoning Workflow for Single-Cell Annotation
by: Han, Dezheng, et al.
Published: (2025)
by: Han, Dezheng, et al.
Published: (2025)
From Source to Target: Leveraging Transfer Learning for Predictive Process Monitoring in Organizations
by: Weinzierl, Sven, et al.
Published: (2025)
by: Weinzierl, Sven, et al.
Published: (2025)
Query Performance Explanation through Large Language Model for HTAP Systems
by: Xiu, Haibo, et al.
Published: (2024)
by: Xiu, Haibo, et al.
Published: (2024)
Duplicate Detection with GenAI
by: Ormesher, Ian
Published: (2024)
by: Ormesher, Ian
Published: (2024)
Pseudo-Label Calibration Semi-supervised Multi-Modal Entity Alignment
by: Wang, Luyao, et al.
Published: (2024)
by: Wang, Luyao, et al.
Published: (2024)
SketchFill: Sketch-Guided Code Generation for Imputing Derived Missing Values
by: Zhang, Yunfan, et al.
Published: (2024)
by: Zhang, Yunfan, et al.
Published: (2024)
Prompt Valuation Based on Shapley Values
by: Liu, Hanxi, et al.
Published: (2023)
by: Liu, Hanxi, et al.
Published: (2023)
Table-LLM-Specialist: Language Model Specialists for Tables using Iterative Generator-Validator Fine-tuning
by: Xing, Junjie, et al.
Published: (2024)
by: Xing, Junjie, et al.
Published: (2024)
Database Entity Recognition with Data Augmentation and Deep Learning
by: Fu, Zikun, et al.
Published: (2025)
by: Fu, Zikun, et al.
Published: (2025)
A Domain-Specific Language for LLM-Driven Trigger Generation in Multimodal Data Collection
by: Reis, Philipp, et al.
Published: (2026)
by: Reis, Philipp, et al.
Published: (2026)
Similar Items
-
Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs
by: Zhou, Wei, et al.
Published: (2026) -
Graph Learning in the Era of LLMs: A Survey from the Perspective of Data, Models, and Tasks
by: Li, Xunkai, et al.
Published: (2024) -
Meta-Tuning LLMs to Leverage Lexical Knowledge for Generalizable Language Style Understanding
by: Guo, Ruohao, et al.
Published: (2023) -
CuTS: Customizable Tabular Synthetic Data Generation
by: Vero, Mark, et al.
Published: (2023) -
Enhancing Temporal Understanding in LLMs for Semi-structured Tables
by: Deng, Irwin, et al.
Published: (2024)