Utilizing Training Data to Improve LLM Reasoning for Tabular Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Chufan, Chen, Jintai, Sun, Jimeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MediTab: Scaling Medical Tabular Data Predictors via Data Consolidation, Enrichment, and Refinement
by: Wang, Zifeng, et al.
Published: (2023)
by: Wang, Zifeng, et al.
Published: (2023)
Cross-Table Pretraining towards a Universal Function Space for Heterogeneous Tabular Data
by: Chen, Jintai, et al.
Published: (2024)
by: Chen, Jintai, et al.
Published: (2024)
Small Models are LLM Knowledge Triggers on Medical Tabular Prediction
by: Yan, Jiahuan, et al.
Published: (2024)
by: Yan, Jiahuan, et al.
Published: (2024)
Language Interaction Network for Clinical Trial Approval Estimation
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
TrialSynth: Generation of Synthetic Sequential Clinical Trial Data
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
Reinforcing Numerical Reasoning in LLMs for Tabular Prediction via Structural Priors
by: Cai, Pengxiang, et al.
Published: (2025)
by: Cai, Pengxiang, et al.
Published: (2025)
Making Pre-trained Language Models Great on Tabular Prediction
by: Yan, Jiahuan, et al.
Published: (2024)
by: Yan, Jiahuan, et al.
Published: (2024)
UniPredict: Large Language Models are Universal Tabular Classifiers
by: Wang, Ruiyu, et al.
Published: (2023)
by: Wang, Ruiyu, et al.
Published: (2023)
Automatically Labeling Clinical Trial Outcomes: A Large-Scale Benchmark for Drug Development
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
ExcelFormer: A neural network surpassing GBDTs on tabular data
by: Chen, Jintai, et al.
Published: (2023)
by: Chen, Jintai, et al.
Published: (2023)
TabReason: A Reinforcement Learning-Enhanced Reasoning LLM for Explainable Tabular Data Prediction
by: Xu, Tommy, et al.
Published: (2025)
by: Xu, Tommy, et al.
Published: (2025)
Team up GBDTs and DNNs: Advancing Efficient and Effective Tabular Prediction with Tree-hybrid MLPs
by: Yan, Jiahuan, et al.
Published: (2024)
by: Yan, Jiahuan, et al.
Published: (2024)
Signal Quality Auditing for Time-series Data
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
From Residuals to Reasons: LLM-Guided Mechanism Inference from Tabular Data
by: Rezaei, Mohammad R., et al.
Published: (2026)
by: Rezaei, Mohammad R., et al.
Published: (2026)
Understanding and Mitigating Memorization in Diffusion Models for Tabular Data
by: Fang, Zhengyu, et al.
Published: (2024)
by: Fang, Zhengyu, et al.
Published: (2024)
Improving LLM Group Fairness on Tabular Data via In-Context Learning
by: Cherepanova, Valeriia, et al.
Published: (2024)
by: Cherepanova, Valeriia, et al.
Published: (2024)
Structural Reasoning Improves Molecular Understanding of LLM
by: Jang, Yunhui, et al.
Published: (2024)
by: Jang, Yunhui, et al.
Published: (2024)
MM-DADM: Multimodal Drug-Aware Diffusion Model for Virtual Clinical Trials
by: Shao, Qian, et al.
Published: (2025)
by: Shao, Qian, et al.
Published: (2025)
LLM Meeting Decision Trees on Tabular Data
by: Ye, Hangting, et al.
Published: (2025)
by: Ye, Hangting, et al.
Published: (2025)
Synthetic Patient-Physician Dialogue Generation from Clinical Notes Using LLM
by: Das, Trisha, et al.
Published: (2024)
by: Das, Trisha, et al.
Published: (2024)
MALT: Improving Reasoning with Multi-Agent LLM Training
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
ClinicalAgent: Clinical Trial Multi-Agent System with Large Language Model-based Reasoning
by: Yue, Ling, et al.
Published: (2024)
by: Yue, Ling, et al.
Published: (2024)
Accurate, Efficient, and Explainable Deep Learning Approaches for Environmental Science Problems
by: Shi, Jimeng
Published: (2026)
by: Shi, Jimeng
Published: (2026)
LLM Embeddings for Deep Learning on Tabular Data
by: Koloski, Boshko, et al.
Published: (2025)
by: Koloski, Boshko, et al.
Published: (2025)
Self-Supervision Improves Diffusion Models for Tabular Data Imputation
by: Liu, Yixin, et al.
Published: (2024)
by: Liu, Yixin, et al.
Published: (2024)
In-Context Bias Propagation in LLM-Based Tabular Data Generation
by: Recasens, Pol G., et al.
Published: (2025)
by: Recasens, Pol G., et al.
Published: (2025)
Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data
by: Yakushev, George, et al.
Published: (2025)
by: Yakushev, George, et al.
Published: (2025)
Rethinking Pre-Training in Tabular Data: A Neighborhood Embedding Perspective
by: Ye, Han-Jia, et al.
Published: (2023)
by: Ye, Han-Jia, et al.
Published: (2023)
Understanding the Effect of Noise in LLM Training Data with Algorithmic Chains of Thought
by: Havrilla, Alex, et al.
Published: (2024)
by: Havrilla, Alex, et al.
Published: (2024)
Improving Deep Tabular Learning
by: Sarafian, Sivan, et al.
Published: (2025)
by: Sarafian, Sivan, et al.
Published: (2025)
Social Determinants of Health Prediction for ICD-9 Code with Reasoning Models
by: Khan, Sharim, et al.
Published: (2025)
by: Khan, Sharim, et al.
Published: (2025)
Understanding Silent Data Corruption in LLM Training
by: Ma, Jeffrey, et al.
Published: (2025)
by: Ma, Jeffrey, et al.
Published: (2025)
Evaluating LLM Understanding via Structured Tabular Decision Simulations
by: Li, Sichao, et al.
Published: (2025)
by: Li, Sichao, et al.
Published: (2025)
Tabular Data Adapters: Improving Outlier Detection for Unlabeled Private Data
by: Herurkar, Dayananda, et al.
Published: (2025)
by: Herurkar, Dayananda, et al.
Published: (2025)
CAST: Cluster-Aware Self-Training for Tabular Data via Reliable Confidence
by: Kim, Minwook, et al.
Published: (2023)
by: Kim, Minwook, et al.
Published: (2023)
CACTI: Leveraging Copy Masking and Contextual Information to Improve Tabular Data Imputation
by: Gorla, Aditya, et al.
Published: (2025)
by: Gorla, Aditya, et al.
Published: (2025)
TTM-RE: Memory-Augmented Document-Level Relation Extraction
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
LLM Empowered Prototype Learning for Zero and Few-Shot Tasks on Tabular Data
by: Wang, Peng, et al.
Published: (2025)
by: Wang, Peng, et al.
Published: (2025)
Measuring LLM Sensitivity in Transformer-based Tabular Data Synthesis
by: R, Maria F. Davila, et al.
Published: (2025)
by: R, Maria F. Davila, et al.
Published: (2025)
TAGAL: Tabular Data Generation using Agentic LLM Methods
by: Ronval, Benoît, et al.
Published: (2025)
by: Ronval, Benoît, et al.
Published: (2025)
Similar Items
-
MediTab: Scaling Medical Tabular Data Predictors via Data Consolidation, Enrichment, and Refinement
by: Wang, Zifeng, et al.
Published: (2023) -
Cross-Table Pretraining towards a Universal Function Space for Heterogeneous Tabular Data
by: Chen, Jintai, et al.
Published: (2024) -
Small Models are LLM Knowledge Triggers on Medical Tabular Prediction
by: Yan, Jiahuan, et al.
Published: (2024) -
Language Interaction Network for Clinical Trial Approval Estimation
by: Gao, Chufan, et al.
Published: (2024) -
TrialSynth: Generation of Synthetic Sequential Clinical Trial Data
by: Gao, Chufan, et al.
Published: (2024)