Generating Realistic Tabular Data with Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Dang, Gupta, Sunil, Do, Kien, Nguyen, Thin, Venkatesh, Svetha |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Language Models for Imbalanced Classification: Diversity makes the difference
by: Nguyen, Dang, et al.
Published: (2025)
by: Nguyen, Dang, et al.
Published: (2025)
Adaptive Acquisition Selection for Bayesian Optimization with Large Language Models
by: Ngo, Giang, et al.
Published: (2026)
by: Ngo, Giang, et al.
Published: (2026)
SPaCe: Unlocking Sample-Efficient Large Language Models Training With Self-Pace Curriculum Learning
by: Do, Dai, et al.
Published: (2025)
by: Do, Dai, et al.
Published: (2025)
Score-based Integrated Gradient for Root Cause Explanations of Outliers
by: Nguyen, Phuoc, et al.
Published: (2026)
by: Nguyen, Phuoc, et al.
Published: (2026)
Variational Flow Models: Flowing in Your Style
by: Do, Kien, et al.
Published: (2024)
by: Do, Kien, et al.
Published: (2024)
Stable Hadamard Memory: Revitalizing Memory-Augmented Agents for Reinforcement Learning
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
Novel Kernel Models and Exact Representor Theory for Neural Networks Beyond the Over-Parameterized Regime
by: Shilton, Alistair, et al.
Published: (2024)
by: Shilton, Alistair, et al.
Published: (2024)
Federated Domain Generalization with Latent Space Inversion
by: Palakkadavath, Ragja, et al.
Published: (2025)
by: Palakkadavath, Ragja, et al.
Published: (2025)
Variable-Agnostic Causal Exploration for Reinforcement Learning
by: Nguyen, Minh Hoang, et al.
Published: (2024)
by: Nguyen, Minh Hoang, et al.
Published: (2024)
Universal Multi-Domain Translation via Diffusion Routers
by: Kieu, Duc, et al.
Published: (2025)
by: Kieu, Duc, et al.
Published: (2025)
Beyond Surprise: Improving Exploration Through Surprise Novelty
by: Le, Hung, et al.
Published: (2023)
by: Le, Hung, et al.
Published: (2023)
Predicting the Reliability of an Image Classifier under Image Distortion
by: Nguyen, Dang, et al.
Published: (2024)
by: Nguyen, Dang, et al.
Published: (2024)
Bidirectional Diffusion Bridge Models
by: Kieu, Duc, et al.
Published: (2025)
by: Kieu, Duc, et al.
Published: (2025)
Uncertainty-Guided Checkpoint Selection for Reinforcement Finetuning of Large Language Models
by: Nguyen, Manh, et al.
Published: (2025)
by: Nguyen, Manh, et al.
Published: (2025)
Reasoning Under 1 Billion: Memory-Augmented Reinforcement Learning for Large Language Models
by: Le, Hung, et al.
Published: (2025)
by: Le, Hung, et al.
Published: (2025)
Enhancing Length Extrapolation in Sequential Models with Pointer-Augmented Neural Memory
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
Multi-Reference Preference Optimization for Large Language Models
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
Causal-Aware Generative Adversarial Networks with Reinforcement Learning
by: Nguyen, Tu Anh Hoang, et al.
Published: (2025)
by: Nguyen, Tu Anh Hoang, et al.
Published: (2025)
Enabling Causal Discovery in Post-Nonlinear Models with Normalizing Flows
by: Hoang, Nu, et al.
Published: (2024)
by: Hoang, Nu, et al.
Published: (2024)
Active Level Set Estimation for Continuous Search Space with Theoretical Guarantee
by: Ngo, Giang, et al.
Published: (2024)
by: Ngo, Giang, et al.
Published: (2024)
Scalable Variational Causal Discovery Unconstrained by Acyclicity
by: Hoang, Nu, et al.
Published: (2024)
by: Hoang, Nu, et al.
Published: (2024)
FASTGEN: Fast and Cost-Effective Synthetic Tabular Data Generation with LLMs
by: Nguyen, Anh, et al.
Published: (2025)
by: Nguyen, Anh, et al.
Published: (2025)
CasTGAN: Cascaded Generative Adversarial Network for Realistic Tabular Data Synthesis
by: Alshantti, Abdallah, et al.
Published: (2023)
by: Alshantti, Abdallah, et al.
Published: (2023)
Revisiting LARS for Large Batch Training Generalization of Neural Networks
by: Do, Khoi, et al.
Published: (2023)
by: Do, Khoi, et al.
Published: (2023)
Large Language Models Prompting With Episodic Memory
by: Do, Dai, et al.
Published: (2024)
by: Do, Dai, et al.
Published: (2024)
$π^2$: Structure-Originated Reasoning Data Improves Long-Context Reasoning Ability of Large Language Models
by: Do, Quyet V., et al.
Published: (2026)
by: Do, Quyet V., et al.
Published: (2026)
BSO: Safety Alignment Is Density Ratio Matching
by: Nguyen, Tien-Phat, et al.
Published: (2026)
by: Nguyen, Tien-Phat, et al.
Published: (2026)
Improving Diversity in Black-box Few-shot Knowledge Distillation
by: Vo, Tri-Nhan, et al.
Published: (2026)
by: Vo, Tri-Nhan, et al.
Published: (2026)
Finding the Trigger: Causal Abductive Reasoning on Video Events
by: Le, Thao Minh, et al.
Published: (2025)
by: Le, Thao Minh, et al.
Published: (2025)
GEM-T: Generative Tabular Data via Fitting Moments
by: Li, Miao, et al.
Published: (2025)
by: Li, Miao, et al.
Published: (2025)
A Note on Statistically Accurate Tabular Data Generation Using Large Language Models
by: Sidorenko, Andrey
Published: (2025)
by: Sidorenko, Andrey
Published: (2025)
Revisiting the Dataset Bias Problem from a Statistical Perspective
by: Do, Kien, et al.
Published: (2024)
by: Do, Kien, et al.
Published: (2024)
How Homogenizing the Channel-wise Magnitude Can Enhance EEG Classification Model?
by: Ngo, Huyen, et al.
Published: (2024)
by: Ngo, Huyen, et al.
Published: (2024)
Diverse Image Priors for Black-box Data-free Knowledge Distillation
by: Vo, Tri-Nhan, et al.
Published: (2026)
by: Vo, Tri-Nhan, et al.
Published: (2026)
Continual Fine-Tuning of Large Language Models via Program Memory
by: Le, Hung, et al.
Published: (2026)
by: Le, Hung, et al.
Published: (2026)
Learning Structural Causal Models from Ordering: Identifiable Flow Models
by: Le, Minh Khoa, et al.
Published: (2024)
by: Le, Minh Khoa, et al.
Published: (2024)
Causal Discovery via Bayesian Optimization
by: Duong, Bao, et al.
Published: (2025)
by: Duong, Bao, et al.
Published: (2025)
MALLM-GAN: Multi-Agent Large Language Model as Generative Adversarial Network for Synthesizing Tabular Data
by: Ling, Yaobin, et al.
Published: (2024)
by: Ling, Yaobin, et al.
Published: (2024)
TFM-Retouche: A Lightweight Input-Space Adapter for Tabular Foundation Models
by: Nguyen, Duong, et al.
Published: (2026)
by: Nguyen, Duong, et al.
Published: (2026)
Large Language Models Engineer Too Many Simple Features For Tabular Data
by: Küken, Jaris, et al.
Published: (2024)
by: Küken, Jaris, et al.
Published: (2024)
Similar Items
-
Large Language Models for Imbalanced Classification: Diversity makes the difference
by: Nguyen, Dang, et al.
Published: (2025) -
Adaptive Acquisition Selection for Bayesian Optimization with Large Language Models
by: Ngo, Giang, et al.
Published: (2026) -
SPaCe: Unlocking Sample-Efficient Large Language Models Training With Self-Pace Curriculum Learning
by: Do, Dai, et al.
Published: (2025) -
Score-based Integrated Gradient for Root Cause Explanations of Outliers
by: Nguyen, Phuoc, et al.
Published: (2026) -
Variational Flow Models: Flowing in Your Style
by: Do, Kien, et al.
Published: (2024)