FeatNavigator: Automatic Feature Augmentation on Tabular Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Liang, Jiaming, Lei, Chuan, Qin, Xiao, Zhang, Jiani, Katsifodimos, Asterios, Faloutsos, Christos, Rangwala, Huzefa |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LAKEGEN: A LLM-based Tabular Corpus Generator for Evaluating Dataset Discovery in Data Lakes
por: Dai, Zhenwei, et al.
Publicado: (2025)
por: Dai, Zhenwei, et al.
Publicado: (2025)
OmniMatch: Effective Self-Supervised Any-Join Discovery in Tabular Data Repositories
por: Koutras, Christos, et al.
Publicado: (2024)
por: Koutras, Christos, et al.
Publicado: (2024)
Mixed-Type Tabular Data Synthesis with Score-based Diffusion in Latent Space
por: Zhang, Hengrui, et al.
Publicado: (2023)
por: Zhang, Hengrui, et al.
Publicado: (2023)
FeatAug: Automatic Feature Augmentation From One-to-Many Relationship Tables
por: Qi, Danrui, et al.
Publicado: (2024)
por: Qi, Danrui, et al.
Publicado: (2024)
CoddLLM: Empowering Large Language Models for Data Analytics
por: Zhang, Jiani, et al.
Publicado: (2025)
por: Zhang, Jiani, et al.
Publicado: (2025)
OpenTab: Advancing Large Language Models as Open-domain Table Reasoners
por: Kong, Kezhi, et al.
Publicado: (2024)
por: Kong, Kezhi, et al.
Publicado: (2024)
SiMa: Effective and Efficient Matching Across Data Silos Using Graph Neural Networks
por: Koutras, Christos, et al.
Publicado: (2022)
por: Koutras, Christos, et al.
Publicado: (2022)
Accelerating Machine Learning Queries with Linear Algebra Query Processing
por: Sun, Wenbo, et al.
Publicado: (2023)
por: Sun, Wenbo, et al.
Publicado: (2023)
FeatInsight: An Online ML Feature Management System on 4Paradigm Sage-Studio Platform
por: Tong, Xin, et al.
Publicado: (2025)
por: Tong, Xin, et al.
Publicado: (2025)
Hippasus: Effective and Efficient Automatic Feature Augmentation for Machine Learning Tasks on Relational Data
por: Papadias, Serafeim, et al.
Publicado: (2026)
por: Papadias, Serafeim, et al.
Publicado: (2026)
4DBInfer: A 4D Benchmarking Toolbox for Graph-Centric Predictive Modeling on Relational DBs
por: Wang, Minjie, et al.
Publicado: (2024)
por: Wang, Minjie, et al.
Publicado: (2024)
Hierarchical Conditional Tabular GAN for Multi-Tabular Synthetic Data Generation
por: Ågren, Wilhelm, et al.
Publicado: (2024)
por: Ågren, Wilhelm, et al.
Publicado: (2024)
Exploring the Heterogeneity of Tabular Data: A Diversity-aware Data Generator via LLMs
por: Tang, Yafeng, et al.
Publicado: (2025)
por: Tang, Yafeng, et al.
Publicado: (2025)
Tabular Data Augmentation for Machine Learning: Progress and Prospects of Embracing Generative AI
por: Cui, Lingxi, et al.
Publicado: (2024)
por: Cui, Lingxi, et al.
Publicado: (2024)
Approximate Nearest Neighbor Search for Modern AI: A Projection-Augmented Graph Approach
por: Lu, Kejing, et al.
Publicado: (2026)
por: Lu, Kejing, et al.
Publicado: (2026)
TabPrep: Closing the Feature Engineering Gap in Tabular Benchmarks
por: Tschalzev, Andrej, et al.
Publicado: (2026)
por: Tschalzev, Andrej, et al.
Publicado: (2026)
RFOD: Random Forest-based Outlier Detection for Tabular Data
por: Ang, Yihao, et al.
Publicado: (2025)
por: Ang, Yihao, et al.
Publicado: (2025)
HybGRAG: Hybrid Retrieval-Augmented Generation on Textual and Relational Knowledge Bases
por: Lee, Meng-Chieh, et al.
Publicado: (2024)
por: Lee, Meng-Chieh, et al.
Publicado: (2024)
TableGPT2: A Large Multimodal Model with Tabular Data Integration
por: Su, Aofeng, et al.
Publicado: (2024)
por: Su, Aofeng, et al.
Publicado: (2024)
Towards Universal Tabular Embeddings: A Benchmark Across Data Tasks
por: Vogel, Liane, et al.
Publicado: (2026)
por: Vogel, Liane, et al.
Publicado: (2026)
Transactional Cloud Applications: Status Quo, Challenges, and Opportunities
por: Laigner, Rodrigo, et al.
Publicado: (2025)
por: Laigner, Rodrigo, et al.
Publicado: (2025)
Auto-FP: An Experimental Study of Automated Feature Preprocessing for Tabular Data
por: Qi, Danrui, et al.
Publicado: (2023)
por: Qi, Danrui, et al.
Publicado: (2023)
SemPipes -- Optimizable Semantic Data Operators for Tabular Machine Learning Pipelines
por: Ovcharenko, Olga, et al.
Publicado: (2026)
por: Ovcharenko, Olga, et al.
Publicado: (2026)
Systematic Assessment of Tabular Data Synthesis
por: Du, Yuntao, et al.
Publicado: (2024)
por: Du, Yuntao, et al.
Publicado: (2024)
C$^{2}$TC: A Training-Free Framework for Efficient Tabular Data Condensation
por: Xu, Sijia, et al.
Publicado: (2026)
por: Xu, Sijia, et al.
Publicado: (2026)
GAN-based Tabular Data Generator for Constructing Synopsis in Approximate Query Processing: Challenges and Solutions
por: Fallahian, Mohammadali, et al.
Publicado: (2022)
por: Fallahian, Mohammadali, et al.
Publicado: (2022)
Graph-Based Feature Augmentation for Predictive Tasks on Relational Datasets
por: Qiao, Lianpeng, et al.
Publicado: (2025)
por: Qiao, Lianpeng, et al.
Publicado: (2025)
Mitra: Mixed Synthetic Priors for Enhancing Tabular Foundation Models
por: Zhang, Xiyuan, et al.
Publicado: (2025)
por: Zhang, Xiyuan, et al.
Publicado: (2025)
CuTS: Customizable Tabular Synthetic Data Generation
por: Vero, Mark, et al.
Publicado: (2023)
por: Vero, Mark, et al.
Publicado: (2023)
Stateful Entities: Object-oriented Cloud Applications as Distributed Dataflows
por: Psarakis, Kyriakos, et al.
Publicado: (2021)
por: Psarakis, Kyriakos, et al.
Publicado: (2021)
Styx: Transactional Stateful Functions on Streaming Dataflows
por: Psarakis, Kyriakos, et al.
Publicado: (2023)
por: Psarakis, Kyriakos, et al.
Publicado: (2023)
LLM-PQA: LLM-enhanced Prediction Query Answering
por: Li, Ziyu, et al.
Publicado: (2024)
por: Li, Ziyu, et al.
Publicado: (2024)
Towards Pattern-aware Data Augmentation for Temporal Knowledge Graph Completion
por: Zhang, Jiasheng, et al.
Publicado: (2024)
por: Zhang, Jiasheng, et al.
Publicado: (2024)
TabularMark: Watermarking Tabular Datasets for Machine Learning
por: Zheng, Yihao, et al.
Publicado: (2024)
por: Zheng, Yihao, et al.
Publicado: (2024)
Towards Synthesizing High-Dimensional Tabular Data with Limited Samples
por: Li, Zuqing, et al.
Publicado: (2025)
por: Li, Zuqing, et al.
Publicado: (2025)
Robust Detection of Synthetic Tabular Data under Schema Variability
por: Kindji, G. Charbel N., et al.
Publicado: (2025)
por: Kindji, G. Charbel N., et al.
Publicado: (2025)
DemoTuner: Automatic Performance Tuning for Database Management Systems Based on Demonstration Reinforcement Learning
por: Dou, Hui, et al.
Publicado: (2025)
por: Dou, Hui, et al.
Publicado: (2025)
Cross-table Synthetic Tabular Data Detection
por: Kindji, G. Charbel N., et al.
Publicado: (2024)
por: Kindji, G. Charbel N., et al.
Publicado: (2024)
PyTorch Frame: A Modular Framework for Multi-Modal Tabular Learning
por: Hu, Weihua, et al.
Publicado: (2024)
por: Hu, Weihua, et al.
Publicado: (2024)
DiffImpute: Tabular Data Imputation With Denoising Diffusion Probabilistic Model
por: Wen, Yizhu, et al.
Publicado: (2024)
por: Wen, Yizhu, et al.
Publicado: (2024)
Ejemplares similares
-
LAKEGEN: A LLM-based Tabular Corpus Generator for Evaluating Dataset Discovery in Data Lakes
por: Dai, Zhenwei, et al.
Publicado: (2025) -
OmniMatch: Effective Self-Supervised Any-Join Discovery in Tabular Data Repositories
por: Koutras, Christos, et al.
Publicado: (2024) -
Mixed-Type Tabular Data Synthesis with Score-based Diffusion in Latent Space
por: Zhang, Hengrui, et al.
Publicado: (2023) -
FeatAug: Automatic Feature Augmentation From One-to-Many Relationship Tables
por: Qi, Danrui, et al.
Publicado: (2024) -
CoddLLM: Empowering Large Language Models for Data Analytics
por: Zhang, Jiani, et al.
Publicado: (2025)