Towards Human-Guided, Data-Centric LLM Co-Pilots
Fuente:
arXiv
Saved in:
| Main Authors: | Saveliev, Evgeny, Liu, Jiashuo, Seedat, Nabeel, Boyd, Anders, van der Schaar, Mihaela |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What's the next frontier for Data-centric AI? Data Savvy Agents
by: Seedat, Nabeel, et al.
Published: (2025)
by: Seedat, Nabeel, et al.
Published: (2025)
Dissecting Sample Hardness: A Fine-Grained Analysis of Hardness Characterization Methods for Data-Centric AI
by: Seedat, Nabeel, et al.
Published: (2024)
by: Seedat, Nabeel, et al.
Published: (2024)
Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search with Granular Feedback
by: Saveliev, Evgeny S., et al.
Published: (2026)
by: Saveliev, Evgeny S., et al.
Published: (2026)
Matchmaker: Self-Improving Large Language Model Programs for Schema Matching
by: Seedat, Nabeel, et al.
Published: (2024)
by: Seedat, Nabeel, et al.
Published: (2024)
DC-Check: A Data-Centric AI checklist to guide the development of reliable machine learning systems
by: Seedat, Nabeel, et al.
Published: (2022)
by: Seedat, Nabeel, et al.
Published: (2022)
Curated LLM: Synergy of LLMs and Data Curation for tabular augmentation in low-data regimes
by: Seedat, Nabeel, et al.
Published: (2023)
by: Seedat, Nabeel, et al.
Published: (2023)
When is Off-Policy Evaluation (Reward Modeling) Useful in Contextual Bandits? A Data-Centric Perspective
by: Sun, Hao, et al.
Published: (2023)
by: Sun, Hao, et al.
Published: (2023)
You can't handle the (dirty) truth: Data-centric insights improve pseudo-labeling
by: Seedat, Nabeel, et al.
Published: (2024)
by: Seedat, Nabeel, et al.
Published: (2024)
Large Language Models to Enhance Bayesian Optimization
by: Liu, Tennison, et al.
Published: (2024)
by: Liu, Tennison, et al.
Published: (2024)
Relaxed Quantile Regression: Prediction Intervals for Asymmetric Noise
by: Pouplin, Thomas, et al.
Published: (2024)
by: Pouplin, Thomas, et al.
Published: (2024)
Self-Healing Machine Learning: A Framework for Autonomous Adaptation in Real-World Environments
by: Rauba, Paulius, et al.
Published: (2024)
by: Rauba, Paulius, et al.
Published: (2024)
Context-Aware Testing: A New Paradigm for Model Testing with Large Language Models
by: Rauba, Paulius, et al.
Published: (2024)
by: Rauba, Paulius, et al.
Published: (2024)
DAGnosis: Localized Identification of Data Inconsistencies using Structures
by: Huynh, Nicolas, et al.
Published: (2024)
by: Huynh, Nicolas, et al.
Published: (2024)
Unlocking Historical Clinical Trial Data with ALIGN: A Compositional Large Language Model System for Medical Coding
by: Seedat, Nabeel, et al.
Published: (2024)
by: Seedat, Nabeel, et al.
Published: (2024)
Towards Automated Knowledge Integration From Human-Interpretable Representations
by: Kobalczyk, Katarzyna, et al.
Published: (2024)
by: Kobalczyk, Katarzyna, et al.
Published: (2024)
Knowledge-Informed Kernel State Reconstruction from Heterogeneous Partial Observations
by: Muscarnera, Luca, et al.
Published: (2026)
by: Muscarnera, Luca, et al.
Published: (2026)
CliMB: An AI-enabled Partner for Clinical Predictive Modeling
by: Saveliev, Evgeny, et al.
Published: (2024)
by: Saveliev, Evgeny, et al.
Published: (2024)
Technical Report: Facilitating the Adoption of Causal Inference Methods Through LLM-Empowered Co-Pilot
by: Berrevoets, Jeroen, et al.
Published: (2025)
by: Berrevoets, Jeroen, et al.
Published: (2025)
Nonparametric LLM Evaluation from Preference Data
by: Frauen, Dennis, et al.
Published: (2026)
by: Frauen, Dennis, et al.
Published: (2026)
Deep Hierarchical Learning with Nested Subspace Networks for Large Language Models
by: Rauba, Paulius, et al.
Published: (2025)
by: Rauba, Paulius, et al.
Published: (2025)
No Equations Needed: Learning System Dynamics Without Relying on Closed-Form ODEs
by: Kacprzyk, Krzysztof, et al.
Published: (2025)
by: Kacprzyk, Krzysztof, et al.
Published: (2025)
Shape Arithmetic Expressions: Advancing Scientific Discovery Beyond Closed-Form Equations
by: Kacprzyk, Krzysztof, et al.
Published: (2024)
by: Kacprzyk, Krzysztof, et al.
Published: (2024)
Why Tabular Foundation Models Should Be a Research Priority
by: van Breugel, Boris, et al.
Published: (2024)
by: van Breugel, Boris, et al.
Published: (2024)
Not All Explanations for Deep Learning Phenomena Are Equally Valuable
by: Jeffares, Alan, et al.
Published: (2025)
by: Jeffares, Alan, et al.
Published: (2025)
Preference Learning for AI Alignment: a Causal Perspective
by: Kobalczyk, Katarzyna, et al.
Published: (2025)
by: Kobalczyk, Katarzyna, et al.
Published: (2025)
Active Timepoint Selection for Learning Measure-Valued Trajectories
by: Huynh, Nicolas, et al.
Published: (2026)
by: Huynh, Nicolas, et al.
Published: (2026)
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
by: Sun, Hao, et al.
Published: (2024)
by: Sun, Hao, et al.
Published: (2024)
Hyperparameter Trajectory Inference with Conditional Lagrangian Optimal Transport
by: Amad, Harry, et al.
Published: (2026)
by: Amad, Harry, et al.
Published: (2026)
LaTable: Towards Large Tabular Models
by: van Breugel, Boris, et al.
Published: (2024)
by: van Breugel, Boris, et al.
Published: (2024)
Automatic Construction of Clinical Scoring Systems with LLM Agents
by: Estévez, Silas Ruhrberg, et al.
Published: (2026)
by: Estévez, Silas Ruhrberg, et al.
Published: (2026)
Decision Tree Induction Through LLMs via Semantically-Aware Evolution
by: Liu, Tennison, et al.
Published: (2025)
by: Liu, Tennison, et al.
Published: (2025)
Automatically Learning Hybrid Digital Twins of Dynamical Systems
by: Holt, Samuel, et al.
Published: (2024)
by: Holt, Samuel, et al.
Published: (2024)
Discovery of Hidden Miscalibration Regimes
by: Kobalczyk, Katarzyna, et al.
Published: (2026)
by: Kobalczyk, Katarzyna, et al.
Published: (2026)
Language Bottleneck Models for Qualitative Knowledge State Modeling
by: Berthon, Antonin, et al.
Published: (2025)
by: Berthon, Antonin, et al.
Published: (2025)
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities
by: Sun, Hao, et al.
Published: (2025)
by: Sun, Hao, et al.
Published: (2025)
On Error Propagation of Diffusion Models
by: Li, Yangming, et al.
Published: (2023)
by: Li, Yangming, et al.
Published: (2023)
Dense Reward for Free in Reinforcement Learning from Human Feedback
by: Chan, Alex J., et al.
Published: (2024)
by: Chan, Alex J., et al.
Published: (2024)
A Study of Posterior Stability for Time-Series Latent Diffusion
by: Li, Yangming, et al.
Published: (2024)
by: Li, Yangming, et al.
Published: (2024)
Why do Random Forests Work? Understanding Tree Ensembles as Self-Regularizing Adaptive Smoothers
by: Curth, Alicia, et al.
Published: (2024)
by: Curth, Alicia, et al.
Published: (2024)
Tiny Autoregressive Recursive Models
by: Rauba, Paulius, et al.
Published: (2026)
by: Rauba, Paulius, et al.
Published: (2026)
Similar Items
-
What's the next frontier for Data-centric AI? Data Savvy Agents
by: Seedat, Nabeel, et al.
Published: (2025) -
Dissecting Sample Hardness: A Fine-Grained Analysis of Hardness Characterization Methods for Data-Centric AI
by: Seedat, Nabeel, et al.
Published: (2024) -
Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search with Granular Feedback
by: Saveliev, Evgeny S., et al.
Published: (2026) -
Matchmaker: Self-Improving Large Language Model Programs for Schema Matching
by: Seedat, Nabeel, et al.
Published: (2024) -
DC-Check: A Data-Centric AI checklist to guide the development of reliable machine learning systems
by: Seedat, Nabeel, et al.
Published: (2022)