Many AI Analysts, One Dataset: Navigating the Agentic Data Science Multiverse
Fuente:
arXiv
Saved in:
| Main Authors: | Bertran, Martin, Fogliato, Riccardo, Wu, Zhiwei Steven |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving LLM Group Fairness on Tabular Data via In-Context Learning
by: Cherepanova, Valeriia, et al.
Published: (2024)
by: Cherepanova, Valeriia, et al.
Published: (2024)
Navigating Explanatory Multiverse Through Counterfactual Path Geometry
by: Sokol, Kacper, et al.
Published: (2023)
by: Sokol, Kacper, et al.
Published: (2023)
Multicalibration for Confidence Scoring in LLMs
by: Detommaso, Gianluca, et al.
Published: (2024)
by: Detommaso, Gianluca, et al.
Published: (2024)
Stronger Neyman Regret Guarantees for Adaptive Experimental Design
by: Noarov, Georgy, et al.
Published: (2025)
by: Noarov, Georgy, et al.
Published: (2025)
Sanity Checks for Agentic Data Science
by: Rewolinski, Zachary T., et al.
Published: (2026)
by: Rewolinski, Zachary T., et al.
Published: (2026)
AIRepr: An Analyst-Inspector Framework for Evaluating Reproducibility of LLMs in Data Science
by: Zeng, Qiuhai, et al.
Published: (2025)
by: Zeng, Qiuhai, et al.
Published: (2025)
An Efficient NAS-based Approach for Handling Imbalanced Datasets
by: Yao, Zhiwei
Published: (2024)
by: Yao, Zhiwei
Published: (2024)
Navigating Dataset Documentations in AI: A Large-Scale Analysis of Dataset Cards on Hugging Face
by: Yang, Xinyu, et al.
Published: (2024)
by: Yang, Xinyu, et al.
Published: (2024)
Persona-Augmented Benchmarking: Evaluating LLMs Across Diverse Writing Styles
by: Truong, Kimberly Le, et al.
Published: (2025)
by: Truong, Kimberly Le, et al.
Published: (2025)
Multiverse: Language-Conditioned Multi-Game Level Blending via Shared Representation
by: Baek, In-Chang, et al.
Published: (2026)
by: Baek, In-Chang, et al.
Published: (2026)
Train Once, Answer All: Many Pretraining Experiments for the Cost of One
by: Bordt, Sebastian, et al.
Published: (2025)
by: Bordt, Sebastian, et al.
Published: (2025)
SciHorizon-DataEVA: An Agentic System for AI-Readiness Evaluation of Heterogeneous Scientific Data
by: Liu, Dianyu, et al.
Published: (2026)
by: Liu, Dianyu, et al.
Published: (2026)
CEDAR: Context Engineering for Agentic Data Science
by: Roy, Rishiraj Saha, et al.
Published: (2026)
by: Roy, Rishiraj Saha, et al.
Published: (2026)
Architectures for Building Agentic AI
by: Nowaczyk, Sławomir
Published: (2025)
by: Nowaczyk, Sławomir
Published: (2025)
Bridging Multicalibration and Out-of-distribution Generalization Beyond Covariate Shift
by: Wu, Jiayun, et al.
Published: (2024)
by: Wu, Jiayun, et al.
Published: (2024)
ADAPTive Input Training for Many-to-One Pre-Training on Time-Series Classification
by: Quinlan, Paul, et al.
Published: (2026)
by: Quinlan, Paul, et al.
Published: (2026)
Multi-group Uncertainty Quantification for Long-form Text Generation
by: Liu, Terrance, et al.
Published: (2024)
by: Liu, Terrance, et al.
Published: (2024)
Leveraging Model Guidance to Extract Training Data from Personalized Diffusion Models
by: Wu, Xiaoyu, et al.
Published: (2024)
by: Wu, Xiaoyu, et al.
Published: (2024)
Agentics 2.0: Logical Transduction Algebra for Agentic Data Workflows
by: Gliozzo, Alfio Massimiliano, et al.
Published: (2026)
by: Gliozzo, Alfio Massimiliano, et al.
Published: (2026)
Reconciling Model Multiplicity for Downstream Decision Making
by: Du, Ally Yalei, et al.
Published: (2024)
by: Du, Ally Yalei, et al.
Published: (2024)
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
by: Wu, Xiaoyu, et al.
Published: (2025)
by: Wu, Xiaoyu, et al.
Published: (2025)
Agentic AI for Mobile Network RAN Management and Optimization
by: Pellejero, Jorge, et al.
Published: (2025)
by: Pellejero, Jorge, et al.
Published: (2025)
Agentic Web: Weaving the Next Web with AI Agents
by: Yang, Yingxuan, et al.
Published: (2025)
by: Yang, Yingxuan, et al.
Published: (2025)
PCS Workflow for Veridical Data Science in the Age of AI
by: Rewolinski, Zachary T., et al.
Published: (2025)
by: Rewolinski, Zachary T., et al.
Published: (2025)
ML Compass: Navigating Capability, Cost, and Compliance Trade-offs in AI Model Deployment
by: Digalakis Jr, Vassilis, et al.
Published: (2025)
by: Digalakis Jr, Vassilis, et al.
Published: (2025)
AISSISTANT: Human-AI Collaborative Review and Perspective Research Workflows in Data Science
by: Gaddipati, Sasi Kiran, et al.
Published: (2025)
by: Gaddipati, Sasi Kiran, et al.
Published: (2025)
Data Interpreter: An LLM Agent For Data Science
by: Hong, Sirui, et al.
Published: (2024)
by: Hong, Sirui, et al.
Published: (2024)
STRIDE: A Systematic Framework for Selecting AI Modalities -- Agentic AI, AI Assistants, or LLM Calls
by: Asthana, Shubhi, et al.
Published: (2025)
by: Asthana, Shubhi, et al.
Published: (2025)
Co-Investigator AI: The Rise of Agentic AI for Smarter, Trustworthy AML Compliance Narratives
by: Naik, Prathamesh Vasudeo, et al.
Published: (2025)
by: Naik, Prathamesh Vasudeo, et al.
Published: (2025)
MarkTune: Improving the Quality-Detectability Trade-off in Open-Weight LLM Watermarking
by: Zhao, Yizhou, et al.
Published: (2025)
by: Zhao, Yizhou, et al.
Published: (2025)
MedAgentGym: A Scalable Agentic Training Environment for Code-Centric Reasoning in Biomedical Data Science
by: Xu, Ran, et al.
Published: (2025)
by: Xu, Ran, et al.
Published: (2025)
ADR: An Agentic Detection System for Enterprise Agentic AI Security
by: Li, Chenning, et al.
Published: (2026)
by: Li, Chenning, et al.
Published: (2026)
Effective Resistance Rewiring: A Simple Topological Correction for Over-Squashing
by: Miquel-Oliver, Bertran, et al.
Published: (2026)
by: Miquel-Oliver, Bertran, et al.
Published: (2026)
Deep Classifier Mimicry without Data Access
by: Braun, Steven, et al.
Published: (2023)
by: Braun, Steven, et al.
Published: (2023)
Frozen Layers: Memory-efficient Many-fidelity Hyperparameter Optimization
by: Carstensen, Timur, et al.
Published: (2025)
by: Carstensen, Timur, et al.
Published: (2025)
AgenticCache: Cache-Driven Asynchronous Planning for Embodied AI Agents
by: Kim, Hojoon, et al.
Published: (2026)
by: Kim, Hojoon, et al.
Published: (2026)
GCA Framework: A GCC Countries-Grounded Dataset and Agentic Pipeline for Climate Decision Support
by: Sheikh, Muhammad Umer, et al.
Published: (2026)
by: Sheikh, Muhammad Umer, et al.
Published: (2026)
Large Language Models Engineer Too Many Simple Features For Tabular Data
by: Küken, Jaris, et al.
Published: (2024)
by: Küken, Jaris, et al.
Published: (2024)
Verifiable Agentic Infrastructure: Proof-Derived Authorization for Sovereign AI Systems
by: He, Jun, et al.
Published: (2026)
by: He, Jun, et al.
Published: (2026)
SmartFlow Reinforcement Learning and Agentic AI for Bike-Sharing Optimisation
by: K, Aditya Sreevatsa, et al.
Published: (2025)
by: K, Aditya Sreevatsa, et al.
Published: (2025)
Similar Items
-
Improving LLM Group Fairness on Tabular Data via In-Context Learning
by: Cherepanova, Valeriia, et al.
Published: (2024) -
Navigating Explanatory Multiverse Through Counterfactual Path Geometry
by: Sokol, Kacper, et al.
Published: (2023) -
Multicalibration for Confidence Scoring in LLMs
by: Detommaso, Gianluca, et al.
Published: (2024) -
Stronger Neyman Regret Guarantees for Adaptive Experimental Design
by: Noarov, Georgy, et al.
Published: (2025) -
Sanity Checks for Agentic Data Science
by: Rewolinski, Zachary T., et al.
Published: (2026)