Agentic-imodels: Evolving agentic interpretability tools via autoresearch

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Singh, Chandan, Tan, Yan Shuo, Xu, Weijia, Gero, Zelalem, Yang, Weiwei, Galley, Michel, Gao, Jianfeng
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910192071868416
author Singh, Chandan
Tan, Yan Shuo
Xu, Weijia
Gero, Zelalem
Yang, Weiwei
Galley, Michel
Gao, Jianfeng
author_facet Singh, Chandan
Tan, Yan Shuo
Xu, Weijia
Gero, Zelalem
Yang, Weiwei
Galley, Michel
Gao, Jianfeng
contents Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards a future where agents conduct the vast majority of data-science work. However, current ADS systems use statistical tools designed to be interpretable by humans, rather than interpretable by agents. To address this, we introduce Agentic-imodels, an agentic autoresearch loop that evolves data-science tools designed to be interpretable by agents. Specifically, it develops a library of scikit-learn-compatible regressors for tabular data that are optimized for both predictive performance and a novel LLM-based interpretability metric. The metric measures a suite of LLM-graded tests that probe whether a fitted model's string representation is "simulatable" by an LLM, i.e. whether the LLM can answer questions about the model's behavior by reading its string output alone. We find that the evolved models jointly improve predictive performance and agent-facing interpretability, generalizing to new datasets and new interpretability tests. Furthermore, these evolved models improve downstream end-to-end ADS, increasing performance for Copilot CLI, Claude Code, and Codex on the BLADE benchmark by up to 73%
format Preprint
id arxiv_https___arxiv_org_abs_2605_03808
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Agentic-imodels: Evolving agentic interpretability tools via autoresearch
Singh, Chandan
Tan, Yan Shuo
Xu, Weijia
Gero, Zelalem
Yang, Weiwei
Galley, Michel
Gao, Jianfeng
Artificial Intelligence
Computation and Language
Machine Learning
Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards a future where agents conduct the vast majority of data-science work. However, current ADS systems use statistical tools designed to be interpretable by humans, rather than interpretable by agents. To address this, we introduce Agentic-imodels, an agentic autoresearch loop that evolves data-science tools designed to be interpretable by agents. Specifically, it develops a library of scikit-learn-compatible regressors for tabular data that are optimized for both predictive performance and a novel LLM-based interpretability metric. The metric measures a suite of LLM-graded tests that probe whether a fitted model's string representation is "simulatable" by an LLM, i.e. whether the LLM can answer questions about the model's behavior by reading its string output alone. We find that the evolved models jointly improve predictive performance and agent-facing interpretability, generalizing to new datasets and new interpretability tests. Furthermore, these evolved models improve downstream end-to-end ADS, increasing performance for Copilot CLI, Claude Code, and Codex on the BLADE benchmark by up to 73%
title Agentic-imodels: Evolving agentic interpretability tools via autoresearch
topic Artificial Intelligence
Computation and Language
Machine Learning
url https://arxiv.org/abs/2605.03808