Scientific Inference With Interpretable Machine Learning: Analyzing Models to Learn About Real-World Phenomena

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Freiesleben, Timo, König, Gunnar, Molnar, Christoph, Tejero-Cantero, Alvaro
Format: Preprint
Published: 2022
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914870662791168
author Freiesleben, Timo
König, Gunnar
Molnar, Christoph
Tejero-Cantero, Alvaro
author_facet Freiesleben, Timo
König, Gunnar
Molnar, Christoph
Tejero-Cantero, Alvaro
contents To learn about real world phenomena, scientists have traditionally used models with clearly interpretable elements. However, modern machine learning (ML) models, while powerful predictors, lack this direct elementwise interpretability (e.g. neural network weights). Interpretable machine learning (IML) offers a solution by analyzing models holistically to derive interpretations. Yet, current IML research is focused on auditing ML models rather than leveraging them for scientific inference. Our work bridges this gap, presenting a framework for designing IML methods-termed 'property descriptors' -- that illuminate not just the model, but also the phenomenon it represents. We demonstrate that property descriptors, grounded in statistical learning theory, can effectively reveal relevant properties of the joint probability distribution of the observational data. We identify existing IML methods suited for scientific inference and provide a guide for developing new descriptors with quantified epistemic uncertainty. Our framework empowers scientists to harness ML models for inference, and provides directions for future IML research to support scientific understanding.
format Preprint
id arxiv_https___arxiv_org_abs_2206_05487
institution arXiv
publishDate 2022
record_format arxiv
spellingShingle Scientific Inference With Interpretable Machine Learning: Analyzing Models to Learn About Real-World Phenomena
Freiesleben, Timo
König, Gunnar
Molnar, Christoph
Tejero-Cantero, Alvaro
Machine Learning
To learn about real world phenomena, scientists have traditionally used models with clearly interpretable elements. However, modern machine learning (ML) models, while powerful predictors, lack this direct elementwise interpretability (e.g. neural network weights). Interpretable machine learning (IML) offers a solution by analyzing models holistically to derive interpretations. Yet, current IML research is focused on auditing ML models rather than leveraging them for scientific inference. Our work bridges this gap, presenting a framework for designing IML methods-termed 'property descriptors' -- that illuminate not just the model, but also the phenomenon it represents. We demonstrate that property descriptors, grounded in statistical learning theory, can effectively reveal relevant properties of the joint probability distribution of the observational data. We identify existing IML methods suited for scientific inference and provide a guide for developing new descriptors with quantified epistemic uncertainty. Our framework empowers scientists to harness ML models for inference, and provides directions for future IML research to support scientific understanding.
title Scientific Inference With Interpretable Machine Learning: Analyzing Models to Learn About Real-World Phenomena
topic Machine Learning
url https://arxiv.org/abs/2206.05487