Identifiability Guarantees for Causal Disentanglement from Purely Observational Data

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Welch, Ryan, Zhang, Jiaqi, Uhler, Caroline
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910760915959808
author Welch, Ryan
Zhang, Jiaqi
Uhler, Caroline
author_facet Welch, Ryan
Zhang, Jiaqi
Uhler, Caroline
contents Causal disentanglement aims to learn about latent causal factors behind data, holding the promise to augment existing representation learning methods in terms of interpretability and extrapolation. Recent advances establish identifiability results assuming that interventions on (single) latent factors are available; however, it remains debatable whether such assumptions are reasonable due to the inherent nature of intervening on latent variables. Accordingly, we reconsider the fundamentals and ask what can be learned using just observational data. We provide a precise characterization of latent factors that can be identified in nonlinear causal models with additive Gaussian noise and linear mixing, without any interventions or graphical restrictions. In particular, we show that the causal variables can be identified up to a layer-wise transformation and that further disentanglement is not possible. We transform these theoretical results into a practical algorithm consisting of solving a quadratic program over the score estimation of the observed data. We provide simulation results to support our theoretical guarantees and demonstrate that our algorithm can derive meaningful causal representations from purely observational data.
format Preprint
id arxiv_https___arxiv_org_abs_2410_23620
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Identifiability Guarantees for Causal Disentanglement from Purely Observational Data
Welch, Ryan
Zhang, Jiaqi
Uhler, Caroline
Machine Learning
Causal disentanglement aims to learn about latent causal factors behind data, holding the promise to augment existing representation learning methods in terms of interpretability and extrapolation. Recent advances establish identifiability results assuming that interventions on (single) latent factors are available; however, it remains debatable whether such assumptions are reasonable due to the inherent nature of intervening on latent variables. Accordingly, we reconsider the fundamentals and ask what can be learned using just observational data. We provide a precise characterization of latent factors that can be identified in nonlinear causal models with additive Gaussian noise and linear mixing, without any interventions or graphical restrictions. In particular, we show that the causal variables can be identified up to a layer-wise transformation and that further disentanglement is not possible. We transform these theoretical results into a practical algorithm consisting of solving a quadratic program over the score estimation of the observed data. We provide simulation results to support our theoretical guarantees and demonstrate that our algorithm can derive meaningful causal representations from purely observational data.
title Identifiability Guarantees for Causal Disentanglement from Purely Observational Data
topic Machine Learning
url https://arxiv.org/abs/2410.23620