ME LGOct 22, 2023

Shortcuts for causal discovery of nonlinear models by score matching

Francesco Montagna, Nicoletta Noceti, Lorenzo Rosasco, Francesco Locatello

arXiv:2310.14246v15.14 citationsh-index: 18

Originality Incremental advance

AI Analysis

This work addresses causal discovery for researchers in machine learning and statistics, but it is incremental as it builds on existing methods and focuses on specific patterns in synthetic data.

The paper tackled the problem of causal discovery in nonlinear models by analyzing score-sortability patterns, showing that these patterns define an identifiable class of bivariate causal models overlapping with nonlinear additive noise models. It demonstrated ScoreSort's statistical efficiency advantages over prior methods and highlighted limitations in current synthetic benchmarks.

The use of simulated data in the field of causal discovery is ubiquitous due to the scarcity of annotated real data. Recently, Reisach et al., 2021 highlighted the emergence of patterns in simulated linear data, which displays increasing marginal variance in the casual direction. As an ablation in their experiments, Montagna et al., 2023 found that similar patterns may emerge in nonlinear models for the variance of the score vector $\nabla \log p_{\mathbf{X}}$, and introduced the ScoreSort algorithm. In this work, we formally define and characterize this score-sortability pattern of nonlinear additive noise models. We find that it defines a class of identifiable (bivariate) causal models overlapping with nonlinear additive noise models. We theoretically demonstrate the advantages of ScoreSort in terms of statistical efficiency compared to prior state-of-the-art score matching-based methods and empirically show the score-sortability of the most common synthetic benchmarks in the literature. Our findings remark (1) the lack of diversity in the data as an important limitation in the evaluation of nonlinear causal discovery approaches, (2) the importance of thoroughly testing different settings within a problem class, and (3) the importance of analyzing statistical properties in causal discovery, where research is often limited to defining identifiability conditions of the model.

View on arXiv PDF

Similar