Search | arXiv e-print repository

Cause-effect inference through spectral independence in linear dynamical systems: theoretical foundations

Authors: Michel Besserve, Naji Shajarisales, Dominik Janzing, Bernhard Schölkopf

Abstract: Distinguishing between cause and effect using time series observational data is a major challenge in many scientific fields. A new perspective has been provided based on the principle of Independence of Causal Mechanisms (ICM), leading to the Spectral Independence Criterion (SIC), postulating that the power spectral density (PSD) of the cause time series is uncorrelated with the squared modulus of… ▽ More Distinguishing between cause and effect using time series observational data is a major challenge in many scientific fields. A new perspective has been provided based on the principle of Independence of Causal Mechanisms (ICM), leading to the Spectral Independence Criterion (SIC), postulating that the power spectral density (PSD) of the cause time series is uncorrelated with the squared modulus of the frequency response of the filter generating the effect. Since SIC rests on methods and assumptions in stark contrast with most causal discovery methods for time series, it raises questions regarding what theoretical grounds justify its use. In this paper, we provide answers covering several key aspects. After providing an information theoretic interpretation of SIC, we present an identifiability result that sheds light on the context for which this approach is expected to perform well. We further demonstrate the robustness of SIC to downsampling - an obstacle that can spoil Granger-based inference. Finally, an invariance perspective allows to explore the limitations of the spectral independence assumption and how to generalize it. Overall, these results support the postulate of Spectral Independence is a well grounded leading principle for causal inference based on empirical time series. △ Less

Submitted 29 October, 2021; originally announced October 2021.

arXiv:2003.01067 [pdf, ps, other]

Learning from Positive and Unlabeled Data by Identifying the Annotation Process

Authors: Naji Shajarisales, Peter Spirtes, Kun Zhang

Abstract: In binary classification, Learning from Positive and Unlabeled data (LePU) is semi-supervised learning but with labeled elements from only one class. Most of the research on LePU relies on some form of independence between the selection process of annotated examples and the features of the annotated class, known as the Selected Completely At Random (SCAR) assumption. Yet the annotation process is… ▽ More In binary classification, Learning from Positive and Unlabeled data (LePU) is semi-supervised learning but with labeled elements from only one class. Most of the research on LePU relies on some form of independence between the selection process of annotated examples and the features of the annotated class, known as the Selected Completely At Random (SCAR) assumption. Yet the annotation process is an important part of the data collection, and in many cases it naturally depends on certain features of the data (e.g., the intensity of an image and the size of the object to be detected in the image). Without any constraints on the model for the annotation process, classification results in the LePU problem will be highly non-unique. So proper, flexible constraints are needed. In this work we incorporate more flexible and realistic models for the annotation process than SCAR, and more importantly, offer a solution for the challenging LePU problem. On the theory side, we establish the identifiability of the properties of the annotation process and the classification function, in light of the considered constraints on the data-generating process. We also propose an inference algorithm to learn the parameters of the model, with successful experimental results on both simulated and real data. We also propose a novel real-world dataset forLePU, as a benchmark dataset for future studies. △ Less

Submitted 2 March, 2020; originally announced March 2020.

Comments: Submitted to UAI 2020

arXiv:1707.06819 [pdf, ps, other]

A central limit like theorem for Fourier sums

Authors: Dominik Janzing, Naji Shajarisales, Michel Besserve

Abstract: We consider the probability distributions of values in the complex plane attained by Fourier sums of the form \sum_{j=1}^n a_j exp(-2πi j nu) /sqrt{n} when the frequency nu is drawn uniformly at random from an interval of length 1. If the coefficients a_j are i.i.d. drawn with finite third moment, the distance of these distributions to an isotropic two-dimensional Gaussian on C converges in probab… ▽ More We consider the probability distributions of values in the complex plane attained by Fourier sums of the form \sum_{j=1}^n a_j exp(-2πi j nu) /sqrt{n} when the frequency nu is drawn uniformly at random from an interval of length 1. If the coefficients a_j are i.i.d. drawn with finite third moment, the distance of these distributions to an isotropic two-dimensional Gaussian on C converges in probability to zero for any pseudometric on the set of distributions for which the distance between empirical distributions and the underlying distribution converges to zero in probability. △ Less

Submitted 21 July, 2017; originally announced July 2017.

Comments: 7 pages

MSC Class: 60Fxx

arXiv:1705.02212 [pdf, other]

Group invariance principles for causal generative models

Authors: Michel Besserve, Naji Shajarisales, Bernhard Schölkopf, Dominik Janzing

Abstract: The postulate of independence of cause and mechanism (ICM) has recently led to several new causal discovery algorithms. The interpretation of independence and the way it is utilized, however, varies across these methods. Our aim in this paper is to propose a group theoretic framework for ICM to unify and generalize these approaches. In our setting, the cause-mechanism relationship is assessed by c… ▽ More The postulate of independence of cause and mechanism (ICM) has recently led to several new causal discovery algorithms. The interpretation of independence and the way it is utilized, however, varies across these methods. Our aim in this paper is to propose a group theoretic framework for ICM to unify and generalize these approaches. In our setting, the cause-mechanism relationship is assessed by comparing it against a null hypothesis through the application of random generic group transformations. We show that the group theoretic view provides a very general tool to study the structure of data generating mechanisms with direct applications to machine learning. △ Less

Submitted 5 May, 2017; originally announced May 2017.

Comments: 16 pages, 6 figures

ACM Class: I.2.6; I.2.10; G.3; I.5.3

arXiv:1503.01299 [pdf, ps, other]

Telling cause from effect in deterministic linear dynamical systems

Authors: Naji Shajarisales, Dominik Janzing, Bernhard Shoelkopf, Michel Besserve

Abstract: Inferring a cause from its effect using observed time series data is a major challenge in natural and social sciences. Assuming the effect is generated by the cause trough a linear system, we propose a new approach based on the hypothesis that nature chooses the "cause" and the "mechanism that generates the effect from the cause" independent of each other. We therefore postulate that the power spe… ▽ More Inferring a cause from its effect using observed time series data is a major challenge in natural and social sciences. Assuming the effect is generated by the cause trough a linear system, we propose a new approach based on the hypothesis that nature chooses the "cause" and the "mechanism that generates the effect from the cause" independent of each other. We therefore postulate that the power spectrum of the time series being the cause is uncorrelated with the square of the transfer function of the linear filter generating the effect. While most causal discovery methods for time series mainly rely on the noise, our method relies on asymmetries of the power spectral density properties that can be exploited even in the context of deterministic systems. We describe mathematical assumptions in a deterministic model under which the causal direction is identifiable with this approach. We also discuss the method's performance under the additive noise model and its relationship to Granger causality. Experiments show encouraging results on synthetic as well as real-world data. Overall, this suggests that the postulate of Independence of Cause and Mechanism is a promising principle for causal inference on empirical time series. △ Less

Submitted 4 March, 2015; originally announced March 2015.

Comments: This article is under review for a peer-reviewed conference

arXiv:1402.2499 [pdf, other]

Justifying Information-Geometric Causal Inference

Authors: Dominik Janzing, Bastian Steudel, Naji Shajarisales, Bernhard Schölkopf

Abstract: Information Geometric Causal Inference (IGCI) is a new approach to distinguish between cause and effect for two variables. It is based on an independence assumption between input distribution and causal mechanism that can be phrased in terms of orthogonality in information space. We describe two intuitive reinterpretations of this approach that makes IGCI more accessible to a broader audience. M… ▽ More Information Geometric Causal Inference (IGCI) is a new approach to distinguish between cause and effect for two variables. It is based on an independence assumption between input distribution and causal mechanism that can be phrased in terms of orthogonality in information space. We describe two intuitive reinterpretations of this approach that makes IGCI more accessible to a broader audience. Moreover, we show that the described independence is related to the hypothesis that unsupervised learning and semi-supervised learning only works for predicting the cause from the effect and not vice versa. △ Less

Submitted 11 February, 2014; originally announced February 2014.

Comments: 3 Figures

arXiv:1304.7411

Multiplicity of 1 in Laplacian Spectra of trees

Authors: Naji Shajarisales

Abstract: In this paper, we interpret the multiplicity of 1 in Laplacian spectra of trees and prove that Faria's inequality turns to an equality in the case of normal trees which yields that in any tree without a vertex of degree 2, Faria equality holds and multiplicity of 1 in Laplacian spectrum will be equal to star degree of the the tree. As a result we introduce a combinatorial procedure for computing t… ▽ More In this paper, we interpret the multiplicity of 1 in Laplacian spectra of trees and prove that Faria's inequality turns to an equality in the case of normal trees which yields that in any tree without a vertex of degree 2, Faria equality holds and multiplicity of 1 in Laplacian spectrum will be equal to star degree of the the tree. As a result we introduce a combinatorial procedure for computing the multiplicity of 1. In the way to prove this results we will introduce many transformation on graphs which are invariant regarding to the multiplicity of 1. We also introduce an inequality for the multiplicity of 0 in adjacency spectrum of graphs and again introduce a procedure to compute this multiplicity in the case of trees. △ Less

Submitted 5 February, 2014; v1 submitted 27 April, 2013; originally announced April 2013.

Comments: This paper has been withdrawn by the author due to incompleteness of the proof for theorem 1 and 2

MSC Class: 05A20; 05C05; 05C50; 05C75

Showing 1–7 of 7 results for author: Shajarisales, N