We Are Not Your Real Parents: Telling Causal from Confounded using MDL

Kaltenpoth, David; Vreeken, Jilles

Computer Science > Machine Learning

arXiv:1901.06950 (cs)

[Submitted on 21 Jan 2019]

Title:We Are Not Your Real Parents: Telling Causal from Confounded using MDL

Authors:David Kaltenpoth, Jilles Vreeken

View PDF

Abstract:Given data over variables $(X_1,...,X_m, Y)$ we consider the problem of finding out whether $X$ jointly causes $Y$ or whether they are all confounded by an unobserved latent variable $Z$. To do so, we take an information-theoretic approach based on Kolmogorov complexity. In a nutshell, we follow the postulate that first encoding the true cause, and then the effects given that cause, results in a shorter description than any other encoding of the observed variables.
The ideal score is not computable, and hence we have to approximate it. We propose to do so using the Minimum Description Length (MDL) principle. We compare the MDL scores under the models where $X$ causes $Y$ and where there exists a latent variables $Z$ confounding both $X$ and $Y$ and show our scores are consistent. To find potential confounders we propose using latent factor modeling, in particular, probabilistic PCA (PPCA).
Empirical evaluation on both synthetic and real-world data shows that our method, CoCa, performs very well -- even when the true generating process of the data is far from the assumptions made by the models we use. Moreover, it is robust as its accuracy goes hand in hand with its confidence.

Comments:	10 pages, 6 figures
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1901.06950 [cs.LG]
	(or arXiv:1901.06950v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1901.06950

Submission history

From: David Kaltenpoth [view email]
[v1] Mon, 21 Jan 2019 15:09:34 UTC (1,197 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2019-01

Change to browse by:

cs
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

David Kaltenpoth
Jilles Vreeken

export BibTeX citation

Computer Science > Machine Learning

Title:We Are Not Your Real Parents: Telling Causal from Confounded using MDL

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:We Are Not Your Real Parents: Telling Causal from Confounded using MDL

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators