kth.sePublikationer KTH
Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
A Distance Covariance-based Kernel for Nonlinear Causal Clustering in Heterogeneous Populations
KTH, Skolan för teknikvetenskap (SCI), Matematik (Inst.), Matematik för Data och AI. Research Group Neuroinformatics, Faculty of Computer Science, University of Vienna.ORCID-id: 0000-0002-5495-1077
Department of Computer Science and Engineering, Indian Institute of Technology Bombay.
Research Group Neuroinformatics, Faculty of Computer Science, University of Vienna, Research Group Neuroinformatics, Faculty of Computer Science, University of Vienna; Research Platform Data Science @ Uni Vienna, Research Platform Data Science @ Uni Vienna; Vienna Cognitive Science Hub, Vienna Cognitive Science Hub.
2022 (Engelska)Ingår i: Proceedings of the 1st Conference on Causal Learning and Reasoning, CLeaR 2022, ML Research Press , 2022, s. 542-558Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

We consider the problem of causal structure learning in the setting of heterogeneous populations, i.e., populations in which a single causal structure does not adequately represent all population members, as is common in biological and social sciences. To this end, we introduce a distance covariance-based kernel designed specifically to measure the similarity between the underlying nonlinear causal structures of different samples. Indeed, we prove that the corresponding feature map is a statistically consistent estimator of nonlinear independence structure, rendering the kernel itself a statistical test for the hypothesis that sets of samples come from different generating causal structures. Even stronger, we prove that the kernel space is isometric to the space of causal ancestral graphs, so that distance between samples in the kernel space is guaranteed to correspond to distance between their generating causal structures. This kernel thus enables us to perform clustering to identify the homogeneous subpopulations, for which we can then learn causal structures using existing methods. Though we focus on the theoretical aspects of the kernel, we also evaluate its performance on synthetic data and demonstrate its use on a real gene expression data set.

Ort, förlag, år, upplaga, sidor
ML Research Press , 2022. s. 542-558
Nyckelord [en]
clustering, distance covariance, graphical causal models, whole-graph embeddings
Nationell ämneskategori
Sannolikhetsteori och statistik Signalbehandling
Identifikatorer
URN: urn:nbn:se:kth:diva-335773Scopus ID: 2-s2.0-85140201859OAI: oai:DiVA.org:kth-335773DiVA, id: diva2:1795477
Konferens
1st Conference on Causal Learning and Reasoning, CLeaR 2022, Eureka, United States of America, Apr 11 2022 - Apr 13 2022
Anmärkning

QC 20230908

Tillgänglig från: 2023-09-08 Skapad: 2023-09-08 Senast uppdaterad: 2023-09-08Bibliografiskt granskad

Open Access i DiVA

Fulltext saknas i DiVA

Scopus

Person

Markham, Alex

Sök vidare i DiVA

Av författaren/redaktören
Markham, Alex
Av organisationen
Matematik för Data och AI
Sannolikhetsteori och statistikSignalbehandling

Sök vidare utanför DiVA

GoogleGoogle Scholar

urn-nbn

Altmetricpoäng

urn-nbn
Totalt: 104 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf