What You Observe Determines How You Identify Causal Effects: Evaluating Causal Models across Observational Views
Read the original on arXiv Machine Learning →The Flow has not summarised this story yet — read it at arXiv Machine Learning.
The Flow has not summarised this story yet — read it at arXiv Machine Learning.
Causal Foundation Models (CFMs) are pretrained neural networks designed to estimate causal quantities—such as the average treatment effect—across new datasets using in‑context learning, eliminating the need for bespoke pipelines or model updates. The paper introduces CFMs, reviews foundational concepts in causal inference and machine learning, and provides practical code examples and Jupyter notebooks to illustrate their application.
CausalArena is a new benchmark designed to evaluate causal discovery methods in the era of foundation models. It unifies synthetic structural causal models (SCMs), semantically grounded SCMs, and formula‑grounded SCMs, while also including real‑world datasets for external validation. Experiments show that performance rankings vary widely across different SCM families and protocols, indicating that strong results on one benchmark do not necessarily transfer to others.
arXiv:2609.37446v1 Announce Type: new Abstract: Supervised causal discovery learns to infer causal structure for a new dataset from training datasets paired with structural labels. These training pai...
arXiv:2507. 14661v2 Announce Type: replace-cross Abstract: Semi-supervised domain adaptation (SSDA) seeks to achieve accurate predictions in a target domain with limited labeled target data by exploiting abundant source and unlabeled target data.
TabCausal is a causal discovery foundation model that learns to map datasets directly to causal graphs by pretraining across diverse causal environments. It uses a dynamic task construction strategy to expose the model to varied graph priors, mechanisms, noise models, dimensions, sample sizes, and intervention regimes, improving transferability from observational and mixed‑interventional data. On large synthetic benchmarks and a new protocol‑guided semantic benchmark, TabCausal outperforms many classical baselines and shows robust structure recovery, especially when interventional evidence is available.
The paper introduces causal foundation models that can bound the effects of interventions and counterfactuals using only observational data. It defines a canonical prior with full support over structural causal models with discrete observables, enabling the translation of counterfactual bounding into learning distributions over functions that map data and structural assumptions to causal queries. This approach extends causal foundational modelling to partially-identifiable causal effects, where unobserved confounding leads to multiple compatible values for the effect.