arXiv AI

Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models

arXiv:2608. 02893v1 Announce Type: cross Abstract: Counterfactual inference approaches for sequential decision-making typically assume deterministic causal models, where all randomness stems from latent variables.

arXiv Machine Learning
Jun 8

Automatic, Debiased, and Invariant Counterfactual Generation under General Interventions

arXiv:2606. 07399v1 Announce Type: cross Abstract: Generative models for counterfactual outcomes have great potential to support decision-making under complex interventions, but existing approaches are limited by unstable estimation, poor generalization across environments, and bias from nuisance model misspecification.

By Raphael C Kim, Jingsen Zhu, Ramin Zabih, Michele Santacatterina
arXiv Machine Learning
Sep 11

Counterfactual Marginalisation: Framework for Evaluating Robustness to Nuisance Variables

The paper introduces counterfactual (CF) marginalisation, a test‑time evaluation method that assesses how robust classification models are to nuisance variables such as age or sex. By using a CF image generator to intervene on these parent variables, the method creates counterfactual versions of each test image and averages predictions over a chosen intervention distribution, yielding intervention‑aware predictions that filter out demographic effects while retaining patient‑specific latent information. These predictions are then used to define metrics for CF risk, calibration, stability, and worst‑case sensitivity, demonstrating the framework’s usefulness for quantitative robustness evaluation.

By Yasin Ibrahim, Hermione Warr, Robin J. Evans, Konstantinos Kamnitsas
arXiv Machine Learning
Sep 4

Semiparametric Inference for Counterfactual Regression under Intervention-Driven Shift

The paper introduces a semiparametric framework for counterfactual regression along a specified incremental‑intervention path. It estimates a finite‑dimensional constrained projection of counterfactual risk using cross‑fitted influence‑function representations, and establishes consistency, local stability, and first‑order expansions for smooth and finite‑dimensional programs. The results provide asymptotically valid inference, including simultaneous confidence bands, and are demonstrated through simulations and an SMS reminder application.

By Kwangho Kim
arXiv Machine Learning
Jun 2

Interaction-Limited Safe Continuous-Time RL for Dynamical Medical Treatment

arXiv:2606. 01051v1 Announce Type: new Abstract: Dynamic medical treatment requires deciding treatment intensity and intervention timing, while patient states evolve continuously and adverse events may occur between clinical interactions.

By Xun Shen, Yuepeng Wang, Akifumi Wachi, Yongqi Zhou, Richard Weiss, Yoshihiko Fujisawa, Ken Kawano, Mehrshad Sadria, Ying Chen, Xin Liu, Sebastien Gros, Xiao Hu, Kyoung-Sook Kim, Mengmou Li, Katsuki Fujisawa, Kenji Wakabayashi