arXiv Machine Learning By Youssef Drissi, Markus Ettl, Shivaram Subramanian, Wei Sun, Zack Xue

Counterfactual Optimal Action Trees (COAT): Interpretable Prescriptive Policies from Observational Data

Read the original on arXiv Machine Learning →

arXiv:2607. 14318v1 Announce Type: new Abstract: We introduce COAT (Counterfactual Optimal Action Tree), a framework for learning interpretable prescriptive policies from observational data.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Computation and Language
Aug 28

Natural-Language Policies to Executable Decisions: An Interpretable Large Language Model Framework

The paper introduces a production‑grade large language model (LLM) pricing system for the tourism industry that separates structured extraction and policy selection from deterministic numeric pricing. Policies are compiled into interpretable condition trees, allowing new clauses and evolving rules to be added without code changes while maintaining auditability. Deployed across 12 business categories and 1,500 operators, the system handled 3,960 orders in six months, cutting the order‑management team from 15‑20 to 3 and reducing per‑order handling time from 10 minutes to under 2 minutes.

By Ziqiang Zhang, Jing Ma, Zilong Wang, Jiayuan Chen, Yi Qiao, Yu He, Wei Zhang, Dai Cheng, Xiaoyu Shen
Hugging Face Trending Papers
Jul 2

Profit-Based Counterfactual Explanations for Product Improvement: A Case Study of Manga Sales in Japan

Counterfactual explanation (CE) is widely used to enhance the interpretability of machine learning models and support data-driven decision-making based on model predictions. However, existing CE methods typically require two exogenously specified inputs: a desired output value (target) and a distance function that quantifies changes in explanatory variables.

arXiv Machine Learning
5d ago

Counterfactual Online Conformal Prediction Under Adaptive Logging

The paper addresses the failure of online conformal prediction when predictions influence actions that determine which outcomes are used for calibration. It introduces Propensity-Weighted Online Conformal Prediction (PW‑OCP), an inverse‑propensity‑weighted recursion that debiases calibration, and a doubly robust variant (DR‑OCP) that further reduces bias. Experiments on synthetic decision tasks, open bandit data, and financial rebalancing demonstrate that PW‑OCP and DR‑OCP improve counterfactual coverage and downstream regret while preserving prediction‑set sharpness.

By Xinyu Qiao, Yichen Lin, Kaihong Ji, Xue Wang, Tao Yao