Causal Rule Sets for Identifying Subgroups with Enhanced Treatment Effects-Reference-Cited by-同舟云学术

Causal Rule Sets for Identifying Subgroups with Enhanced Treatment Effects

Published:2022-05 Issue:3 Volume:34 Page:1626-1643
ISSN:1091-9856
Container-title:INFORMS Journal on Computing
language:en
Short-container-title:INFORMS Journal on Computing

Author:

Wang Tong¹^ORCID,Rudin Cynthia²^ORCID

Affiliation:

1. Tippie College of Business, University of Iowa, Iowa City, Iowa 52242

2. Department of Computer Science, Duke University, Durham, North Carolina 27708

Abstract

A key question in causal inference analyses is how to find subgroups with elevated treatment effects. This paper takes a machine learning approach and introduces a generative model, causal rule sets (CRS), for interpretable subgroup discovery. A CRS model uses a small set of short decision rules to capture a subgroup in which the average treatment effect is elevated. We present a Bayesian framework for learning a causal rule set. The Bayesian model consists of a prior that favors simple models for better interpretability as well as avoiding overfitting and a Bayesian logistic regression that captures the likelihood of data, characterizing the relation between outcomes, attributes, and subgroup membership. The Bayesian model has tunable parameters that can characterize subgroups with various sizes, providing users with more flexible choices of models from the treatment-efficient frontier. We find maximum a posteriori models using iterative discrete Monte Carlo steps in the joint solution space of rules sets and parameters. To improve search efficiency, we provide theoretically grounded heuristics and bounding strategies to prune and confine the search space. Experiments show that the search algorithm can efficiently recover true underlying subgroups. We apply CRS on public and real-world data sets from domains in which interpretability is indispensable. We compare CRS with state-of-the-art rule-based subgroup discovery models. Results show that CRS achieves consistently competitive performance on data sets from various domains, represented by high treatment-efficient frontiers. Summary of Contribution: This paper is motivated by the large heterogeneity of treatment effect in many applications and the need to accurately locate subgroups for enhanced treatment effect. Existing methods either rely on prior hypotheses to discover subgroups or greedy methods, such as tree-based recursive partitioning. Our method adopts a machine learning approach to find an optimal subgroup learned with a carefully global objective. Our model is more flexible in capturing subgroups by using a set of short decision rules compared with tree-based baselines. We evaluate our model using a novel metric, treatment-efficient frontier, that characterizes the trade-off between the subgroup size and achievable treatment effect, and our model demonstrates better performance than baseline models.

Publisher

Institute for Operations Research and the Management Sciences (INFORMS)

Subject

General Engineering

Link

https://pubsonline.informs.org/doi/pdf/10.1287/ijoc.2021.1143

Reference48 articles.

1. Identification of Causal Effects Using Instrumental Variables

2. Subgroup analysis and other (mis)uses of baseline data in clinical trials

3. Subgroup discovery

4. Adjuvant Trastuzumab: A Milestone in the Treatment of HER-2-Positive Early Breast Cancer

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. CURLS: Causal Rule Learning for Subgroups with Significant Treatment Effect;Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining;2024-08-24

2. CRE: An R package for interpretable discovery and inference of heterogeneous treatment effects;Journal of Open Source Software;2023-12-15

3. Improved Inference for Doubly Robust Estimators of Heterogeneous Treatment Effects;Biometrics;2023-02-06

4. Targeting resources efficiently and justifiably by combining causal machine learning and theory;Frontiers in Artificial Intelligence;2022-12-07

5. Detecting heterogeneous treatment effects with instrumental variables and application to the Oregon health insurance experiment;The Annals of Applied Statistics;2022-06-01