Minimax-Optimal Policy Learning Under Unobserved Confounding-Reference-Cited by-同舟云学术

Minimax-Optimal Policy Learning Under Unobserved Confounding

Published:2021-05 Issue:5 Volume:67 Page:2870-2890
ISSN:0025-1909
Container-title:Management Science
language:en
Short-container-title:Management Science

Author:

Kallus Nathan¹^ORCID,Zhou Angela¹^ORCID

Affiliation:

1. Cornell University, New York, New York 10044

Abstract

We study the problem of learning personalized decision policies from observational data while accounting for possible unobserved confounding. Previous approaches, which assume unconfoundedness, that is, that no unobserved confounders affect both the treatment assignment as well as outcome, can lead to policies that introduce harm rather than benefit when some unobserved confounding is present as is generally the case with observational data. Instead, because policy value and regret may not be point-identifiable, we study a method that minimizes the worst-case estimated regret of a candidate policy against a baseline policy over an uncertainty set for propensity weights that controls the extent of unobserved confounding. We prove generalization guarantees that ensure our policy is safe when applied in practice and in fact obtains the best possible uniform control on the range of all possible population regrets that agree with the possible extent of confounding. We develop efficient algorithmic solutions to compute this minimax-optimal policy. Finally, we assess and compare our methods on synthetic and semisynthetic data. In particular, we consider a case study on personalizing hormone replacement therapy based on observational data, in which we validate our results on a randomized experiment. We demonstrate that hidden confounding can hinder existing policy-learning approaches and lead to unwarranted harm although our robust approach guarantees safety and focuses on well-evidenced improvement, a necessity for making personalized treatment policies learned from observational data reliable in practice. This paper was accepted by Hamid Nazerzadeh, big data analytics.

Publisher

Institute for Operations Research and the Management Sciences (INFORMS)

Subject

Management Science and Operations Research,Strategy and Management

Reference54 articles.

1. Generalized random forests

2. Latest evidence on using hormone replacement therapy in the menopause

3. Local Rademacher complexities

4. Optimal classification trees

Cited by 15 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Model-assisted sensitivity analysis for treatment effects under unmeasured confounding via regularized calibrated estimation;Journal of the Royal Statistical Society Series B: Statistical Methodology;2024-05-03

2. Doubly-Valid/Doubly-Sharp Sensitivity Analysis for Causal Inference with Unmeasured Confounding;Journal of the American Statistical Association;2024-04-24

3. Treatment Allocation with Strategic Agents;Management Science;2024-03-27

4. Optimal regimes for algorithm-assisted human decision-making;Biometrika;2024-03-19

5. Policy Learning with Asymmetric Counterfactual Utilities;Journal of the American Statistical Association;2024-02-13