More Efficient Policy Learning via Optimal Retargeting

Kallus, Nathan

Statistics > Machine Learning

arXiv:1906.08611 (stat)

[Submitted on 20 Jun 2019 (v1), last revised 3 Dec 2020 (this version, v2)]

Title:More Efficient Policy Learning via Optimal Retargeting

Authors:Nathan Kallus

View PDF

Abstract:Policy learning can be used to extract individualized treatment regimes from observational data in healthcare, civics, e-commerce, and beyond. One big hurdle to policy learning is a commonplace lack of overlap in the data for different actions, which can lead to unwieldy policy evaluation and poorly performing learned policies. We study a solution to this problem based on retargeting, that is, changing the population on which policies are optimized. We first argue that at the population level, retargeting may induce little to no bias. We then characterize the optimal reference policy and retargeting weights in both binary-action and multi-action settings. We do this in terms of the asymptotic efficient estimation variance of the new learning objective. Extensive empirical results in a simulation study and a case study of personalized job counseling demonstrate that retargeting is a fairly easy way to significantly improve any policy learning procedure applied to observational data.

Subjects:	Machine Learning (stat.ML); Machine Learning (cs.LG); Optimization and Control (math.OC)
Cite as:	arXiv:1906.08611 [stat.ML]
	(or arXiv:1906.08611v2 [stat.ML] for this version)
	https://doi.org/10.48550/arXiv.1906.08611

Submission history

From: Nathan Kallus [view email]
[v1] Thu, 20 Jun 2019 13:40:59 UTC (155 KB)
[v2] Thu, 3 Dec 2020 00:30:29 UTC (120 KB)

Statistics > Machine Learning

Title:More Efficient Policy Learning via Optimal Retargeting

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Machine Learning

Title:More Efficient Policy Learning via Optimal Retargeting

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators