Did we personalize? Assessing personalization by an online reinforcement learning algorithm using resampling
Name
10994_2024_6526_ReferencePDF.pdf
Size
1.21 MB
Format
Adobe PDF
Checksum (MD5)
fe57218e581179fd793b72e6433758d3
Author(s) • • • • • • •
Ghosh, Susobhan
Kim, Raphael
Chhabria, Prasidh
Dwivedi, Raaz
Klasnja, Predrag
Liao, Peng
Zhang, Kelly
Murphy, Susan
Date Issued
April 10, 2024
Journal
Machine Learning
Publisher
Springer US
Citation
Ghosh, S., Kim, R., Chhabria, P. et al. Did we personalize? Assessing personalization by an online reinforcement learning algorithm using resampling. Mach Learn 113, 3961–3997 (2024).
Version
Author's final manuscript
Abstract
There is a growing interest in using reinforcement learning (RL) to personalize sequences of treatments in digital health to support users in adopting healthier behaviors. Such sequential decision-making problems involve decisions about when to treat and how to treat based on the user’s context (e.g., prior activity level, location, etc.). Online RL is a promising data-driven approach for this problem as it learns based on each user’s historical responses and uses that knowledge to personalize these decisions. However, to decide whether the RL algorithm should be included in an “optimized” intervention for real-world deployment, we must assess the data evidence indicating that the RL algorithm is actually personalizing the treatments to its users. Due to the stochasticity in the RL algorithm, one may get a false impression that it is learning in certain states and using this learning to provide specific treatments. We use a working definition of personalization and introduce a resampling-based methodology for investigating whether the personalization exhibited by the RL algorithm is an artifact of the RL algorithm stochasticity. We illustrate our methodology with a case study by analyzing the data from a physical activity clinical trial called HeartSteps, which included the use of an online RL algorithm. We demonstrate how our approach enhances data-driven truth-in-advertising of algorithm personalization both across all users as well as within specific users in the study.
MIT Department
Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
Terms of Use
Article is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.
Persistent DSpace Link
DOI of Published Version
https://doi.org/10.1007/s10994-024-06526-x