Interactive model building for -learning. | Innovative Methods Program for Advancing Clinical Trials (IMPACT)

Title	Interactive model building for -learning.
Publication Type	Journal Article
Year of Publication	2014
Authors	Laber, Eric B., Kristin A. Linn, and Leonard A. Stefanski
Journal	Biometrika
Volume	101
Issue	4
Pagination	831-847
Date Published	2014 Oct 20
ISSN	0006-3444
Abstract	Evidence-based rules for optimal treatment allocation are key components in the quest for efficient, effective health care delivery. Q-learning, an approximate dynamic programming algorithm, is a popular method for estimating optimal sequential decision rules from data. Q-learning requires the modeling of nonsmooth, nonmonotone transformations of the data, complicating the search for adequately expressive, yet parsimonious, statistical models. The default Q-learning working model is multiple linear regression, which is not only provably misspecified under most data-generating models, but also results in nonregular regression estimators, complicating inference. We propose an alternative strategy for estimating optimal sequential decision rules for which the requisite statistical modeling does not depend on nonsmooth, nonmonotone transformed data, does not result in nonregular regression estimators, is consistent under a broader array of data-generation models than Q-learning, results in estimated sequential decision rules that have better sampling properties, and is amenable to established statistical approaches for exploratory data analysis, model building, and validation. We derive the new method, IQ-learning, via an interchange in the order of certain steps in Q-learning. In simulated experiments IQ-learning improves on Q-learning in terms of integrated mean squared error and power. The method is illustrated using data from a study of major depressive disorder.
DOI	10.1093/biomet/asu043
Alternate Journal	Biometrika
Original Publication	Interactive model building for Q-learning.
PubMed ID	25541562
PubMed Central ID	PMC4274394
Grant List	P01 CA142538 / CA / NCI NIH HHS / United States R01 CA085848 / CA / NCI NIH HHS / United States

Project:

Project 2.4