Fetching the paper…
Reading the bibliography…
Due to the recent advancements in wearables and sensing technology, health scientists are increasingly developing mobile health (mHealth) interventions.
Liao, P., Klasjna, P., Tewari, A. & Murphy, S. (2016), ‘Micro-randomized trials in mhealth’, Statistics in Medicine
1944
Earlier work this paper cites.
Howard, R. A. (1960), ‘Dynamic programming and markov processes.’
1960
Earlier work this paper cites.
Donald, S. G., Newey, W. K. et al. (1994), ‘Series estimation of semilinear models’, Journal of Multivariate Analysis
1994
Earlier work this paper cites.
Puterman, M. L. (1994), ‘Markov decision processes: Discrete stochastic dynamic programming’
1994
Earlier work this paper cites.
Bradtke, S. J. & Barto, A. G. (1996), ‘Linear least-squares algorithms for temporal difference learning’, Machine learning
1996
Earlier work this paper cites.
Mahadevan, S. (1996), ‘Average reward reinforcement learning: Foundations, algorithms, and empirical results’, Machine learning
1996
Earlier work this paper cites.
Van Roy, B. (1998), Learning and value function approximation in complex decision processes, PhD thesis, Massachusetts Institute of Technology
1998
Earlier work this paper cites.
Hernández-Lerma, O. & Lasserre, J. B. (1999), Further topics on discrete-time Markov control processes
1999
Earlier work this paper cites.
Van de Geer, S. (2000), Empirical Processes in M-estimation
2000
Earlier work this paper cites.
Murphy, S. A., van der Laan, M. J., Robins, J. M. & Group, C. P. P. R. (2001), ‘Marginal mean models for dynamic regimes’, Journal of the American Statistical Association
2001
Earlier work this paper cites.
Györfi, L., Kohler, M., Krzyzak, A. & Walk, H. (2006), A distribution-free theory of nonparametric regression
2006
Earlier work this paper cites.
Antos, A., Szepesvári, C. & Munos, R. (2008), ‘Learning near-optimal policies with bellman-residual minimization based fitted policy iteration and a single sample path’, Machine Learning
2008
Earlier work this paper cites.
Steinwart, I. & Christmann, A. (2008), Support vector machines
2008
Cited alongside, same era.
Farahmand, A.-m. & Szepesvári, C. (2011), ‘Model selection in reinforcement learning’, Machine learning
2011
Cited alongside, same era.
Ortner, R. & Ryabko, D. (2012), Online regret bounds for undiscounted continuous reinforcement learning, in
2012
Cited alongside, same era.
Chakraborty, B. & Moodie, E. (2013), Statistical methods for dynamic treatment regimes
2013
Cited alongside, same era.
Kizakevich, P. N., Eckhoff, R., Weger, S., Weeks, A., Brown, J., Bryant, S., Bakalov, V., Zhang, Y., Lyden, J. & Spira, J. (2014), ‘A personal health information toolkit for health intervention research’, Stud Health Technol Inform
2014
Cited alongside, same era.
Farajtabar, M., Chow, Y. & Ghavamzadeh, M. (2018), More robust doubly robust off-policy evaluation, in
2018
Later among the works it cites.
Lee, J.-A., Choi, M., Lee, S. A. & Jiang, N. (2018), ‘Effective behavioral intervention strategies using mobile health applications for chronic disease management: a systematic review’, BMC medical informatics and decision making
2018
Later among the works it cites.
Liao, P., Dempsey, W., Sarker, H., Hossain, S. M., al’Absi, M., Klasnja, P. & Murphy, S. (2018), ‘Just-in-time but not too much: Determining treatment timing in mobile health’, Proceedings of the ACM on interactive, mobile, wearable and ubiquitous technologies
2018
Later among the works it cites.
Liu, Q., Li, L., Tang, Z. & Zhou, D. (2018), Breaking the curse of horizon: Infinite-horizon off-policy estimation, in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
Klasnja, P., Hekler, E., Shiffman, S., Boruvka, A., Almirall, D., Tewari, A. & Murphy, S. (2015), ‘Micro-randomized trials: An experimental design for developing just-in-time adaptive interventions.’, Health Psychology
2015
Cited alongside, same era.
Farahmand, A.-m., Ghavamzadeh, M., Szepesvári, C. & Mannor, S. (2016), ‘Regularized policy iteration with nonparametric function spaces’, The Journal of Machine Learning Research
2016
Cited alongside, same era.
2016
Cited alongside, same era.
Thomas, P. & Brunskill, E. (2016), Data-efficient off-policy policy evaluation for reinforcement learning, in
2016
Cited alongside, same era.
Zhao, T., Cheng, G. & Liu, H. (2016), ‘A partially linear framework for massive heterogeneous data’, Annals of statistics
2016
Cited alongside, same era.
Nahum-Shani, I., Smith, S. N., Spring, B. J., Collins, L. M., Witkiewitz, K., Tewari, A. & Murphy, S. A. (2018), ‘Just-in-time adaptive interventions (jitais) in mobile health: key components and design principles for ongoing health behavior support’, Annals of Behavioral Medicine
2018
Later among the works it cites.
Rabbi, M., Kotov, M. P., Cunningham, R., Bonar, E. E., Nahum-Shani, I., Klasnja, P., Walton, M. & Murphy, S. (2018), ‘Toward increasing engagement in substance use data collection: development of the substance abuse research assistant app and protocol for a microrandomized trial using adolescents and emerging adults’, JMIR research protocols
2018
Later among the works it cites.
Sutton, R. S. & Barto, A. G. (2018), Reinforcement learning: An introduction
2018
Later among the works it cites.
Kallus, N. & Uehara, M. (2019), Intrinsically efficient, stable, and bounded off-policy evaluation for reinforcement learning, in
2019
Closest in time.
Klasnja, P., Smith, S., Seewald, N. J., Lee, A., Hall, K., Luers, B., Hekler, E. B. & Murphy, S. A. (2019), ‘Efficacy of contextually tailored suggestions for physical activity: A micro-randomized optimization trial of heartsteps’, Annals of Behavioral Medicine
2019
Closest in time.
Dempsey, W., Liao, P., Kumar, S., Murphy, S. A. et al. (2020), ‘The stratified micro-randomized trial design: sample size considerations for testing nested causal effects of time-varying treatments’, Annals of Applied Statistics
2020
Closest in time.
Luckett, D. J., Laber, E. B., Kahkoska, A. R., Maahs, D. M., Mayer-Davis, E. & Kosorok, M. R. (2020), ‘Estimating dynamic treatment regimes in mobile health using v-learning’, Journal of the American Statistical Association
2020
Closest in time.