Fetching the paper…
Reading the bibliography…
Owe to the recent advancements in Artificial Intelligence especially deep learning, many data-driven decision support systems have been implemented to facilitate medical doctors in delivering personalized care.
Dynamic programming and markov processes
Ronald A Howard · 1960
Earlier work this paper cites.
Self-improving reactive agents based on reinforcement learning, planning and teaching
Long-Ji Lin · 1992
Earlier work this paper cites.
Q-learning
Christopher JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Advantage updating
Leemon C Baird III · 1993
Earlier work this paper cites.
Issues in using function approximation for reinforcement learning
Sebastian Thrun and Anton Schwartz · 1993
Earlier work this paper cites.
Nonlinear and adaptive control design , volume 222
Miroslav Krstic, Ioannis Kanellakopoulos, Petar V Kokotovic, et al · 1995
Earlier work this paper cites.
Multi-player residual advantage learning with general function approximation
Mance E Harmon and Leemon C Baird III · 1996
Earlier work this paper cites.
Reinforcement learning: A survey
Leslie Pack Kaelbling, Michael L Littman, and Andrew W Moore · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
Leslie Pack Kaelbling, Michael L Littman, and Anthony R Cassandra · 1998
Earlier work this paper cites.
Discriminative training of hidden Markov models
Sadik Kapadia · 1998
Earlier work this paper cites.
Introduction to reinforcement learning , volume 135
Richard S Sutton, Andrew G Barto, et al · 1998
Earlier work this paper cites.
Is imitation learning the route to humanoid robots?
Stefan Schaal · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng, Stuart J Russell, et al · 2000
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S Sutton, David A McAllester, Satinder P Singh, and Yishay Mansour · 2000
Earlier work this paper cites.
Finite-time analysis of the multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer · 2002
Earlier work this paper cites.
Multiple model-based reinforcement learning
Kenji Doya, Kazuyuki Samejima, Ken-ichi Katagiri, and Mitsuo Kawato · 2002
Earlier work this paper cites.
Gaussian processes in machine learning
Carl Edward Rasmussen · 2003
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
Neural fitted q iteration–first experiences with a data efficient neural reinforcement learning method
Martin Riedmiller · 2005
Earlier work this paper cites.
Metamap: Mapping text to the umls metathesaurus
Alan R Aronson · 2006
Earlier work this paper cites.
Confidence-based policy learning from demonstration using gaussian mixture models
Sonia Chernova and Manuela Veloso · 2007
Cited alongside, same era.
Finite-time bounds for fitted value iteration
Rémi Munos and Csaba Szepesvári · 2008
Cited alongside, same era.
Large-scale machine learning with stochastic gradient descent
Léon Bottou · 2010
Cited alongside, same era.
Double q-learning
Hado V Hasselt · 2010
Cited alongside, same era.
A survey on transfer learning
Sinno Jialin Pan and Qiang Yang · 2010
Cited alongside, same era.
An empirical evaluation of thompson sampling
Olivier Chapelle and Lihong Li · 2011
Cited alongside, same era.
Pilco: A model-based and data-efficient approach to policy search
Vime: Variational information maximizing exploration
Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel · 2016
Later among the works it cites.
Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation
Tejas D Kulkarni, Karthik Narasimhan, Ardavan Saeedi, and Josh Tenenbaum · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Optimal medication dosing from suboptimal clinical examples: A deep reinforcement learning approach
Shamim Nemati, Mohammad M Ghassemi, and Gari D Clifford · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Marc Deisenroth and Carl E Rasmussen · 2011
Cited alongside, same era.
Insights in reinforcement learning
Hado Philip van Hasselt · 2011
Cited alongside, same era.
Multi-column deep neural networks for image classification
Dan Cireşan, Ueli Meier, and Jürgen Schmidhuber · 2012
Cited alongside, same era.
A survey of actor-critic reinforcement learning: Standard and natural policy gradients
Ivo Grondman, Lucian Busoniu, Gabriel AD Lopes, and Robert Babuska · 2012
Cited alongside, same era.
An application of inverse reinforcement learning to medical records of diabetes treatment
Hideki Asoh, Masanori Shiro1 Shotaro Akaho, Toshihiro Kamishima, Koiti Hasida, Eiji Aramaki, and Takahide Kohro · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Cited alongside, same era.
Inquire and diagnose: Neural symptom checking ensemble using deep reinforcement learning
Kai-Fu Tang, Hao-Cheng Kao, Chun-Nan Chou, and Edward Y Chang · 2016
Later among the works it cites.
A brief survey of deep reinforcement learning
Kai Arulkumaran, Marc Peter Deisenroth, Miles Brundage, and Anil Anthony Bharath · 2017
Later among the works it cites.
Reinforcement learning with deep energy-based policies
Tuomas Haarnoja, Haoran Tang, Pieter Abbeel, and Sergey Levine · 2017
Later among the works it cites.
Diagnostic inferencing via improving clinical concept extraction with deep reinforcement learning: A preliminary study
Yuan Ling, Sadid A Hasan, Vivek Datla, Ashequl Qadir, Kathy Lee, Joey Liu, and Oladimeji Farri · 2017
Later among the works it cites.
Data-efficient reinforcement learning in continuous state-action gaussian-pomdps
Rowan McAllister and Carl Edward Rasmussen · 2017
Later among the works it cites.
A reinforcement learning approach to weaning of mechanical ventilation in intensive care units
Niranjani Prasad, Li-Fang Cheng, Corey Chivers, Michael Draugelis, and Barbara E Engelhardt · 2017
Later among the works it cites.
Continuous state-space models for optimal sepsis treatment-a deep reinforcement learning approach
Aniruddh Raghu, Matthieu Komorowski, Leo Anthony Celi, Peter Szolovits, and Marzyeh Ghassemi · 2017
Later among the works it cites.
Deep reinforcement learning for automated radiation adaptation in lung cancer
Huan-Hsin Tseng, Yi Luo, Sunan Cui, Jen-Tzung Chien, Randall K Ten Haken, and Issam El Naqa · 2017
Later among the works it cites.
Asynchronous advantage actor-critic agent for starcraft ii
Basel Alghanem et al · 2018
Later among the works it cites.
Learning to treat sepsis with multi-output gaussian process deep recurrent q-networks
Joseph Futoma, Anthony Lin, Mark Sendak, Armando Bedoya, Meredith Clement, Cara O’Brien, and Katherine Heller · 2018
Later among the works it cites.
Deep variational reinforcement learning for pomdps
Maximilian Igl, Luisa Zintgraf, Tuan Anh Le, Frank Wood, and Shimon Whiteson · 2018
Later among the works it cites.
Context-aware symptom checking for disease diagnosis using hierarchical reinforcement learning
Hao-Cheng Kao, Kai-Fu Tang, and Edward Y Chang · 2018
Later among the works it cites.
Optimal and autonomous control using reinforcement learning: A survey
Bahare Kiumarsi, Kyriakos G Vamvoudakis, Hamidreza Modares, and Frank L Lewis · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Supervised reinforcement learning with recurrent neural network for dynamic treatment recommendation
Lu Wang, Wei Zhang, Xiaofeng He, and Hongyuan Zha · 2018
Later among the works it cites.