Fetching the paper…
Reading the bibliography…
An important application of interactive machine learning is extending or amplifying the cognitive and physical capabilities of a human.
Rummery, G.A. and Niranjan, M., 1994. On-line Q-learning using connectionist systems
1994
Earlier work this paper cites.
Breazeal, C., 1998, May. Regulating human-robot interaction using ‘emotions’,‘drives’, and facial expressions. In Proceedings of Autonomous Agents (Vol. 98, pp. 14-21)
1998
Earlier work this paper cites.
Sutton, R.S. and Barto, A.G., 1998. Introduction to reinforcement learning (Vol. 135). Cambridge: MIT Press
1998
Earlier work this paper cites.
Liu, K. and Picard, R.W., 2003, April. Subtle expressivity in a robotic computer. In CHI 2003 Workshop on Subtle Expressiveness in Characters and Robots
2003
Earlier work this paper cites.
Kuhlmann, G., Stone, P., Mooney, R. and Shavlik, J., 2004, July. Guiding a reinforcement learner with natural language advice: Initial results in RoboCup soccer. In The AAAI-2004 workshop on supervisory control of learning and adaptive systems
2004
Earlier work this paper cites.
Thomaz, A.L. and Breazeal, C., 2008. Teachable robots: Understanding human teaching behavior to build more effective robot learners. Artificial Intelligence, 172(6), pp.716-737
2008
Earlier work this paper cites.
Knox, W.B., Fasel, I. and Stone, P., 2009, March. Design Principles for Creating Human-Shapable Agents. In AAAI Spring Symposium: Agents that Learn from Human Teachers (pp. 79-86)
2009
Earlier work this paper cites.
Williams, T. W. (2011). Guest editorial: progress on stabilizing and controlling powered upper-limb prostheses. Journal of Rehabilitation Research and Development 48(6): ix-xix
2011
Cited alongside, same era.
Breazeal, C., Love, B.C. and Mooney, R.J., 2012. Learning from human-generated reward
2012
Cited alongside, same era.
Koenig, N.P. and Mataric, M.J., 2012, October. Training Wheels for the Robot: Learning from Demonstration Using Simulation. In AAAI Fall Symposium: Robots Learning Interactively from Human Teachers
2012
Cited alongside, same era.
Pilarski, P.M. and Sutton, R.S., 2012, October. Between Instruction and Reward: Human-Prompted Switching. In AAAI Fall Symposium: Robots Learning Interactively from Human Teachers
2012
Cited alongside, same era.
Alizadeh, T., Calinon, S. and Caldwell, D.G., 2014, May. Learning from demonstrations with partially observable task parameters. In Robotics and Automation (ICRA), 2014 IEEE International Conference on (pp. 3309-3314). IEEE
2014
Later among the works it cites.
Kazemi, V. and Sullivan, J., 2014. One millisecond face alignment with an ensemble of regression trees. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (pp. 1867-1874)
2014
Later among the works it cites.
Edwards, A.L., Dawson, M.R., Hebert, J.S., Sherstan, C., Sutton, R.S., Chan, K.M. and Pilarski, P.M., 2015. Application of real-time machine learning to myoelectric prosthesis control: A case series in adaptive switching. Prosthetics and orthotics international, p.0309364615605373
2015
Later among the works it cites.
Knox, W.B. and Stone, P., 2015. Framing reinforcement learning from human reward: Reward positivity, temporal discounting, episodicity, and performance. Artificial Intelligence, 225, pp.24-50
2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
Mead, R., Atrash, A. and Mataric, M.J., 2013. Automated proxemic feature extraction and behavior recognition: Applications in human-robot interaction. International Journal of Social Robotics, 5(3), pp.367-378
2013
Cited alongside, same era.
Schulman, J., Ho, J., Lee, C. and Abbeel, P., 2013. Learning from demonstrations through the use of non-rigid registration. In Proceedings of the 16th international symposium on robotics research (ISRR)
2013
Cited alongside, same era.
Later among the works it cites.
Pilarski, P.M., Sutton, R.S. and Mathewson, K.W., 2015. Prosthetic Devices as Goal-Seeking Agents
2015
Later among the works it cites.
Loftin, R., Peng, B., MacGlashan, J., Littman, M.L., Taylor, M.E., Huang, J. and Roberts, D.L., 2016. Learning behaviors via human-delivered discrete feedback: modeling implicit feedback strategies to speed up learning. Autonomous Agents and Multi-Agent Systems, 30(1), pp.30-59
2016
Closest in time.
Peng, B., MacGlashan, J., Loftin, R., Littman, M.L., Roberts, D.L. and Taylor, M.E., 2016, May. A Need for Speed: Adapting Agent Action Speed to Improve Task Learning from Non-Expert Humans. In Proceedings of the 2016 International Conference on Autonomous Agents & Multiagent Systems (pp. 957-965). International Foundation for Autonomous Agents and Multiagent Systems
2016
Closest in time.