Fetching the paper…
Reading the bibliography…
Human gaze is known to be an intention-revealing signal in human demonstrations of tasks.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Orienting of attention
Michael I Posner. 1980 · 1980
Earlier work this paper cites.
Constrained differential optimization for neural networks
John C Platt and Alan H Barr. 1988 · 1988
Earlier work this paper cites.
Learning from demonstration. In Advances in neural information processing systems . 1040–1046
Stefan Schaal. 1997 · 1997
Earlier work this paper cites.
A framework for behavioural claning
Michael Bain and Claude Sommut. 1999 · 1999
Earlier work this paper cites.
Convex optimization
Stephen Boyd, Stephen P Boyd, and Lieven Vandenberghe. 2004 · 2004
Earlier work this paper cites.
Eye movements in natural behavior
Mary Hayhoe and Dana Ballard. 2005 · 2005
Earlier work this paper cites.
Task and context determine where you look
Constantin A Rothkopf, Dana H Ballard, and Mary M Hayhoe. 2007 · 2007
Earlier work this paper cites.
A survey of robot learning from demonstration
Brenna D Argall, Sonia Chernova, Manuela Veloso, and Brett Browning. 2009 · 2009
Earlier work this paper cites.
Vision, eye movements, and natural behavior
Michael F Land. 2009 · 2009
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning. In Proceedings of the fourteenth international conference on artificial intelligence and statistics . 627–635
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell. 2011 · 2011
Earlier work this paper cites.
Eye guidance in natural vision: Reinterpreting salience
Benjamin W Tatler, Mary M Hayhoe, Michael F Land, and Dana H Ballard. 2011 · 2011
Earlier work this paper cites.
Improving neural networks by preventing co-adaptation of feature detectors
Geoffrey E Hinton, Nitish Srivastava, Alex Krizhevsky, Ilya Sutskever, and Ruslan R Salakhutdinov. 2012 · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control. In 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 5026–5033
Emanuel Todorov, Tom Erez, and Yuval Tassa. 2012 · 2012
Earlier work this paper cites.
Cheating experience: Guiding novices to adopt the gaze strategies of experts expedites the learning of technical laparoscopic skills
Samuel J Vine, Rich SW Masters, John S McGrath, Elizabeth Bright, and Mark R Wilson. 2012 · 2012
Earlier work this paper cites.
Adadelta: an adaptive learning rate method
Matthew D Zeiler. 2012 · 2012
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling. 2013 · 2013
Earlier work this paper cites.
Methods for comparing scanpaths and saliency maps: strengths and weaknesses
Olivier Le Meur and Thierry Baccino. 2013 · 2013
Earlier work this paper cites.
Selective attention enables action selection: evidence from evolutionary robotics experiments
Giancarlo Petrosino, Domenico Parisi, and Stefano Nolfi. 2013 · 2013
Earlier work this paper cites.
Atari-HEAD: Atari Human Eye-Tracking and Demonstration Dataset
Ruohan Zhang, Calen Walshe, Zhuode Liu, Lin Guan, Karl S Muller, Jake A Whritner, Luxin Zhang, Mary M Hayhoe, and Dana H Ballard. 2020b · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Towards dropout training for convolutional neural networks
Haibing Wu and Xiaodong Gu. 2015 · 2015
Cited alongside, same era.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Visualizing and understanding atari agents. In International Conference on Machine Learning . 1792–1801
Samuel Greydanus, Anurag Koul, Jonathan Dodge, and Alan Fern. 2018 · 2018
Later among the works it cites.
Rainbow: Combining improvements in deep reinforcement learning. In AAAI
Matteo Hessel, Joseph Modayil, Hado Van Hasselt, Tom Schaul, Georg Ostrovski, Will Dabney, Dan Horgan, Bilal Piot, Mohammad Azar, and David Silver. 2018 · 2018
Later among the works it cites.
Reward learning from human preferences and demonstrations in Atari. In Advances in Neural Information Processing Systems . 8011–8023
Borja Ibarz, Jan Leike, Tobias Pohlen, Geoffrey Irving, Shane Legg, and Dario Amodei. 2018 · 2018
Later among the works it cites.
Adversarial imitation via variational inverse reinforcement learning
Ahmed H Qureshi, Byron Boots, and Michael C Yip. 2018 · 2018
Later among the works it cites.
Human gaze following for human-robot interaction. In 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 8615–8621
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning transferable policies for monocular reactive mav control. In International Symposium on Experimental Robotics . Springer, 3–11
Shreyansh Daftry, J Andrew Bagnell, and Martial Hebert. 2016 · 2016
Cited alongside, same era.
Guided cost learning: Deep inverse optimal control via policy optimization. In ICML . 49–58
Chelsea Finn, Sergey Levine, and Pieter Abbeel. 2016 · 2016
Cited alongside, same era.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon. 2016 · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Cited alongside, same era.
OpenAI Baselines
Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov. 2017 · 2017
Cited alongside, same era.
Learning robust rewards with adversarial inverse reinforcement learning
Justin Fu, Katie Luo, and Sergey Levine. 2017 · 2017
Cited alongside, same era.
Rainbow: Combining improvements in deep reinforcement learning
Matteo Hessel, Joseph Modayil, Hado Van Hasselt, Tom Schaul, Georg Ostrovski, Will Dabney, Dan Horgan, Bilal Piot, Mohammad Azar, and David Silver. 2017 · 2017
Cited alongside, same era.
Akanksha Saran, Srinjoy Majumdar, Elaine Schaertl Short, Andrea Thomaz, and Scott Niekum. 2018 · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 2018 · 2018
Later among the works it cites.
Behavioral cloning from observation. In IJCAI . AAAI Press, 4950–4957
Faraz Torabi, Garrett Warnell, and Peter Stone. 2018 · 2018
Later among the works it cites.
Inverse reinforcement learning for video games
Aaron Tucker, Adam Gleave, and Stuart Russell. 2018 · 2018
Later among the works it cites.
An initial attempt of combining visual selective attention with deep reinforcement learning
Liu Yuezhang, Ruohan Zhang, and Dana H Ballard. 2018 · 2018
Later among the works it cites.
Agil: Learning attention from human for visuomotor tasks. In Proceedings of the European Conference on Computer Vision (ECCV) . 663–679
Ruohan Zhang, Zhuode Liu, Luxin Zhang, Jake A Whritner, Karl S Muller, Mary M Hayhoe, and Dana H Ballard. 2018 · 2018
Later among the works it cites.
Extrapolating Beyond Suboptimal Demonstrations via Inverse Reinforcement Learning from Observations
Daniel S Brown, Wonjoon Goo, Prabhat Nagarajan, and Scott Niekum. 2019 · 2019
Later among the works it cites.
Gaze Training by Modulated Dropout Improves Imitation Learning
Yuying Chen, Congcong Liu, Lei Tai, Ming Liu, and Bertram E Shi. 2019 · 2019
Later among the works it cites.
Causal confusion in imitation learning. In Advances in Neural Information Processing Systems . 11698–11709
Pim de Haan, Dinesh Jayaraman, and Sergey Levine. 2019 · 2019
Later among the works it cites.
A gaze model improves autonomous driving. In Proceedings of the 11th ACM Symposium on Eye Tracking Research & Applications . ACM, 33
Congcong Liu, Yuying Chen, Lei Tai, Haoyang Ye, Ming Liu, and Bertram E Shi. 2019 · 2019
Later among the works it cites.
Understanding Teacher Gaze Patterns for Robot Learning
Akanksha Saran, Elaine Schaertl Short, Andrea Thomaz, and Scott Niekum. 2019 · 2019
Later among the works it cites.
Periphery-Fovea Multi-Resolution Driving Model guided by Human Attention
Ye Xia, Jinkyu Kim, John Canny, Karl Zipser, and David Whitney. 2019 · 2019
Later among the works it cites.
Leveraging human guidance for deep reinforcement learning tasks
Ruohan Zhang, Faraz Torabi, Lin Guan, Dana H Ballard, and Peter Stone. 2019 · 2019
Later among the works it cites.
Why are machine learning algorithms hard to tune?
Ira Korshunova Jonas Degrave. 2021 · 2021
Closest in time.