Fetching the paper…
Reading the bibliography…
Many machine learning systems are built to solve the hardest examples of a particular task, which often makes them large and expensive to run---especially with respect to the easier examples, which might require much less computation.
Reasoning about beliefs and actions under computational resource constraints
Eric J. Horvitz · 1988
Earlier work this paper cites.
Principles of metareasoning
Stuart Russell and Eric Wefald · 1991
Earlier work this paper cites.
Principles of metareasoning
Stuart Russell and Eric Wefald · 1991
Earlier work this paper cites.
Function optimization using connectionist reinforcement learning algorithms
Ronald J. Williams and Jing Peng · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J. Williams · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J. Williams · 1992
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Multiple paired forward and inverse models for motor control
D.M. Wolpert and M. Kawato · 1998
Earlier work this paper cites.
Mechanical reasoning by mental simulation
Mary Hegarty · 2004
Earlier work this paper cites.
Efficient selectivity and backup operators in monte-carlo tree search
Rémi Coulom · 2006
Earlier work this paper cites.
States versus rewards: Dissociable neural prediction error signals underlying model-based and model-free reinforcement learning
Jan Gläscher, Nathaniel Daw, Peter Dayan, and John P. O’Doherty · 2010
Earlier work this paper cites.
Mental models and human reasoning
Philip N Johnson-Laird · 2010
Earlier work this paper cites.
Portfolio allocation for Bayesian optimization
Matthew W Hoffman, Eric Brochu, and Nando de Freitas · 2011
Earlier work this paper cites.
Selecting computations: Theory and applications
Nicholas Hay, Stuart J. Russell, David Tolpin, and Solomon Eyal Shimony · 2012
Earlier work this paper cites.
Selecting computations: Theory and applications
Nicholas Hay, Stuart J. Russell, David Tolpin, and Solomon Eyal Shimony · 2012
Cited alongside, same era.
Simulation as an engine of physical scene understanding
Peter W. Battaglia, Jessica B. Hamrick, and Joshua B. Tenenbaum · 2013
Cited alongside, same era.
Deep learning of representations: Looking forward
Yoshua Bengio · 2013
Cited alongside, same era.
On the difficulty of training recurrent neural networks
Razvan Pascanu, Tomas Mikolov, and Yoshua Bengio · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Neural computations underlying arbitration between model-based and model-free learning
Continuous control with deep reinforcement learning
Timothy P. Lillicrap, Jonathan J. Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Later among the works it cites.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, et al · 2015
Later among the works it cites.
Learning continuous control policies by stochastic value gradients
Nicolas Heess, Gregory Wayne, David Silver, Tim Lillicrap, Tom Erez, and Yuval Tassa · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sang Wan Lee, Shinsuke Shimojo, and John P. O’Doherty · 2014
Cited alongside, same era.
Algorithm selection by rational metareasoning as a model of human strategy selection
Falk Lieder, Dillon Plunkett, Jessica B. Hamrick, Stuart J. Russell, Nicholas J. Hay, and Thomas L. Griffiths · 2014
Cited alongside, same era.
An entropy search portfolio for Bayesian optimization
Bobak Shahriari, Ziyu Wang, Matthew W Hoffman, Alexandre Bouchard-Côté, and Nando de Freitas · 2014
Cited alongside, same era.
Deterministic policy gradient algorithms
David Silver, Guy Lever, Nicolas Heess, Thomas Degris, Daan Wierstra, and Martin Riedmiller · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Conditional computation in neural networks for faster models
Emmanuel Bengio, Pierre-Luc Bacon, Joelle Pineau, and Doina Precup · 2015
Cited alongside, same era.
Learning Visual Predictive Models of Physics for Playing Billiards
Katerina Fragkiadaki, Pulkit Agrawal, Sergey Levine, and Jitendra Malik · 2015
Cited alongside, same era.
Marcin Andrychowicz, Misha Denil, Sergio Gomez, Matthew W Hoffman, David Pfau, Tom Schaul, and Nando de Freitas · 2016
Later among the works it cites.
Interaction networks for learning about objects, relations and physics
Peter Battaglia, Razvan Pascanu, Matthew Lai, Danilo Jimenez Rezende, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Adaptive computation time for recurrent neural networks
Alex Graves · 2016
Later among the works it cites.
End-to-end training of deep visuomotor policies
Sergey Levine, Chelsea Finn, Trevor Darrell, and Pieter Abbeel · 2016
Later among the works it cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.
Aviv Tamar, Sergey Levine, and Pieter Abbeel · 2016
Later among the works it cites.
Interaction networks for learning about objects, relations and physics
Peter Battaglia, Razvan Pascanu, Matthew Lai, Danilo Jimenez Rezende, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Adaptive computation time for recurrent neural networks
Alex Graves · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza, Alex Graves, Timothy P. Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Conditional image generation with PixelCNN decoders
Aäron van den Oord, Nal Kalchbrenner, Oriol Vinyals, Lasse Espeholt, Alex Graves, and Koray Kavukcuoglu · 2016
Later among the works it cites.