Fetching the paper…
Reading the bibliography…
While deep neural networks (DNNs) and Gaussian Processes (GPs) are both popularly utilized to solve problems in reinforcement learning, both approaches feature undesirable drawbacks for challenging problems.
The parti-game algorithm for variable resolution reinforcement learning in multidimensional state-spaces
Andrew W Moore · 1994
Earlier work this paper cites.
Priors for infinite networks
Radford M Neal · 1996
Earlier work this paper cites.
Nonlinear adaptive control using nonparametric gaussian process prior models
Roderick Murray-Smith and Daniel Sbarbaro · 2002
Earlier work this paper cites.
Gaussian processes in reinforcement learning
Malte Kuss and Carl E Rasmussen · 2004
Earlier work this paper cites.
Probabilistic non-linear principal component analysis with gaussian process latent variable models
Neil Lawrence · 2005
Earlier work this paper cites.
Gaussian processes for machine learning
Christopher KI Williams and Carl Edward Rasmussen · 2006
Earlier work this paper cites.
Gp-ukf: Unscented kalman filters with gaussian process prediction and observation models
Jonathan Ko, Daniel J Kleint, Dieter Fox, and Dirk Haehnelt · 2007
Earlier work this paper cites.
Local gaussian process regression for real-time model-based robot control
Duy Nguyen-Tuong and Jan Peters · 2008
Earlier work this paper cites.
Kernel methods for deep learning
Youngmin Cho and Lawrence K Saul · 2009
Earlier work this paper cites.
Gaussian process dynamic programming
Marc Peter Deisenroth, Carl Edward Rasmussen, and Jan Peters · 2009
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition
Geoffrey Hinton, Li Deng, Dong Yu, George Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Brian Kingsbury, et al · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Reinforcement learning and feedback control: Using natural decision methods to design optimal adaptive controllers
Frank L Lewis, Draguna Vrabie, and Kyriakos G Vamvoudakis · 2012
Earlier work this paper cites.
Gaussian processes for data-efficient learning in robotics and control
Marc Peter Deisenroth, Dieter Fox, and Carl Edward Rasmussen · 2013
Earlier work this paper cites.
Recurrent continuous translation models
Nal Kalchbrenner and Phil Blunsom · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Cited alongside, same era.
Learning about physical parameters: The importance of model discrepancy
Jennỳ Brynjarsdóttir and Anthony O’Hagan · 2014
Cited alongside, same era.
Convolutional kernel networks
Julien Mairal, Piotr Koniusz, Zaid Harchaoui, and Cordelia Schmid · 2014
Cited alongside, same era.
Dropout as a bayesian approximation: Insights and applications
Yarin Gal and Zoubin Ghahramani · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Deep neural networks as gaussian processes
Jaehoon Lee, Yasaman Bahri, Roman Novak, Samuel S Schoenholz, Jeffrey Pennington, and Jascha Sohl-Dickstein · 2018
Later among the works it cites.
Gaussian process behaviour in wide deep neural networks
Alexander G de G Matthews, Mark Rowland, Jiri Hron, Richard E Turner, and Zoubin Ghahramani · 2018
Later among the works it cites.
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Comprehensive comparison of online adp algorithms for continuous-time optimal control
Yuanheng Zhu and Dongbin Zhao · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Openai gym, 2016
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Cited alongside, same era.
Toward deeper understanding of neural networks: The power of initialization and a dual view on expressivity
Amit Daniely, Roy Frostig, and Yoram Singer · 2016
Cited alongside, same era.
Neural architecture search with reinforcement learning
Barret Zoph and Quoc V. Le · 2016
Cited alongside, same era.
Reinforcement learning and dynamic programming using function approximators
Lucian Busoniu, Robert Babuska, Bart De Schutter, and Damien Ernst · 2017
Cited alongside, same era.
Game theory-based control system algorithms with real-time reinforcement learning: How to solve multiplayer games online
Kyriakos G Vamvoudakis, Hamidreza Modares, Bahare Kiumarsi, and Frank L Lewis · 2017
Cited alongside, same era.
Gpytorch: Blackbox matrix-matrix gaussian process inference with gpu acceleration
Jacob Gardner, Geoff Pleiss, Kilian Q Weinberger, David Bindel, and Andrew G Wilson · 2018
Cited alongside, same era.
Sanjeev Arora, Simon S Du, Wei Hu, Zhiyuan Li, Ruslan Salakhutdinov, and Ruosong Wang · 2019
Later among the works it cites.
Harnessing the power of infinitely wide deep nets on small-data tasks
Sanjeev Arora, Simon S Du, Zhiyuan Li, Ruslan Salakhutdinov, Ruosong Wang, and Dingli Yu · 2019
Later among the works it cites.
A case study competition among methods for analyzing large spatial data
Matthew J Heaton, Abhirup Datta, Andrew O Finley, Reinhard Furrer, Joseph Guinness, Rajarshi Guhaniyogi, Florian Gerber, Robert B Gramacy, Dorit Hammerling, Matthias Katzfuss, et al · 2019
Later among the works it cites.
Wide neural networks of any depth evolve as linear models under gradient descent
Jaehoon Lee, Lechao Xiao, Samuel S Schoenholz, Yasaman Bahri, Jascha Sohl-Dickstein, and Jeffrey Pennington · 2019
Later among the works it cites.
Bayesian deep convolutional neural networks with many channels are gaussian processes
Roman Novak, Lechao Xiao, Jaehoon Lee, Yasaman Bahri, Greg Yang, Dan Abolafia, Jeffrey Pennington, and Jascha Sohl-dickstein · 2019
Later among the works it cites.
Exact gaussian processes on a million data points
Ke Wang, Geoff Pleiss, Jacob Gardner, Stephen Tyree, Kilian Q Weinberger, and Andrew Gordon Wilson · 2019
Later among the works it cites.
Greg Yang · 2019
Later among the works it cites.
A fine-grained spectral perspective on neural networks
Greg Yang and Hadi Salman · 2019
Later among the works it cites.
When gaussian process meets big data: A review of scalable gps
Haitao Liu, Yew-Soon Ong, Xiaobo Shen, and Jianfei Cai · 2020
Closest in time.