Fetching the paper…
Reading the bibliography…
We introduce a deep generative model for functions.
On the likelihood that one unknown probability exceeds another in view of the evidence of two samples
Thompson, William R · 1933
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, Ronald J · 1992
Earlier work this paper cites.
The im algorithm: a variational approach to information maximization
Barber, David and Agakov, Felix · 2003
Earlier work this paper cites.
Pattern Recognition and Machine Learning
Bishop, Christopher M · 2006
Earlier work this paper cites.
Gaussian processes for machine learning
Rasmussen, Carl Edward and Williams, Christopher K. I · 2006
Earlier work this paper cites.
Predicting motor vehicle collisions using bayesian neural network models: An empirical analysis
Xie, Yuanchang, Lord, Dominique, and Zhang, Yunlong · 2007
Earlier work this paper cites.
Bayesian learning for neural networks , volume 118
Neal, Radford M · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, Diederik P and Welling, Max · 2013
Earlier work this paper cites.
Learning with pseudo-ensembles
Bachman, Philip, Alsharif, Ouais, and Precup, Doina · 2014
Earlier work this paper cites.
Weight uncertainty in neural networks
Blundell, Charles, Cornebise, Julien, Kavukcuoglu, Koray, and Wierstra, Daan · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Lillicrap, Timothy P, Hunt, Jonathan J, Pritzel, Alexander, Heess, Nicolas, Erez, Tom, Tassa, Yuval, Silver, David, and Wierstra, Daan · 2015
Earlier work this paper cites.
Beattie, Charles, Leibo, Joel Z, Teplyashin, Denis, Ward, Tom, Wainwright, Marcus, Kuttler, Heinrich, Lefrancq, Andrew, Green, Simon, Valdes, Victor, Sadik, Amir, et al · 2016
Earlier work this paper cites.
Dropout as a bayesian approximation: Representing model uncertainty in deep learning
Gal, Yarin and Ghahramani, Zoubin · 2016
Cited alongside, same era.
Deep exploration via bootstrapped dqn
Osband, Ian, Blundell, Charles, Pritzel, Alexander, and Van Roy, Benjamin · 2016
Cited alongside, same era.
Sharp minima can generalize for deep nets
Dinh, Laurent, Pascanu, Razvan, Bengio, Samy, and Bengio, Yoshua · 2017
Cited alongside, same era.
Stochastic neural networks for hierarchical reinforcement learning
Florensa, Carlos, Duan, Yan, and Abbeel, Pieter · 2017
Cited alongside, same era.
Noisy networks for exploration
Fortunato, Meire, Azar, Mohammad Gheshlaghi, Piot, Bilal, Menick, Jacob, Osband, Ian, Graves, Alex, Mnih, Vlad, Munos, Remi, Hassabis, Demis, Pietquin, Olivier, Blundell, Charles, and Legg, Shane · 2017
Cited alongside, same era.
A unified view of entropy-regularized markov decision processes
Neu, Gergely, Jonsson, Anders, and Gomez, Vicenc · 2017
Later among the works it cites.
Implicit weight uncertainty in neural networks
Pawlowski, Nick, Rajchl, Martin, and Glocker, Ben · 2017
Later among the works it cites.
Parameter space noise for exploration
Plappert, Matthias, Houthooft, Rein, Dhariwal, Prafulla, Sidor, Szymon, Chen, Richard Y, Chen, Xi, Asfour, Tamim, Abbeel, Pieter, and Andrychowicz, Marcin · 2017
Later among the works it cites.
A tutorial on thompson sampling
Russo, Daniel, Roy, Benjamin Van, Kazerouni, Abbas, and Osband, Ian · 2017
Later among the works it cites.
Deep sets
Zaheer, Manzil, Kottur, Satwik, Ravanbakhsh, Siamak, Poczos, Barnabas, Salakhutdinov, Ruslan, and Smola, Alexander · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Meta learning shared hierarchies
Frans, Kevin, Ho, Jonathan, Chen, Xi, Abbeel, Pieter, and Schulman, John · 2017
Cited alongside, same era.
Deep bayesian active learning with image data
Gal, Yarin, Islam, Riashat, and Ghahramani, Zoubin · 2017
Cited alongside, same era.
Inverse reward design
Hadfield-Menell, Dylan, Milli, Smitha, Abbeel, Pieter, Russell, Stuart J, and Dragan, Anca · 2017
Cited alongside, same era.
What uncertainties do we need in bayesian deep learning for computer vision?
Kendall, Alex and Gal, Yarin · 2017
Cited alongside, same era.
Krueger, David, Huang, Chin-Wei, Islam, Riashat, Turner, Ryan, Lacoste, Alexandre, and Courville, Aaron C · 2017
Cited alongside, same era.
Multiplicative normalizing flows for variational bayesian neural networks
Louizos, Christos and Welling, Max · 2017
Cited alongside, same era.
Bridging the gap between value and policy based reinforcement learning
Nachum, Ofir, Norouzi, Mohammad, Xu, Kelvin, and Schuurmans, Dale · 2017
Cited alongside, same era.
Later among the works it cites.
Mine: Mutual information neural estimation
Belghazi, Mohamed Ishmael, Baratin, Aristide, Rajeswar, Sai, Ozair, Sherjil, Bengio, Yoshua, Courville, Aaron, and Hjelm, R Devon · 2018
Closest in time.
Diversity is all you need: Learning skills without a reward function
Eysenbach, Benjamin, Gupta, Abhishek, Ibarz, Julian, and Levine, Sergey · 2018
Closest in time.
Meta-learning and universality: Deep representations and gradient descent can approximate any learning algorithm
Finn, Chelsea and Levine, Sergey · 2018
Closest in time.
Garnelo, Marta, Schwarz, Jonathan, Rosenbaum, Dan, Viola, Fabio, Rezende, Danilo J., Eslami, S.M. Ali, and Teh, Yee Whye · 2018
Closest in time.
Learning an embedding space for transferable robot skills
Hausman, Karol, Springenberg, Jost Tobias, Wang, Ziyu, Heess, Nicolas, and Riedmiller, Martin · 2018
Closest in time.
Minimalistic gridworld environment for openai gym
Maxime Chevalier-Boisvert, Lucas Willems · 2018
Closest in time.
Riquelme, Carlos, Tucker, George, and Snoek, Jasper · 2018
Closest in time.