Fetching the paper…
Reading the bibliography…
We study the cross-entropy method (CEM) for the non-convex optimization of a continuous and parameterized objective function and introduce a differentiable variant that enables us to differentiate the output of CEM with respect to the objective function's parameters.
Learning convex optimization control policies
Agrawal, A., Barratt, S., Boyd, S., and Stellato, B · 1912
Earlier work this paper cites.
Neuronlike adaptive elements that can solve difficult learning control problems
Barto, A. G., Sutton, R. S., and Anderson, C. W · 1983
Earlier work this paper cites.
Convexity properties of entropy functions and analysis of diversity
Rao, C. R · 1984
Earlier work this paper cites.
Receding horizon control of nonlinear systems
Mayne, D. Q. and Michalska, H · 1990
Earlier work this paper cites.
Python reference manual
Van Rossum, G. and Drake Jr, F. L · 1995
Earlier work this paper cites.
Weighted likelihood estimating equations: The discrete case with applications to logistic regression
Markatou, M., Basu, A., and Lindsay, B · 1997
Earlier work this paper cites.
Optimization of computer simulation models with rare events
Rubinstein, R. Y · 1997
Earlier work this paper cites.
Weighted likelihood equations with bootstrap root search
Markatou, M., Basu, A., and Lindsay, B. G · 1998
Earlier work this paper cites.
Genetic optimization using derivatives
Sekhon, J. S. and Mebane, W. R · 1998
Earlier work this paper cites.
The elements of statistical learning , volume 1
Friedman, J., Hastie, T., and Tibshirani, R · 2001
Earlier work this paper cites.
Maximum weighted likelihood estimation
Wang, S. X · 2001
Earlier work this paper cites.
The weighted likelihood
Hu, F. and Zidek, J. V · 2002
Earlier work this paper cites.
A tutorial on the cross-entropy method
De Boer, P.-T., Kroese, D. P., Mannor, S., and Rubinstein, R. Y · 2005
Earlier work this paper cites.
Learning structured prediction models: A large margin approach
Taskar, B., Chatalbashev, V., Koller, D., and Guestrin, C · 2005
Earlier work this paper cites.
A tutorial on energy-based learning
LeCun, Y., Chopra, S., Hadsell, R., Ranzato, M., and Huang, F · 2006
Earlier work this paper cites.
A guide to NumPy , volume 1
Oliphant, T. E · 2006
Earlier work this paper cites.
Matplotlib: A 2d graphics environment
Hunter, J. D · 2007
Earlier work this paper cites.
Python for scientific computing
Oliphant, T. E · 2007
Earlier work this paper cites.
Variational analysis , volume 317
Rockafellar, R. T. and Wets, R. J.-B · 2009
Earlier work this paper cites.
A generalized path integral control approach to reinforcement learning
Theodorou, E., Buchli, J., and Schaal, S · 2010
Earlier work this paper cites.
The numpy array: a structure for efficient numerical computation
Van Der Walt, S., Colbert, S. C., and Varoquaux, G · 2011
Earlier work this paper cites.
Generic methods for optimization-based modeling
Domke, J · 2012
Earlier work this paper cites.
Python for data analysis: Data wrangling with Pandas, NumPy, and IPython
McKinney, W · 2012
Earlier work this paper cites.
Path integral policy improvement with covariance matrix adaptation
Stulp, F. and Sigaud, O · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y · 2012
Earlier work this paper cites.
Active learning of linear embeddings for gaussian processes
Garnett, R., Osborne, M. A., and Hennig, P · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2013
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Earlier work this paper cites.
{ \{ SciPy
Jones, E., Oliphant, T., and Peterson, P · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, D. J., Mohamed, S., and Wierstra, D · 2014
Earlier work this paper cites.
Doubly stochastic variational bayes for non-conjugate inference
Titsias, M. and Lázaro-Gredilla, M · 2014
Earlier work this paper cites.
Embed to control: A locally linear latent dynamics model for control from raw images
Watter, M., Springenberg, J., Boedecker, J., and Riedmiller, M · 2015
Earlier work this paper cites.
Structured prediction energy networks
Belanger, D. and McCallum, A · 2016
Earlier work this paper cites.
Manifold gaussian processes for regression
Calandra, R., Peters, J., Rasmussen, C. E., and Deisenroth, M. P · 2016
Cited alongside, same era.
Gould, S., Fernando, B., Cherian, A., Anderson, P., Cruz, R. S., and Guo, E · 2016
Cited alongside, same era.
Composing graphical models with neural networks for structured representations and fast inference
Johnson, M., Duvenaud, D. K., Wiltschko, A., Adams, R. P., and Datta, S. R · 2016
Cited alongside, same era.
Jupyter notebooks-a publishing format for reproducible computational workflows
Kluyver, T., Ragan-Kelley, B., Pérez, F., Granger, B. E., Bussonnier, M., Frederic, J., Kelley, K., Hamrick, J. B., Grout, J., Corlay, S., et al · 2016
Cited alongside, same era.
Li, K. and Malik, J · 2016
Cited alongside, same era.
Sparse and constrained attention for neural machine translation
Malaviya, C., Ferreira, P., and Martins, A. F · 2018
Later among the works it cites.
Continuous-time gaussian process motion planning via probabilistic inference
Mukadam, M., Dong, J., Yan, X., Dellaert, F., and Boots, B · 2018
Later among the works it cites.
Bock: Bayesian optimization with cylindrical kernels
Oh, C., Gavves, E., and Welling, M · 2018
Later among the works it cites.
Mpc-inspired neural network policies for sequential decision making
Pereira, M., Fan, D. D., An, G. N., and Theodorou, E · 2018
Later among the works it cites.
Meta-learning with latent embedding optimization
Rusu, A. A., Rao, D., Sygnowski, J., Vinyals, O., Pascanu, R., Osindero, S., and Hadsell, R · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unrolled generative adversarial networks
Metz, L., Poole, B., Pfau, D., and Sohl-Dickstein, J · 2016
Cited alongside, same era.
Hyperparameter optimization with approximate gradient
Pedregosa, F · 2016
Cited alongside, same era.
Bayesian optimization in a billion dimensions via random embeddings
Wang, Z., Hutter, F., Zoghi, M., Matheson, D., and de Feitas, N · 2016
Cited alongside, same era.
Optnet: Differentiable optimization as a layer in neural networks
Amos, B. and Kolter, J. Z · 2017
Cited alongside, same era.
Input convex neural networks
Amos, B., Xu, L., and Kolter, J. Z · 2017
Cited alongside, same era.
Robust locally-linear controllable embedding
Banijamali, E., Shu, R., Ghavamzadeh, M., Bui, H., and Ghodsi, A · 2017
Cited alongside, same era.
Goal-driven dynamics learning via Bayesian optimization
Bansal, S., Calandra, R., Xiao, T., Levine, S., and Tomlin, C. J · 2017
Cited alongside, same era.
Later among the works it cites.
Srinivas, A., Jabri, A., Abbeel, P., Levine, S., and Finn, C · 2018
Later among the works it cites.
Tasfi, N. and Capretz, M · 2018
Later among the works it cites.
Tassa, Y., Doron, Y., Muldal, A., Erez, T., Li, Y., Casas, D. d. L., Budden, D., Abdolmaleki, A., Merel, J., Lefrancq, A., et al · 2018
Later among the works it cites.
mwaskom/seaborn: v0.9.0 (july 2018), July 2018
Waskom, M., Botvinnik, O., O’Kane, D., Hobson, P., Ostblom, J., Lukauskas, S., Gemperline, D. C., Augspurger, T., Halchenko, Y., Cole, J. B., Warmenhoven, J., de Ruiter, J., Pye, C., Hoyer, S., Vanderplas, J., Villalba, S., Kunter, G., Quintero, E., Bachant, P., Martin, M., Meyer, K., Miles, A., Ram, Y., Brunner, T., Yarkoni, T., Williams, M. L., Evans, C., Fitzgerald, C., Brian, and Qalieh, A · 2018
Later among the works it cites.
Solar: Deep structured latent representations for model-based reinforcement learning
Zhang, M., Vikram, S., Smith, L., Abbeel, P., Johnson, M. J., and Levine, S · 2018
Later among the works it cites.
The limited multi-label projection layer
Amos, B., Koltun, V., and Kolter, J. Z · 2019
Closest in time.
Unsupervised state representation learning in atari
Anand, A., Racah, E., Ozair, S., Bengio, Y., Côté, M.-A., and Hjelm, R. D · 2019
Closest in time.
Bayesian optimization in variational latent spaces with dynamic compression
Antonova, R., Rai, A., Li, T., and Kragic, D · 2019
Closest in time.
Sequential dimension reduction for learning features of expensive black-box functions
Ben Salem, M., Bachoc, F., Roustant, O., Gamboa, F., and Tomaso, L · 2019
Closest in time.
Learning action representations for reinforcement learning
Chandak, Y., Theocharous, G., Kostas, J., Jordan, S., and Thomas, P. S · 2019
Closest in time.
Deepmdp: Learning continuous latent space models for representation learning
Gelada, C., Kumar, S., Buckman, J., Nachum, O., and Bellemare, M. G · 2019
Closest in time.
Robot motion planning in learned latent spaces
Ichter, B. and Pavone, M · 2019
Closest in time.
Adaptive and safe bayesian optimization in high dimensions via one-dimensional subspaces
Kirschner, J., Mutnỳ, M., Hiller, N., Ischebeck, R., and Krause, A · 2019
Closest in time.
Prediction, consistency, curvature: Representation learning for locally-linear control
Levine, N., Chow, Y., Shu, R., Li, A., Ghavamzadeh, M., and Bui, H · 2019
Closest in time.
Learning latent plans from play
Lynch, C., Khansari, M., Xiao, T., Kumar, V., Tompson, J., Levine, S., and Sermanet, P · 2019
Closest in time.
Disentangled state space representations
Miladinović, Đ., Gondal, M. W., Schölkopf, B., Buhmann, J. M., and Bauer, S · 2019
Closest in time.
Monte carlo gradient estimation in machine learning
Mohamed, S., Rosca, M., Figurnov, M., and Mnih, A · 2019
Closest in time.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., et al · 2019
Closest in time.
Meta-learning with implicit gradients
Rajeswaran, A., Finn, C., Kakade, S., and Levine, S · 2019
Closest in time.
Learning stabilizable nonlinear dynamics with contraction-based regularization
Singh, S., Richards, S. M., Sindhwani, V., Slotine, J.-J. E., and Pavone, M · 2019
Closest in time.
Meta-learning acquisition functions for bayesian optimization
Volpp, M., Fröhlich, L., Doerr, A., Hutter, F., and Daniel, C · 2019
Closest in time.
Exploring model-based planning with policy networks
Wang, T. and Ba, J · 2019
Closest in time.
Hydra - a framework for elegantly configuring complex applications
Yadan, O · 2019
Closest in time.
Unsupervised visuomotor control through distributional planning networks
Yu, T., Shevchuk, G., Sadigh, D., and Finn, C · 2019
Closest in time.
Zhang, Y., Hare, J., and Prügel-Bennett, A · 2019
Closest in time.
Fast context adaptation via meta-learning
Zintgraf, L., Shiarli, K., Kurin, V., Hofmann, K., and Whiteson, S · 2019
Closest in time.
Objective mismatch in model-based reinforcement learning
Lambert, N., Amos, B., Yadan, O., and Calandra, R · 2020
Closest in time.