Fetching the paper…
Reading the bibliography…
We study zeroth-order optimization for convex functions where we further assume that function evaluations are unavailable.
Minimization methods for nonsmooth convex and quasiconvex functions
Yurii E Nesterov · 1984
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Adaptive stochastic approximation by the simultaneous perturbation method
James C Spall · 2000
Earlier work this paper cites.
Association of parameter, software, and hardware variation with large-scale behavior across 57,000 climate models
Christopher G Knight, Sylvia HE Knight, Neil Massey, Tolu Aina, Carl Christensen, Dave J Frame, Jamie A Kettleborough, Andrew Martin, Stephen Pascoe, Ben Sanderson, et al · 2007
Earlier work this paper cites.
1-bit compressive sensing
Petros T Boufounos and Richard G Baraniuk · 2008
Earlier work this paper cites.
Interactively shaping agents via human reinforcement: The tamer framework
W Bradley Knox and Peter Stone · 2009
Earlier work this paper cites.
Interactively optimizing information retrieval systems as a dueling bandits problem
Yisong Yue and Thorsten Joachims · 2009
Earlier work this paper cites.
CoSaMP: Iterative signal recovery from incomplete and inaccurate samples
Deanna Needell and Joel A Tropp · 2009
Earlier work this paper cites.
Concise formulas for the area and volume of a hyperspherical cap
Shengqiao Li · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio · 2012
Earlier work this paper cites.
Robust 1-bit compressed sensing and sparse logistic regression: A convex programming approach
Yaniv Plan and Roman Vershynin · 2012
Earlier work this paper cites.
Preference-based reinforcement learning: a formal framework and a policy iteration algorithm
Johannes Fürnkranz, Eyke Hüllermeier, Weiwei Cheng, and Sang-Hyeun Park · 2012
Earlier work this paper cites.
Generalization of value in reinforcement learning by humans
G Elliott Wimmer, Nathaniel D Daw, and Daphna Shohamy · 2012
Earlier work this paper cites.
Reinforcement learning from simultaneous human and MDP reward
W Bradley Knox and Peter Stone · 2012
Earlier work this paper cites.
Query complexity of derivative-free optimization
Kevin G Jamieson, Robert Nowak, and Ben Recht · 2012
Earlier work this paper cites.
Bandit theory meets compressed sensing for high dimensional stochastic linear bandit
Alexandra Carpentier and Rémi Munos · 2012
Cited alongside, same era.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Cited alongside, same era.
High-dimensional gaussian process bandits
Josip Djolonga, Andreas Krause, and Volkan Cevher · 2013
Cited alongside, same era.
An efficient approach for assessing hyperparameter importance
Frank Hutter, Holger Hoos, and Kevin Leyton-Brown · 2014
Cited alongside, same era.
Restricted strong convexity and its applications to convergence analysis of gradient-type methods in convex optimization
Hui Zhang and Lizhi Cheng · 2015
Cited alongside, same era.
Active subspaces: Emerging ideas for dimension reduction in parameter studies
Paul G Constantine · 2015
Stochastic zeroth-order optimization in high dimensions
Yining Wang, Simon Du, Sivaraman Balakrishnan, and Aarti Singh · 2018
Later among the works it cites.
Zeroth-order (non)-convex stochastic optimization via conditional gradient and gradient updates
Krishnakumar Balasubramanian and Saeed Ghadimi · 2018
Later among the works it cites.
CEM-RL: Combining evolutionary and gradient-based methods for policy search, 2018
Aloïs Pourchot and Olivier Sigaud · 2018
Later among the works it cites.
Derivative-free optimization methods
Jeffrey Larson, Matt Menickelly, and Stefan M Wild · 2019
Later among the works it cites.
Preference-based learning for exoskeleton gait optimization
Maegan Tucker, Ellen Novoseller, Claudia Kann, Yanan Sui, Yisong Yue, Joel Burdick, and Aaron D Ames · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Beyond convexity: Stochastic quasi-convex optimization
Elad Hazan, Kfir Y Levy, and Shai Shalev-Shwartz · 2015
Cited alongside, same era.
Online stochastic linear optimization under one-bit feedback
Lijun Zhang, Tianbao Yang, Rong Jin, Yichi Xiao, and Zhi-Hua Zhou · 2016
Cited alongside, same era.
Bayesian optimization in a billion dimensions via random embeddings
Ziyu Wang, Frank Hutter, Masrour Zoghi, David Matheson, and Nando de Feitas · 2016
Cited alongside, same era.
The power of normalization: Faster evasion of saddle points
Kfir Y Levy · 2016
Cited alongside, same era.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Cited alongside, same era.
A law of comparative judgment
Louis L Thurstone · 2017
Cited alongside, same era.
Minhao Cheng, Simranjit Singh, Patrick Chen, Pin-Yu Chen, Sijia Liu, and Cho-Jui Hsieh · 2019
Later among the works it cites.
Gradientless descent: High-dimensional zeroth-order optimization
Daniel Golovin, John Karro, Greg Kochanski, Chansoo Lee, Xingyou Song, et al · 2019
Later among the works it cites.
From complexity to simplicity: Adaptive es-active subspaces for blackbox optimization
Krzysztof M Choromanski, Aldo Pacchiano, Jack Parker-Holder, Yunhao Tang, and Vikas Sindhwani · 2019
Later among the works it cites.
Hoang Tran and Guannan Zhang · 2020
Closest in time.
A primer on zeroth-order optimization in signal processing and machine learning: Principals, recent advances, and applications
Sijia Liu, Pin-Yu Chen, Bhavya Kailkhura, Gaoyuan Zhang, Alfred O Hero III, and Pramod K Varshney · 2020
Closest in time.
Provably robust blackbox optimization for reinforcement learning
Krzysztof Choromanski, Aldo Pacchiano, Jack Parker-Holder, Yunhao Tang, Deepali Jain, Yuxiang Yang, Atil Iscen, Jasmine Hsu, and Vikas Sindhwani · 2020
Closest in time.
Human preference-based learning for high-dimensional optimization of exoskeleton walking gaits
Maegan Tucker, Myra Cheng, Ellen Novoseller, Richard Cheng, Yisong Yue, Joel W Burdick, and Aaron D Ames · 2020
Closest in time.
Coralia Cartis and Adilet Otemissov · 2020
Closest in time.
Zeroth-order regularized optimization (ZORO): Approximately sparse gradients and adaptive sampling
HanQin Cai, Daniel Mckenzie, Wotao Yin, and Zhenliang Zhang · 2022
Closest in time.