Fetching the paper…
Reading the bibliography…
Balancing performance and safety is crucial to deploying autonomous vehicles in multi-agent environments.
Random optimization
J. Matyas · 1965
Earlier work this paper cites.
Optimization by Vector Space Methods
D. Luenberger · 1969
Earlier work this paper cites.
Monte carlo sampling methods using markov chains and their applications
W. K. Hastings · 1970
Earlier work this paper cites.
Optimization by simulated annealing
S. Kirkpatrick, C. D. Gelatt, and M. P. Vecchi · 1983
Earlier work this paper cites.
Efficient monte carlo procedures for generating points uniformly distributed over bounded regions
R. L. Smith · 1984
Earlier work this paper cites.
Thermodynamical approach to the traveling salesman problem: An efficient simulation algorithm
V. Černỳ · 1985
Earlier work this paper cites.
A survey of some results in stochastic adaptive control
P. R. Kumar · 1985
Earlier work this paper cites.
Replica monte carlo simulation of spin-glasses
R. H. Swendsen and J.-S. Wang · 1986
Earlier work this paper cites.
State-space solutions to standard h 2 h_{2} and h ∞ h_{\infty} control problems
J. Doyle, K. Glover, P. Khargonekar, and B. Francis · 1988
Earlier work this paper cites.
Markov chain monte carlo maximum likelihood
C. J. Geyer · 1991
Earlier work this paper cites.
Implementation of the pure pursuit path tracking algorithm
R. C. Coulter · 1992
Earlier work this paper cites.
Simulated tempering: a new monte carlo scheme
E. Marinari and G. Parisi · 1992
Earlier work this paper cites.
Hit-and-run algorithms for generating multivariate distributions
C. J. Bélisle, H. E. Romeijn, and R. L. Smith · 1993
Earlier work this paper cites.
Simulated annealing: Practice versus theory
L. Ingber · 1993
Earlier work this paper cites.
Frequency domain uncertainty and the graph topology
G. Vinnicombe · 1993
Earlier work this paper cites.
Robust and optimal control
K. Zhou, J. C. Doyle, and K. Glover · 1996
Earlier work this paper cites.
Multitask learning
R. Caruana · 1997
Earlier work this paper cites.
On three-layer architectures
E. Gat, R. P. Bonnasso, R. Murphy, et al · 1998
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra · 1998
Earlier work this paper cites.
Robust model predictive control: A survey
A. Bemporad and M. Morari · 1999
Earlier work this paper cites.
Hit-and-run mixes fast
L. Lovász · 1999
Earlier work this paper cites.
Concentration of measure inequalities for Markov chains and ϕ \phi -mixing processes
P. Samson · 2000
Earlier work this paper cites.
Solving uncertain markov decision processes
J. A. Bagnell, A. Y. Ng, and J. G. Schneider · 2001
Earlier work this paper cites.
Trajectory generation for car-like robots using cubic curvature polynomials
B. Nagy and A. Kelly · 2001
Earlier work this paper cites.
The nonstochastic multiarmed bandit problem
P. Auer, N. Cesa-Bianchi, Y. Freund, and R. E. Schapire · 2002
Earlier work this paper cites.
Reactive nonholonomic trajectory generation via parametric optimal control
A. Kelly and B. Nagy · 2003
Earlier work this paper cites.
Hit-and-run is fast and fun
L. Lovász and S. Vempala · 2003
Earlier work this paper cites.
Inequalities for the l1 deviation of the empirical distribution
T. Weissman, E. Ordentlich, G. Seroussi, S. Verdu, and M. J. Weinberger · 2003
Earlier work this paper cites.
The price of robustness
D. Bertsimas and M. Sim · 2004
Earlier work this paper cites.
Robust control of markov decision processes with uncertain transition matrices
A. Nilim and L. El Ghaoui · 2005
Earlier work this paper cites.
Hit-and-run from a corner
L. Lovász and S. Vempala · 2006
Earlier work this paper cites.
Mapreduce: simplified data processing on large clusters
J. Dean and S. Ghemawat · 2008
Earlier work this paper cites.
Efficient projections onto thel1-ball for learning in high dimensions
J. Duchi, S. Shalev-Shwartz, Y. Singer, and T. Chandra · 2008
Earlier work this paper cites.
Motion planning in urban environments
D. Ferguson, T. M. Howard, and M. Likhachev · 2008
Cited alongside, same era.
Beating the adaptive bandit with high probability
J. Abernethy and A. Rakhlin · 2009
Cited alongside, same era.
Adaptive model-predictive motion planning for navigation in complex environments
T. M. Howard · 2009
Cited alongside, same era.
Automatic steering methods for autonomous automobile path tracking
J. M. Snider et al · 2009
Cited alongside, same era.
Algorithms for adversarial bandit problems with multiple plays
T. Uchiya, A. Nakamura, and M. Kudo · 2010
Cited alongside, same era.
Annealing adaptive search, cross-entropy, and stochastic approximation in global optimization
J. Hu and P. Hu · 2011
Cited alongside, same era.
Commonroad: Composable benchmarks for motion planning on roads
M. Althoff, M. Koschi, and S. Manzinger · 2017
Later among the works it cites.
Emergent complexity via multi-agent competition
T. Bansal, J. Pachocki, S. Sidor, I. Sutskever, and I. Mordatch · 2017
Later among the works it cites.
Population based training of neural networks
M. Jaderberg, V. Dalibard, S. Osindero, W. M. Czarnecki, J. Donahue, A. Razavi, O. Vinyals, T. Green, I. Dunning, K. Simonyan, et al · 2017
Later among the works it cites.
Adversarially robust policy learning: Active construction of physically-plausible perturbations
A. Mandlekar, Y. Zhu, A. Garg, L. Fei-Fei, and S. Savarese · 2017
Later among the works it cites.
Variance regularization with convex objectives
H. Namkoong and J. C. Duchi · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sampling-based algorithms for optimal motion planning
S. Karaman and E. Frazzoli · 2011
Cited alongside, same era.
Parallel algorithms for real-time motion planning
M. McNaughton · 2011
Cited alongside, same era.
Lqg-mp: Optimized path planning for robots with motion uncertainty and imperfect state information
J. Van Den Berg, P. Abbeel, and K. Goldberg · 2011
Cited alongside, same era.
Random search for hyper-parameter optimization
J. Bergstra and Y. Bengio · 2012
Cited alongside, same era.
A survey of monte carlo tree search methods
C. B. Browne, E. Powley, D. Whitehouse, S. M. Lucas, P. I. Cowling, P. Rohlfshagen, S. Tavener, D. Perez, S. Samothrakis, and S. Colton · 2012
Cited alongside, same era.
Regret analysis of stochastic and nonstochastic multi-armed bandit problems
S. Bubeck, N. Cesa-Bianchi, et al · 2012
Cited alongside, same era.
Masked autoregressive flow for density estimation
G. Papamakarios, T. Pavlakou, and I. Murray · 2017
Later among the works it cites.
Robust adversarial reinforcement learning
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta · 2017
Later among the works it cites.
Certifiable distributional robustness with principled adversarial training
A. Sinha, H. Namkoong, and J. Duchi · 2017
Later among the works it cites.
Cddt: Fast approximate 2d ray casting for accelerated localization
C. Walsh and S. Karaman · 2017
Later among the works it cites.
Autonomous racing with autorally vehicles and differential games
G. Williams, B. Goldfain, P. Drews, J. M. Rehg, and E. A. Theodorou · 2017
Later among the works it cites.
Efficient gradient-free variational inference using policy search
O. Arenz, M. Zhong, G. Neumann, et al · 2018
Later among the works it cites.
Regret bounds for robust adaptive control of the linear quadratic regulator
S. Dean, H. Mania, N. Matni, B. Recht, and S. Tu · 2018
Later among the works it cites.
Probabilistic model-agnostic meta-learning
C. Finn, K. Xu, and S. Levine · 2018
Later among the works it cites.
Addressing function approximation error in actor-critic methods
S. Fujimoto, H. Van Hoof, and D. Meger · 2018
Later among the works it cites.
Parallel wavenet: Fast high-fidelity speech synthesis
A. Oord, Y. Li, I. Babuschkin, K. Simonyan, O. Vinyals, K. Kavukcuoglu, G. Driessche, E. Lockhart, L. Cobo, F. Stimberg, et al · 2018
Later among the works it cites.
A general reinforcement learning algorithm that masters chess, shogi, and go through self-play
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel, et al · 2018
Later among the works it cites.
Worst-case analysis of the time-to-react using reachable sets
S. Sontges, M. Koschi, and M. Althoff · 2018
Later among the works it cites.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Later among the works it cites.
Budget-constrained multi-armed bandits with multiple plays
D. P. Zhou and C. J. Tomlin · 2018
Later among the works it cites.
Online control with adversarial disturbances
N. Agarwal, B. Bullins, E. Hazan, S. M. Kakade, and K. Singh · 2019
Later among the works it cites.
Alphastar: An evolutionary computation perspective
K. Arulkumaran, A. Cully, and J. Togelius · 2019
Later among the works it cites.
Dota 2 with large scale deep reinforcement learning
C. Berner, G. Brockman, B. Chan, V. Cheung, P. Dkebiak, C. Dennison, D. Farhi, Q. Fischer, S. Hashme, C. Hesse, et al · 2019
Later among the works it cites.
W. Ding and S. Shen · 2019
Later among the works it cites.
Formula one 2019 results
Federation Internationale de l’Automobile · 2019
Later among the works it cites.
Adversarial policies: Attacking deep reinforcement learning
A. Gleave, M. Dennis, N. Kant, C. Wild, S. Levine, and S. Russell · 2019
Later among the works it cites.
Human-level performance in 3d multiplayer games with population-based reinforcement learning
M. Jaderberg, W. M. Czarnecki, I. Dunning, L. Marris, G. Lever, A. G. Castaneda, C. Beattie, N. C. Rabinowitz, A. S. Morcos, A. Ruderman, et al · 2019
Later among the works it cites.
A noncooperative game approach to autonomous racing
A. Liniger and J. Lygeros · 2019
Later among the works it cites.
Efficient black-box assessment of autonomous vehicle safety
J. Norden, M. O’Kelly, and A. Sinha · 2019
Later among the works it cites.
Technical Report: TunerCar: A Superoptimization Toolchain for Autonomous Racing
M. O’Kelly, H. Zheng, J. Auckley, A. Jain, K. Luong, and R. Mangharam · 2019
Later among the works it cites.
Jointly learnable behavior and trajectory planning for self-driving vehicles
A. Sadat, M. Ren, A. Pokrovsky, Y.-C. Lin, E. Yumer, and R. Urtasun · 2019
Later among the works it cites.
Distributionally robust reinforcement learning
E. Smirnova, E. Dohmatob, and J. Mary · 2019
Later among the works it cites.
Game theoretic motion planning for multi-robot racing
Z. Wang, R. Spica, and M. Schwager · 2019
Later among the works it cites.