Fetching the paper…
Reading the bibliography…
Safety is a critical component of autonomous systems and remains a challenge for learning-based policies to be utilized in the real world.
Integrating grid-based and topological maps for mobile robot navigation,
S. Thrun, A. Bücken, · 1996
Earlier work this paper cites.
K. Zhou, J. C. Doyle, Essentials of robust control, volume 104, Prentice hall Upper Saddle River, NJ, 1998
1998
Earlier work this paper cites.
Some pac-bayesian theorems,
D. A. McAllester, · 1999
Earlier work this paper cites.
(Not) bounding the true error,
J. Langford, R. Caruana, · 2002
Earlier work this paper cites.
Introduction to statistical learning theory,
O. Bousquet, S. Boucheron, G. Lugosi, · 2003
Earlier work this paper cites.
Autonomous vision-based exploration and mapping using hybrid maps and Rao-Blackwellised particle filters,
R. Sim, J. J. Little, · 2006
Earlier work this paper cites.
Visual navigation for mobile robots: A survey,
F. Bonin-Font, A. Ortiz, G. Oliver, · 2008
Earlier work this paper cites.
A tutorial on conformal prediction.,
G. Shafer, V. Vovk, · 2008
Earlier work this paper cites.
2010
Earlier work this paper cites.
Algorithms for cvar optimization in mdps,
Y. Chow, M. Ghavamzadeh, · 2014
Earlier work this paper cites.
Reinforcement learning with multi-fidelity simulators,
M. Cutler, T. J. Walsh, J. P. How, · 2014
Earlier work this paper cites.
A comprehensive survey on safe reinforcement learning,
J. García, F. Fernández, · 2015
Earlier work this paper cites.
Reach-Avoid Problems with Time-Varying Dynamics, Targets and Constraints,
J. F. Fisac, M. Chen, C. J. Tomlin, S. S. Sastry, · 2015
Earlier work this paper cites.
On the uniform convergence of relative frequencies of events to their probabilities,
V. N. Vapnik, A. Y. Chervonenkis, · 2015
Earlier work this paper cites.
Safe controller optimization for quadrotors with gaussian processes,
F. Berkenkamp, A. P. Schoellig, A. Krause, · 2016
Earlier work this paper cites.
Target-driven visual navigation in indoor scenes using deep reinforcement learning,
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, A. Farhadi, · 2017
Earlier work this paper cites.
Domain randomization for transferring deep neural networks from simulation to the real world,
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, P. Abbeel, · 2017
Earlier work this paper cites.
Cad2rl: Real single-image flight without a single real image,
F. Sadeghi, S. Levine, · 2017
Earlier work this paper cites.
Risk-constrained reinforcement learning with percentile risk criteria,
Y. Chow, M. Ghavamzadeh, L. Janson, M. Pavone, · 2017
Earlier work this paper cites.
Funnel libraries for real-time robust feedback motion planning,
A. Majumdar, R. Tedrake, · 2017
Earlier work this paper cites.
Robust online motion planning via contraction theory and convex optimization,
S. Singh, A. Majumdar, J.-J. Slotine, M. Pavone, · 2017
Earlier work this paper cites.
Hamilton-jacobi reachability: A brief overview and recent advances,
S. Bansal, M. Chen, S. Herbert, C. J. Tomlin, · 2017
Earlier work this paper cites.
Computing nonvacuous generalization bounds for deep (stochastic) neural networks with many more parameters than training data,
G. K. Dziugaite, D. M. Roy, · 2017
Earlier work this paper cites.
Cognitive mapping and planning for visual navigation,
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, J. Malik, · 2017
Earlier work this paper cites.
Safe visual navigation via deep learning and novelty detection,
C. Richter, N. Roy, · 2017
Earlier work this paper cites.
Uncertainty-aware reinforcement learning for collision avoidance,
G. Kahn, A. Villaflor, V. Pong, P. Abbeel, S. Levine, · 2017
Cited alongside, same era.
2018
Cited alongside, same era.
Learning-based model predictive control for safe exploration,
T. Koller, F. Berkenkamp, M. Turchetta, A. Krause, · 2018
Cited alongside, same era.
Soft Actor-Critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,
T. Haarnoja, A. Zhou, P. Abbeel, S. Levine, · 2018
Cited alongside, same era.
Safe reinforcement learning via shielding,
M. Alshiekh, R. Bloem, R. Ehlers, B. Könighofer, S. Niekum, U. Topcu, · 2018
Cited alongside, same era.
Active domain randomization,
B. Mehta, M. Diaz, F. Golemo, C. J. Pal, L. Paull, · 2020
Later among the works it cites.
One solution is not all you need: Few-shot extrapolation via structured MaxEnt RL,
S. Kumar, A. Kumar, S. Levine, C. Finn, · 2020
Later among the works it cites.
Dynamics-aware unsupervised discovery of skills,
A. Sharma, S. Gu, S. Levine, V. Kumar, K. Hausman, · 2020
Later among the works it cites.
Coresets via bilevel optimization for continual learning and streaming,
Z. Borsos, M. Mutny, A. Krause, · 2020
Later among the works it cites.
RMA: Rapid Motor Adaptation for Legged Robots,
A. Kumar, Z. Fu, D. Pathak, J. Malik, · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, D. Quillen, · 2018
Cited alongside, same era.
Leave no trace: Learning to reset for safe and autonomous reinforcement learning,
B. Eysenbach, S. Gu, J. Ibarz, S. Levine, · 2018
Cited alongside, same era.
Stronger generalization bounds for deep nets via a compression approach,
S. Arora, R. Ge, B. Neyshabur, Y. Zhang, · 2018
Cited alongside, same era.
A general safety framework for learning-based control in uncertain robotic systems,
J. F. Fisac, A. K. Akametalu, M. N. Zeilinger, S. Kaynama, J. Gillula, C. J. Tomlin, · 2019
Cited alongside, same era.
Bridging hamilton-jacobi safety analysis and reinforcement learning,
J. F. Fisac, N. F. Lugovoy, V. Rubies-Royo, S. Ghosh, C. J. Tomlin, · 2019
Cited alongside, same era.
End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks,
R. Cheng, G. Orosz, R. M. Murray, J. W. Burdick, · 2019
Cited alongside, same era.
Diversity is all you need: Learning skills without a reward function,
B. Eysenbach, A. Gupta, J. Ibarz, S. Levine, · 2019
Cited alongside, same era.
2021
Later among the works it cites.
3D-FRONT: 3D Furnished Rooms With layOuts and semaNTics,
H. Fu, B. Cai, L. Gao, L.-X. Zhang, J. Wang, C. Li, Q. Zeng, C. Sun, R. Jia, B. Zhao, H. Zhang, · 2021
Later among the works it cites.
Safety and liveness guarantees through reach-avoid reinforcement learning,
K.-C. Hsu, V. Rubies-Royo, C. J. Tomlin, J. F. Fisac, · 2021
Later among the works it cites.
Recovery RL: Safe reinforcement learning with learned recovery zones,
B. Thananjeyan, A. Balakrishna, S. Nair, M. Luo, K. Srinivasan, M. Hwang, J. E. Gonzalez, J. Ibarz, C. Finn, K. Goldberg, · 2021
Later among the works it cites.
PAC-Bayes Control: Learning policies that provably generalize to novel environments,
A. Majumdar, A. Farid, A. Sonar, · 2021
Later among the works it cites.
Task-driven out-of-distribution detection with statistical guarantees for robot learning,
A. Farid, S. Veer, A. Majumdar, · 2021
Later among the works it cites.
Probably approximately correct vision-based planning using motion primitives,
S. Veer, A. Majumdar, · 2021
Later among the works it cites.
2021
Later among the works it cites.
Tighter risk certificates for neural networks,
M. Pérez-Ortiz, O. Rivasplata, J. Shawe-Taylor, C. Szepesvári, · 2021
Later among the works it cites.
Generalization guarantees for imitation learning,
A. Z. Ren, S. Veer, A. Majumdar, · 2021
Later among the works it cites.
Learning provably robust motion planners using funnel libraries,
A. E. Gurgen, A. Majumdar, S. Veer, · 2021
Later among the works it cites.
A. Agarwal, S. Veer, A. Z. Ren, A. Majumdar, · 2021
Later among the works it cites.
Planar robot casting with real2sim2real self-supervised learning,
V. Lim, H. Huang, L. Y. Chen, J. Wang, J. Ichnowski, D. Seita, M. Laskey, K. Goldberg, · 2021
Later among the works it cites.
Data-efficient domain randomization with bayesian optimization,
F. Muratore, C. Eilers, M. Gienger, J. Peters, · 2021
Later among the works it cites.
Boston-Dynamics, Inside the Lab: Robotics After Hours, https://www.youtube.com/watch?v=Jq0GknnKvXM , 2022
2022
Closest in time.
Failure prediction with statistical guarantees for vision-based robot control,
A. Farid, D. Snyder, A. Z. Ren, A. Majumdar, · 2022
Closest in time.
FutureCar, A look at how waymo’s self-driving test fleet safely traveled 2.7 million miles in san francisco last year, https://www.futurecar.com/5158/A-Look-at-How-Waymos-Self-Driving-Test-Fleet-Safely-Traveled-2-7-Million-Miles-in-San-Francisco-Last-Year , 2022
2022
Closest in time.
Fogros 2: An adaptive and extensible platform for cloud and fog robotics using ros 2,
J. Ichnowski, K. Chen, K. Dharmarajan, S. Adebola, M. Danielczuk, V. Mayoral-Vilches, H. Zhan, D. Xu, R. Ghassemi, J. Kubiatowicz, et al., · 2022
Closest in time.
Robust h-infinity control for uncertain stochastic systems with state delay,
S. Xu, T. Chen, · 2094
Closest in time.