Fetching the paper…
Reading the bibliography…
The dramatic increase of autonomous systems subject to variable environments has given rise to the pressing need to consider risk in both the synthesis and verification of policies for these systems.
M. M. Flood, The traveling-salesman problem, Operations research 4 (1) (1956) 61–75
1956
Earlier work this paper cites.
G. E. Monahan, State of the art—a survey of partially observable markov decision processes: theory, models, and algorithms, Management science 28 (1) (1982) 1–16
1982
Earlier work this paper cites.
M. L. Puterman, Markov Decision Processes: Discrete Stochastic Dynamic Programming, 1st Edition, John Wiley & Sons, Inc., USA, 1994
1994
Earlier work this paper cites.
M. Heger, Consideration of risk in reinforcement learning, in: Machine Learning Proceedings 1994, Elsevier, 1994, pp. 105–111
1994
Earlier work this paper cites.
D. P. Bertsekas, J. N. Tsitsiklis, Neuro-dynamic programming: an overview, in: Proceedings of 1995 34th IEEE conference on decision and control, Vol. 1, IEEE, 1995, pp. 560–564
1995
Earlier work this paper cites.
L. P. Kaelbling, M. L. Littman, A. W. Moore, Reinforcement learning: A survey, Journal of artificial intelligence research 4 (1996) 237–285
1996
Earlier work this paper cites.
P. Artzner, F. Delbaen, J.-M. Eber, D. Heath, Coherent measures of risk, Mathematical finance 9 (3) (1999) 203–228
1999
Earlier work this paper cites.
O. Mihatsch, R. Neuneier, Risk-sensitive reinforcement learning, Machine learning 49 (2) (2002) 267–290
2002
Earlier work this paper cites.
F. Delbaen, Coherent risk measures on general probability spaces, in: Advances in finance and stochastics, Springer, 2002, pp. 1–37
2002
Earlier work this paper cites.
W. Korb, D. Engel, R. Boesecke, G. Eggers, R. Marmulla, N. O’Sullivan, J. Raczkowsky, S. Hassfeld, Risk analysis for a reliable and safe surgical robot system, in: International Congress Series, Vol. 1256, Elsevier, 2003, pp. 766–770
2003
Earlier work this paper cites.
P. Geibel, F. Wysotzki, Risk-sensitive reinforcement learning applied to control under constraints, Journal of Artificial Intelligence Research 24 (2005) 81–108
2005
Earlier work this paper cites.
A. Corso, R. J. Moss, M. Koren, R. Lee, M. J. Kochenderfer, A survey of algorithms for black-box safety validation, arXiv e-prints (2020) arXiv–2005
2005
Earlier work this paper cites.
D. B. Brown, Large deviations bounds for estimating conditional value-at-risk, Operations Research Letters 35 (6) (2007) 722–730
2007
Earlier work this paper cites.
L. T. Dung, T. Komeda, M. Takagi, Reinforcement learning for pomdp using state classification, Applied Artificial Intelligence 22 (7-8) (2008) 761–779
2008
Earlier work this paper cites.
M. C. Campi, S. Garatti, The exact feasibility of randomized solutions of uncertain convex programs, SIAM Journal on Optimization 19 (3) (2008) 1211–1230
2008
Earlier work this paper cites.
A. Donzé, O. Maler, Robust satisfaction of temporal logic over real-valued signals, in: International Conference on Formal Modeling and Analysis of Timed Systems, Springer, 2010, pp. 92–106
2010
Earlier work this paper cites.
S. Png, J. Pineau, Bayesian reinforcement learning for pomdp-based dialogue systems, in: 2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), IEEE, 2011, pp. 2156–2159
2011
Earlier work this paper cites.
H. A. Taha, Operations research: an introduction, Vol. 790, Pearson/Prentice Hall Upper Saddle River, NJ, USA, 2011
2011
Cited alongside, same era.
E. ISO, 10218: Robots and robotic devices-safety requirements for industrial robots-part 1: Robots, ISO: Geneve, Switzerland (2011)
2011
Cited alongside, same era.
I. ISO, 10218-2: 2011: Robots and robotic devices–safety requirements for industrial robots–part 2: Robot systems and integration, Geneva, Switzerland: International Organization for Standardization 3 (2011)
2011
Cited alongside, same era.
A. Ahmadi-Javid, Entropic value-at-risk: A new coherent risk measure, Journal of Optimization Theory and Applications 155 (3) (2012) 1105–1123
2012
Cited alongside, same era.
A. Hakobyan, G. C. Kim, I. Yang, Risk-aware motion planning and control using cvar-constrained optimization, IEEE Robotics and Automation letters 4 (4) (2019) 3924–3931
2019
Later among the works it cites.
F. Vicentini, M. Askarpour, M. G. Rossi, D. Mandrioli, Safety assessment of collaborative robotics through automated formal verification, IEEE Transactions on Robotics 36 (1) (2019) 42–61
2019
Later among the works it cites.
A. Corso, P. Du, K. Driggs-Campbell, M. J. Kochenderfer, Adaptive stress testing with reward augmentation for autonomous vehicle validatio, in: 2019 IEEE Intelligent Transportation Systems Conference (ITSC), IEEE, 2019, pp. 163–168
2019
Later among the works it cites.
P. Thomas, E. Learned-Miller, Concentration inequalities for conditional value at risk, in: International Conference on Machine Learning, PMLR, 2019, pp. 6225–6233
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
V. Raman, A. Donzé, M. Maasoumy, R. M. Murray, A. Sangiovanni-Vincentelli, S. A. Seshia, Model predictive control with signal temporal logic specifications, in: 53rd IEEE Conference on Decision and Control, IEEE, 2014, pp. 81–87
2014
Cited alongside, same era.
2015
Cited alongside, same era.
X. Xu, P. Tabuada, J. W. Grizzle, A. D. Ames, Robustness of control barrier functions for safety critical control, IFAC-PapersOnLine 48 (27) (2015) 54–61
2015
Cited alongside, same era.
A. D. Ames, X. Xu, J. W. Grizzle, P. Tabuada, Control barrier function based quadratic programs for safety critical systems, IEEE Transactions on Automatic Control 62 (8) (2016) 3861–3876
2016
Cited alongside, same era.
R. Calandra, A. Seyfarth, J. Peters, M. P. Deisenroth, Bayesian optimization for learning gaits under uncertainty, Annals of Mathematics and Artificial Intelligence 76 (1) (2016) 5–23
2016
Cited alongside, same era.
Y. Chow, M. Ghavamzadeh, L. Janson, M. Pavone, Risk-constrained reinforcement learning with percentile risk criteria, The Journal of Machine Learning Research 18 (1) (2017) 6070–6120
2017
Cited alongside, same era.
J. Deshmukh, M. Horvat, X. Jin, R. Majumdar, V. S. Prabhu, Testing cyber-physical systems through bayesian optimization, ACM Transactions on Embedded Computing Systems (TECS) 16 (5s) (2017) 1–18
2017
Cited alongside, same era.
2019
Later among the works it cites.
doi:10.1109/MCS.2019.2949973
S. Wilson, P. Glotfelter, L. Wang, S. Mayya, G. Notomista, M. Mote, M. Egerstedt, The robotarium: Globally impactful opportunities, challenges, and lessons learned in remote-access, distributed control of multirobot systems, IEEE Control Systems Magazine 40 (1) (2020) 26–44 · 2019
Later among the works it cites.
C. Doersch, A. Zisserman, Sim2real transfer learning for 3d human pose estimation: motion to the rescue, Advances in Neural Information Processing Systems 32 (2019)
2019
Later among the works it cites.
S. Bhattacharya, S. Badyal, T. Wheeler, S. Gil, D. Bertsekas, Reinforcement learning for pomdp: Partitioned rollout and policy iteration with application to autonomous sequential repair problems, IEEE Robotics and Automation Letters 5 (3) (2020) 3967–3974
2020
Later among the works it cites.
A. Majumdar, M. Pavone, How should a robot assess risk? towards an axiomatic theory of risk in robotics, in: Robotics Research, Springer, 2020, pp. 75–84
2020
Later among the works it cites.
M. Koren, M. J. Kochenderfer, Adaptive stress testing without domain heuristics using go-explore, in: 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC), IEEE, 2020, pp. 1–6
2020
Later among the works it cites.
Z. Mhammedi, B. Guedj, R. C. Williamson, Pac-bayesian bound for the conditional value at risk, Advances in Neural Information Processing Systems 33 (2020) 17919–17930
2020
Later among the works it cites.
R. Shalloo, S. Dann, J.-N. Gruse, C. Underwood, A. Antoine, C. Arran, M. Backhouse, C. Baird, M. Balcazar, N. Bourgeois, et al., Automation and control of laser wakefield accelerators using bayesian optimization, Nature communications 11 (1) (2020) 1–8
2020
Later among the works it cites.
A. Kadian, J. Truong, A. Gokaslan, A. Clegg, E. Wijmans, S. Lee, M. Savva, S. Chernova, D. Batra, Sim2real predictivity: Does evaluation in simulation predict real-world performance?, IEEE Robotics and Automation Letters 5 (4) (2020) 6670–6677
2020
Later among the works it cites.
2021
Later among the works it cites.
F. Berkenkamp, A. Krause, A. P. Schoellig, Bayesian optimization with safety constraints: safe and automatic parameter tuning in robotics, Machine Learning (2021) 1–35
2021
Later among the works it cites.