Fetching the paper…
Reading the bibliography…
This paper studies the constrained/safe reinforcement learning (RL) problem with sparse indicator signals for constraint violations.
E. Altman, “Constrained markov decision processes with total cost criteria: Lagrangian approach and dual linear program,” Mathematical methods of operations research , vol. 48, no. 3, pp. 387–417, 1998
1998
Earlier work this paper cites.
C. Gaskett, “Reinforcement learning under circumstances beyond its control,” 2003
2003
Earlier work this paper cites.
P. Geibel and F. Wysotzki, “Risk-sensitive reinforcement learning applied to control under constraints,” Journal of Artificial Intelligence Research , vol. 24, pp. 81–108, 2005
2005
Earlier work this paper cites.
Y. Sun, A. K. Wong, and M. S. Kamel, “Classification of imbalanced data: A review,” International journal of pattern recognition and artificial intelligence , vol. 23, no. 04, pp. 687–719, 2009
2009
Earlier work this paper cites.
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
Z. I. Botev, D. P. Kroese, R. Y. Rubinstein, and P. L’Ecuyer, “The cross-entropy method for optimization,” in Handbook of statistics . Elsevier, 2013, vol. 31, pp. 35–59
2013
Earlier work this paper cites.
J. Garcıa and F. Fernández, “A comprehensive survey on safe reinforcement learning,” Journal of Machine Learning Research , vol. 16, no. 1, pp. 1437–1480, 2015
2015
Earlier work this paper cites.
Y. Sui, A. Gotovos, J. Burdick, and A. Krause, “Safe exploration for optimization with gaussian processes,” in International Conference on Machine Learning , 2015, pp. 997–1005
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2017
Earlier work this paper cites.
F. Berkenkamp, M. Turchetta, A. Schoellig, and A. Krause, “Safe model-based reinforcement learning with stability guarantees,” in Advances in neural information processing systems , 2017, pp. 908–918
2017
Cited alongside, same era.
G. Ke, Q. Meng, T. Finley, T. Wang, W. Chen, W. Ma, Q. Ye, and T.-Y. Liu, “Lightgbm: A highly efficient gradient boosting decision tree,” in Advances in neural information processing systems , 2017, pp. 3146–3154
2017
Cited alongside, same era.
P. Drews, G. Williams, B. Goldfain, E. A. Theodorou, and J. M. Rehg, “Aggressive deep driving: Combining convolutional neural networks and model predictive control,” in Conference on Robot Learning , 2017, pp. 133–142
2017
Cited alongside, same era.
T. Koller, F. Berkenkamp, M. Turchetta, and A. Krause, “Learning-based model predictive control for safe exploration,” in 2018 IEEE Conference on Decision and Control (CDC) . IEEE, 2018, pp. 6059–6066
2018
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
Y. Wang, W. Gan, J. Yang, W. Wu, and J. Yan, “Dynamic curriculum learning for imbalanced data classification,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , October 2019
2019
Later among the works it cites.
2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T.-H. Pham, G. De Magistris, and R. Tachibana, “Optlayer-practical constrained optimization for deep reinforcement learning in the real world,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 6236–6243
2018
Cited alongside, same era.
M. Wen and U. Topcu, “Constrained cross-entropy method for safe reinforcement learning,” in Advances in Neural Information Processing Systems , 2018, pp. 7450–7460
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Y. Chow, O. Nachum, E. Duenez-Guzman, and M. Ghavamzadeh, “A lyapunov-based approach to safe reinforcement learning,” in Advances in neural information processing systems , 2018, pp. 8092–8101
2018
Cited alongside, same era.
K. Chua, R. Calandra, R. McAllister, and S. Levine, “Deep reinforcement learning in a handful of trials using probabilistic dynamics models,” in Advances in Neural Information Processing Systems , 2018, pp. 4754–4765
2018
Cited alongside, same era.
A. Nagabandi, G. Kahn, R. S. Fearing, and S. Levine, “Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 7559–7566
2018
Cited alongside, same era.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
A. Wachi and Y. Sui, “Safe reinforcement learning in constrained markov decision processes,” in International Conference on Machine Learning (ICML) , 2020
2020
Closest in time.
2020
Closest in time.
M. Okada and T. Taniguchi, “Variational inference mpc for bayesian model-based reinforcement learning,” in Conference on Robot Learning , 2020, pp. 258–272
2020
Closest in time.