Fetching the paper…
Reading the bibliography…
While reinforcement learning (RL) algorithms have been successfully applied across numerous sequential decision-making problems, their generalization to unforeseen testing environments remains a significant concern.
Stratospheric Aerosol Injection as a Deep Reinforcement Learning Problem
Christian Schroeder de Witt and Thomas Hornigold. 2019 · 1905
Earlier work this paper cites.
Sequential Tests of Statistical Hypotheses
A. Wald. 1945 · 1945
Earlier work this paper cites.
Continuous Inspection Schemes
E. S. Page. 1954 · 1954
Earlier work this paper cites.
Digital spectral analysis with applications
S Lawrence Marple Jr and William M Carey. 1989 · 1989
Earlier work this paper cites.
Isolation forest. In 2008 eighth ieee international conference on data mining . IEEE, 413–422
Fei Tony Liu, Kai Ming Ting, and Zhi-Hua Zhou. 2008 · 2008
Earlier work this paper cites.
Isolation-based anomaly detection
Fei Tony Liu, Kai Ming Ting, and Zhi-Hua Zhou. 2012 · 2012
Earlier work this paper cites.
MuJoCo: A physics engine for model-based control. In 2012 International Conference on Intelligent Robots and Systems
Emanuel Todorov, Tom Erez, and Yuval Tassa. 2012 · 2012
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Earlier work this paper cites.
Optimal sequential detection in multi-stream data
Hock Peng Chan. 2017 · 2017
Earlier work this paper cites.
Simple and scalable predictive uncertainty estimation using deep ensembles
Balaji Lakshminarayanan, Alexander Pritzel, and Charles Blundell. 2017 · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017b · 2017
Cited alongside, same era.
Time series feature extraction on basis of scalable hypothesis tests (tsfresh–a python package)
Maximilian Christ, Nils Braun, Julius Neuffer, and Andreas W Kempa-Liehr. 2018 · 2018
Cited alongside, same era.
Addressing function approximation error in actor-critic methods. In International conference on machine learning . PMLR, 1587–1596
Scott Fujimoto, Herke Hoof, and David Meger. 2018 · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor. In International conference on machine learning . PMLR
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine. 2018 · 2018
Out-of-Distribution Dynamics Detection: RL-Relevant Benchmarks and Results
Mohamad H Danesh and Alan Fern. 2021 · 2021
Later among the works it cites.
Benchmark for out-of-distribution detection in deep reinforcement learning
Aaqib Parvez Mohammed and Matias Valdenegro-Toro. 2021 · 2021
Later among the works it cites.
Expect the unexpected: unsupervised feature selection for automated sensor anomaly detection
Hui Yie Teh, I Kevin, Kai Wang, and Andreas W Kempa-Liehr. 2021 · 2021
Later among the works it cites.
Generalized out-of-distribution detection: A survey
Jingkang Yang, Kaiyang Zhou, Yixuan Li, and Ziwei Liu. 2021 · 2021
Later among the works it cites.
High-dimensional, multiscale online changepoint detection
Yudong Chen, Tengyao Wang, and Richard J Samworth. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Assessing generalization in deep reinforcement learning
Charles Packer, Katelyn Gao, Jernej Kos, Philipp Krähenbühl, Vladlen Koltun, and Dawn Song. 2018 · 2018
Cited alongside, same era.
Measuring the reliability of reinforcement learning algorithms
Stephanie CY Chan, Samuel Fishman, John Canny, Anoop Korattikara, and Sergio Guadarrama. 2019 · 2019
Cited alongside, same era.
Approximate entropy and sample entropy: A comprehensive tutorial
Alfonso Delgado-Bonal and Alexander Marshak. 2019 · 2019
Cited alongside, same era.
Uncertainty-based out-of-distribution classification in deep reinforcement learning
Andreas Sedlmeier, Thomas Gabor, Thomy Phan, Lenz Belzner, and Claudia Linnhoff-Popien. 2019 · 2019
Cited alongside, same era.
Learning dexterous in-hand manipulation
OpenAI: Marcin Andrychowicz, Bowen Baker, Maciek Chociej, Rafal Jozefowicz, Bob McGrew, Jakub Pachocki, Arthur Petron, Matthias Plappert, Glenn Powell, Alex Ray, et al · 2020
Cited alongside, same era.
i GM Ljung, Time series analysis: forecasting and control
GE Box, Gwilym M Jenkins, and Gregory C Reinsel. 2015a
Cited in the paper.
Time series analysis: forecasting and control
George EP Box, Gwilym M Jenkins, Gregory C Reinsel, and Greta M Ljung. 2015b
Cited in the paper.
Later among the works it cites.
Magnetic control of tokamak plasmas through deep reinforcement learning
Jonas Degrave, Federico Felici, Jonas Buchli, Michael Neunert, Brendan Tracey, Francesco Carpanese, Timo Ewalds, Roland Hafner, Abbas Abdolmaleki, Diego de Las Casas, et al · 2022
Later among the works it cites.
Illusionary Attacks on Sequential Decision Makers and Countermeasures
Tim Franzmeyer, João F Henriques, Jakob N Foerster, Philip HS Torr, Adel Bibi, and Christian Schroeder de Witt. 2022 · 2022
Later among the works it cites.
Towards Anomaly Detection in Reinforcement Learning. In Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems . 1799–1803
Robert Müller, Steffen Illium, Thomy Phan, Tom Haider, and Claudia Linnhoff-Popien. 2022 · 2022
Later among the works it cites.
Out-of-Distribution Detection for Reinforcement Learning Agents with Probabilistic Dynamics Models. In Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems . 851–859
Tom Haider, Karsten Roscher, Felippe Schmoeller da Roza, and Stephan Günnemann. 2023 · 2023
Later among the works it cites.
Champion-Level Drone Racing Using Deep Reinforcement Learning
Elia Kaufmann, Leonard Bauersfeld, Antonio Loquercio, Matthias Müller, Vladlen Koltun, and Davide Scaramuzza. 2023 · 2023
Later among the works it cites.