Fetching the paper…
Reading the bibliography…
Optimizing the combustion efficiency of a thermal power generating unit (TPGU) is a highly challenging and critical task in the energy industry.
Reinforcement learning applications
Li, Y. 2019 · 1908
Earlier work this paper cites.
Behavior Regularized Offline Reinforcement Learning
Wu, Y.; Tucker, G.; and Nachum, O. 2019 · 1911
Earlier work this paper cites.
Model predictive control: theory and practice—a survey
Garcia, C. E.; Prett, D. M.; and Morari, M. 1989 · 1989
Earlier work this paper cites.
Constrained Markov decision processes , volume 7
Altman, E. 1999 · 1999
Earlier work this paper cites.
The explicit linear quadratic regulator for constrained systems
Bemporad, A.; Morari, M.; Dua, V.; and Pistikopoulos, E. N. 2002 · 2002
Earlier work this paper cites.
An empirical investigation of the challenges of real-world reinforcement learning
Dulac-Arnold, G.; Levine, N.; Mankowitz, D. J.; Li, J.; Paduraru, C.; Gowal, S.; and Hester, T. 2020 · 2003
Earlier work this paper cites.
Artificial intelligence for the modeling and control of combustion processes: a review
Kalogirou, S. A. 2003 · 2003
Earlier work this paper cites.
A survey of industrial model predictive control technology
Qin, S. J.; and Badgwell, T. A. 2003 · 2003
Earlier work this paper cites.
Convex optimization
Boyd, S.; Boyd, S. P.; and Vandenberghe, L. 2004 · 2004
Earlier work this paper cites.
D4rl: Datasets for deep data-driven reinforcement learning
Fu, J.; Kumar, A.; Nachum, O.; Tucker, G.; and Levine, S. 2020 · 2004
Earlier work this paper cites.
Offline reinforcement learning: Tutorial, review, and perspectives on open problems
Levine, S.; Kumar, A.; Tucker, G.; and Fu, J. 2020 · 2005
Earlier work this paper cites.
Advanced PID control
Åström, K.; and Hägglund, T. 2006 · 2006
Earlier work this paper cites.
Neural network-based modeling for a large-scale power plant
Lee, K. Y.; Heo, J. S.; Hoffman, J. A.; Kim, S.-H.; and Jung, W.-H. 2007 · 2007
Earlier work this paper cites.
Neural network based superheater steam temperature control for a large-scale supercritical boiler unit
Ma, L.; and Lee, K. Y. 2011 · 2011
Cited alongside, same era.
Batch reinforcement learning
Lange, S.; Gabel, T.; and Riedmiller, M. 2012 · 2012
Cited alongside, same era.
Mujoco: A physics engine for model-based control
Todorov, E.; Erez, T.; and Tassa, Y. 2012 · 2012
Cited alongside, same era.
Auto-Encoding Variational Bayes
Kingma, D. P.; and Welling, M. 2014 · 2014
Cited alongside, same era.
Integrating multi-objective optimization with computational fluid dynamics to optimize boiler combustion process of a coal fired power plant
Liu, X.; and Bansal, R. 2014 · 2014
Cited alongside, same era.
Scheduled sampling for sequence prediction with recurrent Neural networks
Bengio, S.; Vinyals, O.; Jaitly, N.; and Shazeer, N. 2015 · 2015
Sensitivity and Generalization in Neural Networks: an Empirical Study
Novak, R.; Bahri, Y.; Abolafia, D. A.; Pennington, J.; and Sohl-Dickstein, J. 2018 · 2018
Later among the works it cites.
Reward constrained policy optimization
Tessler, C.; Mankowitz, D. J.; and Mannor, S. 2018 · 2018
Later among the works it cites.
When to trust your model: Model-based policy optimization
Janner, M.; Fu, J.; Zhang, M.; and Levine, S. 2019 · 2019
Later among the works it cites.
Stabilizing off-policy q-learning via bootstrapping error reduction
Kumar, A.; Fu, J.; Soh, M.; Tucker, G.; and Levine, S. 2019 · 2019
Later among the works it cites.
Virtual-taobao: Virtualizing real-world online retail environment for reinforcement learning
Shi, J.-C.; Yu, Y.; Da, Q.; Chen, S.-Y.; and Zeng, A.-X. 2019 · 2019
Later among the works it cites.
MOReL: Model-Based Offline Reinforcement Learning
Kidambi, R.; Rajeswaran, A.; Netrapalli, P.; and Joachims, T. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V.; Kavukcuoglu, K.; Silver, D.; Rusu, A. A.; Veness, J.; Bellemare, M. G.; Graves, A.; Riedmiller, M.; Fidjeland, A. K.; Ostrovski, G.; et al. 2015 · 2015
Cited alongside, same era.
End-to-end training of deep visuomotor policies
Levine, S.; Finn, C.; Darrell, T.; and Abbeel, P. 2016 · 2016
Cited alongside, same era.
Mastering the game of go without human knowledge
Silver, D.; Schrittwieser, J.; Simonyan, K.; Antonoglou, I.; Huang, A.; Guez, A.; Hubert, T.; Baker, L.; Lai, M.; Bolton, A.; et al. 2017 · 2017
Cited alongside, same era.
Addressing Function Approximation Error in Actor-Critic Methods
Fujimoto, S.; Hoof, H.; and Meger, D. 2018 · 2018
Cited alongside, same era.
Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Haarnoja, T.; Zhou, A.; Abbeel, P.; and Levine, S. 2018 · 2018
Cited alongside, same era.
Data center cooling using model-predictive control
Lazic, N.; Lu, T.; Boutilier, C.; Ryu, M.; Wong, E.; Roy, B.; and Imwalle, G. 2018 · 2018
Cited alongside, same era.
Later among the works it cites.
Conservative Q-Learning for Offline Reinforcement Learning
Kumar, A.; Zhou, A.; Tucker, G.; and Levine, S. 2020 · 2020
Later among the works it cites.
MOPO: Model-based Offline Policy Optimization
Yu, T.; Thomas, G.; Yu, L.; Ermon, S.; Zou, J.; Levine, S.; Finn, C.; and Ma, T. 2020 · 2020
Later among the works it cites.
Model-Based Offline Planning
Argenson, A.; and Dulac-Arnold, G. 2021 · 2021
Closest in time.
Offline Reinforcement Learning with Soft Behavior Regularization
Xu, H.; Zhan, X.; Li, J.; and Yin, H. 2021 · 2021
Closest in time.
Model-based offline planning with trajectory pruning
Zhan, X.; Zhu, X.; and Xu, H. 2021 · 2021
Closest in time.
Constraints Penalized Q-Learning for Safe Offline Reinforcement Learning
Xu, H.; Zhan, X.; and Zhu, X. 2022 · 2022
Closest in time.
Off-policy deep reinforcement learning without exploration
Fujimoto, S.; Meger, D.; and Precup, D. 2019 · 2062
Closest in time.