Fetching the paper…
Reading the bibliography…
Building upon prior research that highlighted the need for standardizing environments for building control research, and inspired by recently introduced challenges for real life reinforcement learning control, here we propose a non-exhaustive set of nine real world challenges for reinforcement learning control in grid-interactive buildings.
Technical Note: Q-Learning
Christopher Watkins and Peter Dayan. 1992 · 1992
Earlier work this paper cites.
Model predictive control: Past, present and future
Manfred Morari and Jay H. Lee. 1999 · 1999
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In 2009 IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 248–255
Jia Deng, Wei Dong, R. Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
Demand Response Architecture-Integration into the Distribution Management System
S Mohagheghi, J Stoupis, Z Wang, and Z Li. 2010 · 2010
Earlier work this paper cites.
Understanding smart cities: An integrative framework
Hafedh Chourabi, Taewoo Nam, Shawn Walker, J. Ramon Gil-Garcia, Sehl Mellouli, Karine Nahon, Theresa A. Pardo, and Hans Jochen Scholl. 2011 · 2012
Earlier work this paper cites.
Building modeling as a crucial part for building predictive control
Samuel Prívara, Jiří Cigler, Zdeněk Váňa, Frauke Oldewurtel, Carina Sagerschnig, and Eva Žáčeková. 2013 · 2012
Earlier work this paper cites.
Jose R Vazquez-Canteli, Sourav Dey, Gregor Henze, and Zoltan Nagy. 2020a · 2012
Earlier work this paper cites.
Short-term demand response of flexible electric heating systems: The need for integrated simulations
Kenneth Bruninx, Dieter Patteeuw, Erik Delarue, Lieve Helsen, and William D’Haeseleer. 2013 · 2013
Earlier work this paper cites.
Demand response and smart grids - A survey
Pierluigi Siano. 2014 · 2013
Earlier work this paper cites.
Impact of residential demand response on power system operation: A Belgian case study
B. Dupont, K. Dietrich, C. De Jonghe, A. Ramos, and R. Belmans. 2014 · 2014
Earlier work this paper cites.
Fifth Assessment Report, Mitigation of Climate Change
O. Lucon and D. Ürge-Vorsatz. 2014 · 2014
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Earlier work this paper cites.
Soft Actor-Critic Algorithms and Applications
Tuomas Haarnoja, Aurick Zhou, Kristian Hartikainen, G. Tucker, Sehoon Ha, Jie Tan, Vikash Kumar, Henry Zhu, Abhishek Gupta, P. Abbeel, and Sergey Levine. 2018b · 2018
Cited alongside, same era.
Simulation-based evaluation and optimization of control strategies in buildings
Georgios D. Kontes, Georgios I. Giannakis, Víctor Sánchez, Pablo de Agustin-Camacho, Ander Romero-Amorrortu, Natalia Panagiotidou, Dimitrios V. Rovas, Simone Steiger, Christopher Mutschler, and Gunnar Gruen. 2018 · 2018
Cited alongside, same era.
Optimal decarbonization pathways for urban residential building energy services
Benjamin D. Leibowicz, Christopher M. Lanham, Max T. Brozynski, Jose R. Vazquez-Canteli, Nicolas Castillo Castejon, and Zoltan Nagy. 2018 · 2018
Cited alongside, same era.
Reinforcement learning for intelligent environments: A Tutorial
Zoltan Nagy, June Young Park, and Jose Vazquez-Canteli. 2018 · 2018
Cited alongside, same era.
Reinforcement Learning, Second Edition An Introduction
Richard S. Sutton and Andrew G. Barto. 2018 · 2018
A Guide for the Design of Benchmark Environments for Building Energy Optimization
David Wölfle, Arun Vishwanath, and Hartmut Schmeck. 2020 · 2020
Later among the works it cites.
COBS: COmprehensive Building Simulator. In Proceedings of the 7th ACM International Conference on Systems for Energy-Efficient Buildings, Cities, and Transportation . ACM, New York, NY, USA, 314–315
Tianyu Zhang and Omid Ardakanian. 2020 · 2020
Later among the works it cites.
Building optimization testing framework (BOPTEST) for simulation-based benchmarking of control strategies in buildings
David Blum, Javier Arroyo, Sen Huang, Ján Drgoňa, Filip Jorissen, Harald Taxt Walnum, Yan Chen, Kyle Benne, Draguna Vrabie, Michael Wetter, and Lieve Helsen. 2021 · 2021
Closest in time.
Exploring the potentialities of deep reinforcement learning for incentive-based demand response in a cluster of small commercial buildings
Davide Deltetto, Davide Coraci, Giuseppe Pinto, Marco Savino Piscitelli, and Alfonso Capozzoli. 2021 · 2021
Closest in time.
A National Roadmap for Grid-Interactive Efficient Buildings
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Reinforcement learning for demand response: A review of algorithms and modeling techniques
Jose R. Vazquez-Canteli and Zoltan Nagy. 2019 · 2018
Cited alongside, same era.
CityLearn v1.0: An OpenAI gym environment for demand response with deep reinforcement learning
J.R. José R. Vázquez-Canteli, Jérôme Kämpf, Gregor Henze, and Zoltan Nagy. 2019 · 2019
Cited alongside, same era.
Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms on a Building Energy Demand Coordination Task
Gauraang Dhamankar, Jose R. Vazquez-Canteli, and Zoltan Nagy. 2020 · 2020
Cited alongside, same era.
All you need to know about model predictive control for buildings
Ján Drgoňa, Javier Arroyo, Iago Cupeiro Figueroa, David Blum, Krzysztof Arendt, Donghun Kim, Enric Perarnau Ollé, Juraj Oravec, Michael Wetter, Draguna L. Vrabie, and Lieve Helsen. 2020 · 2020
Cited alongside, same era.
A Centralised Soft Actor Critic Deep Reinforcement Learning Approach to District Demand Side Management through CityLearn. In Proceedings of the 1st International Workshop on Reinforcement Learning for Energy Management in Buildings & Cities . ACM, New York, NY, USA, 11–14
Anjukan Kathirgamanathan, Kacper Twardowski, Eleni Mangina, and Donal P. Finn. 2020 · 2020
Cited alongside, same era.
MARLISA: Multi-Agent Reinforcement Learning with Iterative Sequential Action Selection for Load Shaping of Grid-Interactive Connected Buildings
Jose R. Vazquez-Canteli, Gregor Henze, and Zoltan Nagy. 2020b · 2020
Cited alongside, same era.
Reinforcement learning for building controls: The opportunities and challenges
Zhe Wang and Tianzhen Hong. 2020 · 2020
Cited alongside, same era.
Department of Energy. 2021 · 2021
Closest in time.
Physics-constrained deep learning of multi-zone building thermal dynamics
Ján Drgoňa, Aaron R. Tuor, Vikas Chandan, and Draguna L. Vrabie. 2021 · 2021
Closest in time.
Challenges of real-world reinforcement learning: definitions, benchmarks and analysis
Gabriel Dulac-Arnold, Nir Levine, Daniel J. Mankowitz, Jerry Li, Cosmin Paduraru, Sven Gowal, and Todd Hester. 2021 · 2021
Closest in time.
Collaborative energy demand response with decentralized actor and centralized critic. In Proceedings of the 8th ACM International Conference on Systems for Energy-Efficient Buildings, Cities, and Transportation . ACM, New York, NY, USA, 333–337
Ruben Glatt, Felipe Leno da Silva, Braden Soper, William A. Dawson, Edward Rusu, and Ryan A. Goldhahn. 2021 · 2021
Closest in time.
Sinergym: a building simulation and control framework for training reinforcement learning agents. In Proceedings of the 8th ACM International Conference on Systems for Energy-Efficient Buildings, Cities, and Transportation . ACM, New York, NY, USA, 319–323
Javier Jiménez-Raboso, Alejandro Campoy-Nieves, Antonio Manjavacas-Lucas, Juan Gómez-Romero, and Miguel Molina-Solana. 2021 · 2021
Closest in time.
The citylearn challenge 2021. In Proceedings of the 8th ACM International Conference on Systems for Energy-Efficient Buildings, Cities, and Transportation . ACM, New York, NY, USA, 218–219
Zoltan Nagy, José R. Vázquez-Canteli, Sourav Dey, and Gregor Henze. 2021 · 2021
Closest in time.
Coordinated energy management for a cluster of buildings through deep reinforcement learning
Giuseppe Pinto, Marco Savino Piscitelli, José Ramón Vázquez-Canteli, Zoltán Nagy, and Alfonso Capozzoli. 2021 · 2021
Closest in time.
NeoRL: A Near Real-World Benchmark for Offline Reinforcement Learning
Rongjun Qin, Songyi Gao, Xingyuan Zhang, Zhen Xu, Shengkai Huang, Zewen Li, Weinan Zhang, and Yang Yu. 2021 · 2021
Closest in time.