Fetching the paper…
Reading the bibliography…
Recent advances in combining deep neural network architectures with reinforcement learning techniques have shown promising potential results in solving complex control problems with high dimensional state and action spaces.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Reinforcement learning for robots using neural networks
L.-J. Lin · 1993
Earlier work this paper cites.
Traffic light control using sarsa with three state representations
T. L. Thorpe and C. W. Anderson · 1996
Earlier work this paper cites.
An analysis of temporal-difference learning with function approximation
J. N. Tsitsiklis, B. Van Roy, et al · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Introduction to Reinforcement Learning
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
R. S. Sutton, D. A. McAllester, S. P. Singh, Y. Mansour, et al · 1999
Earlier work this paper cites.
Multi-agent reinforcement learning for traffic light control
M. Wiering et al · 2000
Earlier work this paper cites.
Experiments with infinite-horizon, policy-gradient estimation
J. Baxter, P. L. Bartlett, and L. Weaver · 2001
Earlier work this paper cites.
Optimizing traffic lights in a cellular automaton model for city traffic
E. Brockfeld, R. Barlovic, A. Schadschneider, and M. Schreckenberg · 2001
Earlier work this paper cites.
Reinforcement learning for true adaptive traffic signal control
B. Abdulhai, R. Pringle, and G. J. Karakoulas · 2003
Earlier work this paper cites.
Recent advances in hierarchical reinforcement learning
A. G. Barto and S. Mahadevan · 2003
Earlier work this paper cites.
Traffic light scheduling using policy-gradient reinforcement learning
S. Ritcher · 2007
Earlier work this paper cites.
Reinforcement learning-based multi-agent system for network traffic signal control
I. Arel, C. Liu, T. Urbanik, and A. Kohls · 2010
Earlier work this paper cites.
Urban traffic signal control using reinforcement learning agents
P. Balaji, X. German, and D. Srinivasan · 2010
Cited alongside, same era.
Deep auto-encoder neural networks in reinforcement learning
S. Lange and M. Riedmiller · 2010
Cited alongside, same era.
Recurrent policy gradients
D. Wierstra, A. Förster, J. Peters, and J. Schmidhuber · 2010
Cited alongside, same era.
Q-learning based traffic optimization in management of signal timing plan
Y. K. Chin, N. Bolong, A. Kiring, S. S. Yang, and K. T. K. Teo · 2011
Cited alongside, same era.
Reinforcement learning with function approximation for traffic signal control
L. Prashanth and S. Bhatnagar · 2011
Cited alongside, same era.
Model-free reinforcement learning with continuous action in practice
T. Degris, P. M. Pilarski, and R. S. Sutton · 2012
Cited alongside, same era.
Automatic abstraction controller in reinforcement learning agent via automata
S. S. Mousavi, B. Ghazanfari, N. Mozayani, and M. R. Jahed-Motlagh · 2014
Later among the works it cites.
Reinforcement learning algorithms with function approximation: Recent advances and applications
X. Xu, L. Zuo, and Z. Huang · 2014
Later among the works it cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Later among the works it cites.
Gradient estimation using stochastic computation graphs
J. Schulman, N. Heess, T. Weber, and P. Abbeel · 2015
Later among the works it cites.
An autonomous network aware vm migration strategy in cloud data centres
M. Duggan, J. Duggan, E. Howley, and E. Barrett · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Recent development and applications of sumo-simulation of urban mobility
D. Krajzewicz, J. Erdmann, M. Behrisch, and L. Bieker · 2012
Cited alongside, same era.
Holonic multi-agent system for traffic signals control
M. Abdoos, N. Mozayani, and A. L. Bazzan · 2013
Cited alongside, same era.
Multiagent reinforcement learning for integrated network of adaptive traffic signal controllers (marlin-atsc): methodology and large-scale application on downtown toronto
S. El-Tantawy, B. Abdulhai, and H. Abdelgawad · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Cited alongside, same era.
A tutorial survey of architectures, algorithms, and applications for deep learning
L. Deng · 2014
Cited alongside, same era.
Enhancing nash q-learning and team q-learning mechanisms by using bottlenecks
B. Ghazanfari and N. Mozayani · 2014
Cited alongside, same era.
A reinforcement learning approach for dynamic selection of virtual machines in cloud data centres
M. Duggan, K. Flesk, J. Duggan, E. Howley, and E. Barrett · 2016
Later among the works it cites.
Using a deep reinforcement learning agent for traffic signal control
W. Genders and S. Razavi · 2016
Later among the works it cites.
Extracting bottlenecks for reinforcement learning agent by holonic concept clustering and attentional functions
B. Ghazanfari and N. Mozayani · 2016
Later among the works it cites.
Traffic signal timing via deep reinforcement learning
L. Li, Y. Lv, and F.-Y. Wang · 2016
Later among the works it cites.
Parallel systems for traffic control: A rethinking
L. Li and D. Wen · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. P. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Later among the works it cites.
Deep reinforcement learning: An overview
S. S. Mousavi, M. Schukat, and E. Howley · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Later among the works it cites.
Coordinated deep reinforcement learners for traffic light control
E. Van der Pol and F. A. Oliehoek · 2016
Later among the works it cites.