Fetching the paper…
Reading the bibliography…
Recently, extensive studies on photonic reinforcement learning to accelerate the process of calculation by exploiting the physical nature of light have been conducted.
“The theory of dynamic programming”
Richard Bellman · 1954
Earlier work this paper cites.
“Dynamic programming”
Richard Bellman · 1966
Earlier work this paper cites.
“Neuronlike adaptive elements that can solve difficult learning control problems”
Andrew Barto, Richard Sutton and Charles Anderson · 1983
Earlier work this paper cites.
“Learning from delayed rewards”
Christopher Watkins · 1989
Earlier work this paper cites.
“Integrated architectures for learning, planning, and reacting based on approximating dynamic programming”
Richard Sutton · 1990
Earlier work this paper cites.
“Exploration and exploitation in organizational learning”
James March · 1991
Earlier work this paper cites.
“Q-learning”
Christopher Watkins and Peter Dayan · 1992
Earlier work this paper cites.
“Realizable higher-dimensional two-particle entanglements via multiport beam splitters”
Marek Żukowski, Anton Zeilinger and Michael Horne · 1997
Earlier work this paper cites.
“Reinforcement learning: An introduction”
Richard Sutton and Andrew Barto · 1998
Earlier work this paper cites.
“Some studies in machine learning using the game of checkers”
Arthur Samuel · 2000
Earlier work this paper cites.
“Three-photon Hong-Ou-Mandel interference at a multiport mixer”
Richard Campos · 2000
Earlier work this paper cites.
“Cortical substrates for exploratory decisions in humans”
Nathaniel Daw et al · 2006
Cited alongside, same era.
“Cognitive medium access: Exploration, exploitation, and competition”
Lifeng Lai, Hesham El, Hai Jiang and H Poor · 2010
Cited alongside, same era.
“Single-photon decision maker”
Makoto Naruse et al · 2015
Cited alongside, same era.
“Generalized multiphoton quantum interference”
Max Tillmann et al · 2015
Cited alongside, same era.
“Mastering the game of Go with deep neural networks and tree search”
David Silver et al · 2016
Cited alongside, same era.
“Multi-armed bandits with application to 5G small cells”
Setareh Maghsudi and Ekram Hossain · 2016
Cited alongside, same era.
“Multi-player bandits revisited”
Lilian Besson and Emilie Kaufmann · 2018
Later among the works it cites.
“Quantum optical neural networks”
Gregory Steinbrecher, Jonathan Olson, Dirk Englund and Jacques Carolan · 2019
Later among the works it cites.
“Entangled-photon decision maker”
Nicolas Chauvet et al · 2019
Later among the works it cites.
“Q-learning algorithms: A comprehensive classification and applications”
Beakcheol Jang, Myeonghwi Kim, Gaspard Harerimana and Jong Kim · 2019
Later among the works it cites.
“Photonic architecture for reinforcement learning”
Fulvio Flamini et al · 2020
Later among the works it cites.
“Entangled N-photon states for fair and optimal social decision making”
Nicolas Chauvet et al · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Harnessing the computational power of fluids for optimization of collective decision making”
Song-Ju Kim, Makoto Naruse and Masashi Aono · 2016
Cited alongside, same era.
“A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play”
David Silver et al · 2018
Cited alongside, same era.
“Reinforcement learning in different phases of quantum control”
Marin Bukov et al · 2018
Cited alongside, same era.
“Reinforcement learning in a large-scale photonic recurrent neural network”
Julian Bueno et al · 2018
Cited alongside, same era.
“Experimental quantum speed-up in reinforcement learning agents”
Valeria Saggio et al · 2021
Later among the works it cites.
“Conflict-free collective stochastic decision making by orbital angular momentum of photons through quantum interference”
Takashi Amakasu et al · 2021
Later among the works it cites.
“Optimal preference satisfaction for conflict-free joint decisions”
Hiroaki Shinkawa et al · 2022
Closest in time.
“Conflict-Free Joint Sampling for Preference Satisfaction through Quantum Interference”
Hiroaki Shinkawa et al · 2022
Closest in time.