Fetching the paper…
Reading the bibliography…
In this paper, we investigate the problem of \textit{episodic reinforcement learning} with quantum oracles for state evolution.
Asymptotically efficient adaptive allocation rules
Tze Leung Lai, Herbert Robbins, et al · 1985
Earlier work this paper cites.
A framework for fast quantum mechanical algorithms
Lov K Grover · 1998
Earlier work this paper cites.
Quantum amplitude amplification and estimation
Gilles Brassard, Peter Hoyer, Michele Mosca, and Alain Tapp · 2002
Earlier work this paper cites.
Near-optimal reinforcement learning in polynomial time
Michael Kearns and Satinder Singh · 2002
Earlier work this paper cites.
Pac model-free reinforcement learning
Alexander L Strehl, Lihong Li, Eric Wiewiora, John Langford, and Michael L Littman · 2006
Earlier work this paper cites.
Near-optimal regret bounds for reinforcement learning
Peter Auer, Thomas Jaksch, and Ronald Ortner · 2008
Earlier work this paper cites.
Quantum reinforcement learning
Daoyi Dong, Chunlin Chen, Hanxiong Li, and Tzyh-Jong Tarn · 2008
Earlier work this paper cites.
Gilles Brassard, Frederic Dupuis, Sebastien Gambs, and Alain Tapp · 2011
Earlier work this paper cites.
Quantum speed-up for unsupervised learning
Esma Aïmeur, Gilles Brassard, and Sébastien Gambs · 2013
Earlier work this paper cites.
(more) efficient reinforcement learning via posterior sampling
Ian Osband, Daniel Russo, and Benjamin Van Roy · 2013
Earlier work this paper cites.
Quantum speedup for active learning agents
Giuseppe Davide Paparo, Vedran Dunjko, Adi Makmal, Miguel Angel Martin-Delgado, and Hans J Briegel · 2014
Earlier work this paper cites.
Quantum support vector machine for big data classification
Patrick Rebentrost, Masoud Mohseni, and Seth Lloyd · 2014
Earlier work this paper cites.
Quantum speedup of monte carlo methods
Ashley Montanaro · 2015
Cited alongside, same era.
Quantum-enhanced machine learning
Vedran Dunjko, Jacob M Taylor, and Hans J Briegel · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Generalization and exploration via randomized value functions
Ian Osband, Benjamin Van Roy, and Zheng Wen · 2016
Cited alongside, same era.
Guest column: A survey of quantum learning theory
Srinivasan Arunachalam and Ronald de Wolf · 2017
Cited alongside, same era.
Minimax regret bounds for reinforcement learning
Mohammad Gheshlaghi Azar, Ian Osband, and Rémi Munos · 2017
Cited alongside, same era.
Reinforcement learning: Theory and algorithms
Alekh Agarwal, Nan Jiang, Sham M Kakade, and Wen Sun · 2019
Later among the works it cites.
Deeppool: Distributed model-free algorithm for ride-sharing using deep reinforcement learning
Abubakr O Al-Abbasi, Arnob Ghosh, and Vaneet Aggarwal · 2019
Later among the works it cites.
Provably efficient q-learning with function approximation via distribution shift error checking oracle
Simon S Du, Yuping Luo, Ruosong Wang, and Hanrui Zhang · 2019
Later among the works it cites.
Near-optimal optimistic reinforcement learning using empirical bernstein inequalities
Aristide Tossou, Debabrota Basu, and Christos Dimitrakakis · 2019
Later among the works it cites.
Provably efficient exploration in policy optimization
Qi Cai, Zhuoran Yang, Chi Jin, and Zhaoran Wang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unifying pac and regret: Uniform pac bounds for episodic reinforcement learning
Christoph Dann, Tor Lattimore, and Emma Brunskill · 2017
Cited alongside, same era.
Advances in quantum reinforcement learning
Vedran Dunjko, Jacob M Taylor, and Hans J Briegel · 2017
Cited alongside, same era.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Cited alongside, same era.
Is q-learning provably efficient?
Chi Jin, Zeyuan Allen-Zhu, Sebastien Bubeck, and Michael I Jordan · 2018
Cited alongside, same era.
David Rohde, Stephen Bonner, Travis Dunlop, Flavian Vasile, and Alexandros Karatzoglou · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Cited alongside, same era.
Balthazar Casalé, Giuseppe Di Molfetta, Hachem Kadri, and Liva Ralaivola · 2020
Later among the works it cites.
Provably efficient reinforcement learning with linear function approximation
Chi Jin, Zhuoran Yang, Zhaoran Wang, and Michael I Jordan · 2020
Later among the works it cites.
Quantum sub-gaussian mean estimator
Yassine Hamoudi · 2021
Later among the works it cites.
Quantum enhancements for deep reinforcement learning in large spaces
Sofiene Jerbi, Lea M Trenkwalder, Hendrik Poulsen Nautrup, Hans J Briegel, and Vedran Dunjko · 2021
Later among the works it cites.
Experimental quantum speed-up in reinforcement learning agents
Valeria Saggio, Beate E Asenbeck, Arne Hamann, Teodor Strömberg, Peter Schiansky, Vedran Dunjko, Nicolai Friis, Nicholas C Harris, Michael Hochberg, Dirk Englund, et al · 2021
Later among the works it cites.
Differentially private regret minimization in episodic markov decision processes
Sayak Ray Chowdhury and Xingyu Zhou · 2022
Later among the works it cites.
Multi-armed quantum bandits: Exploration versus exploitation when learning properties of quantum states
Josep Lumbreras, Erkka Haapasalo, and Marco Tomamichel · 2022
Later among the works it cites.