Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) often requires a large number of data and environment interactions, making the training process time-consuming.
S. Thrun and A. Schwartz, “Issues in using function approximation for reinforcement learning,” in Proceedings of the Fourth Connectionist Models Summer School , vol. 255. Hillsdale, NJ, 1993, p. 263
1993
Earlier work this paper cites.
L. K. Grover, “A fast quantum mechanical algorithm for database search,” in Proceedings of the twenty-eighth annual ACM symposium on Theory of computing , 1996, pp. 212–219
1996
Earlier work this paper cites.
J. Rissanen, “Fisher information and stochastic complexity,” IEEE Transactions on Information Theory , vol. 42, no. 1, pp. 40–47, 1996
1996
Earlier work this paper cites.
D. Dong, C. Chen, and Z. Chen, “Quantum reinforcement learning,” in Advances in Natural Computation: First International Conference, ICNC 2005, Changsha, China, August 27-29, 2005, Proceedings, Part II 1 . Springer, 2005, pp. 686–689
2005
Earlier work this paper cites.
D. Dong, C. Chen, H. Li, and T.-J. Tarn, “Quantum reinforcement learning,” IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics) , vol. 38, no. 5, pp. 1207–1220, 2008
2008
Earlier work this paper cites.
M. A. Nielsen and I. L. Chuang, Quantum computation and quantum information . Cambridge university press, 2010
2010
Earlier work this paper cites.
M. Van Otterlo and M. Wiering, “Reinforcement learning and markov decision processes,” Reinforcement learning: State-of-the-art , pp. 3–42, 2012
2012
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
P. J. O’Malley, R. Babbush, I. D. Kivlichan, J. Romero, J. R. McClean, R. Barends, J. Kelly, P. Roushan, A. Tranter, N. Ding et al. , “Scalable quantum simulation of molecular energies,” Physical Review X , vol. 6, no. 3, p. 031007, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
H. Van Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning,” in Proceedings of the AAAI conference on artificial intelligence , vol. 30, 2016
2016
Earlier work this paper cites.
K. Temme, S. Bravyi, and J. M. Gambetta, “Error mitigation for short-depth quantum circuits,” Physical Review Letters , vol. 119, no. 18, Nov. 2017. [Online]. Available: http://dx.doi.org/10.1103/PhysRevLett.119.180509
2017
Earlier work this paper cites.
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel et al. , “A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,” Science , vol. 362, no. 6419, pp. 1140–1144, 2018
2018
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Earlier work this paper cites.
K. Mitarai, M. Negoro, M. Kitagawa, and K. Fujii, “Quantum circuit learning,” Physical Review A , vol. 98, no. 3, p. 032309, 2018
2018
Earlier work this paper cites.
2019
Cited alongside, same era.
M. Schuld, V. Bergholm, C. Gogolin, J. Izaac, and N. Killoran, “Evaluating analytic gradients on quantum hardware,” Phys. Rev. A , vol. 99, p. 032331, Mar 2019. [Online]. Available: https://link.aps.org/doi/10.1103/PhysRevA.99.032331
2019
Cited alongside, same era.
2019
Cited alongside, same era.
R. Agarwal, D. Schuurmans, and M. Norouzi, “Striving for simplicity in off-policy deep reinforcement learning,” 2020. [Online]. Available: https://openreview.net/forum?id=ryeUg0VFwr
2020
Cited alongside, same era.
A. Abbas, D. Sutter, C. Zoufal, A. Lucchi, A. Figalli, and S. Woerner, “The power of quantum neural networks,” Nature Computational Science , vol. 1, no. 6, pp. 403–409, 2021
2021
Later among the works it cites.
A. Fawzi, M. Balog, A. Huang, T. Hubert, B. Romera-Paredes, M. Barekatain, A. Novikov, F. J. R Ruiz, J. Schrittwieser, G. Swirszcz et al. , “Discovering faster matrix multiplication algorithms with reinforcement learning,” Nature , vol. 610, no. 7930, pp. 47–53, 2022
2022
Later among the works it cites.
O. Kiss, M. Grossi, P. Lougovski, F. Sanchez, S. Vallecorsa, and T. Papenbrock, “Quantum computing of the li 6 nucleus via ordered unitary coupled clusters,” Physical Review C , vol. 106, no. 3, p. 034325, 2022
2022
Later among the works it cites.
M. C. Caro, H. Huang, M. Cerezo et al. , “Generalization in quantum machine learning from few training data,” Nat Commun 13 , vol. 4919, 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Schuld, A. Bocharov, K. M. Svore, and N. Wiebe, “Circuit-centric quantum classifiers,” Physical Review A , vol. 101, no. 3, p. 032308, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
F. J. Gil Vidal and D. O. Theis, “Input redundancy for parameterized quantum circuits,” Frontiers in Physics , vol. 8, p. 297, 2020
2020
Cited alongside, same era.
S. Y.-C. Chen, C.-H. H. Yang, J. Qi, P.-Y. Chen, X. Ma, and H.-S. Goan, “Variational quantum circuits for deep reinforcement learning,” IEEE Access , vol. 8, pp. 141 007–141 024, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
A. Pérez-Salinas, A. Cervera Lierta, E. Gil-Fuster, and J. Latorre, “Data re-uploading for a universal quantum classifier,” Quantum , vol. 4, p. 226, 02 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
A. Blance and M. Spannowsky, “Quantum machine learning for particle physics using a variational quantum classifier,” Journal of High Energy Physics , vol. 2021, no. 2, pp. 1–20, 2021
2021
Cited alongside, same era.
2022
Later among the works it cites.
A. Skolik, S. Jerbi, and V. Dunjko, “Quantum agents in the gym: a variational quantum algorithm for deep q-learning,” Quantum , vol. 6, p. 720, 2022
2022
Later among the works it cites.
I. Miháliková, M. Friák, M. Pivoluska, M. Plesch, M. Saip, and M. Šob, “Best-practice aspects of quantum-computer calculations: A case study of the hydrogen molecule,” Molecules , vol. 27, no. 3, p. 597, 2022
2022
Later among the works it cites.
M. Franz, L. Wolf, M. Periyasamy, C. Ufrecht, D. D. Scherer, A. Plinge, C. Mutschler, and W. Mauerer, “Uncovering instabilities in variational-quantum deep q-networks,” Journal of The Franklin Institute , 2022
2022
Later among the works it cites.
M. Periyasamy, N. Meyer, C. Ufrecht, D. D. Scherer, A. Plinge, and C. Mutschler, “Incremental data-uploading for full-quantum classification,” in 2022 IEEE International Conference on Quantum Computing and Engineering (QCE) . IEEE, 2022, pp. 31–37
2022
Later among the works it cites.
W. J. Yun, J. P. Kim, S. Jung, J.-H. Kim, and J. Kim, “Quantum multi-agent actor-critic neural networks for internet-connected multi-robot coordination in smart factory management,” IEEE Internet of Things Journal , 2023
2023
Closest in time.
2023
Closest in time.
Z. Cheng, K. Zhang, L. Shen, and D. Tao, “Offline quantum reinforcement learning in a conservative manner,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 37, no. 6, 2023, pp. 7148–7156
2023
Closest in time.
N. Meyer, D. Scherer, A. Plinge, C. Mutschler, and M. Hartmann, “Quantum policy gradient algorithm with optimized action decoding,” in International Conference on Machine Learning . PMLR, 2023, pp. 24 592–24 613
2023
Closest in time.
Qiskit contributors, “Qiskit: An open-source framework for quantum computing,” 2023
2023
Closest in time.
S. Fujimoto, D. Meger, and D. Precup, “Off-policy deep reinforcement learning without exploration,” in International conference on machine learning . PMLR, 2019, pp. 2052–2062
2062
Closest in time.