Fetching the paper…
Reading the bibliography…
Transformers are neural network models that utilize multiple layers of self-attention heads and have exhibited enormous potential in natural language processing tasks.
Q-learning
Christopher JCH Watkins and Peter Dayan. 1992 · 1992
Earlier work this paper cites.
Issues in using function approximation for reinforcement learning. In Proceedings of the 4th Connectionist Models Summer School . Hillsdale, NJ, 255–263
Sebastian Thrun and Anton Schwartz. 1993 · 1993
Earlier work this paper cites.
Double Q-learning. In Advances in neural information processing systems 23 . Curran Associates, Inc., Red Hook, NY, 2613–2621
Hado Hasselt. 2010 · 2010
Earlier work this paper cites.
Playing Atari with Deep Reinforcement Learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Incentivizing exploration in reinforcement learning with deep predictive models
Bradly C Stadie, Sergey Levine, and Pieter Abbeel. 2015 · 2015
Earlier work this paper cites.
Deep exploration via bootstrapped DQN
Ian Osband, Charles Blundell, Alexander Pritzel, and Benjamin Van Roy. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Momentum contrast for unsupervised visual representation learning. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 9729–9738
Kaiming He, Haoqi Fan, Yuxin Wu, Saining Xie, and Ross Girshick. 2020 · 2020
Earlier work this paper cites.
Exertion-aware path generation
Wanwan Li, Biao Xie, Yongqi Zhang, Walter Meiss, Haikun Huang, and Lap-Fai Yu. 2020 · 2020
Earlier work this paper cites.
Self-attention with linear complexity
Sinong Wang, Belinda Li, Madian Khabsa, Han Fang, and H Linformer Ma. 2020 · 2020
Cited alongside, same era.
Scene-aware background music synthesis. In Proceedings of the 28th ACM International Conference on Multimedia . 1162–1170
Wang Y., Liang W., Li W., Li D., and Yu L.F. 2020 · 2020
Cited alongside, same era.
Deep reinforcement learning at the edge of the statistical precipice
Rishabh Agarwal, Max Schwarzer, Pablo Samuel Castro, Aaron C Courville, and Marc Bellemare. 2021 · 2021
Cited alongside, same era.
Emerging properties in self-supervised vision transformers. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 9650–9660
Mathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou, Julien Mairal, Piotr Bojanowski, and Armand Joulin. 2021 · 2021
Cited alongside, same era.
Decision transformer: Reinforcement learning via sequence modeling
Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Misha Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch. 2021a · 2021
Cited alongside, same era.
Expert Q-learning: Deep Q-learning With State Values From Expert Examples
Li Meng, Anis Yazidi, Morten Goodwin, and Paal Engelstad. 2021 · 2021
Later among the works it cites.
Improving Sample Efficiency of Value Based Models Using Attention and Vision Transformers
Amir Ardalan Kalantari, Mohammad Amini, Sarath Chandar, and Doina Precup. 2022 · 2022
Closest in time.
Interactive augmented reality storytelling guided by scene semantics
Changyang Li, Wanwan Li, Haikun Huang, and Lap-Fai Yu. 2022 · 2022
Closest in time.
Musical Instrument Performance in Augmented Virtuality. In Proceedings of the 6th International Conference on Digital Signal Processing . 91–97
Wanwan Li. 2022 · 2022
Closest in time.
Evaluating Vision Transformer Methods for Deep Reinforcement Learning from Pixels
Tianxin Tao, Daniele Reda, and Michiel van de Panne. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
An empirical study of training self-supervised vision transformers. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 9640–9649
Xinlei Chen, Saining Xie, and Kaiming He. 2021b · 2021
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2021
Cited alongside, same era.
Vision transformer for learning driving policies in complex multi-agent environments
Eshagh Kargar and Ville Kyrki. 2021 · 2021
Cited alongside, same era.
Image Synthesis and Editing with Generative Adversarial Networks (GANs): A Review. In 2021 Fifth World Conference on Smart Trends in Systems Security and Sustainability (WorldS4) . IEEE, 65–70
Wanwan Li. 2021 · 2021
Cited alongside, same era.
Swin transformer: Hierarchical vision transformer using shifted windows. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 10012–10022
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. 2021 · 2021
Cited alongside, same era.
Closest in time.
Online decision transformer. In International Conference on Machine Learning . PMLR, 27042–27059
Qinqing Zheng, Amy Zhang, and Aditya Grover. 2022 · 2022
Closest in time.
Synthesizing 3D VR Sketch Using Generative Adversarial Neural Network. In Proceedings of the 2023 7th International Conference on Big Data and Internet of Things . 122–128
Wanwan Li. 2023a · 2023
Closest in time.
Terrain synthesis for treadmill exergaming in virtual reality. In 2023 IEEE Conference on Virtual Reality and 3D User Interfaces Abstracts and Workshops (VRW) . IEEE, 263–269
Wanwan Li. 2023b · 2023
Closest in time.
Location-Aware Adaptation of Augmented Reality Narratives. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems . 1–15
Wanwan Li, Changyang Li, Minyoung Kim, Haikun Huang, and Lap-Fai Yu. 2023 · 2023
Closest in time.
Image Generation of Egyptian Hieroglyphs. In ICMLC 2024
B. Hui S. Gao and Li W. 2024 · 2024
Closest in time.