Fetching the paper…
Reading the bibliography…
Transformers have significantly impacted domains like natural language processing, computer vision, and robotics, where they improve performance compared to other neural networks.
Encoding Musical Style with Transformer Autoencoders. In Proceedings of the 37th International Conference on Machine Learning, ICML , Vol. 119. PMLR, 1899–1908
Kristy Choi, Curtis Hawthorne, Ian Simon, Monica Dinculescu, and Jesse H. Engel. 2020 · 1908
Earlier work this paper cites.
Stabilizing Voltage in Power Distribution Networks via Multi-Agent Reinforcement Learning with Transformer. In KDD ’22: The 28th ACM SIGKDD Conference on Knowledge Discovery and Data Mining, Washington, DC, USA, August 14 - 18, 2022 . ACM, 1899–1909
Minrui Wang, Mingxiao Feng, Wengang Zhou, and Houqiang Li. 2022b · 1909
Earlier work this paper cites.
Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning
Ronald J. Williams. 1992 · 1992
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
The Vanishing Gradient Problem During Learning Recurrent Neural Nets and Problem Solutions
Sepp Hochreiter. 1998 · 1998
Earlier work this paper cites.
Reinforcement learning - an introduction
Richard S. Sutton and Andrew G. Barto. 1998 · 1998
Earlier work this paper cites.
Computational challenges in portfolio management
Martin B. Haugh and Andrew W. Lo. 2001 · 2001
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation. In Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics . ACL, 311–318
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Recent Advances in Hierarchical Reinforcement Learning
Andrew G. Barto and Sridhar Mahadevan. 2003 · 2003
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks. In Proceedings of the Thirteenth International Conference on Artificial Intelligence and Statistics, AISTATS 2010, Chia Laguna Resort, Sardinia, Italy, May 13-15, 2010 (JMLR Proceedings, Vol. 9) , Yee Whye Teh and D. Mike Titterington (Eds.). JMLR.org, 249–256
Xavier Glorot and Yoshua Bengio. 2010 · 2010
Earlier work this paper cites.
Random Search for Hyper-Parameter Optimization
James Bergstra and Yoshua Bengio. 2012 · 2012
Earlier work this paper cites.
Playing Atari with Deep Reinforcement Learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller. 2013 · 2013
Earlier work this paper cites.
On the difficulty of training recurrent neural networks. In Proceedings of the 30th International Conference on Machine Learning, ICML , Vol. 28. JMLR.org, 1310–1318
Razvan Pascanu, Tomás Mikolov, and Yoshua Bengio. 2013 · 2013
Earlier work this paper cites.
Rigid body dynamics algorithms
Roy Featherstone. 2014 · 2014
Earlier work this paper cites.
Generative Adversarial Networks
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron C. Courville, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Large-Scale Video Classification with Convolutional Neural Networks. In IEEE Conference on Computer Vision and Pattern Recognition, CVPR . IEEE Computer Society, 1725–1732
Andrej Karpathy, George Toderici, Sanketh Shetty, Thomas Leung, Rahul Sukthankar, and Li Fei-Fei. 2014 · 2014
Earlier work this paper cites.
The Arcade Learning Environment: An Evaluation Platform for General Agents (Extended Abstract). In Proceedings of the Twenty-Fourth International Joint Conference on Artificial Intelligence, IJCAI 2015, Buenos Aires, Argentina, July 25-31, 2015 , Qiang Yang and Michael J. Wooldridge (Eds.). AAAI Press, 4148–4152
Marc G. Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling. 2015 · 2015
Earlier work this paper cites.
Scalable Bayesian Optimization Using Deep Neural Networks. In Proceedings of the 32nd ICML , Vol. 37. JMLR.org, 2171–2180
Jasper Snoek, Oren Rippel, Kevin Swersky, Ryan Kiros, Nadathur Satish, Narayanan Sundaram, Md. Mostofa Ali Patwary, Prabhat, and Ryan P. Adams. 2015 · 2015
Earlier work this paper cites.
CIDEr: Consensus-based image description evaluation. In IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2015, Boston, MA, USA, June 7-12, 2015 . IEEE Computer Society, 4566–4575
Ramakrishna Vedantam, C. Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Earlier work this paper cites.
A survey of scheduling problems with no-wait in process
Ali Allahverdi. 2016 · 2016
Earlier work this paper cites.
Learning to Communicate with Deep Multi-Agent Reinforcement Learning. In NeurIPS . 2137–2145
Jakob N. Foerster, Yannis M. Assael, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Earlier work this paper cites.
Building machines that learn and think like people
Brenden M. Lake, Tomer David Ullman, Joshua B. Tenenbaum, and Samuel J. Gershman. 2016 · 2016
Earlier work this paper cites.
Continuous control with deep reinforcement learning. In 4th International Conference on Learning Representations, ICLR 2016, San Juan, Puerto Rico, May 2-4, 2016, Conference Track Proceedings
Timothy P. Lillicrap, Jonathan J. Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra. 2016 · 2016
Earlier work this paper cites.
Equity forecast: Predicting long term stock price movement using machine learning
Nikola Milosevic. 2016 · 2016
Earlier work this paper cites.
Deep Reinforcement Learning: A Brief Survey
Kai Arulkumaran, Marc Peter Deisenroth, Miles Brundage, and Anil Anthony Bharath. 2017 · 2017
Earlier work this paper cites.
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks. In Proceedings of the 34th ICML , Vol. 70. PMLR, 1126–1135
Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017 · 2017
Earlier work this paper cites.
Survey on machine learning based scheduling in cloud computing. In Proceedings of the 2017 International Conference on Intelligent Systems, Metaheuristics & Swarm Intelligence . 57–61
Naveen Kumar Gondhi and Ayushi Gupta. 2017 · 2017
Earlier work this paper cites.
Deep Multimodal Learning: A Survey on Recent Advances and Trends
Dhanesh Ramachandram and Graham W. Taylor. 2017 · 2017
Earlier work this paper cites.
Proximal Policy Optimization Algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Earlier work this paper cites.
Attention is All you Need. In 30th Annual Conference on Neural Information Processing Systems . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
StarCraft II: A New Challenge for Reinforcement Learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John P. Agapiou, Julian Schrittwieser, John Quan, Stephen Gaffney, Stig Petersen, Karen Simonyan, Tom Schaul, Hado van Hasselt, David Silver, Timothy P. Lillicrap, Kevin Calderone, Paul Keet, Anthony Brunasso, David Lawrence, Anders Ekermo, Jacob Repp, and Rodney Tsing. 2017 · 2017
Earlier work this paper cites.
Multi-channel LSTM-CNN model for Vietnamese sentiment analysis. In 9th International Conference on Knowledge and Systems Engineering, KSE 2017, Hue, Vietnam, October 19-21, 2017 . IEEE, 24–29
Quan-Hoang Vo, Huy-Tien Nguyen, Bac Le, and Minh-Le Nguyen. 2017 · 2017
Earlier work this paper cites.
Scalable lifelong reinforcement learning
Yusen Zhan, Haitham Bou-Ammar, and Matthew E. Taylor. 2017 · 2017
Earlier work this paper cites.
Relational inductive biases, deep learning, and graph networks
Peter W. Battaglia, Jessica B. Hamrick, Victor Bapst, Alvaro Sanchez-Gonzalez, Vinícius Flores Zambaldi, Mateusz Malinowski, Andrea Tacchetti, David Raposo, Adam Santoro, Ryan Faulkner, Çaglar Gülçehre, H. Francis Song, Andrew J. Ballard, Justin Gilmer, George E. Dahl, Ashish Vaswani, Kelsey R. Allen, Charles Nash, Victoria Langston, Chris Dyer, Nicolas Heess, Daan Wierstra, Pushmeet Kohli, Matthew M. Botvinick, Oriol Vinyals, Yujia Li, and Razvan Pascanu. 2018 · 2018
Earlier work this paper cites.
Speech-Transformer: A No-Recurrence Sequence-to-Sequence Model for Speech Recognition. In 2018 IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2018, Calgary, AB, Canada, April 15-20, 2018 . IEEE, 5884–5888
Linhao Dong, Shuang Xu, and Bo Xu. 2018 · 2018
Earlier work this paper cites.
David Ha and Jürgen Schmidhuber. 2018 · 2018
Earlier work this paper cites.
Optimizing Agent Behavior over Long Time Scales by Transporting Value
Chia-Chun Hung, Timothy P. Lillicrap, Josh Abramson, Yan Wu, Mehdi Mirza, Federico Carnevale, Arun Ahuja, and Greg Wayne. 2018 · 2018
Earlier work this paper cites.
Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation. In 2nd Annual Conference on Robot Learning, CoRL 2018, Zürich , Vol. 87. PMLR, 651–673
Dmitry Kalashnikov, Alex Irpan, Peter Pastor, Julian Ibarz, Alexander Herzog, Eric Jang, Deirdre Quillen, Ethan Holly, Mrinal Kalakrishnan, Vincent Vanhoucke, and Sergey Levine. 2018 · 2018
Earlier work this paper cites.
State representation learning for control: An overview
Timothée Lesort, Natalia Díaz Rodríguez, Jean-François Goudou, and David Filliat. 2018 · 2018
Earlier work this paper cites.
Learning Beam Search Policies via Imitation Learning. In NeurIPS . 10675–10684
Renato Negrinho, Matthew R. Gormley, and Geoffrey J. Gordon. 2018 · 2018
Earlier work this paper cites.
Improving stability in deep reinforcement learning with weight averaging. In Uncertainty in artificial intelligence workshop on uncertainty in Deep learning
Evgenii Nikishin, Pavel Izmailov, Ben Athiwaratkun, Dmitrii Podoprikhin, Timur Garipov, Pavel Shvechikov, Dmitry Vetrov, and Andrew Gordon Wilson. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al · 2018
Earlier work this paper cites.
NerveNet: Learning Structured Policy with Graph Neural Networks. In 6th International Conference on Learning Representations, ICLR
Tingwu Wang, Renjie Liao, Jimmy Ba, and Sanja Fidler. 2018 · 2018
Earlier work this paper cites.
Transformer-XL: Attentive Language Models beyond a Fixed-Length Context. In Proceedings of the 57th Conference of the Association for Computational Linguistics, ACL . Association for Computational Linguistics, 2978–2988
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime G. Carbonell, Quoc Viet Le, and Ruslan Salakhutdinov. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2019, Minneapolis, MN, USA, June 2-7, 2019, Volume 1 (Long and Short Papers) . Association for Computational Linguistics, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
A Primer on PAC-Bayesian Learning
Benjamin Guedj. 2019 · 2019
Earlier work this paper cites.
On Inductive Biases in Deep Reinforcement Learning
Matteo Hessel, Hado van Hasselt, Joseph Modayil, and David Silver. 2019 · 2019
Earlier work this paper cites.
Axial Attention in Multidimensional Transformers
Jonathan Ho, Nal Kalchbrenner, Dirk Weissenborn, and Tim Salimans. 2019 · 2019
Earlier work this paper cites.
An Introduction to Variational Autoencoders
Diederik P. Kingma and Max Welling. 2019 · 2019
Earlier work this paper cites.
Neural network based reinforcement learning for audio-visual gaze control in human-robot interaction
Stéphane Lathuilière, Benoit Massé, Pablo Mesejo, and Radu Horaud. 2019 · 2019
Earlier work this paper cites.
Enhancing the Locality and Breaking the Memory Bottleneck of Transformer on Time Series Forecasting. In NeurIPS . 5244–5254
Shiyang Li, Xiaoyong Jin, Yao Xuan, Xiyou Zhou, Wenhu Chen, Yu-Xiang Wang, and Xifeng Yan. 2019 · 2019
Earlier work this paper cites.
Taming MAML: Efficient unbiased meta-reinforcement learning. In Proceedings of the 36th ICML , Vol. 97. PMLR, 4061–4071
Hao Liu, Richard Socher, and Caiming Xiong. 2019 · 2019
Earlier work this paper cites.
Reinforcement Learning with Attention that Works: A Self-Supervised Approach. In Neural Information Processing - 26th International Conference, ICONIP , Vol. 1143. Springer, 223–230
Anthony Manchin, Ehsan Abbasnejad, and Anton van den Hengel. 2019 · 2019
Earlier work this paper cites.
Solving Rubik’s Cube with a Robot Hand
OpenAI, Ilge Akkaya, Marcin Andrychowicz, Maciek Chociej, Mateusz Litwin, Bob McGrew, Arthur Petron, Alex Paino, Matthias Plappert, Glenn Powell, Raphael Ribas, Jonas Schneider, Nikolas Tezak, Jerry Tworek, Peter Welinder, Lilian Weng, Qiming Yuan, Wojciech Zaremba, and Lei Zhang. 2019 · 2019
Earlier work this paper cites.
Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables. In Proceedings of the 36th International Conference on Machine Learning, ICML , Vol. 97. PMLR, 5331–5340
Kate Rakelly, Aurick Zhou, Chelsea Finn, Sergey Levine, and Deirdre Quillen. 2019 · 2019
Earlier work this paper cites.
Habitat: A Platform for Embodied AI Research. In IEEE/CVF International Conference on Computer Vision, ICCV . IEEE, 9338–9346
Manolis Savva, Jitendra Malik, Devi Parikh, Dhruv Batra, Abhishek Kadian, Oleksandr Maksymets, Yili Zhao, Erik Wijmans, Bhavana Jain, Julian Straub, Jia Liu, and Vladlen Koltun. 2019 · 2019
Earlier work this paper cites.
Reinforcement Learning Upside Down: Don’t Predict Rewards - Just Map Them to Actions
Jürgen Schmidhuber. 2019 · 2019
Earlier work this paper cites.
Rewards Prediction-Based Credit Assignment for Reinforcement Learning With Sparse Binary Rewards
Minah Seo, Luiz Felipe Vecchietti, Sangkeum Lee, and Dongsoo Har. 2019 · 2019
Earlier work this paper cites.
Is Attention Interpretable?. In Proceedings of the 57th Conference of the Association for Computational Linguistics, ACL . Association for Computational Linguistics, 2931–2951
Sofia Serrano and Noah A. Smith. 2019 · 2019
Earlier work this paper cites.
A Survey of Deep Reinforcement Learning in Video Games
Kun Shao, Zhentao Tang, Yuanheng Zhu, Nannan Li, and Dongbin Zhao. 2019a · 2019
Earlier work this paper cites.
StarCraft Micromanagement With Reinforcement Learning and Curriculum Transfer Learning
Kun Shao, Yuanheng Zhu, and Dongbin Zhao. 2019b · 2019
Earlier work this paper cites.
Felix M. Strnad, Wolfram Barfuss, Jonathan F. Donges, and Jobst Heitzig. 2019 · 2019
Earlier work this paper cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M. Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H. Choi, Richard Powell, Timo Ewalds, Petko Georgiev, Junhyuk Oh, Dan Horgan, Manuel Kroiss, Ivo Danihelka, Aja Huang, Laurent Sifre, Trevor Cai, John P. Agapiou, Max Jaderberg, Alexander Sasha Vezhnevets, Rémi Leblond, Tobias Pohlen, Valentin Dalibard, David Budden, Yury Sulsky, James Molloy, Tom Le Paine, Çaglar Gülçehre, Ziyu Wang, Tobias Pfaff, Yuhuai Wu, Roman Ring, Dani Yogatama, Dario Wünsch, Katrina McKinney, Oliver Smith, Tom Schaul, Timothy P. Lillicrap, Koray Kavukcuoglu, Demis Hassabis, Chris Apps, and David Silver. 2019 · 2019
Earlier work this paper cites.
Reinforced Transformer for Medical Image Captioning. In Machine Learning in Medical Imaging - 10th International Workshop, MLMI, MICCAI , Vol. 11861. Springer, 673–680
Yuxuan Xiong, Bo Du, and Pingkun Yan. 2019 · 2019
Earlier work this paper cites.
Interpretable, Verifiable, and Robust Reinforcement Learning via Program Synthesis. In xxAI - Beyond Explainable AI - International Workshop, Held in Conjunction with ICML , Vol. 13200. Springer, 207–228
Osbert Bastani, Jeevana Priya Inala, and Armando Solar-Lezama. 2020 · 2020
Earlier work this paper cites.
On the Computational Power of Transformers and Its Implications in Sequence Modeling. In Proceedings of the 24th Conference on Computational Natural Language Learning, CoNLL 2020, Online, November 19-20, 2020 . Association for Computational Linguistics, 455–475
Satwik Bhattamishra, Arkil Patel, and Navin Goyal. 2020 · 2020
Earlier work this paper cites.
Language Models are Few-Shot Learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Earlier work this paper cites.
VisualHints: A Visual-Lingual Environment for Multimodal Reinforcement Learning
Thomas Carta, Subhajit Chaudhury, Kartik Talamadupula, and Michiaki Tatsubori. 2020 · 2020
Earlier work this paper cites.
Learning To Explore Using Active Neural SLAM. In 8th International Conference on Learning Representations, ICLR
Devendra Singh Chaplot, Dhiraj Gandhi, Saurabh Gupta, Abhinav Gupta, and Ruslan Salakhutdinov. 2020 · 2020
Cited alongside, same era.
Meshed-Memory Transformer for Image Captioning. In IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR . Computer Vision Foundation / IEEE, 10575–10584
Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi, and Rita Cucchiara. 2020 · 2020
Cited alongside, same era.
A Generalization of Transformer Networks to Graphs
Vijay Prakash Dwivedi and Xavier Bresson. 2020 · 2020
Cited alongside, same era.
Representations for Stable Off-Policy Reinforcement Learning. In Proceedings of the 37th ICML , Vol. 119. PMLR, 3556–3565
Dibya Ghosh and Marc G. Bellemare. 2020 · 2020
Cited alongside, same era.
Deep Reinforcement Learning for Stock Portfolio Optimization
All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL
Kai Arulkumaran, Dylan R. Ashley, Jürgen Schmidhuber, and Rupesh Kumar Srivastava. 2022 · 2022
Later among the works it cites.
CoBERL: Contrastive BERT for Reinforcement Learning. In The Tenth International Conference on Learning Representations, ICLR
Andrea Banino, Adrià Puigdomènech Badia, Jacob C. Walker, Tim Scholtes, Jovana Mitrovic, and Charles Blundell. 2022 · 2022
Later among the works it cites.
Contextualize Me - The Case for Context in Reinforcement Learning
Carolin Benjamins, Theresa Eimer, Frederik Schubert, Aditya Mohan, André Biedenkapp, Bodo Rosenhahn, Frank Hutter, and Marius Lindauer. 2022 · 2022
Later among the works it cites.
TransDreamer: Reinforcement Learning with Transformer World Models
Chang Chen, Yi-Fu Wu, Jaesik Yoon, and Sungjin Ahn. 2022a · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Le Trung Hieu. 2020 · 2020
Cited alongside, same era.
The act of remembering: a study in partially observable reinforcement learning
Rodrigo Toro Icarte, Richard Anthony Valenzano, Toryn Q. Klassen, Phillip J. K. Christoffersen, Amir-massoud Farahmand, and Sheila A. McIlraith. 2020 · 2020
Cited alongside, same era.
Scaling Laws for Neural Language Models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020 · 2020
Cited alongside, same era.
Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention. In Proceedings of the 37th ICML , Vol. 119. PMLR, 5156–5165
Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, and François Fleuret. 2020 · 2020
Cited alongside, same era.
CURL: Contrastive Unsupervised Representations for Reinforcement Learning. In Proceedings of the 37th ICML , Vol. 119. PMLR, 5639–5650
Michael Laskin, Aravind Srinivas, and Pieter Abbeel. 2020 · 2020
Cited alongside, same era.
Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Sergey Levine, Aviral Kumar, George Tucker, and Justin Fu. 2020a · 2020
Cited alongside, same era.
Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Sergey Levine, Aviral Kumar, George Tucker, and Justin Fu. 2020b · 2020
Cited alongside, same era.
Understanding the Difficulty of Training Transformers. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing, EMNLP 2020, Online, November 16-20, 2020 . Association for Computational Linguistics, 5747–5763
Liyuan Liu, Xiaodong Liu, Jianfeng Gao, Weizhu Chen, and Jiawei Han. 2020 · 2020
Cited alongside, same era.
Wei Chen, Cheng Zhong, Jiajie Peng, and Zhongyu Wei. 2022b · 2022
Later among the works it cites.
Dynamic Planning in Open-Ended Dialogue using Reinforcement Learning
Deborah Cohen, Moonkyung Ryu, Yinlam Chow, Orgad Keller, Ido Greenberg, Avinatan Hassidim, Michael Fink, Yossi Matias, Idan Szpektor, Craig Boutilier, and Gal Elidan. 2022 · 2022
Later among the works it cites.
Deep Transformer Q-Networks for Partially Observable Reinforcement Learning
Kevin Esslinger, Robert Platt, and Christopher Amato. 2022 · 2022
Later among the works it cites.
Object Memory Transformer for Object Goal Navigation
Rui Fukushima, Kei Ota, Asako Kanezaki, Yoko Sasaki, and Yusuke Yoshiyasu. 2022 · 2022
Later among the works it cites.
Manuel Goulão and Arlindo L. Oliveira. 2022 · 2022
Later among the works it cites.
Multi-agent deep reinforcement learning: a survey
Sven Gronauer and Klaus Diepold. 2022 · 2022
Later among the works it cites.
Perceiver IO: A General Architecture for Structured Inputs & Outputs
Andrew Jaegle, Sebastian Borgeaud, Jean-Baptiste Alayrac, Carl Doersch, Catalin Ionescu, David Ding, Skanda Koppula, Daniel Zoran, Andrew Brock, Evan Shelhamer, Olivier J. Hénaff, Matthew M. Botvinick, Andrew Zisserman, Oriol Vinyals, and João Carreira. 2022 · 2022
Later among the works it cites.
Selective Token Generation for Few-shot Natural Language Generation. In Proceedings of the 29th International Conference on Computational Linguistics, COLING 2022, Gyeongju, Republic of Korea, October 12-17, 2022 . International Committee on Computational Linguistics, 5837–5856
DaeJin Jo, Taehwan Kwon, Eun-Sol Kim, and Sungwoong Kim. 2022 · 2022
Later among the works it cites.
Vision Transformer for Learning Driving Policies in Complex and Dynamic Environments. In IEEE Intelligent Vehicles Symposium . IEEE, 1558–1564
Eshagh Kargar and Ville Kyrki. 2022 · 2022
Later among the works it cites.
Transformer-Based Value Function Decomposition for Cooperative Multi-Agent Reinforcement Learning in StarCraft
Muhammad Junaid Khan, Syed Hammad Ahmed, and Gita Sukthankar. 2022 · 2022
Later among the works it cites.
Deep Reinforcement Learning for Autonomous Driving: A Survey
B. Ravi Kiran, Ibrahim Sobh, Victor Talpaert, Patrick Mannion, Ahmad A. Al Sallab, Senthil Kumar Yogamani, and Patrick Pérez. 2022 · 2022
Later among the works it cites.
Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning. In 10th International Conference on Learning Representations, ICLR
Jakub Grudzien Kuba, Ruiqing Chen, Muning Wen, Ying Wen, Fanglei Sun, Jun Wang, and Yaodong Yang. 2022 · 2022
Later among the works it cites.
Multi-Game Decision Transformers. In NeurIPS
Kuang-Huei Lee, Ofir Nachum, Mengjiao Yang, Lisa Lee, Daniel Freeman, Sergio Guadarrama, Ian Fischer, Winnie Xu, Eric Jang, Henryk Michalewski, and Igor Mordatch. 2022 · 2022
Later among the works it cites.
Learning to Navigate in Interactive Environments with the Transformer-based Memory
Weiyuan Li, Ruoxin Hong, Jiwei Shen, and Yue Lu. 2022a · 2022
Later among the works it cites.
A survey of transformers
Tianyang Lin, Yuxin Wang, Xiangyang Liu, and Xipeng Qiu. 2022 · 2022
Later among the works it cites.
Haochen Liu, Zhiyu Huang, Xiaoyu Mo, and Chen Lv. 2022b · 2022
Later among the works it cites.
Challenges and Opportunities in Offline Reinforcement Learning from Visual Observations
Cong Lu, Philip J. Ball, Tim G. J. Rudner, Jack Parker-Holder, Michael A. Osborne, and Yee Whye Teh. 2022 · 2022
Later among the works it cites.
Transformers are Meta-Reinforcement Learners. In International Conference on Machine Learning, ICML , Vol. 162. PMLR, 15340–15359
Luckeciano C. Melo. 2022 · 2022
Later among the works it cites.
Transformers are Sample Efficient World Models
Vincent Micheli, Eloi Alonso, and François Fleuret. 2022 · 2022
Later among the works it cites.
A Survey of Explainable Reinforcement Learning
Stephanie Milani, Nicholay Topin, Manuela Veloso, and Fei Fang. 2022 · 2022
Later among the works it cites.
Vehicle routing problems over time: a survey
Andrea Mor and Maria Grazia Speranza. 2022 · 2022
Later among the works it cites.
Comparing BERT-based Reward Functions for Deep Reinforcement Learning in Machine Translation. In Proceedings of the 9th Workshop on Asian Translation, WAT@COLING 2022, Gyeongju, Republic of Korea, October 17, 2022 . International Conference on Computational Linguistics, 37–43
Yuki Nakatani, Tomoyuki Kajiwara, and Takashi Ninomiya. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback. In NeurIPS
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F. Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Later among the works it cites.
You Can’t Count on Luck: Why Decision Transformers and RvS Fail in Stochastic Environments
Keiran Paster, Sheila A. McIlraith, and Jimmy Ba. 2022 · 2022
Later among the works it cites.
Interpretable Navigation Agents Using Attention-Augmented Memory. In IEEE International Conference on Systems, Man, and Cybernetics, SMC 2022, Prague, Czech Republic, October 9-12, 2022 . IEEE, 2575–2582
Jia Qu, Shotaro Miwa, and Yukiyasu Domae. 2022 · 2022
Later among the works it cites.
Masked World Models for Visual Control. In Conference on Robot Learning, CoRL , Vol. 205. PMLR, 1332–1344
Younggyo Seo, Danijar Hafner, Hao Liu, Fangchen Liu, Stephen James, Kimin Lee, and Pieter Abbeel. 2022 · 2022
Later among the works it cites.
StARformer: Transformer with State-Action-Reward Representations for Visual Reinforcement Learning. In Computer Vision - ECCV - 17th European Conference , Vol. 13699. Springer, 462–479
Jinghuan Shang, Kumara Kahatapitiya, Xiang Li, and Michael S. Ryoo. 2022 · 2022
Later among the works it cites.
Attention-Based Learning for Combinatorial Optimization
Carson Smith. 2022 · 2022
Later among the works it cites.
Evaluating Vision Transformer Methods for Deep Reinforcement Learning from Pixels
Tianxin Tao, Daniele Reda, and Michiel van de Panne. 2022 · 2022
Later among the works it cites.
Natural language processing with transformers
Lewis Tunstall, Leandro von Werra, and Thomas Wolf. 2022 · 2022
Later among the works it cites.
A Distributed Vehicle-assisted Computation Offloading Scheme based on DRL in Vehicular Networks. In 22nd IEEE International Symposium on Cluster, Cloud and Internet Computing, CCGrid 2022, Taormina, Italy, May 16-19, 2022 . IEEE, 200–209
Jiayue Wang, Hongbo Zhao, Haoqiang Liu, Liwei Geng, and Zebin Sun. 2022d · 2022
Later among the works it cites.
A Deep Reinforcement Learning Algorithm Using A New Graph Transformer Model for Routing Problems. In Intelligent Systems and Applications - Proceedings of the Intelligent Systems Conference, IntelliSys , Vol. 544. Springer, 365–379
Yang Wang and Zhibin Chen. 2022 · 2022
Later among the works it cites.
Multi-Agent Reinforcement Learning is a Sequence Modeling Problem. In NeurIPS
Muning Wen, Jakub Grudzien Kuba, Runji Lin, Weinan Zhang, Ying Wen, Jun Wang, and Yaodong Yang. 2022 · 2022
Later among the works it cites.
Learning Improvement Heuristics for Solving Routing Problems
Yaoxin Wu, Wen Song, Zhiguang Cao, Jie Zhang, and Andrew Lim. 2022 · 2022
Later among the works it cites.
Multimodal Learning with Transformers: A Survey
Peng Xu, Xiatian Zhu, and David A. Clifton. 2022c · 2022
Later among the works it cites.
Taku Yamagata, Ahmed Khalil, and Raúl Santos-Rodríguez. 2022 · 2022
Later among the works it cites.
Multi-granularity scenarios understanding network for trajectory prediction
Biao Yang, Jicheng Yang, Rongrong Ni, Changchun Yang, and Xiaofeng Liu. 2022 · 2022
Later among the works it cites.
Learning Efficient Multi-agent Cooperative Visual Exploration. In Computer Vision - ECCV - 17th European Conference , Vol. 13699. Springer, 497–515
Chao Yu, Xinyi Yang, Jiaxuan Gao, Huazhong Yang, Yu Wang, and Yi Wu. 2022 · 2022
Later among the works it cites.
Exploiting Transformer in Reinforcement Learning for Interpretable Temporal Logic Motion Planning
Hao Zhang, Hao Wang, and Zhen Kan. 2022c · 2022
Later among the works it cites.
Visual Representation Learning with Transformer: A Sequence-to-Sequence Perspective
Li Zhang, Sixiao Zheng, Jiachen Lu, Xinxuan Zhao, Xiatian Zhu, Yanwei Fu, Tao Xiang, and Jianfeng Feng. 2022d · 2022
Later among the works it cites.
TVENet: Transformer-Based Visual Exploration Network for Mobile Robot in Unseen Environment
Tianyao Zhang, Xiaoguang Hu, Jin Xiao, and Guofeng Zhang. 2022a · 2022
Later among the works it cites.
Hyperparameter Search for Machine Learning Algorithms for Optimizing the Computational Complexity
Yasser A Ali, Emad Mahrous Awwad, Muna Al-Razgan, and Ali Maarouf. 2023 · 2023
Closest in time.
A Deep Reinforcement Learning Framework Based on an Attention Mechanism and Disjunctive Graph Embedding for the Job-Shop Scheduling Problem
Ruiqi Chen, Wenxin Li, and Hongbing Yang. 2023 · 2023
Closest in time.
Credit assignment with predictive contribution measurement in multi-agent reinforcement learning
Renlong Chen and Ying Tan. 2023 · 2023
Closest in time.
Reward modeling for mitigating toxicity in transformer-based language models
Farshid Faal, Ketra A. Schmitt, and Jia Yuan Yu. 2023 · 2023
Closest in time.
Unsupervised Learning of Temporal Abstractions With Slot-Based Transformers
Anand Gopalakrishnan, Kazuki Irie, Jürgen Schmidhuber, and Sjoerd van Steenkiste. 2023 · 2023
Closest in time.
On The Computational Complexity of Self-Attention. In International Conference on Algorithmic Learning Theory , Vol. 201. PMLR, 597–619
Feyza Duman Keles, Pruthuvi Mahesakya Wijewardena, and Chinmay Hegde. 2023 · 2023
Closest in time.
Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning
Taylor W. Killian, Sonali Parbhoo, and Marzyeh Ghassemi. 2023 · 2023
Closest in time.
Preference Transformer: Modeling Human Preferences using Transformers for RL
Changyeon Kim, Jongjin Park, Jinwoo Shin, Honglak Lee, Pieter Abbeel, and Kimin Lee. 2023 · 2023
Closest in time.
A survey on deep reinforcement learning for audio-based applications
Siddique Latif, Heriberto Cuayáhuitl, Farrukh Pervez, Fahad Shamshad, Hafiz Shehbaz Ali, and Erik Cambria. 2023 · 2023
Closest in time.
BATFormer: Towards Boundary-Aware Lightweight Transformer for Efficient Medical Image Segmentation
Xian Lin, Li Yu, Kwang-Ting Cheng, and Zengqiang Yan. 2023 · 2023
Closest in time.
DrugEx v3: scaffold-constrained drug design with graph transformer-based reinforcement learning
Xuhan Liu, Kai Ye, Herman W. T. van Vlijmen, Adriaan P. IJzerman, and Gerard J. P. van Westen. 2023 · 2023
Closest in time.
Foundation models for generalist medical artificial intelligence
Michael Moor, Oishi Banerjee, Zahra Shakeri Hossein Abad, Harlan M Krumholz, Jure Leskovec, Eric J Topol, and Pranav Rajpurkar. 2023 · 2023
Closest in time.
POPGym: Benchmarking Partially Observable Reinforcement Learning
Steven D. Morad, Ryan Kortvelesy, Matteo Bettini, Stephan Liwicki, and Amanda Prorok. 2023 · 2023
Closest in time.
Deep reinforcement learning for optimal well control in subsurface systems with uncertain geology
Yusuf Nasir and Louis J. Durlofsky. 2023 · 2023
Closest in time.
A review of cooperative multi-agent deep reinforcement learning
Afshin Oroojlooy and Davood Hajinezhad. 2023 · 2023
Closest in time.
Structured learning based heuristics to solve the single machine scheduling problem with release times and sum of completion times
Axel Parmentier and Vincent T’kindt. 2023 · 2023
Closest in time.
Transformer-based World Models Are Happy With 100k Interactions
Jan Robine, Marc Höftmann, Tobias Uelwer, and Stefan Harmeling. 2023 · 2023
Closest in time.
Solving combinatorial optimization problems over graphs with BERT-Based Deep Reinforcement Learning
Qi Wang, Kenneth H. Lai, and Chunlei Tang. 2023 · 2023
Closest in time.
ConvNeXt V2: Co-designing and Scaling ConvNets with Masked Autoencoders
Sanghyun Woo, Shoubhik Debnath, Ronghang Hu, Xinlei Chen, Zhuang Liu, In So Kweon, and Saining Xie. 2023 · 2023
Closest in time.
Reinforcement Learning in Healthcare: A Survey
Chao Yu, Jiming Liu, Shamim Nemati, and Guosheng Yin. 2023 · 2023
Closest in time.
One for all: One-stage referring expression comprehension with dynamic reasoning
Zhipeng Zhang, Zhimin Wei, Zhongzhen Huang, Rui Niu, and Peng Wang. 2023 · 2023
Closest in time.
A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT
Ce Zhou, Qian Li, Chen Li, Jun Yu, Yixin Liu, Guangjing Wang, Kai Zhang, Cheng Ji, Qiben Yan, Lifang He, Hao Peng, Jianxin Li, Jia Wu, Ziwei Liu, Pengtao Xie, Caiming Xiong, Jian Pei, Philip S. Yu, and Lichao Sun. 2023 · 2023
Closest in time.