Fetching the paper…
Reading the bibliography…
Recent advances in reinforcement learning have shown its potential to tackle complex real-life tasks.
A synopsis of linguistic theory, 1930-1955
John R Firth · 1957
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Approximation with artificial neural networks
Balázs Csanád Csáji · 2001
Earlier work this paper cites.
Learning embedded maps of markov processes
Yaakov Engel and Shie Mannor · 2001
Earlier work this paper cites.
Predictive representations of state
Michael L Littman and Richard S Sutton · 2002
Earlier work this paper cites.
Language: An introduction to the study of speech
Edward Sapir · 2004
Earlier work this paper cites.
Approximate Dynamic Programming: Solving the curses of dimensionality , volume 703
Warren B Powell · 2007
Earlier work this paper cites.
Natural language processing (almost) from scratch
Ronan Collobert, Jason Weston, Léon Bottou, Michael Karlen, Koray Kavukcuoglu, and Pavel Kuksa · 2011
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Cited alongside, same era.
Deep learning
Ian Goodfellow, Yoshua Bengio, and Aaron Courville · 2016
Cited alongside, same era.
Vizdoom: A doom-based ai research platform for visual reinforcement learning
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Emergence of invariance and disentanglement in deep representations
Alessandro Achille and Stefano Soatto · 2018
Later among the works it cites.
Textworld: A learning environment for text-based games
Marc-Alexandre Côté, Ákos Kádár, Xingdi Yuan, Ben Kybartas, Tavian Barnes, Emery Fine, James Moore, Matthew Hausknecht, Layla El Asri, Mahmoud Adada, et al · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Later among the works it cites.
Do latent tree learning models identify meaningful structure in sentences?
Adina Williams, Andrew Drozdov, and Samuel R Bowman · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Michał Kempka, Marek Wydmuch, Grzegorz Runc, Jakub Toczek, and Wojciech Jaśkowski · 2016
Cited alongside, same era.
Ask me anything: Dynamic memory networks for natural language processing
Ankit Kumar, Ozan Irsoy, Peter Ondruska, Mohit Iyyer, James Bradbury, Ishaan Gulrajani, Victor Zhong, Romain Paulus, and Richard Socher · 2016
Cited alongside, same era.
A decomposable attention model for natural language inference
Ankur Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit · 2016
Cited alongside, same era.
The state of the art in semantic representation
Omri Abend and Ari Rappoport · 2017
Cited alongside, same era.
Wojciech Samek, Thomas Wiegand, and Klaus-Robert Müller · 2017
Cited alongside, same era.
Later among the works it cites.
A comprehensive survey of deep learning for image captioning
MD Hossain, Ferdous Sohel, Mohd Fairuz Shiratuddin, and Hamid Laga · 2019
Closest in time.
Hierarchical decision making by generating and following natural language instructions
Hengyuan Hu, Denis Yarats, Qucheng Gong, Yuandong Tian, and Mike Lewis · 2019
Closest in time.
Adaptive discretization for episodic reinforcement learning in metric spaces
Sean R Sinclair, Siddhartha Banerjee, and Christina Lee Yu · 2019
Closest in time.
The natural language of actions
Guy Tennenholtz and Shie Mannor · 2019
Closest in time.
Action assembly: Sparse imitation learning for text based games with combinatorial action spaces
Chen Tessler, Tom Zahavy, Deborah Cohen, Daniel J Mankowitz, and Shie Mannor · 2019
Closest in time.