Fetching the paper…
Reading the bibliography…
Inattentional blindness is the psychological phenomenon that causes one to miss things in plain sight.
Learning to generalize from sparse and underspecified rewards
Rishabh Agarwal, Chen Liang, Dale Schuurmans, and Mohammad Norouzi. 2019 · 1902
Earlier work this paper cites.
Investigating generalisation in continuous deep reinforcement learning
Chenyang Zhao, Olivier Siguad, Freek Stulp, and Timothy M Hospedales. 2019 · 1902
Earlier work this paper cites.
Attention augmented convolutional networks. In
Irwan Bello, Barret Zoph, Ashish Vaswani, Jonathon Shlens, and Quoc V Le. 2019 · 1904
Earlier work this paper cites.
Yuning Chai. 2019 · 1904
Earlier work this paper cites.
Local relation networks for image recognition. In
Han Hu, Zheng Zhang, Zhenda Xie, and Stephen Lin. 2019 · 1904
Earlier work this paper cites.
Skin Lesion Classification Using CNNs with Patch-Based Attention and Diagnosis-Guided Loss Weighting
Nils Gessert, Thilo Sentker, Frederic Madesta, Rudiger Schmitz, Helge Kniep, Ivo Baltruschat, Rene Werner, and Alexander Schlaefer. 2019 · 1905
Earlier work this paper cites.
Deepmdp: Learning continuous latent space models for representation learning
Carles Gelada, Saurabh Kumar, Jacob Buckman, Ofir Nachum, and Marc G Bellemare. 2019 · 1906
Earlier work this paper cites.
The Animal-AI Environment: Training and Testing Animal-Like Artificial Cognition
Benjamin Beyret, José Hernández-Orallo, Lucy Cheke, Marta Halina, Murray Shanahan, and Matthew Crosby. 2019 · 1909
Earlier work this paper cites.
Recurrent independent mechanisms
Anirudh Goyal, Alex Lamb, Jordan Hoffmann, Shagun Sodhani, Sergey Levine, Yoshua Bengio, and Bernhard Schölkopf. 2019 · 1909
Earlier work this paper cites.
Emergent systematic generalization in a situated agent
Felix Hill, Andrew Lampinen, Rosalia Schneider, Stephen Clark, Matthew Botvinick, James L McClelland, and Adam Santoro. 2019 · 1910
Earlier work this paper cites.
On the Relationship between Self-Attention and Convolutional Layers
Jean-Baptiste Cordonnier, Andreas Loukas, and Martin Jaggi. 2019 · 1911
Earlier work this paper cites.
Improving Semantic Segmentation of Aerial Images Using Patch-based Attention
Lei Ding, Hao Tang, and Lorenzo Bruzzone. 2019 · 1911
Earlier work this paper cites.
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model
Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert, Karen Simonyan, Laurent Sifre, Simon Schmitt, Arthur Guez, Edward Lockhart, Demis Hassabis, Thore Graepel, Timothy P. Lillicrap, and David Silver. 2019 · 1911
Earlier work this paper cites.
Leveraging Procedural Generation to Benchmark Reinforcement Learning
Karl Cobbe, Christopher Hesse, Jacob Hilton, and John Schulman. 2019 · 1912
Earlier work this paper cites.
The organization of behavior
Donald O Hebb. 1949 · 1949
Earlier work this paper cites.
Learning to generate artificial fovea trajectories for target detection
Juergen Schmidhuber and Rudolf Huber. 1991 · 1991
Earlier work this paper cites.
A ‘self-referential’ weight matrix. In
Juergen Schmidhuber. 1993 · 1993
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Discovering neural nets with low Kolmogorov complexity and high generalization capability
Juergen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Inattentional blindness
Arien Mack, Irvin Rock, et al · 1998
Earlier work this paper cites.
Improving Deep Neuroevolution via Deep Innovation Protection
Sebastian Risi and Kenneth O. Stanley. 2020 · 2001
Earlier work this paper cites.
Rotation, Translation, and Cropping for Zero-Shot Generalization
Chang Ye, Ahmed Khalifa, Philip Bontrager, and Julian Togelius. 2020 · 2001
Earlier work this paper cites.
A taxonomy for artificial embryogeny
Kenneth O Stanley and Risto Miikkulainen. 2003 · 2003
Earlier work this paper cites.
Real-time video annotations for augmented reality. In
Edward Rosten, Gerhard Reitmayr, and Tom Drummond. 2005 · 2005
Earlier work this paper cites.
The CMA Evolution Strategy: A Comparing Review
Nikolaus Hansen. 2006 · 2006
Earlier work this paper cites.
Core knowledge
E. S Spelke and K. D. Kinzler. 2007 · 2007
Earlier work this paper cites.
Compositional pattern producing networks: A novel abstraction of development
Kenneth O Stanley. 2007 · 2007
Earlier work this paper cites.
Evolving coordinated quadruped gaits with the HyperNEAT generative encoding. In
Jeff Clune, Benjamin E Beckmann, Charles Ofria, and Robert T Pennock. 2009 · 2009
Earlier work this paper cites.
Matrix factorization techniques for recommender systems
Yehuda Koren, Robert Bell, and Chris Volinsky. 2009 · 2009
Earlier work this paper cites.
A hypercube-based encoding for evolving large-scale neural networks
Kenneth O Stanley, David B D’Ambrosio, and Jason Gauci. 2009 · 2009
Earlier work this paper cites.
Attention as inference: selection is probabilistic; responses are all-or-none samples
Edward Vul, Deborah Hanus, and Nancy Kanwisher. 2009 · 2009
Earlier work this paper cites.
On the performance of indirect encoding across the continuum of regularity
Jeff Clune, Kenneth O Stanley, Robert T Pennock, and Charles Ofria. 2011 · 2011
Earlier work this paper cites.
Metamers of the ventral stream
Jeremy Freeman and Eero P Simoncelli. 2011 · 2011
Earlier work this paper cites.
Thinking, fast and slow
Daniel Kahneman. 2011 · 2011
Earlier work this paper cites.
Vision: Processing Information
2012 · 2012
Earlier work this paper cites.
HyperNEAT-GGP: A HyperNEAT-based Atari general game player. In
Matthew Hausknecht, Piyush Khandelwal, Risto Miikkulainen, and Peter Stone. 2012 · 2012
Earlier work this paper cites.
An enhanced hypercube-based encoding for evolving the placement, density, and connectivity of neurons
Sebastian Risi and Kenneth O Stanley. 2012 · 2012
Cited alongside, same era.
Evolving large-scale neural networks for vision-based reinforcement learning. In
Jan Koutník, Giuseppe Cuccu, Jürgen Schmidhuber, and Faustino Gomez. 2013 · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller. 2013 · 2013
Cited alongside, same era.
Confronting the challenge of learning a flexible neural controller for a diversity of morphologies. In
Sebastian Risi and Kenneth O Stanley. 2013 · 2013
Cited alongside, same era.
Low-rank matrix factorization for deep neural network training with high-dimensional output targets. In
Tara N Sainath, Brian Kingsbury, Vikas Sindhwani, Ebru Arisoy, and Bhuvana Ramabhadran. 2013 · 2013
Learning latent dynamics for planning from pixels
Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson. 2018 · 2018
Later among the works it cites.
Psychlab: a psychology laboratory for deep reinforcement learning agents
Joel Z Leibo, Cyprien de Masson d’Autume, Daniel Zoran, David Amos, Charles Beattie, Keith Anderson, Antonio García Castañeda, Manuel Sanchez, Simon Green, Audrunas Gruslys, et al · 2018
Later among the works it cites.
Simple random search of static linear policies is competitive for reinforcement learning. In
Horia Mania, Aurelia Guy, and Benjamin Recht. 2018 · 2018
Later among the works it cites.
Differentiable plasticity: training plastic neural networks with backpropagation
Thomas Miconi, Jeff Clune, and Kenneth O Stanley. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multiple object recognition with visual attention
Jimmy Ba, Volodymyr Mnih, and Koray Kavukcuoglu. 2014 · 2014
Cited alongside, same era.
Consciousness and the brain: Deciphering how the brain codes our thoughts
Stanislas Dehaene. 2014 · 2014
Cited alongside, same era.
Recurrent models of visual attention. In
Volodymyr Mnih, Nicolas Heess, Alex Graves, et al · 2014
Cited alongside, same era.
Deep networks with internal selective attention through feedback connections. In
Marijn F Stollenga, Jonathan Masci, Faustino Gomez, and Jürgen Schmidhuber. 2014 · 2014
Cited alongside, same era.
Fat Cat Fail Vines Compilation
Extraordinary Epic. 2015 · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Deep Attention Recurrent Q-Network
Ivan Sorokin, Alexey Seleznev, Mikhail Pavlov, Aleksandr Fedorov, and Anastasiia Ignateva. 2015 · 2015
Cited alongside, same era.
Nils Müller and Tobias Glasmachers. 2018 · 2018
Later among the works it cites.
Assessing generalization in deep reinforcement learning
Charles Packer, Katelyn Gao, Jernej Kos, Philipp Krähenbühl, Vladlen Koltun, and Dawn Song. 2018 · 2018
Later among the works it cites.
ViZDoom Competitions: Playing Doom From Pixels
Marek Wydmuch, Michal Kempka, and Wojciech Jaskowski. 2019 · 2018
Later among the works it cites.
Natural environment benchmarks for reinforcement learning
Amy Zhang, Yuxin Wu, and Joelle Pineau. 2018 · 2018
Later among the works it cites.
Transformers from Scratch
Peter Bloem. 2019 · 2019
Later among the works it cites.
Playing atari with six neurons. In
Giuseppe Cuccu, Julian Togelius, and Philippe Cudré-Mauroux. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Saccader: Improving Accuracy of Hard Attention Models for Vision
Gamaleldin Elsayed, Simon Kornblith, and Quoc V Le. 2019 · 2019
Later among the works it cites.
Learning to Predict Without Looking Ahead: World Models Without Forward Prediction. In
Daniel Freeman, David Ha, and Luke Metz. 2019 · 2019
Later among the works it cites.
Weight agnostic neural networks. In
Adam Gaier and David Ha. 2019 · 2019
Later among the works it cites.
CMA-ES/pycma on Github
Nikolaus Hansen, Youhei Akimoto, and Petr Baudis. 2019 · 2019
Later among the works it cites.
Obstacle tower: A generalization challenge in vision, control, and planning
Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges, Jonathan Harper, Ervin Teng, Hunter Henry, Adam Crespi, Julian Togelius, and Danny Lange. 2019 · 2019
Later among the works it cites.
Model-based reinforcement learning for atari
Lukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski, Roy H Campbell, Konrad Czechowski, Dumitru Erhan, Chelsea Finn, Piotr Kozakowski, Sergey Levine, et al · 2019
Later among the works it cites.
Unsupervised learning of object keypoints for perception and control. In
Tejas D Kulkarni, Ankush Gupta, Catalin Ionescu, Sebastian Borgeaud, Malcolm Reynolds, Andrew Zisserman, and Volodymyr Mnih. 2019 · 2019
Later among the works it cites.
Towards Interpretable Reinforcement Learning Using Attention Augmented Agents. In
Alexander Mott, Daniel Zoran, Mike Chrzanowski, Daan Wierstra, and Danilo Jimenez Rezende. 2019 · 2019
Later among the works it cites.
Stand-Alone Self-Attention in Vision Models. In
Niki Parmar, Prajit Ramachandran, Ashish Vaswani, Irwan Bello, Anselm Levskaya, and Jon Shlens. 2019 · 2019
Later among the works it cites.
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Deep neuroevolution of recurrent and discrete world models. In
Sebastian Risi and Kenneth O. Stanley. 2019 · 2019
Later among the works it cites.
Gencer Sumbul and Begüm Demir. 2019 · 2019
Later among the works it cites.
How to run evolution strategies on Google Kubernetes Engine
Yujin Tang and David Ha. 2019 · 2019
Later among the works it cites.
A critique of pure learning and what artificial neural networks can learn from animal brains
Anthony M Zador. 2019 · 2019
Later among the works it cites.
Deep reinforcement learning with relational inductive biases. In
Vinicius Zambaldi, David Raposo, Adam Santoro, Victor Bapst, Yujia Li, Igor Babuschkin, Karl Tuyls, David Reichert, Timothy Lillicrap, Edward Lockhart, Murray Shanahan, Victoria Langston, Razvan Pascanu, Matthew Botvinick, Oriol Vinyals, and Peter Battaglia. 2019 · 2019
Later among the works it cites.
Direct Fit to Nature: An Evolutionary Perspective on Biological and Artificial Neural Networks
Uri Hasson, Samuel A Nastase, and Ariel Goldstein. 2020 · 2020
Closest in time.
CarRacing-v0
Oleg Klimov. 2016 · 2020
Closest in time.
Network Randomization: A Simple Technique for Generalization in Deep Reinforcement Learning. In
Kimin Lee, Kibok Lee, Jinwoo Shin, and Honglak Lee. 2020 · 2020
Closest in time.
Car Racing with PyTorch
Xiaoteng Ma. 2019 · 2020
Closest in time.
DoomTakeCover-v0
Philip Paquette. 2017 · 2020
Closest in time.
Observational Overfitting in Reinforcement Learning. In
Xingyou Song, Yiding Jiang, Stephen Tu, Yilun Du, and Behnam Neyshabur. 2020 · 2020
Closest in time.
Continual learning with hypernetworks. In
Johannes von Oswald, Christian Henning, João Sacramento, and Benjamin F. Grewe. 2020 · 2020
Closest in time.
Discovery of latent 3d keypoints via end-to-end geometric reasoning. In
Supasorn Suwajanakorn, Noah Snavely, Jonathan J Tompson, and Mohammad Norouzi. 2018 · 2070
Closest in time.
Reproducible, Reusable, and Robust Reinforcement Learning - NeurIPS 2018
Joelle Pineau. 2018 · 2077
Closest in time.