Fetching the paper…
Reading the bibliography…
In complex systems, we often observe complex global behavior emerge from a collection of agents interacting with each other in their environment, with each individual agent acting only on locally available information, without knowing the full picture.
Theory of self-reproducing automata
J. Neumann, A. W. Burks, et al · 1966
Earlier work this paper cites.
Cellular automata
E. F. Codd · 1968
Earlier work this paper cites.
Vision substitution by tactile image projection
P. Bach-y Rita, C. C. Collins, F. A. Saunders, B. White, and L. Scadden · 1969
Earlier work this paper cites.
The game of life
J. Conway · 1970
Earlier work this paper cites.
Cellular automata as models of complexity
S. Wolfram · 1984
Earlier work this paper cites.
Cellular neural networks: Theory
L. O. Chua and L. Yang · 1988
Earlier work this paper cites.
Learning to control fast-weight memories: An alternative to dynamic recurrent networks
J. Schmidhuber · 1992
Earlier work this paper cites.
Reducing the ratio between learning complexity and number of time varying variables in fully recurrent nets
J. Schmidhuber · 1993
Earlier work this paper cites.
A ‘self-referential’weight matrix
J. Schmidhuber · 1993
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Cellular automata
B. Chopard and M. Droz · 1998
Earlier work this paper cites.
Mosaic model for sensorimotor learning and control
M. Haruno, D. M. Wolpert, and M. Kawato · 2001
Earlier work this paper cites.
Learning to learn using gradient descent
S. Hochreiter, A. S. Younger, and P. R. Conwell · 2001
Earlier work this paper cites.
Sensory substitution and the human–machine interface
P. Bach-y Rita and S. W. Kercel · 2003
Earlier work this paper cites.
The cma evolution strategy: a comparing review
N. Hansen · 2006
Earlier work this paper cites.
Visualizing data using t-sne
L. Van der Maaten and G. Hinton · 2008
Earlier work this paper cites.
Pilco: A model-based and data-efficient approach to policy search
M. Deisenroth and C. E. Rasmussen · 2011
Earlier work this paper cites.
Unshackling evolution: evolving soft robots with multiple materials and a powerful generative encoding
N. Cheney, R. MacCurdy, J. Clune, and H. Lipson · 2014
Earlier work this paper cites.
A. Graves, G. Wayne, and I. Danihelka · 2014
Earlier work this paper cites.
Effective approaches to attention-based neural machine translation
M.-T. Luong, H. Pham, and C. D. Manning · 2015
Earlier work this paper cites.
Deep attention recurrent q-network
I. Sorokin, A. Seleznev, M. Pavlov, A. Fedorov, and A. Ignateva · 2015
Earlier work this paper cites.
Using fast weights to attend to the recent past
J. Ba, G. Hinton, V. Mnih, J. Z. Leibo, and C. Ionescu · 2016
Earlier work this paper cites.
J. L. Ba, J. R. Kiros, and G. E. Hinton · 2016
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Earlier work this paper cites.
Pybullet, a python module for physics simulation for games, robotics and machine learning, 2016
E. Coumans and Y. Bai · 2016
Earlier work this paper cites.
Rl2: Fast reinforcement learning via slow reinforcement learning
Y. Duan, J. Schulman, X. Chen, P. L. Bartlett, I. Sutskever, and P. Abbeel · 2016
Earlier work this paper cites.
Improving PILCO with Bayesian neural network dynamics models
Y. Gal, R. McAllister, and C. E. Rasmussen · 2016
Earlier work this paper cites.
Permutation-equivariant neural networks applied to dynamics prediction
N. Guttenberg, N. Virgo, O. Witkowski, H. Aoki, and R. Kanai · 2016
Earlier work this paper cites.
D. Ha, A. Dai, and Q. V. Le · 2016
Earlier work this paper cites.
Carracing-v0, 2016
O. Klimov · 2016
Cited alongside, same era.
Learning to reinforcement learn
J. X. Wang, Z. Kurth-Nelson, D. Tirumala, H. Soyer, J. Z. Leibo, R. Munos, C. Blundell, D. Kumaran, and M. Botvinick · 2016
Cited alongside, same era.
Multi-focus attention network for efficient deep reinforcement learning
J. Choi, B.-J. Lee, and B.-T. Zhang · 2017
Cited alongside, same era.
Evolving stable strategies
D. Ha · 2017
Cited alongside, same era.
Geometric deep learning on graphs and manifolds using mixture model cnns
F. Monti, D. Boscaini, J. Masci, E. Rodola, J. Svoboda, and M. M. Bronstein · 2017
Cited alongside, same era.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. u. Kaiser, and I. Polosukhin · 2017
Deep reinforcement learning with relational inductive biases
V. Zambaldi, D. Raposo, A. Santoro, V. Bapst, Y. Li, I. Babuschkin, K. Tuyls, D. Reichert, T. Lillicrap, E. Lockhart, M. Shanahan, V. Langston, R. Pascanu, M. Botvinick, O. Vinyals, and P. Battaglia · 2019
Later among the works it cites.
Language models are few-shot learners
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Later among the works it cites.
Extracting training data from large language models
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, U. Erlingsson, et al · 2020
Later among the works it cites.
Decentralized reinforcement learning: Global decision-making via local economic transactions
M. Chang, S. Kaushik, S. M. Weinberg, T. Griffiths, and S. Levine · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
P. Veličković, G. Cucurull, A. Casanova, A. Romero, P. Lio, and Y. Bengio · 2017
Cited alongside, same era.
M. Zaheer, S. Kottur, S. Ravanbakhsh, B. Poczos, R. Salakhutdinov, and A. Smola · 2017
Cited alongside, same era.
Differentiable mpc for end-to-end planning and control
B. Amos, I. D. J. Rodriguez, J. Sacks, B. Boots, and J. Z. Kolter · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Cited alongside, same era.
Investigating human priors for playing video games
R. Dubey, P. Agrawal, D. Pathak, T. L. Griffiths, and A. A. Efros · 2018
Cited alongside, same era.
Som-vae: Interpretable discrete representation learning on time series
V. Fortuin, M. Hüser, F. Locatello, H. Strathmann, and G. Rätsch · 2018
Cited alongside, same era.
K. Choromanski, V. Likhosherstov, D. Dohan, X. Song, A. Gane, T. Sarlos, P. Hawkins, J. Davis, A. Mohiuddin, L. Kaiser, et al · 2020
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, et al · 2020
Later among the works it cites.
Livewired: The inside story of the ever-changing brain
D. Eagleman · 2020
Later among the works it cites.
Taming transformers for high-resolution image synthesis
P. Esser, R. Rombach, and B. Ommer · 2020
Later among the works it cites.
One policy to control them all: Shared modular policies for agent-agnostic control
W. Huang, I. Mordatch, and D. Pathak · 2020
Later among the works it cites.
Transformers are graph neural networks
C. Joshi · 2020
Later among the works it cites.
Meta learning backpropagation and improving it
L. Kirsch and J. Schmidhuber · 2020
Later among the works it cites.
Pic: permutation invariant critic for multi-agent deep reinforcement learning
I.-J. Liu, R. A. Yeh, and A. G. Schwing · 2020
Later among the works it cites.
Backpropamine: training self-modifying neural networks with differentiable neuromodulated plasticity
T. Miconi, A. Rawal, J. Clune, and K. O. Stanley · 2020
Later among the works it cites.
Growing neural cellular automata
A. Mordvintsev, E. Randazzo, E. Niklasson, and M. Levin · 2020
Later among the works it cites.
Meta-learning through hebbian plasticity in random networks
E. Najarro and S. Risi · 2020
Later among the works it cites.
Giving up control: Neurons as reinforcement learning agents
J. Ott · 2020
Later among the works it cites.
Self-classifying mnist digits
E. Randazzo, A. Mordvintsev, E. Niklasson, M. Levin, and S. Greydanus · 2020
Later among the works it cites.
Neuroevolution of self-interpretable agents
Y. Tang, D. Nguyen, and D. Ha · 2020
Later among the works it cites.
Linformer: Self-attention with linear complexity
S. Wang, B. Li, M. Khabsa, H. Fang, and H. Ma · 2020
Later among the works it cites.
A comprehensive survey on graph neural networks
Z. Wu, S. Pan, F. Chen, G. Long, C. Zhang, and S. Y. Philip · 2020
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell · 2021
Closest in time.
Recurrent independent mechanisms
A. Goyal, A. Lamb, J. Hoffmann, S. Sodhani, S. Levine, Y. Bengio, and B. Schölkopf · 2021
Closest in time.
Perceiver: General perception with iterative attention
A. Jaegle, F. Gimeno, A. Brock, A. Zisserman, O. Vinyals, and J. Carreira · 2021
Closest in time.
A gentle introduction to graph neural networks
B. Sanchez-Lengeling, E. Reif, A. Pearce, and A. Wiltschko · 2021
Closest in time.
Meta-learning bidirectional update rules
M. Sandler, M. Vladymyrov, A. Zhmoginov, N. Miller, A. Jackson, T. Madams, et al · 2021
Closest in time.
Growing 3d artefacts and functional machines with neural cellular automata
S. Sudhakaran, D. Grbic, S. Li, A. Katona, E. Najarro, C. Glanois, and S. Risi · 2021
Closest in time.
Nyströmformer: A nyström-based algorithm for approximating self-attention
Y. Xiong, Z. Zeng, R. Chakraborty, M. Tan, G. Fung, Y. Li, and V. Singh · 2021
Closest in time.
Learning to generate 3d shapes with generative cellular automata
D. Zhang, C. Choi, J. Kim, and Y. M. Kim · 2021
Closest in time.