Fetching the paper…
Reading the bibliography…
A common vision from science fiction is that robots will one day inhabit our physical spaces, sense the world as we do, assist our physical labours, and communicate with us through natural language.
The Republic of Plato
Adam, J. et al. (1902) · 1902
Earlier work this paper cites.
Information asymmetry in kl-regularized rl
Galashov, A., Jayakumar, S. M., Hasenclever, L., Tirumala, D., Schwarz, J., Desjardins, G., Czarnecki, W. M., Teh, Y. W., Pascanu, R., and Heess, N. (2019) · 1905
Earlier work this paper cites.
Data-efficient image recognition with contrastive predictive coding
Hénaff, O. J., Srinivas, A., De Fauw, J., Razavi, A., Doersch, C., Eslami, S., and Oord, A. v. d. (2019) · 1905
Earlier work this paper cites.
Scaling data-driven robotics with reward sketching and batch reinforcement learning
Cabi, S., Gómez Colmenarejo, S., Novikov, A., Konyushkova, K., Reed, S., Jeong, R., Zolna, K., Aytar, Y., Budden, D., Vecerik, M., et al. (2019) · 1909
Earlier work this paper cites.
Task-relevant adversarial imitation learning
Zolna, K., Reed, S., Novikov, A., Colmenarej, S. G., Budden, D., Cabi, S., Denil, M., de Freitas, N., and Wang, Z. (2019) · 1910
Earlier work this paper cites.
On the measure of intelligence
Chollet, F. (2019) · 1911
Earlier work this paper cites.
Extending machine language models toward human-level language understanding
McClelland, J. L., Hill, F., Rudolph, M., Baldridge, J., and Schütze, H. (2019) · 1912
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Mnih, V., Badia, A. P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K. (2016) · 1937
Earlier work this paper cites.
Computing machinery and intelligence
Turing, A. M. (1950) · 1950
Earlier work this paper cites.
Prediction and entropy of printed English
Shannon, C. E. (1951) · 1951
Earlier work this paper cites.
Philosophische Untersuchungen, von Ludwig Wittgenstein.-Philosophical investigations, by Ludwig Wittgenstein. Translated by GEM Anscombe
Wittgenstein, L. (1953) · 1953
Earlier work this paper cites.
Chomsky, n. 1959. a review of bf skinner’s verbal behavior. language, 35 (1), 26–58
Chomsky, N. (1959) · 1959
Earlier work this paper cites.
Understanding natural language
Winograd, T. (1972) · 1972
Earlier work this paper cites.
The psychology of human-computer interaction
Card, S. K., Moran, T. P., and Newell, A. (1983) · 1983
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
Pomerleau, D. A. (1989) · 1989
Earlier work this paper cites.
The symbol grounding problem
Harnad, S. (1990) · 1990
Earlier work this paper cites.
Coevolution of neocortical size, group size and language in humans
Dunbar, R. I. (1993) · 1993
Earlier work this paper cites.
Social learning in animals: the roots of culture
Heyes, C. M. and Galef Jr, B. G. (1996) · 1996
Earlier work this paper cites.
Learning by imitation: A hierarchical approach
Byrne, R. W. and Russon, A. E. (1998) · 1998
Earlier work this paper cites.
Is imitation learning the route to humanoid robots?
Schaal, S. (1999) · 1999
Earlier work this paper cites.
Scaling laws for neural language models
Kaplan, J., McCandlish, S., Henighan, T., Brown, T. B., Chess, B., Child, R., Gray, S., Radford, A., Wu, J., and Amodei, D. (2020) · 2001
Earlier work this paper cites.
On the sample complexity of reinforcement learning
Kakade, S. M. et al. (2003) · 2003
Earlier work this paper cites.
Social learning strategies
Laland, K. N. (2004) · 2004
Earlier work this paper cites.
Language models are few-shot learners
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al. (2020) · 2005
Earlier work this paper cites.
Learning a similarity metric discriminatively, with application to face verification
Chopra, S., Hadsell, R., and LeCun, Y. (2005) · 2005
Earlier work this paper cites.
Lynch, C. and Sermanet, P. (2020) · 2005
Cited alongside, same era.
The human speechome project
Roy, D., Patel, R., DeCamp, P., Kubat, R., Fleischman, M., Roy, B., Mavridis, N., Tellex, S., Salata, A., Guinness, J., et al. (2006) · 2006
Cited alongside, same era.
Shifting viewpoints: Artificial intelligence and human–computer interaction
Winograd, T. (2006) · 2006
Cited alongside, same era.
Word meaning in minds and machines
Lake, B. M. and Murphy, G. L. (2020) · 2008
Cited alongside, same era.
What’s in view for toddlers? using a head camera to study visual experience
Yoshida, H. and Smith, L. B. (2008) · 2008
Cited alongside, same era.
Animal imitation
Mask r-cnn
He, K., Gkioxari, G., Dollár, P., and Girshick, R. (2017) · 2017
Later among the works it cites.
Grounded language learning in a simulated 3d world
Hermann, K. M., Hill, F., Green, S., Wang, F., Faulkner, R., Soyer, H., Szepesvari, D., Czarnecki, W. M., Jaderberg, M., Teplyashin, D., et al. (2017) · 2017
Later among the works it cites.
In-datacenter performance analysis of a tensor processing unit
Jouppi, N. P., Young, C., Patil, N., Patterson, D., Agrawal, G., Bajwa, R., Bates, S., Bhatia, S., Boden, N., Borchers, A., et al. (2017) · 2017
Later among the works it cites.
Infogail: Interpretable imitation learning from visual demonstrations
Li, Y., Song, J., and Ermon, S. (2017) · 2017
Later among the works it cites.
Learning human behaviors from motion capture by adversarial imitation
Merel, J., Tassa, Y., TB, D., Srinivasan, S., Lemmon, J., Wang, Z., Wayne, G., and Heess, N. (2017) · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Byrne, R. W. (2009) · 2009
Cited alongside, same era.
How intelligence happens
Duncan, J. (2010) · 2010
Cited alongside, same era.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Gutmann, M. and Hyvärinen, A. (2010) · 2010
Cited alongside, same era.
Origins of human communication
Tomasello, M. (2010) · 2010
Cited alongside, same era.
Modeling purposeful adaptive behavior with the principle of maximum causal entropy
Ziebart, B. D. (2010) · 2010
Cited alongside, same era.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G., and Bagnell, D. (2011) · 2011
Cited alongside, same era.
Understanding natural language commands for robotic navigation and mobile manipulation
Tellex, S. A., Kollar, T. F., Dickerson, S. R., Walter, M. R., Banerjee, A., Teller, S., and Roy, N. (2011) · 2011
Cited alongside, same era.
Third-person imitation learning
Stadie, B. C., Abbeel, P., and Sutskever, I. (2017) · 2017
Later among the works it cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I. (2017) · 2017
Later among the works it cites.
Playing hard exploration games by watching youtube
Aytar, Y., Pfaff, T., Budden, D., Paine, T., Wang, Z., and de Freitas, N. (2018) · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2018) · 2018
Later among the works it cites.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., et al. (2018) · 2018
Later among the works it cites.
An algorithmic perspective on imitation learning
Osa, T., Pajarinen, J., Neumann, G., Bagnell, J. A., Abbeel, P., and Peters, J. (2018) · 2018
Later among the works it cites.
Film: Visual reasoning with a general conditioning layer
Perez, E., Strub, F., de Vries, H., Dumoulin, V., and Courville, A. C. (2018) · 2018
Later among the works it cites.
Self-attention with relative position representations
Shaw, P., Uszkoreit, J., and Vaswani, A. (2018) · 2018
Later among the works it cites.
Representation learning with contrastive predictive coding
van den Oord, A., Li, Y., and Vinyals, O. (2018) · 2018
Later among the works it cites.
Hierarchical decision making by generating and following natural language instructions
Hu, H., Yarats, D., Gong, Q., Tian, Y., and Lewis, M. (2019) · 2019
Later among the works it cites.
Tsm: Temporal shift module for efficient video understanding
Lin, J., Gan, C., and Han, S. (2019) · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., and Sutskever, I. (2019) · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al. (2019) · 2019
Later among the works it cites.
Randaugment: Practical automated data augmentation with a reduced search space
Cubuk, E. D., Zoph, B., Shlens, J., and Le, Q. V. (2020) · 2020
Closest in time.
A divergence minimization perspective on imitation learning methods
Ghasemipour, S. K. S., Zemel, R., and Gu, S. (2020) · 2020
Closest in time.
Wiki-40b: Multilingual language model dataset
Guo, M., Dai, Z., Vrandecic, D., and Al-Rfou, R. (2020) · 2020
Closest in time.
Haiku: Sonnet for JAX
Hennigan, T., Cai, T., Norman, T., and Babuschkin, I. (2020) · 2020
Closest in time.
Saycam: A large, longitudinal audiovisual dataset recorded from the infant’s perspective
Sullivan, J., Mei, M., Perfors, A., Wojcik, E. H., and Frank, M. C. (2020) · 2020
Closest in time.
Embodied question answering
Das, A., Datta, S., Gkioxari, G., Lee, S., Parikh, D., and Batra, D. (2018) · 2063
Closest in time.