Fetching the paper…
Reading the bibliography…
Human language learners are exposed to a trickle of informative, context-sensitive language, but a flood of raw sensory data.
Thought and language
Lev S Vygotsky · 1934
Earlier work this paper cites.
Short-term memory for word sequences as a function of acoustic, semantic and formal similarity
Alan D Baddeley · 1966
Earlier work this paper cites.
Speech-like coding of pictures in short-term memory
Diane J Schiano and Michael J Watkins · 1981
Earlier work this paper cites.
The development of verbal control over motor behavior: A replication and extension of luria’s findings
Virginia S Tinsley and Harriet Salatas Waters · 1982
Earlier work this paper cites.
The role of private speech in the transition from collaborative to independent task performance in young children
Adam Winsler, Rafael M Diaz, and Ignacio Montero · 1997
Earlier work this paper cites.
Private speech on an executive task: Relations with task difficulty and task performance
Charles Fernyhough and Emma Fradley · 2005
Earlier work this paper cites.
Private speech, executive functioning, and the development of verbal self-regulation
Adam Ed Winsler, Charles Ed Fernyhough, and Ignacio Ed Montero · 2009
Earlier work this paper cites.
Imitation improves language comprehension
Patti Adank, Peter Hagoort, and Harold Bekkering · 2010
Earlier work this paper cites.
Private and inner speech and the regulation of social speech communication
Conchi San Martín Martínez, Humbert Boada i Calbet, and Peter Feigenbaum · 2011
Earlier work this paper cites.
Self-talk and sports performance: A meta-analysis
Antonis Hatzigeorgiadis, Nikos Zourbanos, Evangelos Galanis, and Yiannis Theodorakis · 2011
Earlier work this paper cites.
Semi-supervised learning with deep generative models
Durk P Kingma, Shakir Mohamed, Danilo Jimenez Rezende, and Max Welling · 2014
Earlier work this paper cites.
Notes on noise contrastive estimation and negative sampling
Chris Dyer · 2014
Earlier work this paper cites.
Inner speech: development, cognitive functions, phenomenology, and neurobiology
Ben Alderson-Day and Charles Fernyhough · 2015
Earlier work this paper cites.
Cider: Consensus-based image description evaluation
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli · 2015
Earlier work this paper cites.
Unsupervised feature extraction by time-contrastive learning and nonlinear ica
Aapo Hyvarinen and Hiroshi Morioka · 2016
Earlier work this paper cites.
Language as a latent variable: Discrete generative models for sentence compression
Yishu Miao and Phil Blunsom · 2016
Earlier work this paper cites.
A semi-supervised framework for image captioning
Wenhu Chen, Aurelien Lucchi, and Thomas Hofmann · 2016
Earlier work this paper cites.
Generative adversarial text to image synthesis
Scott Reed, Zeynep Akata, Xinchen Yan, Lajanugen Logeswaran, Bernt Schiele, and Honglak Lee · 2016
Earlier work this paper cites.
The phonological loop as a language learning device
Alan D Baddeley, Susan E Gathercole, and Costanza Papagno · 2017
Earlier work this paper cites.
Equivalence between policy gradients and soft q-learning
John Schulman, Xi Chen, and Pieter Abbeel · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Cited alongside, same era.
Grounded language learning in a simulated 3d world
Karl Moritz Hermann, Felix Hill, Simon Green, Fumin Wang, Ryan Faulkner, Hubert Soyer, David Szepesvari, Wojciech Marian Czarnecki, Max Jaderberg, Denis Teplyashin, et al · 2017
Cited alongside, same era.
One-shot visual imitation learning via meta-learning
Chelsea Finn, Tianhe Yu, Tianhao Zhang, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
One-shot imitation learning
Yan Duan, Marcin Andrychowicz, Bradly Stadie, OpenAI Jonathan Ho, Jonas Schneider, Ilya Sutskever, Pieter Abbeel, and Wojciech Zaremba · 2017
Cited alongside, same era.
Semi-supervised image captioning via reconstruction
Bicheng Xu, Weirui Kong, and Jiaxuan Chen · 2017
Cited alongside, same era.
Reinforcement learning and control as probabilistic inference: Tutorial and review
Human instruction-following with deep reinforcement learning via transfer-learning from text
Felix Hill, Sona Mokra, Nathaniel Wong, and Tim Harley · 2020
Later among the works it cites.
Language conditioned imitation learning over unstructured data
Corey Lynch and Pierre Sermanet · 2020
Later among the works it cites.
A benchmark for systematic generalization in grounded language understanding
Laura Ruis, Jacob Andreas, Marco Baroni, Diane Bouchacourt, and Brenden M Lake · 2020
Later among the works it cites.
Angeliki Lazaridou, Anna Potapenko, and Olivier Tieleman · 2020
Later among the works it cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sergey Levine · 2018
Cited alongside, same era.
Representation learning with contrastive predictive coding
Aaron Van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Cited alongside, same era.
Taku Kudo and John Richardson · 2018
Cited alongside, same era.
Improved adam optimizer for deep neural networks
Zijun Zhang · 2018
Cited alongside, same era.
Babyai: A platform to study the sample efficiency of grounded language learning
Maxime Chevalier-Boisvert, Dzmitry Bahdanau, Salem Lahlou, Lucas Willems, Chitwan Saharia, Thien Huu Nguyen, and Yoshua Bengio · 2018
Cited alongside, same era.
Encoding spatial relations from natural language
Tiago Ramalho, Tomáš Kočiskỳ, Frederic Besse, SM Eslami, Gábor Melis, Fabio Viola, Phil Blunsom, and Karl Moritz Hermann · 2018
Cited alongside, same era.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
Tao Xu, Pengchuan Zhang, Qiuyuan Huang, Han Zhang, Zhe Gan, Xiaolei Huang, and Xiaodong He · 2018
Cited alongside, same era.
Later among the works it cites.
Improved techniques for training score-based generative models
Yang Song and Stefano Ermon · 2020
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Scaling up visual and vision-language representation learning with noisy text supervision
Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig · 2021
Later among the works it cites.
Creating multimodal interactive agents with imitation and self-supervised learning
DeepMind Interactive Agents Team · 2021
Later among the works it cites.
Scaling language models: Methods, analysis & insights from training gopher
Jack W Rae, Sebastian Borgeaud, Trevor Cai, Katie Millican, Jordan Hoffmann, Francis Song, John Aslanides, Sarah Henderson, Roman Ring, Susannah Young, et al · 2021
Later among the works it cites.
Combined scaling for zero-shot transfer learning
Hieu Pham, Zihang Dai, Golnaz Ghiasi, Hanxiao Liu, Adams Wei Yu, Minh-Thang Luong, Mingxing Tan, and Quoc V Le · 2021
Later among the works it cites.
Hierarchical few-shot imitation with skill transition models
Kourosh Hakhamaneshi, Ruihan Zhao, Albert Zhan, Pieter Abbeel, and Michael Laskin · 2021
Later among the works it cites.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
Later among the works it cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Alex Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob McGrew, Ilya Sutskever, and Mark Chen · 2021
Later among the works it cites.
Cross-modal contrastive learning for text-to-image generation
Han Zhang, Jing Yu Koh, Jason Baldridge, Honglak Lee, and Yinfei Yang · 2021
Later among the works it cites.
Improving text-to-image synthesis using contrastive learning
Hui Ye, Xiulong Yang, Martin Takac, Rajshekhar Sunderraman, and Shihao Ji · 2021
Later among the works it cites.
Cloud tensor processing units (tpus), 2022
Google Cloud · 2022
Closest in time.
Zero experience required: Plug & play modular transfer learning for semantic visual navigation
Ziad Al-Halah, Santhosh K Ramakrishnan, and Kristen Grauman · 2022
Closest in time.
Learning language-conditioned robot behavior from offline data and crowd-sourced annotation
Suraj Nair, Eric Mitchell, Kevin Chen, Silvio Savarese, Chelsea Finn, et al · 2022
Closest in time.
Bc-z: Zero-shot task generalization with robotic imitation learning
Eric Jang, Alex Irpan, Mohi Khansari, Daniel Kappler, Frederik Ebert, Corey Lynch, Sergey Levine, and Chelsea Finn · 2022
Closest in time.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen · 2022
Closest in time.