Fetching the paper…
Reading the bibliography…
Human intelligence has the remarkable ability to adapt to new tasks and environments quickly.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
Procedures as a representation for data in a computer program for understanding natural language
Terry Winograd · 1971
Earlier work this paper cites.
The lunar sciences natural language information system: Final report
W. A. Woods, Ronald M Kaplan, and Bonnie L. Webber · 1972
Earlier work this paper cites.
Seven steps to rendezvous with the casual user
Edgar F Codd · 1974
Earlier work this paper cites.
Developing a natural language interface to complex data
Gary G Hendrix, Earl D Sacerdoti, Daniel Sagalowicz, and Jonathan Slocum · 1978
Earlier work this paper cites.
Imitation, Objects, Tools, and the Rudiments of Language in Human Ontogeny, February 1988
Meltzoff An · 1988
Earlier work this paper cites.
The HCRC Map Task corpus: natural dialogue for speech recognition
Henry S. Thompson, Anne Anderson, Ellen Gurman Bard, Gwyneth Doherty-Sneddon, Alison Newlands, and Cathy Sotillo · 1993
Earlier work this paper cites.
How People Learn: Brain, Mind, Experience, and School: Expanded Edition
National Research Council · 1999
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
Natural pedagogy
Gergely Csibra and György Gergely · 2009
Earlier work this paper cites.
Robocup@ home: Demonstrating everyday manipulation skills in robocup@ home
Jorg Stuckler, Dirk Holz, and Sven Behnke · 2012
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Natural language communication with robots
Yonatan Bisk, Deniz Yuret, and Daniel Marcu · 2016
Earlier work this paper cites.
Chia-Wei Liu, Ryan Lowe, Iulian V Serban, Michael Noseworthy, Laurent Charlin, and Joelle Pineau · 2016
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Iterative policy learning in end-to-end trainable task-oriented neural dialog models
Bing Liu and Ian Lane · 2017
Earlier work this paper cites.
Towards problem solving agents that communicate and learn
Anjali Narayan-Chen, Colin Graber, Mayukh Das, Md Rakibul Islam, Soham Dan, Sriraam Natarajan, Janardhan Rao Doppa, Julia Hockenmaier, Martha Palmer, and Dan Roth · 2017
Cited alongside, same era.
Hierarchical and interpretable skill acquisition in multi-task reinforcement learning
Tianmin Shu, Caiming Xiong, and Richard Socher · 2017
Cited alongside, same era.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Rainbow: Combining improvements in deep reinforcement learning
Matteo Hessel, Joseph Modayil, Hado Van Hasselt, Tom Schaul, Georg Ostrovski, Will Dabney, Dan Horgan, Bilal Piot, Mohammad Azar, and David Silver · 2018
Cited alongside, same era.
Convai3: Generating clarifying questions for open-domain dialogue systems (clariq)
Mohammad Aliannejadi, Julia Kiseleva, Aleksandr Chuklin, Jeff Dalton, and Mikhail Burtsev · 2020
Later among the works it cites.
Unleashing the power of AI for education
Dan Ayoub · 2020
Later among the works it cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Later among the works it cites.
Ask your humans: Using human instructions to improve generalization in reinforcement learning
Valerie Chen, Abhinav Gupta, and Kenneth Marino · 2020
Later among the works it cites.
Electra: Pre-training text encoders as discriminators rather than generators
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adversarial learning of task-oriented neural dialog models
Bing Liu and Ian Lane · 2018
Cited alongside, same era.
Sudha Rao and Hal Daumé III · 2018
Cited alongside, same era.
Babyai: First steps towards grounded language learning with a human in the loop
Maxime Chevalier-Boisvert, Dzmitry Bahdanau, Salem Lahlou, Lucas Willems, Chitwan Saharia, Thien Huu Nguyen, and Yoshua Bengio · 2019
Cited alongside, same era.
Self-educated language agent with hindsight experience replay for instruction following
Geoffrey Cideron, Mathieu Seurin, Florian Strub, and Olivier Pietquin · 2019
Cited alongside, same era.
Hierarchical decision making by generating and following natural language instructions
Hengyuan Hu, Denis Yarats, Qucheng Gong, Yuandong Tian, and Mike Lewis · 2019
Cited alongside, same era.
Tom m. mitchell, simon garrod, john e. laird, stephen c. levinson, and kenneth r. koedinger
Stephen C Levinson · 2019
Cited alongside, same era.
Collaborative dialogue in minecraft
Anjali Narayan-Chen, Prashant Jayannavar, and Julia Hockenmaier · 2019
Cited alongside, same era.
Kevin Clark, Minh-Thang Luong, Quoc V Le, and Christopher D Manning · 2020
Later among the works it cites.
The second conversational intelligence challenge (convai2)
Emily Dinan, Varvara Logacheva, Valentin Malykh, Alexander Miller, Kurt Shuster, Jack Urbanek, Douwe Kiela, Arthur Szlam, Iulian Serban, Ryan Lowe, et al · 2020
Later among the works it cites.
Speak to your parser: Interactive text-to-SQL with natural language feedback
Ahmed Elgohary, Saghar Hosseini, and Ahmed Hassan Awadallah · 2020
Later among the works it cites.
Learning to execute instructions in a Minecraft dialogue
Prashant Jayannavar, Anjali Narayan-Chen, and Julia Hockenmaier · 2020
Later among the works it cites.
Interactive task learning from GUI-grounded natural language instructions and demonstrations
Toby Jia-Jun Li, Tom Mitchell, and Brad Myers · 2020
Later among the works it cites.
Sample factory: Egocentric 3d control from pixels at 100000 fps with asynchronous reinforcement learning
Aleksei Petrenko, Zhehui Huang, Tushar Kumar, Gaurav Sukhatme, and Vladlen Koltun · 2020
Later among the works it cites.
Recipes for building an open-domain chatbot
Stephen Roller, Emily Dinan, Naman Goyal, Da Ju, Mary Williamson, Yinhan Liu, Jing Xu, Myle Ott, Kurt Shuster, Eric M Smith, et al · 2020
Later among the works it cites.
Building and evaluating open-domain dialogue corpora with clarifying questions
Mohammad Aliannejadi, Julia Kiseleva, Aleksandr Chuklin, Jeff Dalton, and Mikhail Burtsev · 2021
Later among the works it cites.
Julia Kiseleva, Ziming Li, Mohammad Aliannejadi, Shrestha Mohanty, Maartje ter Hoeve, Mikhail Burtsev, Alexey Skrynnik, Artem Zholus, Aleksandr Panov, Kavya Srinet, Arthur Szlam, Yuxuan Sun, Katja Hofmann, Michel Galley, and Ahmed Awadallah · 2021
Later among the works it cites.
Minerl diamond 2021 competition: Overview, results, and lessons learned, 2022
Anssi Kanervisto, Stephanie Milani, Karolis Ramanauskas, Nicholay Topin, Zichuan Lin, Junyou Li, Jianing Shi, Deheng Ye, Qiang Fu, Wei Yang, Weijun Hong, Zhongyue Huang, Haicheng Chen, Guangjun Zeng, Yue Lin, Vincent Micheli, Eloi Alonso, François Fleuret, Alexander Nikulin, Yury Belousov, Oleg Svidchenko, and Aleksei Shpilman · 2022
Closest in time.
Interactive grounded language understanding in a collaborative environment: Iglu 2021
Julia Kiseleva, Ziming Li, Mohammad Aliannejadi, Shrestha Mohanty, Maartje ter Hoeve, Mikhail Burtsev, Alexey Skrynnik, Artem Zholus, Aleksandr Panov, Kavya Srinet, et al · 2022
Closest in time.
Improving intrinsic exploration with language abstractions, 2022
Jesse Mu, Victor Zhong, Roberta Raileanu, Minqi Jiang, Noah Goodman, Tim Rocktäschel, and Edward Grefenstette · 2022
Closest in time.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Closest in time.