Fetching the paper…
Reading the bibliography…
We propose a novel task, G4C, to study teacher-student natural language interactions in a goal-driven and grounded environment.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, et al. 2019 · 1910
Earlier work this paper cites.
Logic and conversation
Herbert P Grice. 1975 · 1975
Earlier work this paper cites.
Linguistic communication as action and cooperation
Jens Allwood. 1976 · 1976
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff. 1978 · 1978
Earlier work this paper cites.
Contributing to discourse
Herbert H Clark and Edward F Schaefer. 1989 · 1989
Earlier work this paper cites.
Exploration of the autistic child’s theory of mind: Knowledge, belief, and communication
Josef Perner, Uta Frith, Alan M Leslie, and Susan R Leekam. 1989 · 1989
Earlier work this paper cites.
Grounding in communication
Herbert H Clark and Susan E Brennan. 1991 · 1991
Earlier work this paper cites.
Communicative competence and theory of mind in autism: A test of relevance theory
Francesca GE Happé. 1993 · 1993
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Shrimai Prabhumoye, Margaret Li, Jack Urbanek, Emily Dinan, Douwe Kiela, Jason Weston, and Arthur Szlam. 2020 · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Predicting pragmatic reasoning in language games
Michael C Frank and Noah D Goodman. 2012 · 2012
Earlier work this paper cites.
A rational account of pedagogical reasoning: Teaching by, and learning from, examples
Patrick Shafto, Noah D Goodman, and Thomas L Griffiths. 2014 · 2014
Earlier work this paper cites.
Pragmatic language interpretation as probabilistic inference
Noah D Goodman and Michael C Frank. 2016 · 2016
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Earlier work this paper cites.
Deep dungeons and dragons: Learning character-action interactions from role-playing game transcripts
Annie Louis and Charles Sutton. 2018 · 2018
Cited alongside, same era.
Dungeons and dqns: Toward reinforcement learning agents that play tabletop roleplaying games
Lara J Martin, Srijan Sood, and Mark O Riedl. 2018 · 2018
Cited alongside, same era.
Evaluating theory of mind in question answering
Aida Nematzadeh, Kaylee Burns, Erin Grant, Alison Gopnik, and Tom Griffiths. 2018 · 2018
Cited alongside, same era.
Quality signals in generated stories
Manasvi Sagarkar, John Wieting, Lifu Tu, and Kevin Gimpel. 2018 · 2018
Cited alongside, same era.
Revisiting the evaluation of theory of mind through question answering
Matthew Le, Y-Lan Boureau, and Maximilian Nickel. 2019 · 2019
Cited alongside, same era.
Collaborative dialogue in minecraft
Anjali Narayan-Chen, Prashant Jayannavar, and Julia Hockenmaier. 2019 · 2019
Decoding methods for neural narrative generation
Alexandra DeLucia, Aaron Mueller, Xiang Lisa Li, and João Sedoc. 2021 · 2021
Later among the works it cites.
Telling stories through multi-user dialogue by modeling character relations
Wai Man Si, Prithviraj Ammanabrolu, and Mark Riedl. 2021 · 2021
Later among the works it cites.
Situated dialogue learning through procedural environment generation
Prithviraj Ammanabrolu, Renee Jia, and Mark O Riedl. 2022 · 2022
Closest in time.
Video pretraining (vpt): Learning to act by watching unlabeled online videos
Bowen Baker, Ilge Akkaya, Peter Zhokhov, Joost Huizinga, Jie Tang, Adrien Ecoffet, Brandon Houghton, Raul Sampedro, and Jeff Clune. 2022 · 2022
Closest in time.
Human-level play in the game of diplomacy by combining language models with strategic reasoning
Anton Bakhtin, Noam Brown, Emily Dinan, Gabriele Farina, Colin Flaherty, Daniel Fried, Andrew Goff, Jonathan Gray, Hengyuan Hu, Athul Paul Jacob, Mojtaba Komeili, Karthik Konath, Minae Kwon, Adam Lerer, Mike Lewis, Alexander H. Miller, Sasha Mitts, Adithya Renduchintala, Stephen Roller, Dirk Rowe, Weiyan Shi, Joe Spisak, Alexander Wei, David Wu, Hugh Zhang, and Markus Zijlstra. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Learning to speak and act in a fantasy text adventure game
Jack Urbanek, Angela Fan, Siddharth Karamcheti, Saachi Jain, Samuel Humeau, Emily Dinan, Tim Rocktäschel, Douwe Kiela, Arthur Szlam, and Jason Weston. 2019 · 2019
Cited alongside, same era.
D4rl: Datasets for deep data-driven reinforcement learning
Justin Fu, Aviral Kumar, Ofir Nachum, George Tucker, and Sergey Levine. 2020 · 2020
Cited alongside, same era.
Program synthesis with pragmatic communication
Yewen Pu, Kevin Ellis, Marta Kryven, Josh Tenenbaum, and Armando Solar-Lezama. 2020 · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, Peter J Liu, et al. 2020 · 2020
Cited alongside, same era.
Learning knowledge graph-based world models of textual environments
Prithviraj Ammanabrolu and Mark Riedl. 2021 · 2021
Cited alongside, same era.
Closest in time.
Dungeons and dragons as a dialog challenge for artificial intelligence
Chris Callison-Burch, Gaurav Singh Tomar, Lara Martin, Daphne Ippolito, Suma Bailis, and David Reitter. 2022 · 2022
Closest in time.
Pragmatics in grounded language learning: Phenomena, tasks, and modeling approaches
Daniel Fried, Nicholas Tomlin, Jennifer Hu, Roma Patel, and Aida Nematzadeh. 2022 · 2022
Closest in time.
Piotr Mirowski, Kory W Mathewson, Jaylen Pittman, and Richard Evans. 2022 · 2022
Closest in time.
Generating descriptive and rules-adhering spells for dungeons & dragons fifth edition
Pax Newman and Yudong Liu. 2022 · 2022
Closest in time.
Chatgpt: Optimizing language models for dialogue
OpenAI. 2022 · 2022
Closest in time.
Rajkumar Ramamurthy, Prithviraj Ammanabrolu, Kianté Brantley, Jack Hessel, Rafet Sifa, Christian Bauckhage, Hannaneh Hajishirzi, and Yejin Choi. 2022 · 2022
Closest in time.
Neural theory-of-mind? on the limits of social intelligence in large lms
Maarten Sap, Ronan LeBras, Daniel Fried, and Yejin Choi. 2022 · 2022
Closest in time.
Reflect not reflex: Inference-based common ground improves dialogue response quality
Pei Zhou, Hyundong J. Cho, Pegah Jandaghi, Dong-Ho Lee, Bill Yuchen Lin, Jay Pujara, and Xiang Ren. 2022 · 2022
Closest in time.
Few-shot language coordination by modeling theory of mind
Hao Zhu, Graham Neubig, and Yonatan Bisk. 2021 · 2022
Closest in time.
Teach: Task-driven embodied agents that chat
Aishwarya Padmakumar, Jesse Thomason, Ayush Shrivastava, Patrick Lange, Anjali Narayan-Chen, Spandana Gella, Robinson Piramuthu, Gokhan Tur, and Dilek Hakkani-Tur. 2022 · 2025
Closest in time.