Fetching the paper…
Reading the bibliography…
Keeping track of how states of entities change as a text or dialog unfolds is a key prerequisite to discourse understanding.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D. Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 1901
Earlier work this paper cites.
Procedures as a representation for data in a computer program for understanding natural language
Terry Winograd. 1971 · 1971
Earlier work this paper cites.
Discourse referents
Lauri Karttunen. 1976 · 1976
Earlier work this paper cites.
Dynamic predicate logic
Jeroen Groenendijk and Martin Stokhof. 1991 · 1991
Earlier work this paper cites.
Word frequencies in written and spoken English: based on the British National Corpus
Geoffrey Leech, Paul. Rayson, and Andrew Wilson. 2001 · 2001
Earlier work this paper cites.
File change semantics and the familiarity theory of definiteness
Irene Heim. 2002 · 2002
Earlier work this paper cites.
ACE 2005 multilingual training corpus
Christopher Walker, Stephanie Strassel, Julie Medero, and Kazuaki Maeda. 2006 · 2005
Earlier work this paper cites.
When peanuts fall in love: N400 evidence for the power of discourse
Mante S. Nieuwland and Jos J. A. Van Berkum. 2006 · 2006
Earlier work this paper cites.
Coreference resolution in a modular, entity-centered model
Aria Haghighi and Dan Klein. 2010 · 2010
Earlier work this paper cites.
Discourse Representation Theory , pages 125–394. Springer Netherlands, Dordrecht
Hans Kamp, Josef Van Genabith, and Uwe Reyle. 2011 · 2011
Earlier work this paper cites.
Stanford’s multi-pass sieve coreference resolution system at the CoNLL-2011 shared task
Heeyoung Lee, Yves Peirsman, Angel Chang, Nathanael Chambers, Mihai Surdeanu, and Dan Jurafsky. 2011 · 2011
Earlier work this paper cites.
CoNLL-2012 shared task: Modeling multilingual unrestricted coreference in OntoNotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Olga Uryupina, and Yuchen Zhang. 2012 · 2012
Earlier work this paper cites.
Resolving complex cases of definite pronouns: The Winograd schema challenge
Altaf Rahman and Vincent Ng. 2012 · 2012
Earlier work this paper cites.
Mise en place: Unsupervised interpretation of instructional recipes
Chloé Kiddon, Ganesa Thandavam Ponnuraj, Luke Zettlemoyer, and Yejin Choi. 2015 · 2015
Earlier work this paper cites.
Towards AI-complete question answering: a set of prerequisite toy tasks
Jason Weston, Antoine Bordes, Sumit Chopra, Alexander M Rush, Bart Van Merriënboer, Armand Joulin, and Tomas Mikolov. 2015 · 2015
Earlier work this paper cites.
The Goldilocks principle: Reading children’s books with explicit memory representations
Felix Hill, Antoine Bordes, Sumit Chopra, and Jason Weston. 2016 · 2016
Earlier work this paper cites.
Simpler context-dependent logical forms via model projections
Reginald Long, Panupong Pasupat, and Percy Liang. 2016 · 2016
Earlier work this paper cites.
Tracking the world state with recurrent entity networks
Mikael Henaff, Jason Weston, Arthur Szlam, Antoine Bordes, and Yann LeCun. 2017 · 2017
Earlier work this paper cites.
Dynamic entity representations in neural language models
Yangfeng Ji, Chenhao Tan, Sebastian Martschat, Yejin Choi, and Noah A. Smith. 2017 · 2017
Earlier work this paper cites.
End-to-end neural coreference resolution
Kenton Lee, Luheng He, Mike Lewis, and Luke Zettlemoyer. 2017 · 2017
Cited alongside, same era.
Simulating action dynamics with neural process networks
Antoine Bosselut, Omer Levy, Ari Holtzman, Corin Ennis, Dieter Fox, and Yejin Choi. 2018 · 2018
Cited alongside, same era.
PreCo: A large-scale dataset in preschool vocabulary for coreference resolution
Hong Chen, Zhenhua Fan, Hao Lu, Alan Yuille, and Shu Rong. 2018 · 2018
Cited alongside, same era.
What do entity-centric models learn? insights from entity linking in multi-party dialogue
Laura Aina, Carina Silberer, Ionut-Teodor Sorodoc, Matthijs Westera, and Gemma Boleda. 2019 · 2019
Cited alongside, same era.
What does BERT look at? an analysis of BERT’s attention
Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D. Manning. 2019 · 2019
Cited alongside, same era.
Textworld: A learning environment for text-based games
Annotating a broad range of anaphoric phenomena, in a variety of genres: the arrau corpus
Olga Uryupina, Ron Artstein, Antonella Bristot, Federica Cavicchio, Francesca Delogu, Kepa J Rodriguez, and Massimo Poesio. 2020 · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Later among the works it cites.
Implicit representations of meaning in neural language models
Belinda Z. Li, Maxwell Nye, and Jacob Andreas. 2021 · 2021
Later among the works it cites.
Provable limitations of acquiring meaning from ungrounded form: What will future language models understand?
William Merrill, Yoav Goldberg, Roy Schwartz, and Noah A. Smith. 2021 · 2021
Later among the works it cites.
CorefQA: Coreference resolution as query-based span prediction
Wei Wu, Fei Wang, Arianna Yuan, Fei Wu, and Jiwei Li. 2020 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Marc-Alexandre Côté, Ákos Kádár, Xingdi Yuan, Ben Kybartas, Tavian Barnes, Emery Fine, James Moore, Matthew Hausknecht, Layla El Asri, Mahmoud Adada, Wendy Tay, and Adam Trischler. 2019 · 2019
Cited alongside, same era.
Everything happens for a reason: Discovering the purpose of actions in procedural text
Bhavana Dalvi, Niket Tandon, Antoine Bosselut, Wen-tau Yih, and Peter Clark. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Effective use of transformer networks for entity tracking
Aditya Gupta and Greg Durrett. 2019a · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Cited alongside, same era.
BERT rediscovers the classical NLP pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick. 2019 · 2019
Cited alongside, same era.
An annotated dataset of coreference in English literature
David Bamman, Olivia Lewke, and Anya Mansoor. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
How does GPT-3 perform on COVR-10 splits with in-context learning?
Ben Bogin. 2022 · 2022
Later among the works it cites.
Unobserved local structures make compositional generalization hard
Ben Bogin, Shivanshu Gupta, and Jonathan Berant. 2022 · 2022
Later among the works it cites.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Alex Castro-Ros, Marie Pellat, Kevin Robinson, Dasha Valter, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Zhao, Yanping Huang, Andrew Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei. 2022 · 2022
Later among the works it cites.
Andrew Kyle Lampinen. 2022 · 2022
Later among the works it cites.
New or old? exploring how pre-trained language models represent discourse entities
Sharid Loáiciga, Anne Beyer, and David Schlangen. 2022 · 2022
Later among the works it cites.
Language models of code are few-shot commonsense learners
Aman Madaan, Shuyan Zhou, Uri Alon, Yiming Yang, and Graham Neubig. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
Neural theory-of-mind? on the limits of social intelligence in large LMs
Maarten Sap, Ronan Le Bras, Daniel Fried, and Yejin Choi. 2022 · 2022
Later among the works it cites.
When a sentence does not introduce a discourse entity, transformer-based models still sometimes refer to it
Sebastian Schuster and Tal Linzen. 2022 · 2022
Later among the works it cites.
Chess as a testbed for language model state tracking
Shubham Toshniwal, Sam Wiseman, Karen Livescu, and Kevin Gimpel. 2022 · 2022
Later among the works it cites.
Emergent world representations: Exploring a sequence model trained on a synthetic task
Kenneth Li, Aspen K Hopkins, David Bau, Fernanda Viégas, Hanspeter Pfister, and Martin Wattenberg. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Can large language models play text games well? current state-of-the-art and open questions
Chen Feng Tsai, Xiaochen Zhou, Sierra S Liu, Jing Li, Mo Yu, and Hongyuan Mei. 2023 · 2023
Closest in time.