Fetching the paper…
Reading the bibliography…
Despite the recent success of contextualized language models on various NLP tasks, language model itself cannot capture textual coherence of a long, multi-sentence document (e.g., a paragraph).
Bert has a mouth, and it must speak: Bert as a markov random field language model
Alex Wang and Kyunghyun Cho. 2019 · 1902
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Maxwell Forbes, and Yejin Choi. 2019 · 1904
Earlier work this paper cites.
A multiscale visualization of attention in the transformer model
Jesse Vig. 2019 · 1906
Earlier work this paper cites.
CTRL - A Conditional Transformer Language Model for Controllable Generation
Nitish Shirish Keskar, Bryan McCann, Lav Varshney, Caiming Xiong, and Richard Socher. 2019 · 1909
Earlier work this paper cites.
Select and attend: Towards controllable content selection in text generation
Xiaoyu Shen, Jun Suzuki, Kentaro Inui, Hui Su, Dietrich Klakow, and Satoshi Sekine. 2019 · 1909
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R’emi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
Inset: Sentence infilling with inter-sentential generative pre-training
Yichen Huang, Yizhe Zhang, Oussama Elachqar, and Yu Cheng. 2019 · 1911
Earlier work this paper cites.
Script theory: Differential magnification of affects
Silvan S Tomkins. 1978 · 1978
Earlier work this paper cites.
Teaching writing skills
Donn Byrne. 1979 · 1979
Earlier work this paper cites.
Planning natural-language utterances to satisfy multiple goals
Douglas E Appelt. 1982 · 1982
Earlier work this paper cites.
Discourse strategies for generating natural-language text
Kathleen R McKeown. 1985 · 1985
Earlier work this paper cites.
Pragmatics and natural language generation
Eduard H Hovy. 1990 · 1990
Earlier work this paper cites.
Approaches to the planning of coherent text
Eduard H Hovy. 1991 · 1991
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
The science of scientific writing
Judith A. Swan. 2002 · 2002
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Cited alongside, same era.
Modeling local coherence: An entity-based approach
Regina Barzilay and Mirella Lapata. 2008 · 2008
Cited alongside, same era.
Unsupervised learning of narrative event chains
Nathanael Chambers and Dan Jurafsky. 2008 · 2008
Cited alongside, same era.
Automatic keyword extraction from individual documents
Stuart Rose, Dave Engel, Nick Cramer, and Wendy Cowley. 2010 · 2010
Cited alongside, same era.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer. 2011 · 2011
Cited alongside, same era.
Scripts, plans, goals, and understanding: An inquiry into human knowledge structures
Roger C Schank and Robert P Abelson. 2013 · 2013
Detecting and explaining causes from text for a time series event
Dongyeop Kang, Varun Gangal, Ang Lu, Zheng Chen, and Eduard Hovy. 2017 · 2017
Later among the works it cites.
Get to the point: Summarization with pointer-generator networks
Abigail See, Peter J Liu, and Christopher D Manning. 2017 · 2017
Later among the works it cites.
A hierarchical latent variable encoder-decoder model for generating dialogues
Iulian Vlad Serban, Alessandro Sordoni, Ryan Lowe, Laurent Charlin, Joelle Pineau, Aaron C Courville, and Yoshua Bengio. 2017 · 2017
Later among the works it cites.
Maskgan: better text generation via filling in the_
William Fedus, Ian Goodfellow, and Andrew M Dai. 2018 · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
A hierarchical recurrent encoder-decoder for generative context-aware query suggestion
Alessandro Sordoni, Yoshua Bengio, Hossein Vahabi, Christina Lioma, Jakob Grue Simonsen, and Jian-Yun Nie. 2015 · 2015
Cited alongside, same era.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books
Yukun Zhu, Ryan Kiros, Richard S. Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Cited alongside, same era.
Smart reply: Automated response suggestion for email
Anjuli Kannan, Karol Kurach, Sujith Ravi, Tobias Kaufmann, Andrew Tomkins, Balint Miklos, Greg Corrado, László Lukács, Marina Ganea, Peter Young, et al. 2016 · 2016
Cited alongside, same era.
Chia-Wei Liu, Ryan Lowe, Iulian V Serban, Michael Noseworthy, Laurent Charlin, and Joelle Pineau. 2016 · 2016
Cited alongside, same era.
Later among the works it cites.
Strategies for structuring story generation
Angela Fan, Mike Lewis, and Yann Dauphin. 2019 · 2019
Later among the works it cites.
Paraphrase generation with latent bag of words
Yao Fu, Yansong Feng, and John P Cunningham. 2019 · 2019
Later among the works it cites.
Sentence-level content planning and style specification for neural text generation
Xinyu Hua and Lu Wang. 2019 · 2019
Later among the works it cites.
Linguistic versus latent relations for modeling coherent flow in paragraphs
Dongyeop Kang, Hiroaki Hayashi, Alan W Black, and Eduard Hovy. 2019 · 2019
Later among the works it cites.
Selecting, planning, and rewriting: A modular approach for data-to-document generation and translation
Lesly Miculicich, Marc Marone, and Hany Hassan. 2019 · 2019
Later among the works it cites.
Step-by-step: Separating planning from realization in neural data-to-text generation
Amit Moryossef, Yoav Goldberg, and Ido Dagan. 2019 · 2019
Later among the works it cites.
Data-to-text generation with content selection and planning
Ratish Puduppully, Li Dong, and Mirella Lapata. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Yake! keyword extraction from single documents using multiple local features
Ricardo Campos, Vítor Mangaravite, Arian Pasquali, Alípio Jorge, Célia Nunes, and Adam Jatowt. 2020 · 2020
Closest in time.
Linguistically Informed Language Generation: A Multifaceted Approach
Dongyeop Kang. 2020 · 2020
Closest in time.