Fetching the paper…
Reading the bibliography…
Recent advance in deep learning has led to the rapid adoption of machine learning-based NLP models in a wide range of applications.
Improving semantic parsing for task oriented dialog
Arash Einolghozati, Panupong Pasupat, Sonal Gupta, Rushin Shah, Mrinal Mohit, Mike Lewis, and Luke Zettlemoyer. 2019 · 1902
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Julian Salazar, Davis Liang, Toan Q Nguyen, and Katrin Kirchhoff. 2019 · 1910
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Learn++: An incremental learning algorithm for supervised neural networks
Robi Polikar, Lalita Upda, Satish S Upda, and Vasant Honavar. 2001 · 2001
Earlier work this paper cites.
Discriminative reranking for machine translation
Libin Shen, Anoop Sarkar, and Franz Josef Och. 2004 · 2004
Earlier work this paper cites.
Discriminative reranking for natural language parsing
Michael Collins and Terry Koo. 2005 · 2005
Earlier work this paper cites.
Model compression
Cristian Buciluǎ, Rich Caruana, and Alexandru Niculescu-Mizil. 2006 · 2006
Earlier work this paper cites.
Dependency parsing
Sandra Kübler, Ryan McDonald, and Joakim Nivre. 2009 · 2009
Earlier work this paper cites.
Class-incremental learning: survey and performance evaluation
Marc Masana, Xialei Liu, Bartlomiej Twardowski, Mikel Menta, Andrew D. Bagdanov, and J. Weijer. 2020 · 2010
Earlier work this paper cites.
Linguistic structure prediction
Noah A Smith. 2011 · 2011
Earlier work this paper cites.
Parsing with compositional vector grammars
Richard Socher, John Bauer, Christopher D. Manning, and Andrew Y. Ng. 2013 · 2013
Earlier work this paper cites.
Do deep nets really need to be deep?
Lei Jimmy Ba and Rich Caruana. 2014 · 2014
Earlier work this paper cites.
The inside-outside recursive neural network model for dependency parsing
Phong Le and Willem Zuidema. 2014 · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015 · 2015
Earlier work this paper cites.
Incremental learning algorithms and applications
Alexander Gepperth and Barbara Hammer. 2016 · 2016
Cited alongside, same era.
Sequence-level knowledge distillation
Yoon Kim and Alexander M. Rush. 2016 · 2016
Cited alongside, same era.
A diversity-promoting objective function for neural conversation models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan. 2016 · 2016
Cited alongside, same era.
Deep biaffine attention for neural dependency parsing
Timothy Dozat and Christopher D. Manning. 2017 · 2017
Cited alongside, same era.
Lifelong machine learning
Zhiyuan Chen and Bing Liu. 2018 · 2018
Cited alongside, same era.
Hierarchical neural story generation
Angela Fan, Mike Lewis, and Yann Dauphin. 2018 · 2018
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Later among the works it cites.
Simple and effective noisy channel modeling for neural machine translation
Kyra Yee, Yann Dauphin, and Michael Auli. 2019 · 2019
Later among the works it cites.
Reranking for neural semantic parsing
Pengcheng Yin and Graham Neubig. 2019 · 2019
Later among the works it cites.
Continual lifelong learning in natural language processing: A survey
Magdalena Biesialska, Katarzyna Biesialska, and Marta R. Costa-jussà. 2020 · 2020
Later among the works it cites.
Neural reranking for dependency parsing: An evaluation
Bich-Ngoc Do and Ines Rehbein. 2020 · 2020
Later among the works it cites.
Stanza: A python natural language processing toolkit for many human languages
Peng Qi, Yuhao Zhang, Yuhui Zhang, Jason Bolton, and Christopher D. Manning. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Semantic parsing for task oriented dialog using hierarchical representations
Sonal Gupta, Rushin Shah, Mrinal Mohit, Anuj Kumar, and Mike Lewis. 2018 · 2018
Cited alongside, same era.
Stack-pointer networks for dependency parsing
Xuezhe Ma, Zecong Hu, Jingzhou Liu, Nanyun Peng, Graham Neubig, and Eduard Hovy. 2018 · 2018
Cited alongside, same era.
Ranking generated summaries by correctness: An interesting but challenging application for natural language inference
Tobias Falke, Leonardo F. R. Ribeiro, Prasetya Ajie Utama, Ido Dagan, and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Fairness-aware ranking in search & recommendation systems with application to linkedin talent search
Sahin Cem Geyik, Stuart Ambler, and Krishnaram Kenthapadi. 2019 · 2019
Cited alongside, same era.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2019 · 2019
Cited alongside, same era.
Complexity-weighted loss and diverse reranking for sentence simplification
Reno Kriz, João Sedoc, Marianna Apidianaki, Carolina Zheng, Gaurav Kumar, Eleni Miltsakaki, and Chris Callison-Burch. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Don’t parse, generate! a sequence to sequence architecture for task-oriented semantic parsing
Subendhu Rongali, Luca Soldaini, Emilio Monti, and Wael Hamza. 2020 · 2020
Later among the works it cites.
Towards backward-compatible representation learning
Yantao Shen, Yuanjun Xiong, Wei Xia, and Stefano Soatto. 2020 · 2020
Later among the works it cites.
Please mind the root: Decoding arborescences for dependency parsing
Ran Zmigrod, Tim Vieira, and Ryan Cotterell. 2020 · 2020
Later among the works it cites.
Backward-compatible prediction updates: A probabilistic approach
Frederik Träuble, Julius von Kügelgen, Matthäus Kleindessner, Francesco Locatello, Bernhard Schölkopf, and Peter V. Gehler. 2021 · 2021
Later among the works it cites.
Structural knowledge distillation: Tractably distilling information for structured predictor
Xinyu Wang, Yong Jiang, Zhaohui Yan, Zixia Jia, Nguyen Bach, Tao Wang, Zhongqiang Huang, Fei Huang, and Kewei Tu. 2021 · 2021
Later among the works it cites.
Regression bugs are in your model! measuring, reducing and analyzing regressions in NLP model updates
Yuqing Xie, Yi-An Lai, Yuanjun Xiong, Yi Zhang, and Stefano Soatto. 2021 · 2021
Later among the works it cites.
Positive-congruent training: Towards regression-free model updates
Sijie Yan, Yuanjun Xiong, Kaustav Kundu, Shuo Yang, Siqi Deng, Meng Wang, Wei Xia, and Stefano Soatto. 2021 · 2021
Later among the works it cites.
A universal part-of-speech tagset
Slav Petrov, Dipanjan Das, and Ryan McDonald. 2012 · 2096
Closest in time.