Fetching the paper…
Reading the bibliography…
Large-scale pre-trained language models have shown impressive results on language understanding benchmarks like GLUE and SuperGLUE, improving considerably over other pre-training methods like distributed representations (GloVe) and purely supervised approaches.
BERT for joint intent classification and slot filling
Qian Chen, Zhu Zhuo, and Wen Wang. 2019 · 1902
Earlier work this paper cites.
Benchmarking natural language understanding services for building conversational agents
Xingkun Liu, Arash Eshghi, Pawel Swietojanski, and Verena Rieser. 2019b · 1903
Earlier work this paper cites.
To tune or not to tune? adapting pretrained representations to diverse tasks
Matthew E. Peters, Sebastian Ruder, and Noah A. Smith. 2019 · 1903
Earlier work this paper cites.
Docbert: BERT for document classification
Ashutosh Adhikari, Achyudh Ram, Raphael Tang, and Jimmy Lin. 2019 · 1904
Earlier work this paper cites.
A repository of conversational datasets
Matthew Henderson, Pawel Budzianowski, Iñigo Casanueva, Sam Coope, Daniela Gerz, Girish Kumar, Nikola Mrksic, Georgios Spithourakis, Pei-Hao Su, Ivan Vulic, and Tsung-Hsien Wen. 2019a · 1904
Earlier work this paper cites.
Xiaodong Liu, Pengcheng He, Weizhu Chen, and Jianfeng Gao. 2019a · 1904
Earlier work this paper cites.
Videobert: A joint model for video and language representation learning
Chen Sun, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid. 2019a · 1904
Earlier work this paper cites.
Attention is (not) all you need for commonsense reasoning
Tassilo Klein and Moin Nabi. 2019 · 1905
Earlier work this paper cites.
How to fine-tune BERT for text classification?
Chi Sun, Xipeng Qiu, Yige Xu, and Xuanjing Huang. 2019b · 1905
Earlier work this paper cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, Nikita Nangia, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R. Bowman. 2019 · 1905
Earlier work this paper cites.
ERNIE: enhanced language representation with informative entities
Zhengyan Zhang, Xu Han, Zhiyuan Liu, Xin Jiang, Maosong Sun, and Qun Liu. 2019 · 1905
Earlier work this paper cites.
Training neural response selection for task-oriented dialogue systems
Matthew Henderson, Ivan Vulić, Daniela Gerz, Iñigo Casanueva, Paweł Budzianowski, Sam Coope, Georgios Spithourakis, Tsung-Hsien Wen, Nikola Mrkšić, and Pei-Hao Su. 2019c · 1906
Earlier work this paper cites.
Patentbert: Patent classification with fine-tuning a pre-trained BERT model
Jieh-Sheng Lee and Jieh Hsiang. 2019 · 1906
Earlier work this paper cites.
Energy and policy considerations for deep learning in nlp
Emma Strubell, Ananya Ganesh, and Andrew McCallum. 2019 · 1906
Earlier work this paper cites.
A novel bi-directional interrelated model for joint intent detection and slot filling
Ee Haihong, Peiqing Niu, Zhongfu Chen, and Meina Song. 2019 · 1907
Earlier work this paper cites.
Structured fusion networks for dialog
Shikib Mehri, Tejas Srinivasan, and Maxine Eskenazi. 2019 · 1907
Earlier work this paper cites.
Vladimir Vlasov, Johannes EM Mosig, and Alan Nichol. 2019 · 1910
Cited alongside, same era.
Convert: Efficient and accurate conversational representations from transformers
Matthew Henderson, Iñigo Casanueva, Nikola Mrkšić, Pei-Hao Su, Ivan Vulić, et al. 2019b · 1911
Cited alongside, same era.
“cloze procedure”: A new tool for measuring readability
Wilson L Taylor. 1953 · 1953
Cited alongside, same era.
The ATIS spoken language systems pilot corpus
Charles T. Hemphill, John J. Godfrey, and George R. Doddington. 1990 · 1990
Cited alongside, same era.
Text chunking using transformation-based learning
Lance A. Ramshaw and Mitchell P. Marcus. 1995 · 1995
Cited alongside, same era.
The berkeley framenet project
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Jason D Williams, Kavosh Asadi, and Geoffrey Zweig. 2017 · 2017
Later among the works it cites.
Starspace: Embed all the things!
Ledell Wu, Adam Fisch, Sumit Chopra, Keith Adams, Antoine Bordes, and Jason Weston. 2017 · 2017
Later among the works it cites.
Alice Coucke, Alaa Saade, Adrien Ball, Théodore Bluche, Alexandre Caulier, David Leroy, Clément Doumouro, Thibault Gisselbrecht, Francesco Caltagirone, Thibaut Lavril, Maël Primet, and Joseph Dureau. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Collin F. Baker, Charles J. Fillmore, and John B. Lowe. 1998 · 1998
Cited alongside, same era.
Conditional random fields: Probabilistic models for segmenting and labeling sequence data
John Lafferty, Andrew McCallum, and Fernando CN Pereira. 2001 · 2001
Cited alongside, same era.
The class imbalance problem: A systematic study
Nathalie Japkowicz and Shaju Stephen. 2002 · 2002
Cited alongside, same era.
Partially observable markov decision processes for spoken dialog systems
Jason D Williams and Steve Young. 2007 · 2007
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2014 · 2015
Cited alongside, same era.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Gregory S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian J. Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Józefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dan Mané, Rajat Monga, Sherry Moore, Derek Gordon Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul A. Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda B. Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng. 2016 · 2016
Cited alongside, same era.
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Later among the works it cites.
Slot-gated modeling for joint slot filling and intent prediction
Chih-Wen Goo, Guang Gao, Yun-Kai Hsu, Chih-Li Huo, Tsung-Chieh Chen, Keng-Wei Hsu, and Yun-Nung Chen. 2018 · 2018
Later among the works it cites.
Fine-tuned language models for text classification
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Later among the works it cites.
Facebook messenger passes 300,000 bots
Khari Johnson. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
Improving language understanding by generative pre-training
Alec Radford. 2018 · 2018
Later among the works it cites.
Self-attention with relative position representations
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani. 2018 · 2018
Later among the works it cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R. Bowman. 2018 · 2018
Later among the works it cites.
Classification-reconstruction learning for open-set recognition
Ryota Yoshihashi, Wen Shao, Rei Kawakami, Shaodi You, Makoto Iida, and Takeshi Naemura. 2018 · 2018
Later among the works it cites.
Hierarchical multi-task natural language understanding for cross-domain conversational ai: Hermit nlu
Andrea Vanzo, Emanuele Bastianelli, and Oliver Lemon. 2019 · 2019
Later among the works it cites.
Efficient intent detection with dual sentence encoders
Iñigo Casanueva, Tadas Temčinas, Daniela Gerz, Matthew Henderson, and Ivan Vulić. 2020 · 2020
Closest in time.
Bidirectional lstm joint model for intent classification and named entity recognition in natural language understanding
Akson Sam Varghese, Saleha Sarang, Vipul Yadav, Bharat Karotra, and Niketa Gandhi. 2020 · 2020
Closest in time.