Fetching the paper…
Reading the bibliography…
At the heart of text based neural models lay word representations, which are powerful but occupy a lot of memory making it challenging to deploy to devices with memory constraints such as mobile phones, watches and IoT.
Multi-task deep neural networks for natural language understanding
Xiaodong Liu, Pengcheng He, Weizhu Chen, and Jianfeng Gao. 2019a · 1901
Earlier work this paper cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime G. Carbonell, Ruslan Salakhutdinov, and Quoc V. Le. 2019 · 1906
Earlier work this paper cites.
Massively multilingual neural machine translation in the wild: Findings and challenges
Naveen Arivazhagan, Ankur Bapna, Orhan Firat, Dmitry Lepikhin, Melvin Johnson, Maxim Krikun, Mia Xu Chen, Yuan Cao, George Foster, Colin Cherry, Wolfgang Macherey, Zhifeng Chen, and Yonghui Wu. 2019 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
Pruning a bert-based question answering model
J. Scott McCarley. 2019 · 1910
Earlier work this paper cites.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
The ICSI meeting recorder dialog act (MRDA) corpus
Elizabeth Shriberg, Rajdip Dhillon, Sonali Bhagat, Jeremy Ang, and Hannah Carvey. 2004 · 2004
Earlier work this paper cites.
Yukun Zhu, Ryan Kiros, Richard S. Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2004
Earlier work this paper cites.
What is left to be understood in atis?
Gökhan Tür, Dilek Hakkani-Tür, and Larry P. Heck. 2010 · 2010
Earlier work this paper cites.
Multi-domain joint semantic frame parsing using bi-directional rnn-lstm
Dilek Hakkani-Tur, Gokhan Tur, Asli Celikyilmaz, Yun-Nung Vivian Chen, Jianfeng Gao, Li Deng, and Ye-Yi Wang. 2016 · 2016
Cited alongside, same era.
Fasttext.zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski, Matthijs Douze, Hervé Jégou, and Tomas Mikolov. 2016 · 2016
Cited alongside, same era.
Sequential short-text classification with recurrent and convolutional neural networks
Ji Young Lee and Franck Dernoncourt. 2016 · 2016
Cited alongside, same era.
Attention-based recurrent neural network models for joint intent detection and slot filling
Bing Liu and Ian Lane. 2016 · 2016
Cited alongside, same era.
Neural-based context representation learning for dialog act classification
Daniel Ortega and Ngoc Thang Vu. 2017 · 2017
Cited alongside, same era.
Slot-gated modeling for joint slot filling and intent prediction
Chih-Wen Goo, Guang Gao, Yun-Kai Hsu, Chih-Li Huo, Tsung-Chieh Chen, Keng-Wei Hsu, and Yun-Nung Chen. 2018 · 2018
Later among the works it cites.
Self-governing neural networks for on-device short text classification
Sujith Ravi and Zornitsa Kozareva. 2018 · 2018
Later among the works it cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R. Bowman. 2018 · 2018
Later among the works it cites.
ProSeqo: Projection sequence networks for on-device text classification
Zornitsa Kozareva and Sujith Ravi. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sujith Ravi. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Neural graph learning: Training neural networks using graphs
Thang D. Bui, Sujith Ravi, and Vivek Ramavajjala. 2018 · 2018
Cited alongside, same era.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015a
Cited in the paper.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015b
Cited in the paper.
Efficient on-device models using neural projections
Sujith Ravi. 2019 · 2019
Later among the works it cites.
On-device structured and context partitioned projection networks
Sujith Ravi and Zornitsa Kozareva. 2019 · 2019
Later among the works it cites.
Transferable neural projection representations
Chinnadhurai Sankar, Sujith Ravi, and Zornitsa Kozareva. 2019 · 2019
Later among the works it cites.
Dialogue act classification in domain-independent conversations using a deep recurrent neural network
Hamed Khanpour, Nishitha Guntakandla, and Rodney Nielsen. 2016 · 2021
Closest in time.