Fetching the paper…
Reading the bibliography…
Speech-based virtual assistants, such as Amazon Alexa, Google assistant, and Apple Siri, typically convert users' audio signals to text data through automatic speech recognition (ASR) and feed the text to downstream dialog models for natural language understanding and response generation.
EDA: easy data augmentation techniques for boosting performance on text classification tasks
Jason W. Wei and Kai Zou. 2019 · 1901
Earlier work this paper cites.
Low resource text classification with ulmfit and backtranslation
Sam Shleifer. 2019 · 1903
Earlier work this paper cites.
Spoken language intent detection using confusion2vec
Prashanth Gurunath Shivakumar, Mu Yang, and Panayiotis G. Georgiou. 2019 · 1904
Earlier work this paper cites.
Improving robustness of task oriented dialog systems
Arash Einolghozati, Sonal Gupta, Mrinal Mohit, and Rushin Shah. 2019 · 1911
Earlier work this paper cites.
Maryam Fazel-Zarandi, Longshaokan Wang, Aditya Tiwari, and Spyros Matsoukas. 2019 · 1911
Earlier work this paper cites.
Introduction to the CoNLL-2003 shared task: Language-independent named entity recognition
Erik F. Tjong Kim Sang and Fien De Meulder. 2003 · 2003
Earlier work this paper cites.
The fisher corpus: a resource for the next generations of speech-to-text
C. Cieri, D. Miller, and K. Walker. 2004 · 2004
Earlier work this paper cites.
Beyond asr 1-best: Using word confusion networks in spoken language understanding
Dilek Hakkani-Tür, Frédéric Béchet, Giuseppe Riccardic, and Gokhan Tur. 2006 · 2006
Earlier work this paper cites.
Error simulation for training statistical dialogue systems
Jost Schatzmann, Blaise Thomson, and Steve Young. 2007 · 2007
Earlier work this paper cites.
Speech and Language Processing (2nd Edition)
Daniel Jurafsky and James H. Martin. 2009 · 2009
Cited alongside, same era.
Data collection in a wizard-of-oz experiment
Verena Rieser and Oliver Lemon. 2011 · 2011
Cited alongside, same era.
The second dialog state tracking challenge
Matthew Henderson, Blaise Thomson, and Jason D. Williams. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D. Manning. 2014 · 2014
Cited alongside, same era.
Hyperopt: a python library for model selection and hyperparameter optimization
James Bergstra, Brent Komer, Chris Eliasmith, Dan Yamins, and David D Cox. 2015 · 2015
Cited alongside, same era.
Acoustic word embeddings for asr error detection
Sahar Ghannay, Yannick Estève, Nathalie Camelin, and Paul Deléglise. 2016 · 2016
Cited alongside, same era.
From audio to semantics: Approaches to end-to-end spoken language understanding
Parisa Haghani, Arun Narayanan, Michiel Bacchiani, Galen Chuang, Neeraj Gaur, Pedro J. Moreno, Rohit Prabhavalkar, Zhongdi Qu, and Austin Waters. 2018 · 2018
Later among the works it cites.
Incorporating asr errors with attention-based, jointly trained rnn for intent detection and slot filling
R. Schumann and P. Angkititrakul. 2018 · 2018
Later among the works it cites.
Towards end-to-end spoken language understanding
Dmitriy Serdyuk, Yongqiang Wang, Christian Fuegen, Anuj Kumar, Baiyang Liu, and Yoshua Bengio. 2018 · 2018
Later among the works it cites.
Confusion2vec: Towards enriching vector space word representations with representational ambiguities
Prashanth Gurunath Shivakumar and Panayiotis G. Georgiou. 2018 · 2018
Later among the works it cites.
Neural error corrective language models for automatic speech recognition
Tomohiro Tanaka, Ryo Masumura, Hirokazu Masataki, and Yushi Aono. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Latticernn: Recurrent neural networks over lattices
Faisal Ladhak, Ankur Gandhe, Markus Dreyer, Lambert Mathias, Ariya Rastrow, and Björn Hoffmeister. 2016 · 2016
Cited alongside, same era.
End-to-end sequence labeling via bi-directional lstm-cnns-crf
Xuezhe Ma and Eduard H. Hovy. 2016 · 2016
Cited alongside, same era.
Contextual string embeddings for sequence labeling
Alan Akbik, Duncan Blythe, and Roland Vollgraf. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
A survey on image data augmentation for deep learning
Connor Shorten and Taghi M. Khoshgoftaar. 2019 · 2019
Later among the works it cites.
Joint contextual modeling for asr correction and language understanding
Yue Weng, Sai Sumanth Miryala, Chandra Khatri, Runze Wang, Huaixiu Zheng, Piero Molino, Mahdi Namazifar, Alexandros Papangelis, Hugh Williams, Franziska Bell, and Gokhan Tur. 2020 · 2019
Later among the works it cites.