Fetching the paper…
Reading the bibliography…
Spoken language understanding (SLU) tasks involve mapping from speech audio signals to semantic labels.
Probability of error of some adaptive pattern-recognition machines
Henry Scudder. 1965 · 1965
Earlier work this paper cites.
Unsupervised word sense disambiguation rivaling supervised methods
David Yarowsky. 1995 · 1995
Earlier work this paper cites.
Automatically generating extraction patterns from untagged text
Ellen Riloff. 1996 · 1996
Earlier work this paper cites.
A rule-based named entity recognition system for speech input
Ji-Hwan Kim and Philip C Woodland. 2000 · 2000
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino J. Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
Incorporating speech recognition confidence into discriminative named entity recognition of speech data
Katsuhito Sudoh, Hajime Tsukada, and Hideki Isozaki. 2006 · 2006
Earlier work this paper cites.
A survey of named entity recognition and classification
David Nadeau and Satoshi Sekine. 2007 · 2007
Earlier work this paper cites.
Design challenges and misconceptions in named entity recognition
Lev Ratinov and Dan Roth. 2009 · 2009
Earlier work this paper cites.
OOV sensitive named-entity recognition in speech
Carolina Parada, Mark Dredze, and Frederick Jelinek. 2011 · 2011
Earlier work this paper cites.
Towards robust linguistic analysis using OntoNotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Hwee Tou Ng, Anders Björkelund, Olga Uryupina, Yuchen Zhang, and Zhi Zhong. 2013 · 2013
Earlier work this paper cites.
Robust tree-structured named entities recognition from speech
Christian Raymond. 2013 · 2013
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2014 · 2014
Earlier work this paper cites.
How to evaluate ASR output for named entity recognition?
Mohamed Ameur Ben Jannet, Olivier Galibert, Martine Adda-Decker, and Sophie Rosset. 2015 · 2015
Earlier work this paper cites.
Librispeech: An ASR corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur. 2015 · 2015
Earlier work this paper cites.
Deep speech 2 : End-to-end speech recognition in english and mandarin
Dario Amodei, Sundaram Ananthanarayanan, Rishita Anubhai, Jingliang Bai, Eric Battenberg, Carl Case, Jared Casper, Bryan Catanzaro, Jingdong Chen, Mike Chrzanowski, Adam Coates, Greg Diamos, Erich Elsen, Jesse H. Engel, Linxi Fan, Christopher Fougner, Awni Y. Hannun, Billy Jun, Tony Han, Patrick LeGresley, Xiangang Li, Libby Lin, Sharan Narang, Andrew Y. Ng, Sherjil Ozair, Ryan Prenger, Sheng Qian, Jonathan Raiman, Sanjeev Satheesh, David Seetapun, Shubho Sengupta, Chong Wang, Yi Wang, Zhiqian Wang, Bo Xiao, Yan Xie, Dani Yogatama, Jun Zhan, and Zhenyao Zhu. 2016 · 2016
Cited alongside, same era.
Reading Wikipedia to answer open-domain questions
Danqi Chen, Adam Fisch, Jason Weston, and Antoine Bordes. 2017 · 2017
Cited alongside, same era.
End-to-end named entity and semantic concept extraction from speech
Sahar Ghannay, Antoine Caubrière, Yannick Estève, Nathalie Camelin, Edwin Simonnet, Antoine Laurent, and Emmanuel Morin. 2018 · 2018
Cited alongside, same era.
TED-LIUM 3: Twice as Much Data and Corpus Repartition for Experiments on Speaker Adaptation
François Hernandez, Vincent Nguyen, Sahar Ghannay, Natalia Tomashenko, and Yannick Estève. 2018 · 2018
Cited alongside, same era.
SLURP: A spoken language understanding resource package
Emanuele Bastianelli, Andrea Vanzo, Pawel Swietojanski, and Verena Rieser. 2020 · 2020
Later among the works it cites.
Where are we in named entity recognition from speech?
Antoine Caubrière, Sophie Rosset, Yannick Estève, Antoine Laurent, and Emmanuel Morin. 2020 · 2020
Later among the works it cites.
Large-scale transfer learning for low-resource spoken language understanding
Xueli Jia, Jianzong Wang, Zhiyong Zhang, Ning Cheng, and Jing Xiao. 2020 · 2020
Later among the works it cites.
Improved noisy student training for automatic speech recognition
Daniel S. Park, Yu Zhang, Ye Jia, Wei Han, Chung-Cheng Chiu, Bo Li, Yonghui Wu, and Quoc V. Le. 2020 · 2020
Later among the works it cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, et al. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Samuel Louvan and Bernardo Magnini. 2018 · 2018
Cited alongside, same era.
A survey on recent advances in named entity recognition from deep learning models
Vikas Yadav and Steven Bethard. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Speech model pre-training for end-to-end spoken language understanding
Loren Lugosch, Mirco Ravanelli, Patrick Ignoto, Vikrant Singh Tomar, and Yoshua Bengio. 2019 · 2019
Cited alongside, same era.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Cited alongside, same era.
Specaugment: A simple data augmentation method for automatic speech recognition
Daniel S. Park, William Chan, Yu Zhang, Chung-Cheng Chiu, Barret Zoph, Ekin D. Cubuk, and Quoc V. Le. 2019 · 2019
Cited alongside, same era.
Lessons from building acoustic models with a million hours of speech
Sree Hari Krishnan Parthasarathi and Nikko Strom. 2019 · 2019
Cited alongside, same era.
wav2vec: Unsupervised pre-training for speech recognition
Steffen Schneider, Alexei Baevski, Ronan Collobert, and Michael Auli. 2019 · 2019
Cited alongside, same era.
Iterative pseudo-labeling for speech recognition
Qiantong Xu, Tatiana Likhomanenko, Jacob Kahn, Awni Hannun, Gabriel Synnaeve, and Ronan Collobert. 2020 · 2020
Later among the works it cites.
End-to-end named entity recognition from english speech
Hemant Yadav, Sreyan Ghosh, Yi Yu, and Rajiv Ratn Shah. 2020 · 2020
Later among the works it cites.
Do we still need automatic speech recognition for spoken language understanding?
Lasse Borgholt, Jakob Drachmann Havtorn, Mostafa Abdou, Joakim Edin, Lars Maaløe, Anders Søgaard, and Christian Igel. 2021 · 2021
Closest in time.
Deberta: decoding-enhanced bert with disentangled attention
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen. 2021 · 2021
Closest in time.
Layer-wise analysis of a self-supervised speech representation model
Ankita Pasad, Ju-Chieh Chou, and Karen Livescu. 2021 · 2021
Closest in time.
Self-training and pre-training are complementary for speech recognition
Qiantong Xu, Alexei Baevski, Tatiana Likhomanenko, Paden Tomasello, Alexis Conneau, Ronan Collobert, Gabriel Synnaeve, and Michael Auli. 2021 · 2021
Closest in time.
SUPERB: speech processing universal performance benchmark
Shu-Wen Yang, Po-Han Chi, Yung-Sung Chuang, Cheng-I Jeff Lai, Kushal Lakhotia, Yist Y. Lin, Andy T. Liu, Jiatong Shi, Xuankai Chang, Guan-Ting Lin, Tzu-Hsien Huang, Wei-Cheng Tseng, Ko-tik Lee, Da-Rong Liu, Zili Huang, Shuyan Dong, Shang-Wen Li, Shinji Watanabe, Abdelrahman Mohamed, and Hung-yi Lee. 2021 · 2021
Closest in time.
SLUE: New benchmark tasks for spoken language understanding evaluation on natural speech
Suwon Shon, Ankita Pasad, Felix Wu, Pablo Brusco, Yoav Artzi, Karen Livescu, and Kyu J Han. 2022 · 2022
Closest in time.