Fetching the paper…
Reading the bibliography…
As more users across the world are interacting with dialog agents in their daily life, there is a need for better speech understanding that calls for renewed attention to the dynamics between research in automatic speech recognition (ASR) and natural language understanding (NLU).
Fazel-Zarandi, Maryam, Longshaokan Wang, Aditya Tiwari, and Spyros Matsoukas. 2019 · 1911
Earlier work this paper cites.
Combinatory grammars and parasitic gaps
Steedman, Mark. 1987 · 1987
Earlier work this paper cites.
Evaluation of spoken language systems: the ATIS domain
Price, P. J. 1990 · 1990
Earlier work this paper cites.
Switchboard-1 Release 2 LDC97S62
Godfrey, John and Edward Holliman. 1993 · 1993
Earlier work this paper cites.
Preliminaries to a theory of speech disfluencies
Schriberg, Elisabeth. 1994 · 1994
Earlier work this paper cites.
Improving spoken language understanding by exploiting asr n-best hypotheses
Li, Mingda, Weitong Ruan, Xinyue Liu, Luca Soldaini, Wael Hamza, and Chengwei Su. 2020b · 2001
Earlier work this paper cites.
Joint contextual modeling for ASR correction and language understanding
Weng, Yue, Sai Sumanth Miryala, Chandra Khatri, Runze Wang, Huaixiu Zheng, Piero Molino, Mahdi Namazifar, Alexandros Papangelis, Hugh Williams, Franziska Bell, and Gökhan Tür. 2020 · 2002
Earlier work this paper cites.
Sphinx-4: A flexible open source framework for speech recognition
Walker, Willie, Paul Lamere, Philip Kwok, Bhiksha Raj, Rita Singh, Evandro Gouvea, Peter Wolf, and Joe Woelfel. 2004 · 2004
Earlier work this paper cites.
The intention behind web queries
Baeza-Yates, Ricardo, Liliana Calderón-Benavides, and Cristina González-Caro. 2006 · 2006
Earlier work this paper cites.
Beyond asr 1-best: Using word confusion networks in spoken language understanding
Hakkani-Tür, Dilek, Frédéric Béchet, Giuseppe Riccardi, and Gokhan Tur. 2006 · 2006
Earlier work this paper cites.
Enriching speech recognition with automatic detection of sentence boundaries and disfluencies
Liu, Yang, Elizabeth Shriberg, Andreas Stolcke, Dustin Hillard, Mari Ostendorf, and Mary Harper. 2006 · 2006
Earlier work this paper cites.
Learning noun phrase query segmentation
Bergsma, Shane and Qin Iris Wang. 2007 · 2007
Earlier work this paper cites.
The linguistic structure of English web-search queries
Barr, Cory, Rosie Jones, and Moira Regelson. 2008 · 2008
Earlier work this paper cites.
Spoken language understanding
De Mori, R., F. Bechet, D. Hakkani-Tur, M. McTear, G. Riccardi, and G. Tur. 2008 · 2008
Earlier work this paper cites.
Speech segmentation and spoken document processing
Ostendorf, Mari, Benoît Favre, Ralph Grishman, Dilek Hakkani-Tur, Mary Harper, Dustin Hillard, Julia Hirschberg, Heng Ji, Jeremy G Kahn, Yang Liu, et al. 2008 · 2008
Earlier work this paper cites.
Optimizing endpointing thresholds using dialogue features in a spoken dialogue system
Raux, Antoine and Maxine Eskenazi. 2008 · 2008
Earlier work this paper cites.
End-to-end spoken language understanding without full transcripts
Kuo, Hong-Kwang Jeff, Zoltán Tüske, Samuel Thomas, Yinghui Huang, Kartik Audhkhasi, Brian Kingsbury, Gakuto Kurata, Zvi Kons, Ron Hoory, and Luis A. Lastras. 2020 · 2009
Earlier work this paper cites.
Semantic tagging of web search queries
Manshadi, Mehdi and Xiao Li. 2009 · 2009
Earlier work this paper cites.
From keywords to semantic queries—incremental query construction on the semantic web
Zenz, Gideon, Xuan Zhou, Enrico Minack, Wolf Siberski, and Wolfgang Nejdl. 2009 · 2009
Cited alongside, same era.
Why word error rate is not a good metric for speech recognizer training for the speech translation task?
He, Xiaodong, Li Deng, and Alex Acero. 2011 · 2011
Cited alongside, same era.
Unsupervised query segmentation using only query logs
Mishra, Nikita, Rishiraj Saha Roy, Niloy Ganguly, Srivatsan Laxman, and Monojit Choudhury. 2011 · 2011
Cited alongside, same era.
Spoken Language Understanding: Systems for Extracting Semantic Information from Speech
Tur, G. and R. De Mori. 2011 · 2011
Cited alongside, same era.
Joint decoding for speech recognition and semantic tagging
Deoras, Anoop, Ruhi Sarikaya, Gökhan Tür, and Dilek Hakkani-Tür. 2012 · 2012
Cited alongside, same era.
Abstract Meaning Representation for sembanking
The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english
Livingstone, Steven R. and Frank A. Russo. 2018 · 2018
Later among the works it cites.
Towards end-to-end spoken language understanding
Serdyuk, Dmitriy, Yongqiang Wang, Christian Fuegen, Anuj Kumar, Baiyang Liu, and Yoshua Bengio. 2018 · 2018
Later among the works it cites.
Drcd: a chinese machine reading comprehension dataset
Shao, Chih Chieh, Trois Liu, Yuting Lai, Yiying Tseng, and Sam Tsai. 2018 · 2018
Later among the works it cites.
Informing the design of spoken conversational search: Perspective paper
Trippas, Johanne R., Damiano Spina, Lawrence Cavedon, Hideo Joho, and Mark Sanderson. 2018 · 2018
Later among the works it cites.
Semantic lattice processing in contextual automatic speech recognition for google assistant
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Banarescu, Laura, Claire Bonial, Shu Cai, Madalina Georgescu, Kira Griffitt, Ulf Hermjakob, Kevin Knight, Philipp Koehn, Martha Palmer, and Nathan Schneider. 2013 · 2013
Cited alongside, same era.
Semantic parsing using word confusion networks with conditional random fields
Tür, Gökhan, Anoop Deoras, and Dilek Hakkani-Tür. 2013 · 2013
Cited alongside, same era.
The second dialog state tracking challenge
Henderson, Matthew, Blaise Thomson, and Jason D. Williams. 2014 · 2014
Cited alongside, same era.
Mobile query reformulations
Shokouhi, Milad, Rosie Jones, Umut Ozertem, Karthik Raghunathan, and Fernando Diaz. 2014 · 2014
Cited alongside, same era.
Latticernn: Recurrent neural networks over lattices
Ladhak, Faisal, Ankur Gandhe, Markus Dreyer, Lambert Mathias, Ariya Rastrow, and Björn Hoffmeister. 2016 · 2016
Cited alongside, same era.
SQuAD: 100,000+ questions for machine comprehension of text
Rajpurkar, Pranav, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
Glasmachers, Tobias. 2017 · 2017
Cited alongside, same era.
Velikovich, Leonid, Ian Williams, Justin Scheiner, Petar Aleksic, Pedro Moreno, and Michael Riley. 2018 · 2018
Later among the works it cites.
A survey on semantic parsing
Kamath, Aishwarya and Rajarshi Das. 2019 · 2019
Later among the works it cites.
Corpora generation for grammatical error correction
Lichtarge, Jared, Chris Alberti, Shankar Kumar, Noam Shazeer, Niki Parmar, and Simon Tong. 2019 · 2019
Later among the works it cites.
Speech model pre-training for end-to-end spoken language understanding
Lugosch, Loren, Mirco Ravanelli, Patrick Ignoto, Vikrant Singh Tomar, and Yoshua Bengio. 2019 · 2019
Later among the works it cites.
CoQA: A conversational question answering challenge
Reddy, Siva, Danqi Chen, and Christopher D. Manning. 2019 · 2019
Later among the works it cites.
Contextual recovery of out-of-lattice named entities in automatic speech recognition
Serrino, Jack, Leonid Velikovich, Petar Aleksic, and Cyril Allauzen. 2019 · 2019
Later among the works it cites.
Contextual error correction in automatic speech recognition
Faruqui, Manaal and Janara Christensen. 2020 · 2020
Later among the works it cites.
Survey of automatic spelling correction
Hládek, Daniel, Ján Staš, and Matúš Pleva. 2020 · 2020
Later among the works it cites.
Data augmentation for training dialog models robust to speech recognition errors
Wang, Longshaokan, Maryam Fazel-Zarandi, Aditya Tiwari, Spyros Matsoukas, and Lazaros Polymenakos. 2020 · 2020
Later among the works it cites.
Robustness testing of language understanding in task-oriented dialog
Liu, Jiexi, Ryuichi Takanobu, Jiaxin Wen, Dazhen Wan, Hongguang Li, Weiran Nie, Cheng Li, Wei Peng, and Minlie Huang. 2021 · 2021
Closest in time.
RADDLE: An evaluation benchmark and analysis platform for robust task-oriented dialog systems
Peng, Baolin, Chunyuan Li, Zhu Zhang, Chenguang Zhu, Jinchao Li, and Jianfeng Gao. 2021 · 2021
Closest in time.
NoiseQA: Challenge Set Evaluation for User-Centric Question Answering
Ravichander, Abhilasha, Siddharth Dalmia, Maria Ryskina, Florian Metze, Eduard Hovy, and Alan W Black. 2021 · 2021
Closest in time.
Towards data distillation for end-to-end spoken conversational question answering
You, Chenyu, Nuo Chen, Fenglin Liu, Dongchao Yang, Zhiyang Xu, and Yuexian Zou. 2021 · 2021
Closest in time.