Fetching the paper…
Reading the bibliography…
Code-switching speech recognition has attracted an increasing interest recently, but the need for expert linguistic knowledge has always been a big issue.
“Speech recognition on code-switching among the chinese dialects,”
Dau-Cheng Lyu, Ren-Yuan Lyu, Yuang-Chin Chiang, and Chun-Nan Hsu, · 2006
Earlier work this paper cites.
“Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,”
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber, · 2006
Earlier work this paper cites.
“An approach to mixed language automatic speech recognition,”
K Bhuvanagiri and Sunil Kopparapu, · 2010
Earlier work this paper cites.
“Seame: a mandarin-english code-switching speech corpus in south-east asia,”
Dau-Cheng Lyu, Tien Ping Tan, Chng Eng Siong, and Haizhou Li, · 2010
Earlier work this paper cites.
Foundations of bilingual education and bilingualism
Colin Baker, · 2011
Earlier work this paper cites.
“A first speech recognition system for mandarin-english code-switch conversational speech,”
Ngoc Thang Vu, Dau-Cheng Lyu, Jochen Weiner, Dominic Telaar, Tim Schlippe, Fabian Blaicher, Eng-Siong Chng, Tanja Schultz, and Haizhou Li, · 2012
Earlier work this paper cites.
“Integration of language identification into a recognition system for spoken conversations containing code-switches,”
Jochen Weiner, Ngoc Thang Vu, Dominic Telaar, Florian Metze, Tanja Schultz, Dau-Cheng Lyu, Eng-Siong Chng, and Haizhou Li, · 2012
Earlier work this paper cites.
Code-switching in conversation: Language, interaction and identity
Peter Auer, · 2013
Earlier work this paper cites.
“Combination of recurrent neural networks and factored language models for code-switching language modeling,”
Heike Adel, Ngoc Thang Vu, and Tanja Schultz, · 2013
Earlier work this paper cites.
“Features for factored language models for code-switching speech,”
Heike Adel, Katrin Kirchhoff, Dominic Telaar, Ngoc Thang Vu, Tim Schlippe, and Tanja Schultz, · 2014
Earlier work this paper cites.
“Neural machine translation by jointly learning to align and translate,”
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio, · 2014
Cited alongside, same era.
“Syntactic and semantic features for code-switching factored language models,”
Heike Adel, Ngoc Thang Vu, Katrin Kirchhoff, Dominic Telaar, and Tanja Schultz, · 2015
Cited alongside, same era.
“Attention-based models for speech recognition,”
Jan K Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, Kyunghyun Cho, and Yoshua Bengio, · 2015
Cited alongside, same era.
“Neural machine translation of rare words with subword units,”
Rico Sennrich, Barry Haddow, and Alexandra Birch, · 2015
Cited alongside, same era.
“Deep speech 2: End-to-end speech recognition in english and mandarin,”
Dario Amodei, Sundaram Ananthanarayanan, Rishita Anubhai, Jingliang Bai, Eric Battenberg, Carl Case, Jared Casper, Bryan Catanzaro, Qiang Cheng, Guoliang Chen, et al., · 2016
“Joint ctc-attention based end-to-end speech recognition using multi-task learning,”
Suyoun Kim, Takaaki Hori, and Shinji Watanabe, · 2017
Later among the works it cites.
“A comparison of sequence-to-sequence models for speech recognition,”
Rohit Prabhavalkar, Kanishka Rao, Tara N Sainath, Bo Li, Leif Johnson, and Navdeep Jaitly, · 2017
Later among the works it cites.
“Exploring architectures, data and units for streaming end-to-end speech recognition with rnn-transducer,”
Kanishka Rao, Haşim Sak, and Rohit Prabhavalkar, · 2017
Later among the works it cites.
“Subword and crossword units for ctc acoustic models,”
Thomas Zenkel, Ramon Sanabria, Florian Metze, and Alex Waibel, · 2017
Later among the works it cites.
“Joint ctc/attention decoding for end-to-end speech recognition,”
Takaaki Hori, Shinji Watanabe, and John R. Hershey, · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“On online attention-based speech recognition and joint mandarin character-pinyin training,”
William Chan and Ian Lane, · 2016
Cited alongside, same era.
“Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,”
William Chan, Navdeep Jaitly, Quoc V. Le, and Oriol Vinyals, · 2016
Cited alongside, same era.
“End-to-end attention-based large vocabulary speech recognition,”
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk, Philemon Brakel, and Yoshua Bengio, · 2016
Cited alongside, same era.
“Towards better decoding and language model integration in sequence to sequence models,”
Jan Chorowski and Navdeep Jaitly, · 2016
Cited alongside, same era.
“Tensorflow: Large-scale machine learning on heterogeneous distributed systems,”
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, et al., · 2016
Cited alongside, same era.
Chung-Cheng Chiu, Tara N Sainath, Yonghui Wu, Rohit Prabhavalkar, Patrick Nguyen, Zhifeng Chen, Anjuli Kannan, Ron J Weiss, Kanishka Rao, Katya Gonina, et al., · 2017
Later among the works it cites.
“An end-to-end language-tracking speech recognizer for mixed-language speech,”
Hiroshi Seki, Shinji Watanabe, Takaaki Hori, Jonathan Le Roux, and John R. Hershey, · 2018
Closest in time.
“A comparable study of modeling units for end-to-end mandarin speech recognition,”
Wei Zou, Dongwei Jiang, Shuaijiang Zhao, and Xiangang Li, · 2018
Closest in time.
Shiyu Zhou, Linhao Dong, Shuang Xu, and Bo Xu, · 2018
Closest in time.
Taku Kudo and John Richardson, · 2018
Closest in time.