Fetching the paper…
Reading the bibliography…
Speech recognition in mixed language has difficulties to adapt end-to-end framework due to the lack of data and overlapping phone sets, for example in words such as "one" in English and "w\`an" in Chinese.
“A method for solving the convex programming problem with convergence rate o (1/kˆ 2),”
Yurii E Nesterov, · 1983
Earlier work this paper cites.
“Gradient-based learning applied to document recognition,”
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner, · 1998
Earlier work this paper cites.
“Codeswitching: An examination of naturally occurring conversation,”
Rosamina Lowi, · 2005
Earlier work this paper cites.
“Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,”
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber, · 2006
Earlier work this paper cites.
“Hkust/mts: A very large scale mandarin telephone speech corpus,”
Yi Liu, Pascale Fung, Yongsheng Yang, Christopher Cieri, Shudong Huang, and David Graff, · 2006
Earlier work this paper cites.
“Code-switch language model with inversion constraints for mixed language speech recognition,”
Ying Li and Pascale Fung, · 2012
Earlier work this paper cites.
“A first speech recognition system for mandarin-english code-switch conversational speech,”
Ngoc Thang Vu, Dau-Cheng Lyu, Jochen Weiner, Dominic Telaar, Tim Schlippe, Fabian Blaicher, Eng-Siong Chng, Tanja Schultz, and Haizhou Li, · 2012
Earlier work this paper cites.
“Recurrent neural network language modeling for code switching conversational speech,”
Heike Adel, Ngoc Thang Vu, Franziska Kraus, Tim Schlippe, Haizhou Li, and Tanja Schultz, · 2013
Earlier work this paper cites.
“Multi-task learning in deep neural networks for improved phoneme recognition,”
Michael L Seltzer and Jasha Droppo, · 2013
Cited alongside, same era.
“Scalable modified kneser-ney language model estimation,”
Kenneth Heafield, Ivan Pouzyrevsky, Jonathan H Clark, and Philipp Koehn, · 2013
Cited alongside, same era.
“Code switch language modeling with functional head constraint,”
Ying Li and Pascale Fung, · 2014
Cited alongside, same era.
“Learning phrase representations using rnn encoder–decoder for statistical machine translation,”
Kyunghyun Cho, Bart van Merrienboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio, · 2014
Cited alongside, same era.
“The Stanford CoreNLP natural language processing toolkit,”
Christopher D. Manning, Mihai Surdeanu, John Bauer, Jenny Finkel, Steven J. Bethard, and David McClosky, · 2014
Cited alongside, same era.
“To switch or not to switch: Code-switching in a multilingual country,”
“Mandarin-english code-switching in south-east asia ldc2015s04. web download. philadelphia: Linguistic data consortium,” 2015
Universiti Sains Malaysia Nanyang Technological University, · 2015
Later among the works it cites.
“Deep speech 2: End-to-end speech recognition in english and mandarin,”
Dario Amodei, Sundaram Ananthanarayanan, Rishita Anubhai, Jingliang Bai, Eric Battenberg, Carl Case, Jared Casper, Bryan Catanzaro, Qiang Cheng, Guoliang Chen, et al., · 2016
Later among the works it cites.
“Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,”
William Chan, Navdeep Jaitly, Quoc Le, and Oriol Vinyals, · 2016
Later among the works it cites.
“Advances in joint ctc-attention based end-to-end speech recognition with a deep cnn encoder and rnn-lm,”
Takaaki Hori, Shinji Watanabe, Yu Zhang, and William Chan, · 2017
Later among the works it cites.
“Multilingual speech recognition with a single end-to-end model,”
Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss, Bo Li, Pedro J. Moreno, Eugene Weinstein, and Kanishka Rao, · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Orit Shay, · 2015
Cited alongside, same era.
“Effective approaches to attention-based neural machine translation,”
Thang Luong, Hieu Pham, and Christopher D Manning, · 2015
Cited alongside, same era.
Closest in time.
“Phone merging for code-switched speech recognition,”
Sunit Sivasankaran, Brij Mohan Lal Srivastava, Sunayana Sitaram, Kalika Bali, and Monojit Choudhury, · 2018
Closest in time.
“Code-switching language modeling using syntax-aware multi-task learning,”
Genta Indra Winata, Andrea Madotto, Chien-Sheng Wu, and Pascale Fung, · 2018
Closest in time.