H. Scudder, “Probability of error of some adaptive pattern-recognition machines,” IEEE Transactions on Information Theory , vol. 11, no. 3, pp. 363–371, 1965
1965
Earlier work this paper cites.
J. Godfrey and E. Holliman, “Switchboard-1 release 2 LDC97S62,” Philadelphia: LDC , 1993
1993
Earlier work this paper cites.
D. Yarowsky, “Unsupervised word sense disambiguation rivaling supervised methods,” in 33rd annual meeting of the association for computational linguistics , 1995, pp. 189–196
1995
Earlier work this paper cites.
LDC et al. , “2000 hub5 english evaluation speech LDC2002S09 and transcripts LDC2002T43,” Web Download. Philadelphia: LDC , 2002
2002
Earlier work this paper cites.
C. Cieri, , D. Graff, O. Kimball, D. Miller, and K. Walker, “Fisher english training speech parts 1 and 2 transcripts LDC200{4,5}T19,” Philadelphia: LDC , 2004, 2005
2005
Earlier work this paper cites.
C. Cieri, D. Miller, and K. Walker, “Fisher english training speech parts 1 and 2 LDC200{4,5}S13,” Philadelphia: LDC , 2004, 2005
2005
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in Proceedings of the 23rd international conference on Machine learning , 2006, pp. 369–376
2006
Earlier work this paper cites.
D. McClosky, E. Charniak, and M. Johnson, “Effective self-training for parsing,” in Proceedings of the Human Language Technology Conference of the NAACL, Main Conference , 2006, pp. 152–159
2006
Earlier work this paper cites.
N. Ueffing, “Using monolingual source-language data to improve mt performance,” in International Workshop on Spoken Language Translation (IWSLT) 2006 , 2006
2006
Earlier work this paper cites.
R. Reichart and A. Rappoport, “Self-training for enhancement and domain adaptation of statistical parsers trained on small datasets,” in Proceedings of the 45th Annual Meeting of the Association of Computational Linguistics , 2007, pp. 616–623
2007
Earlier work this paper cites.
J. G. Fiscus et al. , “2003 nist rich transcription evaluation data LDC2007S10,” Web Download. Philadelphia: LDC , 2007
2007
Earlier work this paper cites.
Z. Huang and M. Harper, “Self-training pcfg grammars with latent annotations across languages,” in Proceedings of the 2009 conference on empirical methods in natural language processing , 2009, pp. 832–841
2009
Earlier work this paper cites.
S. Novotney and R. Schwartz, “Analysis of low-resource acoustic model self-training,” in Tenth Annual Conference of the International Speech Communication Association , 2009
2009
Earlier work this paper cites.
O. Chapelle, B. Schölkopf, and A. Zien, Semi-supervised Learning . Mit Press, 2010
2010
Earlier work this paper cites.
J. Duchi, E. Hazan, and Y. Singer, “Adaptive subgradient methods for online learning and stochastic optimization,” Journal of machine learning research , vol. 12, no. Jul, pp. 2121–2159, 2011
2011
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz et al. , “The kaldi speech recognition toolkit,” in IEEE 2011 workshop on automatic speech recognition and understanding , no. CONF. IEEE Signal Processing Society, 2011
2011
Earlier work this paper cites.
D.-H. Lee, “Pseudo-label: The simple and efficient semi-supervised learning method for deep neural networks,” in Workshop on challenges in representation learning, ICML , vol. 3, no. 2, 2013
2013
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2015, pp. 5206–5210
2015
Earlier work this paper cites.