Fetching the paper…
Reading the bibliography…
This paper proposes a first attempt to build an end-to-end speech-to-text translation system, which does not use source language transcription during learning or decoding.
Towards speech Translation of Non Written Languages
Besacier, L., Zhou, B., and Gao, Y. (2006) · 2006
Earlier work this paper cites.
Corpus-Based Concatenative Synthesis: Assembling sounds by content-based selection of units from large sound databases
Schwarz, D. (2007) · 2007
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Kingma, D. and Ba, J. (2014) · 2014
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks
Sutskever, I., Vinyals, O., and Le, Q. V. (2014) · 2014
Earlier work this paper cites.
Recurrent neural network regularization
Zaremba, W., Sutskever, I., and Vinyals, O. (2014) · 2014
Earlier work this paper cites.
TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G. S., Davis, A., Dean, J., Devin, M., Ghemawat, S., Goodfellow, I., Harp, A., Irving, G., Isard, M., Jia, Y., Jozefowicz, R., Kaiser, Ł., Kudlur, M., Levenberg, J., Mané, D., Monga, R., Moore, S., Murray, D., Olah, C., Schuster, M., Shlens, J., Steiner, B., Sutskever, I., Talwar, K., Tucker, P., Vanhoucke, V., Vasudevan, V., Viégas, F., Vinyals, O., Warden, P., Wattenberg, M., Wicke, M., Yu, Y., and Zheng, X. (2015) · 2015
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate
Bahdanau, D., Cho, K., and Bengio, Y. (2015) · 2015
Cited alongside, same era.
Attention-Based Models for Speech Recognition
Chorowski, J. K., Bahdanau, D., Serdyuk, D., Cho, K., and Bengio, Y. (2015) · 2015
Cited alongside, same era.
Grammar as a Foreign Language
Vinyals, O., Kaiser, Ł., Koo, T., Petrov, S., Sutskever, I., and Hinton, G. (2015) · 2015
Cited alongside, same era.
An Unsupervised Probability Model for Speech-to-Translation Alignment of Low-Resource Languages
Anastasopoulos, A., Chiang, D., and Duong, L. (2016) · 2016
Cited alongside, same era.
Listen, Attend and Spell: A Neural Network for Large Vocabulary Conversational Speech Recognition
Chan, W., Jaitly, N., Le, Q. V., and Vinyals, O. (2016) · 2016
Cited alongside, same era.
An Attentional Model for Speech Translation Without Transcription
Duong, L., Anastasopoulos, A., Chiang, D., Bird, S., and Cohn, T. (2016) · 2016
Closest in time.
Preliminary Experiments on Unsupervised Word Discovery in Mboshi
Godard, P., Adda, G., Adda-Decker, M., Allauzen, A., Besacier, L., Bonneau-Maynard, H., Kouarata, G.-N., Löser, K., Rialland, A., and Yvon, F. (2016) · 2016
Closest in time.
Multi-task Sequence to Sequence Learning
Luong, M.-T., Le, Q. V., Sutskever, I., Vinyals, O., and Kaiser, Ł. (2016) · 2016
Closest in time.
Multi-Source Neural Translation
Zoph, B. and Knight, K. (2016) · 2016
Closest in time.
Advances in All-Neural Speech Recognition
Zweig, G., Yu, C., Droppo, J., and Stolcke, A. (2016) · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…