Fetching the paper…
Reading the bibliography…
We develop an algorithm which can learn from partially labeled and unsegmented sequential data.
Learning with multiple labels
Jin, R. and Ghahramani, Z · 2002
Earlier work this paper cites.
Temporal classification: extending the classification paradigm to multivariate time series
Kadous, M. W · 2002
Earlier work this paper cites.
The iam-database: an english sentence database for offline handwriting recognition
Marti, U.-V. and Bunke, H · 2002
Earlier work this paper cites.
Probabilistic finite-state machines-part i and ii
Vidal, E., Thollard, F., De La Higuera, C., Casacuberta, F., and Carrasco, R. C · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A., Fernández, S., Gomez, F., and Schmidhuber, J · 2006
Earlier work this paper cites.
Speech recognition with weighted finite-state transducers
Mohri, M., Pereira, F., and Riley, M · 2008
Earlier work this paper cites.
Semi-supervised learning (chapelle, o. et al., eds.; 2006)[book reviews]
Chapelle, O., Scholkopf, B., and Zien, A · 2009
Earlier work this paper cites.
Weighted automata algorithms
Mohri, M · 2009
Earlier work this paper cites.
Learning from partial labels
Cour, T., Sapp, B., and Taskar, B · 2011
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
Duchi, J., Hazan, E., and Singer, Y · 2011
Earlier work this paper cites.
Multi-instance multi-label learning
Zhou, Z.-H., Zhang, M.-L., Huang, S.-J., and Li, Y.-F · 2011
Earlier work this paper cites.
Multiple instance classification: Review, taxonomy and comparative study
Amores, J · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Earlier work this paper cites.
Learnability of the superset label learning problem
Liu, L. and Dietterich, T · 2014
Cited alongside, same era.
Eesen: End-to-end speech recognition using deep rnn models and wfst-based decoding
Miao, Y., Gowayyed, M., and Metze, F · 2015
Cited alongside, same era.
Librispeech: an asr corpus based on public domain audio books
Panayotov, V., Chen, G., Povey, D., and Khudanpur, S · 2015
Cited alongside, same era.
Wav2letter: an end-to-end convnet-based speech recognition system
Collobert, R., Puhrsch, C., and Synnaeve, G · 2016
Cited alongside, same era.
Connectionist temporal modeling for weakly supervised action labeling
Huang, D.-A., Fei-Fei, L., and Niebles, J. C · 2016
Cited alongside, same era.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Word-level speech recognition with a letter to word encoder
Collobert, R., Hannun, A., and Synnaeve, G · 2020
Later among the works it cites.
Conformer: Convolution-augmented Transformer for Speech Recognition
Gulati, A., Qin, J., Chiu, C.-C., Parmar, N., Zhang, Y., Yu, J., Han, W., Wang, S., Zhang, Z., Wu, Y., and Pang, R · 2020
Later among the works it cites.
Differentiable weighted finite-state transducers
Hannun, A., Pratap, V., Kahn, J., and Hsu, W.-N · 2020
Later among the works it cites.
Pay attention to what you read: Non-recurrent handwritten text-line recognition, 2020
Kang, L., Riba, P., Rusiñol, M., Fornés, A., and Villegas, M · 2020
Later among the works it cites.
End-to-end asr: from supervised to semi-supervised learning with modern architectures, 2020
Synnaeve, G., Xu, Q., Kahn, J., Likhomanenko, T., Grave, E., Pratap, V., Sriram, A., Liptchinsky, V., and Collobert, R · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Lead2gold: Towards exploiting the full potential of noisy transcriptions for speech recognition
Dufraux, A., Vincent, E., Hannun, A., Brun, A., and Douze, M · 2019
Cited alongside, same era.
Reducing transformer depth on demand with structured dropout
Fan, A., Grave, E., and Joulin, A · 2019
Cited alongside, same era.
Evaluating sequence-to-sequence models for handwritten text recognition
Michael, J., Labahn, R., Grüning, T., and Zöllner, J · 2019
Cited alongside, same era.
Specaugment: A simple data augmentation method for automatic speech recognition
Park, D. S., Chan, W., Zhang, Y., Chiu, C.-C., Zoph, B., Cubuk, E. D., and Le, Q. V · 2019
Cited alongside, same era.
Wav2letter++: A fast open-source speech recognition system
Pratap, V., Hannun, A., Xu, Q., Cai, J., Kahn, J., Synnaeve, G., Liptchinsky, V., and Collobert, R · 2019
Cited alongside, same era.
Multi-label connectionist temporal classification
Wigington, C., Price, B., and Cohen, S · 2019
Cited alongside, same era.
Origaminet: Weakly-supervised, segmentation-free, one-step, full page textrecognition by learning to unfold
Yousef, M. and Bishop, T. E · 2020
Later among the works it cites.
Accurate, data-efficient, unconstrained text recognition with convolutional neural networks
Yousef, M., Hussain, K. F., and Mohammed, U. S · 2020
Later among the works it cites.
https://github.com/k2-fsa/k2 , 2021
k2-fsa · 2021
Later among the works it cites.
CTC variations through new wfst topologies
Laptev, A., Majumdar, S., and Ginsburg, B · 2021
Later among the works it cites.
Rethinking Evaluation in ASR: Are Our Models Robust Enough?
Likhomanenko, T., Xu, Q., Pratap, V., Tomasello, P., Kahn, J., Avidov, G., Collobert, R., and Synnaeve, G · 2021
Later among the works it cites.
Semi-supervised speech recognition via graph-based temporal classification
Moritz, N., Hori, T., and Le Roux, J · 2021
Later among the works it cites.
Word order does not matter for speech recognition
Pratap, V., Xu, Q., Likhomanenko, T., Synnaeve, G., and Collobert, R · 2021
Later among the works it cites.
W-CTC: a connectionist temporal classification loss with wild cards
Anonymous · 2022
Closest in time.