Fetching the paper…
Reading the bibliography…
Text recognition is a long-standing research problem for document digitalization.
Unified Language Model Pre-training for Natural Language Understanding and Generation
Dong, L.; Yang, N.; Wang, W.; Wei, F.; Liu, X.; Wang, Y.; Gao, J.; Zhou, M.; and Hon, H.-W. 2019 · 1905
Earlier work this paper cites.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
Real-time Scene Text Detection with Differentiable Binarization
Liao, M.; Wan, Z.; Yao, C.; Chen, K.; and Bai, X. 2019 · 1911
Earlier work this paper cites.
Bidirectional scene text recognition with a single decoder
Bleeker, M.; and de Rijke, M. 2019 · 1912
Earlier work this paper cites.
Minilm: Deep self-attention distillation for task-agnostic compression of pre-trained transformers
Wang, W.; Wei, F.; Dong, L.; Bao, H.; Yang, N.; and Zhou, M. 2020b · 2002
Earlier work this paper cites.
Pay attention to what you read: Non-recurrent handwritten text-line recognition
Kang, L.; Riba, P.; Rusiñol, M.; Fornés, A.; and Villegas, M. 2020 · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A.; Fernández, S.; Gomez, F.; and Schmidhuber, J. 2006 · 2006
Earlier work this paper cites.
Offline handwriting recognition with multidimensional recurrent neural networks
Graves, A.; and Schmidhuber, J. 2008 · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J.; Dong, W.; Socher, R.; Li, L.-J.; Li, K.; and Fei-Fei, L. 2009 · 2009
Earlier work this paper cites.
End-to-end scene text recognition
Wang, K.; Babenko, B.; and Belongie, S. 2011 · 2011
Earlier work this paper cites.
Top-down and bottom-up cues for scene text recognition
Mishra, A.; Alahari, K.; and Jawahar, C. 2012 · 2012
Earlier work this paper cites.
ICDAR 2013 robust reading competition
Karatzas, D.; Shafait, F.; Uchida, S.; Iwamura, M.; i Bigorda, L. G.; Mestre, S. R.; Mas, J.; Mota, D. F.; Almazan, J. A.; and De Las Heras, L. P. 2013 · 2013
Earlier work this paper cites.
Recognizing text with perspective distortion in natural scenes
Phan, T. Q.; Shivakumara, P.; Tian, S.; and Tan, C. L. 2013 · 2013
Earlier work this paper cites.
Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition
Jaderberg, M.; Simonyan, K.; Vedaldi, A.; and Zisserman, A. 2014 · 2014
Earlier work this paper cites.
Dropout improves recurrent neural networks for handwriting recognition
Pham, V.; Bluche, T.; Kermorvant, C.; and Louradour, J. 2014 · 2014
Earlier work this paper cites.
A robust arbitrary text detection system for natural scene images
Risnumawan, A.; Shivakumara, P.; Chan, C. S.; and Tan, C. L. 2014 · 2014
Earlier work this paper cites.
Accurate scene text recognition based on recurrent neural network
Su, B.; and Lu, S. 2014 · 2014
Earlier work this paper cites.
ICDAR 2015 competition on robust reading
Karatzas, D.; Gomez-Bigorda, L.; Nicolaou, A.; Ghosh, S.; Bagdanov, A.; Iwamura, M.; Matas, J.; Neumann, L.; Chandrasekhar, V. R.; Lu, S.; et al. 2015 · 2015
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Sennrich, R.; Haddow, B.; and Birch, A. 2015 · 2015
Earlier work this paper cites.
Joint line segmentation and transcription for end-to-end handwritten paragraph recognition
Bluche, T. 2016 · 2016
Earlier work this paper cites.
Synthetic data for text localisation in natural images
Gupta, A.; Vedaldi, A.; and Zisserman, A. 2016 · 2016
Earlier work this paper cites.
Generating Synthetic Data for Text Recognition
Krishnan, P.; and Jawahar, C. V. 2016 · 2016
Earlier work this paper cites.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
Shi, B.; Bai, X.; and Yao, C. 2016 · 2016
Cited alongside, same era.
Robust scene text recognition with automatic rectification
Shi, B.; Wang, X.; Lyu, P.; Yao, C.; and Bai, X. 2016 · 2016
Cited alongside, same era.
Handwriting recognition with large multidimensional long short-term memory recurrent neural networks
Voigtlaender, P.; Doetsch, P.; and Ney, H. 2016 · 2016
Cited alongside, same era.
Gated convolutional recurrent neural networks for multilingual handwriting recognition
Bluche, T.; and Messina, R. 2017 · 2017
Cited alongside, same era.
Are multidimensional recurrent layers really necessary for handwritten text recognition?
Puigcerver, J. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Plugnet: Degradation aware scene text recognition supervised by a pluggable super-resolution unit
Mou, Y.; Tan, L.; Yang, H.; Chen, J.; Liu, L.; Yan, R.; and Huang, Y. 2020 · 2020
Later among the works it cites.
Textscanner: Reading characters in order for robust scene text recognition
Wan, Z.; He, M.; Chen, H.; Bai, X.; and Yao, C. 2020 · 2020
Later among the works it cites.
Towards Accurate Scene Text Recognition With Semantic Reasoning Networks
Yu, D.; Li, X.; Zhang, C.; Liu, T.; Han, J.; Liu, J.; and Ding, E. 2020 · 2020
Later among the works it cites.
Robustscanner: Dynamically enhancing positional clues for robust text recognition
Yue, X.; Kuang, Z.; Lin, C.; Sun, H.; and Zhang, W. 2020 · 2020
Later among the works it cites.
AutoSTR: Efficient backbone search for scene text recognition
Zhang, H.; Yao, Q.; Yang, M.; Xu, Y.; and Bai, X. 2020a · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
An efficient end-to-end neural model for handwritten text recognition
Chowdhury, A.; and Vig, L. 2018 · 2018
Cited alongside, same era.
Kudo, T.; and Richardson, J. 2018 · 2018
Cited alongside, same era.
What is wrong with scene text recognition model comparisons? dataset and model analysis
Baek, J.; Kim, G.; Lee, J.; Park, S.; Han, D.; Yun, S.; Oh, S. J.; and Lee, H. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Cited alongside, same era.
Reading scene text with fully convolutional sequence modeling
Gao, Y.; Chen, Y.; Wang, J.; Tang, M.; and Lu, H. 2019 · 2019
Cited alongside, same era.
A scalable handwritten text recognition system
Ingle, R. R.; Fujii, Y.; Deselaers, T.; Baccash, J.; and Popat, A. C. 2019 · 2019
Cited alongside, same era.
Atienza, R. 2021 · 2021
Closest in time.
What if We Only Use Real Datasets for Scene Text Recognition? Toward Scene Text Recognition With Fewer Labels
Baek, J.; Matsui, Y.; and Aizawa, K. 2021 · 2021
Closest in time.
BEiT: BERT Pre-Training of Image Transformers
Bao, H.; Dong, L.; and Wei, F. 2021 · 2021
Closest in time.
Revisiting Classification Perspective on Scene Text Recognition
Cai, H.; Sun, J.; and Xiong, Y. 2021 · 2021
Closest in time.
Representation and Correlation Enhanced Encoder-Decoder Framework for Scene Text Recognition
Cui, M.; Wang, W.; Zhang, J.; and Wang, L. 2021 · 2021
Closest in time.
Rethinking Text Line Recognition Models
Diaz, D. H.; Qin, S.; Ingle, R.; Fujii, Y.; and Bissacco, A. 2021 · 2021
Closest in time.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; Uszkoreit, J.; and Houlsby, N. 2021 · 2021
Closest in time.
Read Like Humans: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Recognition
Fang, S.; Xie, H.; Wang, Y.; Mao, Z.; and Zhang, Y. 2021 · 2021
Closest in time.
Zero-shot text-to-image generation
Ramesh, A.; Pavlov, M.; Goh, G.; Gray, S.; Voss, C.; Radford, A.; Chen, M.; and Sutskever, I. 2021 · 2021
Closest in time.
Training data-efficient image transformers & distillation through attention
Touvron, H.; Cord, M.; Douze, M.; Massa, F.; Sablayrolles, A.; and Jégou, H. 2021 · 2021
Closest in time.
From Two to One: A New Scene Text Recognizer With Visual Language Modeling Network
Wang, Y.; Xie, H.; Fang, S.; Wang, J.; Zhu, S.; and Zhang, Y. 2021 · 2021
Closest in time.
Primitive Representation Learning for Scene Text Recognition
Yan, R.; Peng, L.; Xiao, S.; and Yao, G. 2021 · 2021
Closest in time.
Scene Text Recognition with Permuted Autoregressive Sequence Models
Bautista, D.; and Atienza, R. 2022 · 2022
Closest in time.
MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining
Lyu, P.; Zhang, C.; Liu, S.; Qiao, M.; Xu, Y.; Wu, L.; Yao, K.; Han, J.; Ding, E.; and Wang, J. 2022 · 2022
Closest in time.
Aster: An attentional scene text recognizer with flexible rectification
Shi, B.; Yang, M.; Wang, X.; Lyu, P.; Yao, C.; and Bai, X. 2018 · 2048
Closest in time.
Esir: End-to-end scene text recognition via iterative image rectification
Zhan, F.; and Lu, S. 2019 · 2068
Closest in time.