Fetching the paper…
Reading the bibliography…
In the last decade, the blossom of deep learning has witnessed the rapid development of scene text recognition.
Gestalt psychology
Köhler, W. 1967 · 1967
Earlier work this paper cites.
Handwritten Korean character image database PE92
KIM, D.-H.; Hwang, Y.-S.; Park, S.-T.; Kim, E.-J.; Paek, S.-H.; and BANG, S.-Y. 1996 · 1996
Earlier work this paper cites.
Super-resolution enhancement of text image sequences
Capel, D.; and Zisserman, A. 2000 · 2000
Earlier work this paper cites.
ICDAR 2003 robust reading competitions: entries, results, and future directions
Lucas, S. M.; Panaretos, A.; Sosa, L.; Tang, A.; Wong, S.; Young, R.; Ashida, K.; Nagai, H.; Okamoto, M.; Yamamoto, H.; et al. 2005 · 2003
Earlier work this paper cites.
Single-frame text super-resolution: A bayesian approach
Dalley, G.; Freeman, B.; and Marks, J. 2004 · 2004
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A.; Fernández, S.; Gomez, F.; and Schmidhuber, J. 2006 · 2006
Earlier work this paper cites.
Word spotting in the wild
Wang, K.; and Belongie, S. 2010 · 2010
Earlier work this paper cites.
End-to-end scene text recognition
Wang, K.; Babenko, B.; and Belongie, S. 2011 · 2011
Earlier work this paper cites.
Real-time scene text localization and recognition
Neumann, L.; and Matas, J. 2012 · 2012
Earlier work this paper cites.
Adadelta: an adaptive learning rate method
Zeiler, M. D. 2012 · 2012
Earlier work this paper cites.
Online and offline handwritten Chinese character recognition: benchmarking on new databases
Liu, C.-L.; Yin, F.; Wang, D.-H.; and Wang, Q.-F. 2013 · 2013
Earlier work this paper cites.
Recognizing text with perspective distortion in natural scenes
Phan, T. Q.; Shivakumara, P.; Tian, S.; and Tan, C. L. 2013 · 2013
Earlier work this paper cites.
ICDAR 2013 Chinese handwriting recognition competition
Yin, F.; Wang, Q.-F.; Zhang, X.-Y.; and Liu, C.-L. 2013 · 2013
Earlier work this paper cites.
Learning a deep convolutional network for image super-resolution
Dong, C.; Loy, C. C.; He, K.; and Tang, X. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P.; and Ba, J. 2014 · 2014
Cited alongside, same era.
A robust arbitrary text detection system for natural scene images
Risnumawan, A.; Shivakumara, P.; Chan, C. S.; and Tan, C. L. 2014 · 2014
Cited alongside, same era.
ICDAR 2015 competition on robust reading
Karatzas, D.; Gomez-Bigorda, L.; Nicolaou, A.; Ghosh, S.; Bagdanov, A.; Iwamura, M.; Matas, J.; Neumann, L.; Chandrasekhar, V. R.; Lu, S.; et al. 2015 · 2015
Cited alongside, same era.
Synthetic data for text localisation in natural images
Gupta, A.; Vedaldi, A.; and Zisserman, A. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Cited alongside, same era.
Reading text in the wild with convolutional neural networks
Toward real-world single image super-resolution: A new benchmark and a new model
Cai, J.; Zeng, H.; Yong, H.; Cao, Z.; and Zhang, L. 2019 · 2019
Later among the works it cites.
Moran: A multi-object rectified attention network for scene text recognition
Luo, C.; Jin, L.; and Sun, Z. 2019 · 2019
Later among the works it cites.
Citizen Id Card Detection using Image Processing and Optical Character Recognition
Satyawan, W.; Pratama, M. O.; Jannati, R.; Muhammad, G.; Fajar, B.; Hamzah, H.; Fikri, R.; and Kristian, K. 2019 · 2019
Later among the works it cites.
NRTR: A no-recurrence sequence-to-sequence model for scene text recognition
Sheng, F.; Chen, Z.; and Xu, B. 2019 · 2019
Later among the works it cites.
Zoom to learn, learn to zoom
Zhang, X.; Chen, Q.; Ng, R.; and Koltun, V. 2019 · 2019
Later among the works it cites.
Seed: Semantics enhanced encoder-decoder framework for scene text recognition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jaderberg, M.; Simonyan, K.; Vedaldi, A.; and Zisserman, A. 2016 · 2016
Cited alongside, same era.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
Shi, B.; Bai, X.; and Yao, C. 2016 · 2016
Cited alongside, same era.
Focusing attention: Towards accurate text recognition in natural images
Cheng, Z.; Bai, F.; Xu, Y.; Zheng, G.; Pu, S.; and Zhou, S. 2017 · 2017
Cited alongside, same era.
Photo-realistic single image super-resolution using a generative adversarial network
Ledig, C.; Theis, L.; Huszár, F.; Caballero, J.; Cunningham, A.; Acosta, A.; Aitken, A.; Tejani, A.; Totz, J.; Wang, Z.; et al. 2017 · 2017
Cited alongside, same era.
Enhanced deep residual networks for single image super-resolution
Lim, B.; Son, S.; Kim, H.; Nah, S.; and Mu Lee, K. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
Learning to super-resolve blurry face and text images
Xu, X.; Sun, D.; Pan, J.; Zhang, Y.; Pfister, H.; and Yang, M.-H. 2017 · 2017
Cited alongside, same era.
Qiao, Z.; Zhou, Y.; Yang, D.; Zhou, Y.; and Wang, W. 2020 · 2020
Later among the works it cites.
ST-SiameseNet: Spatio-Temporal Siamese Networks for Human Mobility Signature Identification
Ren, H.; Pan, M.; Li, Y.; Zhou, X.; and Luo, J. 2020 · 2020
Later among the works it cites.
Scene text image super-resolution in the wild
Wang, W.; Xie, E.; Liu, X.; Wang, W.; Liang, D.; Shen, C.; and Bai, X. 2020 · 2020
Later among the works it cites.
PlugNet: Degradation Aware Scene Text Recognition Supervised by a Pluggable Super-Resolution Unit
Yan, R.; and Huang, Y. 2020 · 2020
Later among the works it cites.
A holistic representation guided attention network for scene text recognition
Yang, L.; Wang, P.; Li, H.; Li, Z.; and Zhang, Y. 2020 · 2020
Later among the works it cites.
Text recognition in the wild: A survey
Chen, X.; Jin, L.; Zhu, Y.; Luo, C.; and Wang, T. 2021 · 2021
Closest in time.
Text Prior Guided Scene Text Image Super-resolution
Ma, J.; Guo, S.; and Zhang, L. 2021 · 2021
Closest in time.
Spatial transformer networks
Jaderberg, M.; Simonyan, K.; Zisserman, A.; et al. 2015 · 2025
Closest in time.
Aster: An attentional scene text recognizer with flexible rectification
Shi, B.; Yang, M.; Wang, X.; Lyu, P.; Yao, C.; and Bai, X. 2018 · 2048
Closest in time.