Fetching the paper…
Reading the bibliography…
Irregular scene text, which has complex layout in 2D space, is challenging to most previous scene text recognizers.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, P. Haffner, et al · 1998
Earlier work this paper cites.
ICDAR 2003 robust reading competitions: entries, results, and future directions
S. M. Lucas, A. Panaretos, L. Sosa, A. Tang, S. Wong, R. Young, K. Ashida, H. Nagai, M. Okamoto, H. Yamamoto, H. Miyao, J. Zhu, W. Ou, C. Wolf, J. Jolion, L. Todoran, M. Worring, and X. Lin · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
A. Graves, S. Fernández, F. J. Gomez, and J. Schmidhuber · 2006
Earlier work this paper cites.
End-to-end scene text recognition
K. Wang, B. Babenko, and S. J. Belongie · 2011
Earlier work this paper cites.
Scene text recognition using higher order language priors
A. Mishra, K. Alahari, and C. V. Jawahar · 2012
Earlier work this paper cites.
Top-down and bottom-up cues for scene text recognition
A. Mishra, K. Alahari, and C. V. Jawahar · 2012
Earlier work this paper cites.
Real-time scene text localization and recognition
L. Neumann and J. Matas · 2012
Earlier work this paper cites.
End-to-end text recognition with convolutional neural networks
T. Wang, D. J. Wu, A. Coates, and A. Y. Ng · 2012
Earlier work this paper cites.
ADADELTA: an adaptive learning rate method
M. D. Zeiler · 2012
Earlier work this paper cites.
Photoocr: Reading text in uncontrolled conditions
A. Bissacco, M. Cummins, Y. Netzer, and H. Neven · 2013
Earlier work this paper cites.
ICDAR 2013 robust reading competition
D. Karatzas, F. Shafait, S. Uchida, M. Iwamura, L. G. i Bigorda, S. R. Mestre, J. Mas, D. F. Mota, J. Almazán, and L. de las Heras · 2013
Earlier work this paper cites.
Recognizing text with perspective distortion in natural scenes
T. Q. Phan, P. Shivakumara, S. Tian, and C. L. Tan · 2013
Earlier work this paper cites.
Word spotting and recognition with embedded attributes
J. Almazán, A. Gordo, A. Fornés, and E. Valveny · 2014
Earlier work this paper cites.
Deep structured output learning for unconstrained text recognition
M. Jaderberg, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Synthetic data and artificial neural networks for natural scene text recognition
M. Jaderberg, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Deep features for text spotting
M. Jaderberg, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Region-based discriminative feature pooling for scene text recognition
C. Lee, A. Bhardwaj, W. Di, V. Jagadeesh, and R. Piramuthu · 2014
Cited alongside, same era.
A robust arbitrary text detection system for natural scene images
A. Risnumawan, P. Shivakumara, C. S. Chan, and C. L. Tan · 2014
Cited alongside, same era.
Accurate scene text recognition based on recurrent neural network
B. Su and S. Lu · 2014
Cited alongside, same era.
A unified framework for multioriented text detection and recognition
C. Yao, X. Bai, and W. Liu · 2014
Cited alongside, same era.
Strokelets: A learned multi-scale representation for scene text recognition
C. Yao, X. Bai, B. Shi, and W. Liu · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2015
Focusing attention: Towards accurate text recognition in natural images
Z. Cheng, F. Bai, Y. Xu, G. Zheng, S. Pu, and S. Zhou · 2017
Later among the works it cites.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
B. Shi, X. Bai, and C. Yao · 2017
Later among the works it cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Later among the works it cites.
Learning to read irregular text with attention mechanisms
X. Yang, D. He, Z. Zhou, D. Kifer, and C. L. Giles · 2017
Later among the works it cites.
Edit probability for scene text recognition
F. Bai, Z. Cheng, Y. Niu, S. Pu, and S. Zhou · 2018
Later among the works it cites.
AON: towards arbitrarily-oriented text recognition
Z. Cheng, Y. Xu, F. Bai, Y. Niu, S. Pu, and S. Zhou · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Attention-based models for speech recognition
J. Chorowski, D. Bahdanau, D. Serdyuk, K. Cho, and Y. Bengio · 2015
Cited alongside, same era.
Supervised mid-level features for word image representation
A. Gordo · 2015
Cited alongside, same era.
Spatial transformer networks
M. Jaderberg, K. Simonyan, A. Zisserman, and K. Kavukcuoglu · 2015
Cited alongside, same era.
ICDAR 2015 competition on robust reading
D. Karatzas, L. Gomez-Bigorda, A. Nicolaou, S. K. Ghosh, A. D. Bagdanov, M. Iwamura, J. Matas, L. Neumann, V. R. Chandrasekhar, S. Lu, F. Shafait, S. Uchida, and E. Valveny · 2015
Cited alongside, same era.
Label embedding: A frugal baseline for text recognition
J. A. Rodríguez-Serrano, A. Gordo, and F. Perronnin · 2015
Cited alongside, same era.
Synthetic data for text localisation in natural images
A. Gupta, A. Vedaldi, and A. Zisserman · 2016
Cited alongside, same era.
Later among the works it cites.
BERT: pre-training of deep bidirectional transformers for language understanding
J. Devlin, M. Chang, K. Lee, and K. Toutanova · 2018
Later among the works it cites.
An end-to-end textspotter with explicit alignment and attention
T. He, Z. Tian, W. Huang, C. Shen, Y. Qiao, and C. Sun · 2018
Later among the works it cites.
Relation networks for object detection
H. Hu, J. Gu, Z. Zhang, J. Dai, and Y. Wei · 2018
Later among the works it cites.
Scene text recognition from two-dimensional perspective
M. Liao, J. Zhang, Z. Wan, F. Xie, J. Liang, P. Lyu, C. Yao, and X. Bai · 2018
Later among the works it cites.
Char-net: A character-aware neural network for distorted scene text recognition
W. Liu, C. Chen, and K. K. Wong · 2018
Later among the works it cites.
FOTS: fast oriented text spotting with a unified network
X. Liu, D. Liang, S. Yan, D. Chen, Y. Qiao, and J. Yan · 2018
Later among the works it cites.
Mask textspotter: An end-to-end trainable neural network for spotting text with arbitrary shapes
P. Lyu, M. Liao, C. Yao, W. Wu, and X. Bai · 2018
Later among the works it cites.
Aster: An attentional scene text recognizer with flexible rectification
B. Shi, M. Yang, X. Wang, P. Lyu, C. Yao, and X. Bai · 2018
Later among the works it cites.
Show, attend and read: A simple and strong baseline for irregular text recognition
H. Li, P. Wang, C. Shen, and G. Zhang · 2019
Closest in time.
Moran: A multi-object rectified attention network for scene text recognition
C. Luo, L. Jin, and Z. Sun · 2019
Closest in time.