Fetching the paper…
Reading the bibliography…
Attention-based scene text recognizers have gained huge success, which leverages a more compact intermediate representation to learn 1d- or 2d- attention by a RNN-based encoder-decoder architecture.
Y. Cao, J. Xu, S. Lin, F. Wei, H. Hu, GCNet: Non-local networks meet squeeze-excitation networks and beyond, in: Proceedings of International Conference on Computer Vision Workshop, 2019, pp. 1971–1980
1980
Earlier work this paper cites.
F. L. Bookstein, Principal warps: thin-plate splines and the decomposition of deformations, IEEE Trans. Pattern Anal. Mach. Intell. 11 (6) (1989) 567–585
1989
Earlier work this paper cites.
S. M. Lucas, A. Panaretos, L. Sosa, A. Tang, S. Wong, R. Young, K. Ashida, H. Nagai, M. Okamoto, H. Yamamoto, H. Miyao, J. Zhu, W. Ou, C. Wolf, J.-M. Jolion, L. Todoran, M. Worring, X. Lin, ICDAR 2003 robust reading competitions: entries, results, and future directions, International Journal of Document Analysis and Recognition 7 (2004) 105–122
2004
Earlier work this paper cites.
P. Shivakumara, S. Bhowmick, B. Su, C. L. Tan, U. Pal, A New Gradient Based Character Segmentation Method for Video Text Recognition, in: Proceedings of International Conference on Document Analysis and Recognition, 2011, pp. 126–130
2011
Earlier work this paper cites.
K. Wang, B. Babenko, S. J. Belongie, End-to-end scene text recognition, in: Proceedings of International Conference on Computer Vision, 2011, pp. 1457–1464
2011
Earlier work this paper cites.
A. Mishra, K. Alahari, C. V. Jawahar, Top-down and bottom-up cues for scene text recognition, in: Proceedings of Computer Vision and Pattern Recognition, 2012, pp. 2687–2694
2012
Earlier work this paper cites.
D. Karatzas, F. Shafait, S. Uchida, M. Iwamura, L. G. i Bigorda, S. R. Mestre, J. M. Romeu, D. F. Mota, J. Almazán, L.-P. de las Heras, ICDAR 2013 Robust Reading Competition, in: Proceedings of International Conference on Document Analysis and Recognition, 2013, pp. 1484–1493
2013
Earlier work this paper cites.
T. Q. Phan, P. Shivakumara, S. Tian, C. L. Tan, Recognizing text with perspective distortion in natural scenes, in: Proceedings of International Conference on Computer Vision, 2013, pp. 569–576
2013
Earlier work this paper cites.
M. Jaderberg, K. Simonyan, A. Vedaldi, A. Zisserman, Synthetic data and artificial neural networks for natural scene text recognition, in: Proceedings of Advances in Neural Information Processing Systems Workshop, 2014
2014
Earlier work this paper cites.
A. Risnumawan, P. Shivakumara, C. S. Chan, C. L. Tan, A robust arbitrary text detection system for natural scene images, Expert Syst. Appl. 41 (2014) 8027–8048
2014
Earlier work this paper cites.
J. Gu, Z. Wang, J. Kuen, L. Ma, A. Shahroudy, B. Shuai, T. Liu, X. Wang, G. Wang, J. Cai, T. Chen, Recent advances in convolutional neural networks, Pattern Recognit. 77 (2015) 354–377
2015
Earlier work this paper cites.
B. Shi, X. Bai, C. Yao, An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition, IEEE Trans. Pattern Anal. Mach. Intell. 39 (2015) 2298–2304
2015
Earlier work this paper cites.
Q. Ye, D. Doermann, Text detection and recognition in imagery: A survey, IEEE Trans. Pattern Anal. Mach. Intell. 37 (7) (2015) 1480–1500
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, J. Sun, Deep residual learning for image recognition, in: Proceedings of Computer Vision and Pattern Recognition, 2015, pp. 770–778
2015
Earlier work this paper cites.
D. P. Kingma, J. Ba, Adam: A method for stochastic optimization, in: International Conference for Learning Representations, 2015
2015
Earlier work this paper cites.
Y. Zhu, C. Yao, X. Bai, Scene text detection and recognition: recent advances and future trends, Frontiers of Computer Science 10 (1) (2016) 19–36
2016
Earlier work this paper cites.
A. Mishra, K. Alahari, C. Jawahar, Enhancing energy minimization framework for scene text recognition with top-down cues, Computer Vision and Image Understanding 145 (2016) 30–42
2016
Earlier work this paper cites.
L. G. i Bigorda, D. Karatzas, Textproposals: A text-specific selective search algorithm for word spotting in the wild, Pattern Recognit. 70 (2016) 60–74
2016
Earlier work this paper cites.
B. Shi, X. Wang, P. Lyu, C. Yao, X. Bai, Robust scene text recognition with automatic rectification, in: Proceedings of Computer Vision and Pattern Recognition, 2016, pp. 4168–4176
2016
Earlier work this paper cites.
J. Ba, J. R. Kiros, G. E. Hinton, Layer normalization, in: Advances in neural information processing systems (NIPS), 2016
2016
Cited alongside, same era.
A. Gupta, A. Vedaldi, A. Zisserman, Synthetic data for text localisation in natural images, in: Proceedings of Computer Vision and Pattern Recognition, 2016
2016
Cited alongside, same era.
W. Liu, C. Chen, K.-Y. K. Wong, Z. Su, J. Han, STAR-Net: A spatial attention residue network for scene text recognition, in: Proceedings of British Machine Vision Conference, 2016
2016
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, I. Polosukhin, Attention is all you need, in: Proceedings of Advances in Neural Information Processing Systems, 2017, pp. 5998–6008
2017
Cited alongside, same era.
D. N. Van, S. Lu, S. Tian, N. Ouarti, M. Mokhtari, A pooling based scene text proposal technique for scene text reading in the wild, Pattern Recognit. 87 (2019) 118–129
2019
Closest in time.
H. Li, P. Wang, C. Shen, G. Zhang, Show, Attend and Read: A simple and strong baseline for irregular text recognition, in: Proceedings of Association for the Advancement of Artificial Intelligence, 2019, pp. 8610–8617
2019
Closest in time.
M. Yang, Y. Guan, M. Liao, X. He, K. Bian, S. Bai, C. Yao, X. Bai, Symmetry-constrained rectification network for scene text recognition, in: 2019 IEEE/CVF International Conference on Computer Vision (ICCV), 2019, pp. 9146–9155
2019
Closest in time.
M. Liao, J. Zhang, Z. Wan, F. Xie, J. Liang, P. Lyu, C. Yao, X. Bai, Scene text recognition from two-dimensional perspective, in: AAAI, 2019
2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
H. Hu, J. Gu, Z. Zhang, J. Dai, Y. Wei, Relation networks for object detection, in: Proceedings of Computer Vision and Pattern Recognition, 2017, pp. 3588–3597
2017
Cited alongside, same era.
J. Hu, L. Shen, G. Sun, Squeeze-and-excitation networks, in: Proceedings of Computer Vision and Pattern Recognition, 2017, pp. 7132–7141
2017
Cited alongside, same era.
X.-Y. Zhang, Y. Bengio, C.-L. Liu, Online and offline handwritten chinese character recognition: A comprehensive study and new benchmark, Pattern Recognit. 61 (2017) 348–360
2017
Cited alongside, same era.
A. P. Giotis, G. Sfikas, B. Gatos, C. Nikou, A survey of document image word spotting techniques, Pattern Recognit. 68 (2017) 310–332
2017
Cited alongside, same era.
B. Su, S. Lu, Accurate recognition of words in scenes without character segmentation using recurrent neural network, Pattern Recognit. 63 (2017) 397–405
2017
Cited alongside, same era.
Z. Cheng, F. Bai, Y. Xu, G. Zheng, S. Pu, S. Zhou, Focusing Attention: Towards accurate text recognition in natural images, in: Proceedings of International Conference on Computer Vision, 2017, pp. 5086–5094
2017
Cited alongside, same era.
J. Wang, X. Hu, Gated recurrent convolution neural network for ocr, in: Proceedings of Advances in Neural Information Processing Systems, 2017, pp. 335–344
2017
Cited alongside, same era.
H. Xie, S. Fang, Z. Zha, Y. Yang, Y. Li, Y. Zhang, Convolutional attention networks for scene text recognition, ACM Trans. Multim. Comput. Commun. Appl. 15 (2019) 3:1–3:17
2019
Closest in time.
J. Devlin, M.-W. Chang, K. Lee, K. Toutanova, BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding, in: Proceedings of The North American Chapter of the Association for Computational Linguistics, 2019
2019
Closest in time.
Z. Yang, Z. Dai, Y. Yang, J. G. Carbonell, R. Salakhutdinov, Q. V. Le, Xlnet: Generalized autoregressive pretraining for language understanding, in: Proceedings of Advances in Neural Information Processing Systems, 2019
2019
Closest in time.
Y. Gao, Y. Chen, J. Wang, M. Tang, H. Lu, Reading scene text with fully convolutional sequence modeling, Neurocomputing 339 (2019) 161–170
2019
Closest in time.
Y. Zhang, S. Nie, W. Liu, X. Xu, D. Zhang, H. T. Shen, Sequence-to-sequence domain adaptation network for robust text image recognition, in: Proceedings of Computer Vision and Pattern Recognition, 2019, pp. 2740–2749
2019
Closest in time.
C. Luo, L. Jin, Z. Sun, MORAN: A Multi-Object Rectified Attention Network for scene text recognition, Pattern Recognit. 90 (2019) 109–118
2019
Closest in time.
M. Liao, P. Lyu, M. He, C. Yao, W. Wu, X. Bai, Mask TextSpotter: An end-to-end trainable neural network for spotting text with arbitrary shapes, IEEE Trans. Pattern Anal. Mach. Intell. (2019)
2019
Closest in time.
2019
Closest in time.
J. Lee, S. Park, J. Baek, S. J. Oh, S. Kim, H. Lee, On recognizing texts of arbitrary shapes with 2d self-attention, in: 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), 2020, pp. 2326–2335
2020
Closest in time.
L. Yang, P. Wang, H. Li, Z. Li, Y. Zhang, A holistic representation guided attention network for scene text recognition, Neurocomputing (2020)
2020
Closest in time.
M. Jaderberg, K. Simonyan, A. Zisserman, K. Kavukcuoglu, Spatial transformer networks, in: Proceedings of Advances in Neural Information Processing Systems, 2015, pp. 2017–2025
2025
Closest in time.
B. Shi, M. Yang, X. Wang, P. Lyu, C. Yao, X. Bai, ASTER: An attentional scene text recognizer with flexible rectification, IEEE Trans. Pattern Anal. Mach. Intell. 41 (2018) 2035–2048
2048
Closest in time.
K. Xu, J. Ba, R. Kiros, K. Cho, A. C. Courville, R. Salakhutdinov, R. S. Zemel, Y. Bengio, Show, Attend and Tell: Neural Image Caption Generation with Visual Attention, in: Proceedings of International Conference on Machine Learning, 2015, pp. 2048–2057
2057
Closest in time.
F. Zhan, S. Lu, ESIR: End-to-end scene text recognition via iterative image rectification, in: Proceedings of Computer Vision and Pattern Recognition, 2019, pp. 2059–2068
2068
Closest in time.