Fetching the paper…
Reading the bibliography…
Existing Scene Text Recognition (STR) methods typically use a language model to optimize the joint probability of the 1D character sequence predicted by a visual recognition (VR) model, which ignore the 2D spatial context of visual semantics within and between character instances, making them not generalize well to arbitrary shape scene text.
Lee, J.; Lee, I.; and Kang, J. 2019 · 1904
Earlier work this paper cites.
Decoupled Attention Network for Text Recognition
Wang, T.; Zhu, Y.; Jin, L.; Luo, C.; Chen, X.; Wu, Y.; Wang, Q.; and Cai, M. 2020 · 1912
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A.; Fernández, S.; Gomez, F.; and Schmidhuber, J. 2006 · 2006
Earlier work this paper cites.
End-to-end scene text recognition
Wang, K.; Babenko, B.; and Belongie, S. 2011 · 2011
Earlier work this paper cites.
Scene text recognition using higher order language priors
Mishra, A.; Alahari, K.; and Jawahar, C. 2012 · 2012
Earlier work this paper cites.
ICDAR 2013 robust reading competition
Karatzas, D.; Shafait, F.; Uchida, S.; Iwamura, M.; i Bigorda, L. G.; Mestre, S. R.; Mas, J.; Mota, D. F.; Almazan, J. A.; and De Las Heras, L. P. 2013 · 2013
Earlier work this paper cites.
Recognizing text with perspective distortion in natural scenes
Phan, T. Q.; Shivakumara, P.; Tian, S.; and Tan, C. L. 2013 · 2013
Earlier work this paper cites.
A robust arbitrary text detection system for natural scene images
Risnumawan, A.; Shivakumara, P.; Chan, C. S.; and Tan, C. L. 2014 · 2014
Earlier work this paper cites.
ICDAR 2015 competition on robust reading
Karatzas, D.; Gomez-Bigorda, L.; Nicolaou, A.; Ghosh, S.; Bagdanov, A.; Iwamura, M.; Matas, J.; Neumann, L.; Chandrasekhar, V. R.; Lu, S.; et al. 2015 · 2015
Earlier work this paper cites.
Synthetic data for text localisation in natural images
Gupta, A.; Vedaldi, A.; and Zisserman, A. 2016 · 2016
Earlier work this paper cites.
Reading text in the wild with convolutional neural networks
Jaderberg, M.; Simonyan, K.; Vedaldi, A.; and Zisserman, A. 2016 · 2016
Earlier work this paper cites.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
Shi, B.; Bai, X.; and Yao, C. 2016 · 2016
Earlier work this paper cites.
Enriching word vectors with subword information
Bojanowski, P.; Grave, E.; Joulin, A.; and Mikolov, T. 2017 · 2017
Earlier work this paper cites.
Focusing attention: Towards accurate text recognition in natural images
Cheng, Z.; Bai, F.; Xu, Y.; Zheng, G.; Pu, S.; and Zhou, S. 2017 · 2017
Earlier work this paper cites.
Semi-Supervised Classification with Graph Convolutional Networks
Kipf, T.; and Welling, M. 2017 · 2017
Cited alongside, same era.
Icdar2017 robust reading challenge on multi-lingual scene text detection and script identification-rrc-mlt
Nayef, N.; Yin, F.; Bizid, I.; Choi, H.; Feng, Y.; Karatzas, D.; Luo, Z.; Pal, U.; Rigaud, C.; Chazalon, J.; et al. 2017 · 2017
Cited alongside, same era.
Mean Teachers are Better Role Models: Weight-Averaged Consistency Targets Improve Semi-supervised Deep Learning Results
Tarvainen, A.; and Valpola, H. 2017 · 2017
Cited alongside, same era.
Attention-based extraction of structured information from street view imagery
Wojna, Z.; Gorban, A. N.; Lee, D.-S.; Murphy, K.; Yu, Q.; Li, Y.; and Ibarz, J. 2017 · 2017
Cited alongside, same era.
Learning to Read Irregular Text with Attention Mechanisms
Yang, X.; He, D.; Zhou, Z.; Kifer, D.; and Giles, C. L. 2017 · 2017
Cited alongside, same era.
Gtc: Guided training of ctc towards efficient and accurate scene text recognition
Hu, W.; Cai, X.; Hou, J.; Yi, S.; and Lin, Z. 2020 · 2020
Later among the works it cites.
Scatter: selective context attentional scene text recognizer
Litman, R.; Anschel, O.; Tsiper, S.; Litman, R.; Mazor, S.; and Manmatha, R. 2020 · 2020
Later among the works it cites.
Seed: Semantics enhanced encoder-decoder framework for scene text recognition
Qiao, Z.; Zhou, Y.; Yang, D.; Zhou, Y.; and Wang, W. 2020 · 2020
Later among the works it cites.
Towards accurate scene text recognition with semantic reasoning networks
Yu, D.; Li, X.; Zhang, C.; Liu, T.; Han, J.; Liu, J.; and Ding, E. 2020 · 2020
Later among the works it cites.
RobustScanner: Dynamically Enhancing Positional Clues for Robust Text Recognition
Yue, X.; Kuang, Z.; Lin, C.; Sun, H.; and Zhang, W. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Aon: Towards arbitrarily-oriented text recognition
Cheng, Z.; Xu, Y.; Bai, F.; Niu, Y.; Pu, S.; and Zhou, S. 2018 · 2018
Cited alongside, same era.
Recurrent Calibration Network for Irregular Text Recognition
Gao, Y.; Chen, Y.; Wang, J.; Lei, Z.; Zhang, X.; and Lu, H. 2018 · 2018
Cited alongside, same era.
What is wrong with scene text recognition model comparisons? dataset and model analysis
Baek, J.; Kim, G.; Lee, J.; Park, S.; Han, D.; Yun, S.; Oh, S. J.; and Lee, H. 2019 · 2019
Cited alongside, same era.
Graph-based global reasoning networks
Chen, Y.; Rohrbach, M.; Yan, Z.; Shuicheng, Y.; Feng, J.; and Kalantidis, Y. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Cited alongside, same era.
Show, attend and read: A simple and strong baseline for irregular text recognition
Li, H.; Wang, P.; Shen, C.; and Zhang, G. 2019 · 2019
Cited alongside, same era.
Scene text recognition from two-dimensional perspective
Liao, M.; Zhang, J.; Wan, Z.; Xie, F.; Liang, J.; Lyu, P.; Yao, C.; and Bai, X. 2019 · 2019
Cited alongside, same era.
Empowering things with intelligence: a survey of the progress, challenges, and opportunities in artificial intelligence of things
Zhang, J.; and Tao, D. 2020 · 2020
Later among the works it cites.
Deep relational reasoning graph network for arbitrary shape text detection
Zhang, S.-X.; Zhu, X.; Hou, J.-B.; Liu, C.; Yang, C.; Wang, H.; and Yin, X.-C. 2020 · 2020
Later among the works it cites.
What If We Only Use Real Datasets for Scene Text Recognition? Toward Scene Text Recognition With Fewer Labels
Baek, J.; Matsui, Y.; and Aizawa, K. 2021 · 2021
Closest in time.
Fang, S.; Xie, H.; Wang, Y.; Mao, Z.; and Zhang, Y. 2021 · 2021
Closest in time.
ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive Bias
Xu, Y.; Zhang, Q.; Zhang, J.; and Tao, D. 2021 · 2021
Closest in time.
Primitive Representation Learning for Scene Text Recognition
Yan, R.; Peng, L.; Xiao, S.; and Yao, G. 2021 · 2021
Closest in time.
I3CL: Intra-and Inter-Instance Collaborative Learning for Arbitrary-shaped Scene Text Detection
Ye, J.; Zhang, J.; Liu, J.; Du, B.; and Tao, D. 2021 · 2021
Closest in time.
Spatial transformer networks
Jaderberg, M.; Simonyan, K.; Zisserman, A.; and Kavukcuoglu, K. 2015 · 2025
Closest in time.
Aster: An attentional scene text recognizer with flexible rectification
Shi, B.; Yang, M.; Wang, X.; Lyu, P.; Yao, C.; and Bai, X. 2018 · 2048
Closest in time.