Fetching the paper…
Reading the bibliography…
Dominant scene text recognition models commonly contain two building blocks, a visual model for feature extraction and a sequence model for text transcription.
End-to-end scene text recognition
K. Wang, B. Babenko, and S. Belongie · 2011
Earlier work this paper cites.
Scene text recognition using higher order language priors
A. Mishra, A. Karteek, and C. V. Jawahar · 2012
Earlier work this paper cites.
Icdar 2013 robust reading competition
D. KaratzasAU, F. ShafaitAU, S. UchidaAU, M. IwamuraAU, L. G. i. BigordaAU, S. R. MestreAU, J. MasAU, D. F. MotaAU, J. A. AlmazànAU, and L. P. de las Heras · 2013
Earlier work this paper cites.
Recognizing text with perspective distortion in natural scenes
T. Q. Phan, P. Shivakumara, S. Tian, and C. L. Tan · 2013
Earlier work this paper cites.
A robust arbitrary text detection system for natural scene images
R. Anhar, S. Palaiahnakote, C. S. Chan, and C. L. Tan · 2014
Earlier work this paper cites.
Synthetic data and artificial neural networks for natural scene text recognition
M. Jaderberg, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Reading text in the wild with convolutional neural networks
M. Jaderberg, K. Simonyan, A. Vedaldi, and A. Zisserman · 2015
Earlier work this paper cites.
Icdar 2015 competition on robust reading
D. Karatzas, L. Gomez-Bigorda, A. Nicolaou, S. Ghosh, A. Bagdanov, M. Iwamura, J. Matas, L. Neumann, V. R. Chandrasekhar, S. Lu, F. Shafait, S. Uchida, and E. Valveny · 2015
Earlier work this paper cites.
Synthetic data for text localisation in natural images
A. Gupta, A. Vedaldi, and A. Zisserman · 2016
Earlier work this paper cites.
Chinese image text recognition with blstm-ctc: A segmentation-free method
C. Zhai, Z. Chen, J. Li, and B. Xu · 2016
Earlier work this paper cites.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
B. Shi, X. Bai, and C. Yao · 2017
Cited alongside, same era.
Rosetta: Large scale system for text detection and recognition in images
F. Borisyuk, A. Gordo, and V. Sivakumar · 2018
Cited alongside, same era.
What is wrong with scene text recognition model comparisons? dataset and model analysis
J. Baek, G. Kim, J. Lee, S. Park, D. Han, S. Yun, S. J. Oh, and H. Lee · 2019
Cited alongside, same era.
Show, attend and read: A simple and strong baseline for irregular text recognition
H. Li, P. Wang, C. Shen, and G. Zhang · 2019
Cited alongside, same era.
A multi-object rectified attention network for scene text recognition
C. Luo, L. Jin, and Z. Sun · 2019
Cited alongside, same era.
Nrtr: A no-recurrence sequence-to-sequence model for scene text recognition
Vision transformer for fast and efficient scene text recognition
R. Atienza · 2021
Later among the works it cites.
Benchmarking chinese text recognition: Datasets, baselines, and an empirical study
J. Chen, H. Yu, J. Ma, M. Guan, X. Xu, X. Wang, S. Qu, B. Li, and X. Xue · 2021
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, J. Uszkoreit, and N. Houlsby · 2021
Later among the works it cites.
Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition
S. Fang, H. Xie, Y. Wang, Z. Mao, and Y. Zhang · 2021
Later among the works it cites.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Sheng, Z. Chen, and B. Xu · 2019
Cited alongside, same era.
Aster: An attentional scene text recognizer with flexible rectification
B. Shi, M. Yang, X. Wang, P. Lyu, C. Yao, and X. Bai · 2019
Cited alongside, same era.
Gtc: Guided training of ctc towards efficient and accurate scene text recognition
W. Hu, X. Cai, J. Hou, S. Yi, and Z. Lin · 2020
Cited alongside, same era.
Towards accurate scene text recognition with semantic reasoning networks
D. Yu, X. Li, C. Zhang, T. Liu, J. Han, J. Liu, and E. Ding · 2020
Cited alongside, same era.
Autostr: Efficient backbone search for scene text recognition
H. Zhang, Q. Yao, M. Yang, Y. Xu, and X. Bai · 2020
Cited alongside, same era.
Later among the works it cites.
From two to one: A new scene text recognizer with visual language modeling network
Y. Wang, H. Xie, S. Fang, J. Wang, S. Zhu, and Y. Zhang · 2021
Later among the works it cites.
Primitive representation learning for scene text recognition
R. Yan, L. Peng, S. Xiao, and G. Yao · 2021
Later among the works it cites.
Cdistnet: Perceiving multi-domain character distance for robust text recognition
T. Zheng, Z. Chen, S. Fang, H. Xie, and Y. Jiang · 2021
Later among the works it cites.
Visual-semantic transformer for scene text recognition
X. Tang, Y. Lai, Y. Liu, Y. Fu, and R. Fang · 2022
Closest in time.