Fetching the paper…
Reading the bibliography…
Unifying text detection and text recognition in an end-to-end training fashion has become a new trend for reading text in the wild, as these two tasks are highly relevant and complementary.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
ICDAR 2003 robust reading competitions
S. M. Lucas, A. Panaretos, L. Sosa, A. Tang, S. Wong, and R. Young · 2003
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
A. Graves, S. Fernández, F. J. Gomez, and J. Schmidhuber · 2006
Earlier work this paper cites.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Earlier work this paper cites.
A method for text localization and recognition in real-world images
L. Neumann and J. Matas · 2010
Earlier work this paper cites.
End-to-end scene text recognition
K. Wang, B. Babenko, and S. Belongie · 2011
Earlier work this paper cites.
Scene text recognition using higher order language priors
A. Mishra, K. Alahari, and C. V. Jawahar · 2012
Earlier work this paper cites.
Top-down and bottom-up cues for scene text recognition
A. Mishra, K. Alahari, and C. V. Jawahar · 2012
Earlier work this paper cites.
Real-time scene text localization and recognition
L. Neumann and J. Matas · 2012
Earlier work this paper cites.
End-to-end text recognition with convolutional neural networks
T. Wang, D. J. Wu, A. Coates, and A. Y. Ng · 2012
Earlier work this paper cites.
Photoocr: Reading text in uncontrolled conditions
A. Bissacco, M. Cummins, Y. Netzer, and H. Neven · 2013
Earlier work this paper cites.
Icdar 2013 robust reading competition
D. Karatzas, F. Shafait, S. Uchida, M. Iwamura, L. G. i Bigorda, S. R. Mestre, J. Mas, D. F. Mota, J. A. Almazan, and L. P. de las Heras · 2013
Earlier work this paper cites.
Recognizing text with perspective distortion in natural scenes
T. Quy Phan, P. Shivakumara, S. Tian, and C. Lim Tan · 2013
Earlier work this paper cites.
Word spotting and recognition with embedded attributes
J. Almazán, A. Gordo, A. Fornés, and E. Valveny · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. B. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Earlier work this paper cites.
Robust scene text detection with convolution neural network induced MSER trees
W. Huang, Y. Qiao, and X. Tang · 2014
Earlier work this paper cites.
Synthetic data and artificial neural networks for natural scene text recognition
M. Jaderberg, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Deep features for text spotting
M. Jaderberg, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
A robust arbitrary text detection system for natural scene images
A. Risnumawan, P. Shivakumara, C. S. Chan, and C. L. Tan · 2014
Earlier work this paper cites.
Accurate scene text recognition based on recurrent neural network
B. Su and S. Lu · 2014
Earlier work this paper cites.
A unified framework for multioriented text detection and recognition
C. Yao, X. Bai, and W. Liu · 2014
Earlier work this paper cites.
Strokelets: A learned multi-scale representation for scene text recognition
C. Yao, X. Bai, B. Shi, and W. Liu · 2014
Earlier work this paper cites.
Long-term recurrent convolutional networks for visual recognition and description
J. Donahue, L. Anne Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell · 2015
Earlier work this paper cites.
Fast R-CNN
R. B. Girshick · 2015
Earlier work this paper cites.
Supervised mid-level features for word image representation
A. Gordo · 2015
Earlier work this paper cites.
Deep structured output learning for unconstrained text recognition
M. Jaderberg, K. Simonyan, A. Vedaldi, and A. Zisserman · 2015
Earlier work this paper cites.
Spatial transformer networks
M. Jaderberg, K. Simonyan, A. Zisserman, et al · 2015
Earlier work this paper cites.
ICDAR 2015 competition on robust reading
D. Karatzas, L. Gomez-Bigorda, A. Nicolaou, S. K. Ghosh, A. D. Bagdanov, M. Iwamura, J. Matas, L. Neumann, V. R. Chandrasekhar, S. Lu, F. Shafait, S. Uchida, and E. Valveny · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
Label embedding: A frugal baseline for text recognition
J. A. Rodríguez-Serrano, A. Gordo, and F. Perronnin · 2015
Cited alongside, same era.
Multi-scale context aggregation by dilated convolutions
F. Yu and V. Koltun · 2015
Cited alongside, same era.
Instance-sensitive fully convolutional networks
J. Dai, K. He, Y. Li, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Synthetic data for text localisation in natural images
A. Gupta, A. Vedaldi, and A. Zisserman · 2016
Towards end-to-end text spotting with convolutional recurrent neural networks
H. Li, P. Wang, and C. Shen · 2017
Later among the works it cites.
Fully convolutional instance-aware semantic segmentation
Y. Li, H. Qi, J. Dai, X. Ji, and Y. Wei · 2017
Later among the works it cites.
Textboxes: A fast text detector with a single deep neural network
M. Liao, B. Shi, X. Bai, X. Wang, and W. Liu · 2017
Later among the works it cites.
Feature pyramid networks for object detection
T. Lin, P. Dollár, R. B. Girshick, K. He, B. Hariharan, and S. J. Belongie · 2017
Later among the works it cites.
ICDAR2017 robust reading challenge on multi-lingual scene text detection and script identification - RRC-MLT
N. Nayef, F. Yin, I. Bizid, H. Choi, Y. Feng, D. Karatzas, Z. Luo, U. Pal, C. Rigaud, J. Chazalon, W. Khlif, M. M. Luqman, J. Burie, C. Liu, and J. Ogier · 2017
Later among the works it cites.
Faster R-CNN: towards real-time object detection with region proposal networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Reading scene text in deep convolutional sequences
P. He, W. Huang, Y. Qiao, C. C. Loy, and X. Tang · 2016
Cited alongside, same era.
Reading text in the wild with convolutional neural networks
M. Jaderberg, K. Simonyan, A. Vedaldi, and A. Zisserman · 2016
Cited alongside, same era.
Recursive recurrent nets with attention modeling for OCR in the wild
C. Lee and S. Osindero · 2016
Cited alongside, same era.
SSD: single shot multibox detector
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. E. Reed, C. Fu, and A. C. Berg · 2016
Cited alongside, same era.
Real-time lexicon-free scene text localization and recognition
L. Neumann and J. Matas · 2016
Cited alongside, same era.
S. Ren, K. He, R. B. Girshick, and J. Sun · 2017
Later among the works it cites.
Detecting oriented text in natural images by linking segments
B. Shi, X. Bai, and S. J. Belongie · 2017
Later among the works it cites.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
B. Shi, X. Bai, and C. Yao · 2017
Later among the works it cites.
Accurate recognition of words in scenes without character segmentation using recurrent neural network
B. Su and S. Lu · 2017
Later among the works it cites.
Attention-based extraction of structured information from street view imagery
Z. Wojna, A. N. Gorban, D.-S. Lee, K. Murphy, Q. Yu, Y. Li, and J. Ibarz · 2017
Later among the works it cites.
Learning to read irregular text with attention mechanisms
X. Yang, D. He, Z. Zhou, D. Kifer, and C. L. Giles · 2017
Later among the works it cites.
Pyramid scene parsing network
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia · 2017
Later among the works it cites.
EAST: an efficient and accurate scene text detector
X. Zhou, C. Yao, H. Wen, Y. Wang, S. Zhou, W. He, and J. Liang · 2017
Later among the works it cites.
Edit probability for scene text recognition
F. Bai, Z. Cheng, Y. Niu, S. Pu, and S. Zhou · 2018
Later among the works it cites.
Integrating scene text and visual appearance for fine-grained image classification
X. Bai, M. Yang, P. Lyu, Y. Xu, and J. Luo · 2018
Later among the works it cites.
Aon: Towards arbitrarily-oriented text recognition
Z. Cheng, Y. Xu, F. Bai, Y. Niu, S. Pu, and S. Zhou · 2018
Later among the works it cites.
Fused text segmentation networks for multi-oriented scene text detection
Y. Dai, Z. Huang, Y. Gao, Y. Xu, K. Chen, J. Guo, and W. Qiu · 2018
Later among the works it cites.
An end-to-end textspotter with explicit alignment and attention
T. He, Z. Tian, W. Huang, C. Shen, Y. Qiao, and C. Sun · 2018
Later among the works it cites.
Textboxes++: A single-shot oriented scene text detector
M. Liao, B. Shi, and X. Bai · 2018
Later among the works it cites.
Rotation-sensitive regression for oriented scene text detection
M. Liao, Z. Zhu, B. Shi, G.-s. Xia, and X. Bai · 2018
Later among the works it cites.
Fots: Fast oriented text spotting with a unified network
X. Liu, D. Liang, S. Yan, D. Chen, Y. Qiao, and J. Yan · 2018
Later among the works it cites.
Textsnake: A flexible representation for detecting text of arbitrary shapes
S. Long, J. Ruan, W. Zhang, X. He, W. Wu, and C. Yao · 2018
Later among the works it cites.
Mask textspotter: An end-to-end trainable neural network for spotting text with arbitrary shapes
P. Lyu, M. Liao, C. Yao, W. Wu, and X. Bai · 2018
Later among the works it cites.
Multi-oriented scene text detection via corner localization and region segmentation
P. Lyu, C. Yao, W. Wu, S. Yan, and X. Bai · 2018
Later among the works it cites.
E2E-MLT - an unconstrained end-to-end method for multi-language scene text
Y. Patel, M. Busta, and J. Matas · 2018
Later among the works it cites.
Aster: An attentional scene text recognizer with flexible rectification
B. Shi, M. Yang, X. Wang, P. Lyu, C. Yao, and X. Bai · 2018
Later among the works it cites.
Accurate scene text detection through border semantics awareness and bootstrapping
C. Xue, S. Lu, and F. Zhan · 2018
Later among the works it cites.