Fetching the paper…
Reading the bibliography…
We propose to improve text recognition from a new perspective by separating the text content from complex backgrounds.
Otsu N (1979) A threshold selection method from gray-level histograms. IEEE Transactions on Systems, Man, and Cybernetics (TSMC) 9(1):62–66
1979
Earlier work this paper cites.
Casey RG, Lecolinet E (1996) A survey of methods and strategies in character segmentation. IEEE Transactions on Pattern Analysis and Machine Intelligence 18(7):690–706
1996
Earlier work this paper cites.
Lucas SM, Panaretos A, Sosa L, Tang A, Wong S, Young R (2003) ICDAR 2003 robust reading competitions. In: International Conference on Document Analysis and Recognition (ICDAR), pp 682–687
2003
Earlier work this paper cites.
Dalal N, Triggs B (2005) Histograms of oriented gradients for human detection. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 886–893
2005
Earlier work this paper cites.
Wang K, Babenko B, Belongie S (2011) End-to-end scene text recognition. In: IEEE International Conference on Computer Vision (ICCV), pp 1457–1464
2011
Earlier work this paper cites.
Mishra A, Alahari K, Jawahar C (2012) Scene text recognition using higher order language priors. In: British Machine Vision Conference (BMVC), pp 1–11
2012
Earlier work this paper cites.
Neumann L, Matas J (2012) Real-time scene text localization and recognition. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 3538–3545
2012
Earlier work this paper cites.
Wang T, Wu DJ, Coates A, Ng AY (2012) End-to-end text recognition with convolutional neural networks. In: IEEE International Conference on Pattern Recognition (ICPR), pp 3304–3308
2012
Earlier work this paper cites.
Bissacco A, Cummins M, Netzer Y, Neven H (2013) PhotoOCR: Reading text in uncontrolled conditions. In: IEEE International Conference on Computer Vision (ICCV), pp 785–792
2013
Earlier work this paper cites.
Karatzas D, Shafait F, Uchida S, Iwamura M, i Bigorda LG, Mestre SR, Mas J, Mota DF, Almazan JA, De Las Heras LP (2013) ICDAR 2013 robust reading competition. In: International Conference on Document Analysis and Recognition (ICDAR), pp 1484–1493
2013
Earlier work this paper cites.
Quy Phan T, Shivakumara P, Tian S, Lim Tan C (2013) Recognizing text with perspective distortion in natural scenes. In: IEEE International Conference on Computer Vision (ICCV), pp 569–576
2013
Earlier work this paper cites.
Goodfellow I, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, Ozair S, Courville A, Bengio Y (2014) Generative adversarial nets. In: Neural Information Processing Systems (NeurIPS), pp 2672–2680
2014
Earlier work this paper cites.
Risnumawan A, Shivakumara P, Chan CS, Tan CL (2014) A robust arbitrary text detection system for natural scene images. Expert Systems with Applications 41(18):8027–8048
2014
Earlier work this paper cites.
Su B, Lu S (2014) Accurate scene text recognition based on recurrent neural network. In: Asian Conference on Computer Vision (ACCV), pp 35–48
2014
Earlier work this paper cites.
Sutskever I, Vinyals O, Le QV (2014) Sequence to sequence learning with neural networks. In: Neural Information Processing Systems (NeurIPS), pp 3104–3112
2014
Earlier work this paper cites.
Yao C, Bai X, Shi B, Liu W (2014) Strokelets: A learned multi-scale representation for scene text recognition. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 4042–4049
2014
Earlier work this paper cites.
Bahdanau D, Cho K, Bengio Y (2015) Neural machine translation by jointly learning to align and translate. In: International Conference on Learning Representations (ICLR)
2015
Earlier work this paper cites.
Gordo A (2015) Supervised mid-level features for word image representation. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 2956–2964
2015
Earlier work this paper cites.
Jaderberg M, Simonyan K, Vedaldi A, Zisserman A (2015) Deep structured output learning for unconstrained text recognition. In: International Conference on Learning Representations (ICLR)
2015
Earlier work this paper cites.
Karatzas D, Gomez-Bigorda L, Nicolaou A, Ghosh S, Bagdanov A, Iwamura M, Matas J, Neumann L, Chandrasekhar VR, Lu S, et al. (2015) ICDAR 2015 competition on robust reading. In: International Conference on Document Analysis and Recognition (ICDAR), pp 1156–1160
2015
Cited alongside, same era.
Kingma D, Ba L, et al. (2015) Adam: A method for stochastic optimization. In: International Conference on Learning Representations (ICLR)
2015
Cited alongside, same era.
Rodriguez-Serrano JA, Gordo A, Perronnin F (2015) Label embedding: A frugal baseline for text recognition. International Journal of Computer Vision (IJCV) 113(3):193–207
2015
Cited alongside, same era.
Ye Q, Doermann D (2015) Text detection and recognition in imagery: A survey. IEEE Transactions on Pattern Analysis and Machine Intelligence 37(7):1480–1500
2015
Cited alongside, same era.
Odena A, Olah C, Shlens J (2017) Conditional image synthesis with auxiliary classifier GANs. In: International Conference on Machine Learning (ICML), pp 2642–2651
2017
Later among the works it cites.
Paszke A, Gross S, Chintala S, Chanan G, Yang E, DeVito Z, Lin Z, Desmaison A, Antiga L, Lerer A (2017) Automatic differentiation in PyTorch. In: Neural Information Processing Systems (NeurIPS) Autodiff Workshop
2017
Later among the works it cites.
Shi B, Bai X, Yao C (2017) An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence 39(11):2298–2304
2017
Later among the works it cites.
Yang X, He D, Zhou Z, Kifer D, Giles CL (2017) Learning to read irregular text with attention mechanisms. In: International Joint Conferences on Artificial Intelligence (IJCAI), pp 3280–3286
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gupta A, Vedaldi A, Zisserman A (2016) Synthetic data for text localisation in natural images. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 2315–2324
2016
Cited alongside, same era.
Jaderberg M, Simonyan K, Vedaldi A, Zisserman A (2016) Reading text in the wild with convolutional neural networks. International Journal of Computer Vision (IJCV) 116(1):1–20
2016
Cited alongside, same era.
Johnson J, Alahi A, Fei-Fei L (2016) Perceptual losses for real-time style transfer and super-resolution. In: European Conference on Computer Vision (ECCV), pp 694–711
2016
Cited alongside, same era.
Lee CY, Osindero S (2016) Recursive recurrent nets with attention modeling for OCR in the wild. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 2231–2239
2016
Cited alongside, same era.
Liu W, Chen C, Wong KYK, Su Z, Han J (2016) STAR-Net: A spatial attention residue network for scene text recognition. In: British Machine Vision Conference (BMVC), pp 7–7
2016
Cited alongside, same era.
Salimans T, Goodfellow I, Zaremba W, Cheung V, Radford A, Chen X (2016) Improved techniques for training GANs. In: Neural Information Processing Systems (NeurIPS), pp 2234–2242
2016
Cited alongside, same era.
Shi B, Wang X, Lyu P, Yao C, Bai X (2016) Robust scene text recognition with automatic rectification. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 4168–4176
2016
Cited alongside, same era.
Zhu Y, Yao C, Bai X (2016) Scene text detection and recognition: Recent advances and future trends. Frontiers of Computer Science 10(1):19–36
2016
Cited alongside, same era.
Zhu JY, Park T, Isola P, Efros AA (2017) Unpaired Image-to-Image Translation Using Cycle-Consistent Adversarial Networks. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 2242–2251
2017
Later among the works it cites.
Azadi S, Fisher M, Kim VG, Wang Z, Shechtman E, Darrell T (2018) Multi-content gan for few-shot font style transfer. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 7564–7573
2018
Later among the works it cites.
Bai F, Cheng Z, Niu Y, Pu S, Zhou S (2018) Edit probability for scene text recognition. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 1508–1516
2018
Later among the works it cites.
Cheng Z, Xu Y, Bai F, Niu Y, Pu S, Zhou S (2018) AON: Towards arbitrarily-oriented text recognition. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 5571–5579
2018
Later among the works it cites.
Shi B, Yang M, Wang X, Lyu P, Yao C, Bai X (2018) ASTER: An attentional scene text recognizer with flexible rectification. IEEE Transactions on Pattern Analysis and Machine Intelligence
2018
Later among the works it cites.
Bau D, Zhu JY, Wulff J, Peebles W, Strobelt H, Zhou B, Torralba A (2019) Seeing what a gan cannot generate. In: IEEE International Conference on Computer Vision (ICCV), pp 4502–4511
2019
Later among the works it cites.
Cheng MM, Liu XC, Wang J, Lu SP, Lai YK, Rosin PL (2019) Structure-Preserving Neural Style Transfer. IEEE Transactions on Image Processing (TIP) 29:909–920
2019
Later among the works it cites.
Cong F, Hu W, Huo Q, Guo L (2019) A Comparative Study of Attention-Based Encoder-Decoder Approaches to Natural Scene Text Recognition. In: International Conference on Document Analysis and Recognition (ICDAR), pp 916–921
2019
Later among the works it cites.
Fang S, Xie H, Chen J, Tan J, Zhang Y (2019) Learning to Draw Text in Natural Images with Conditional Adversarial Networks. In: International Joint Conferences on Artificial Intelligence (IJCAI)
2019
Later among the works it cites.
Jing Y, Yang Y, Feng Z, Ye J, Yu Y, Song M (2019) Neural Style Transfer: A Review. IEEE Transactions on Visualization and Computer Graphics (TVCG)
2019
Later among the works it cites.
Li H, Wang P, Shen C, Zhang G (2019) Show, attend and read: A simple and strong baseline for irregular text recognition. In: AAAI Conference on Artificial Intelligence (AAAI)
2019
Later among the works it cites.
Luo C, Jin L, Sun Z (2019) MORAN: A multi-object rectified attention network for scene text recognition. Pattern Recognition 90:109–118
2019
Later among the works it cites.
Wu L, Zhang C, Liu J, Han J, Liu J, Ding E, Bai X (2019) Editing Text in the Wild. In: ACM International Conference on Multimedia (ACM MM), pp 1500–1508
2019
Later among the works it cites.
Zhang Y, Nie S, Liu W, Xu X, Zhang D, Shen HT (2019) Sequence-to-sequence domain adaptation network for robust text image recognition. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 2740–2749
2019
Later among the works it cites.
Zhan F, Lu S (2019) ESIR: End-to-end scene text recognition via iterative image rectification. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 2059–2068
2068
Closest in time.