Fetching the paper…
Reading the bibliography…
In this paper, we propose a Text-Degradation Invariant Auto Encoder (Text-DIAE), a self-supervised model designed to tackle two tasks, text recognition (handwritten or scene-text) and document image enhancement.
Revisiting self-supervised visual representation learning
Kolesnikov, A.; Zhai, X.; and Beyer, L. 2019 · 1929
Earlier work this paper cites.
The anatomy of conscious vision: an fMRI study of visual hallucinations
Howard, R.; Brammer, M.; David, A.; Woodruff, P.; Williams, S.; et al. 1998 · 1998
Earlier work this paper cites.
Adaptive document image binarization
Sauvola, J.; and Pietikäinen, M. 2000 · 2000
Earlier work this paper cites.
The IAM-database: an English sentence database for offline handwriting recognition
Marti, U.-V.; and Bunke, H. 2002 · 2002
Earlier work this paper cites.
Distance-reciprocal distortion measure for binary document images
Lu, H.; Kot, A. C.; and Shi, Y. Q. 2004 · 2004
Earlier work this paper cites.
Pay attention to what you read: Non-recurrent handwritten text-line recognition
Kang, L.; Riba, P.; Rusiñol, M.; Fornés, A.; and Villegas, M. 2020a · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A.; Fernández, S.; Gomez, F.; and Schmidhuber, J. 2006 · 2006
Earlier work this paper cites.
An overview of the Tesseract OCR engine
Smith, R. 2007 · 2007
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
Vincent, P.; Larochelle, H.; Bengio, Y.; and Manzagol, P.-A. 2008 · 2008
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Glorot, X.; and Bengio, Y. 2010 · 2010
Earlier work this paper cites.
ICDAR 2011 document image binarization contest (DIBCO 2011)
I. Pratikakis, K. N., B. Gatos. 2011 · 2011
Earlier work this paper cites.
Scene text recognition using higher order language priors
Mishra, A.; Alahari, K.; and Jawahar, C. 2012 · 2012
Earlier work this paper cites.
Performance evaluation methodology for historical document image binarization
Ntirogiannis, K.; Gatos, B.; and Pratikakis, I. 2012 · 2012
Earlier work this paper cites.
ICFHR 2012 competition on handwritten document image binarization (H-DIBCO 2012)
Pratikakis, I.; Gatos, B.; and Ntirogiannis, K. 2012 · 2012
Earlier work this paper cites.
ICDAR 2013 robust reading competition
Karatzas, D.; Shafait, F.; Uchida, S.; Iwamura, M.; i Bigorda, L. G.; Mestre, S. R.; Mas, J.; Mota, D. F.; Almazan, J. A.; and De Las Heras, L. P. 2013 · 2013
Earlier work this paper cites.
Cvl-database: An off-line database for writer retrieval, writer identification and word spotting
Kleber, F.; Fiel, S.; Diem, M.; and Sablatnig, R. 2013 · 2013
Earlier work this paper cites.
Consciousness and the brain: Deciphering how the brain codes our thoughts
Dehaene, S. 2014 · 2014
Earlier work this paper cites.
Deep neural networks for large vocabulary handwritten text recognition
Bluche, T. 2015 · 2015
Earlier work this paper cites.
Unsupervised visual representation learning by context prediction
Doersch, C.; Gupta, A.; and Efros, A. A. 2015 · 2015
Earlier work this paper cites.
Convolutional neural networks for direct text deblurring
Hradiš, M.; Kotera, J.; Zemcık, P.; and Šroubek, F. 2015 · 2015
Earlier work this paper cites.
Ba, J. L.; Kiros, J. R.; and Hinton, G. E. 2016 · 2016
Earlier work this paper cites.
Joint line segmentation and transcription for end-to-end handwritten paragraph recognition
Bluche, T. 2016 · 2016
Earlier work this paper cites.
ICFHR2016 competition on the analysis of handwritten text in images of balinese palm leaf manuscripts
Burie, J.-C.; Coustaty, M.; Hadi, S.; Kesiman, M. W. A.; Ogier, J.-M.; Paulus, E.; Sok, K.; Sunarya, I. M. G.; and Valy, D. 2016 · 2016
Earlier work this paper cites.
Synthetic data for text localisation in natural images
Gupta, A.; Vedaldi, A.; and Zisserman, A. 2016 · 2016
Earlier work this paper cites.
Unsupervised learning of visual representations by solving jigsaw puzzles
Noroozi, M.; and Favaro, P. 2016 · 2016
Cited alongside, same era.
Context encoders: Feature learning by inpainting
Pathak, D.; Krahenbuhl, P.; Donahue, J.; Darrell, T.; and Efros, A. A. 2016 · 2016
Cited alongside, same era.
ICFHR2016 handwritten document image binarization contest (H-DIBCO 2016)
Pratikakis, I.; Zagoris, K.; Barlas, G.; and Gatos, B. 2016 · 2016
Cited alongside, same era.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
Shi, B.; Bai, X.; and Yao, C. 2016 · 2016
Cited alongside, same era.
Robust scene text recognition with automatic rectification
Shi, B.; Wang, X.; Lyu, P.; Yao, C.; and Bai, X. 2016 · 2016
Cited alongside, same era.
A survey on handwritten character recognition (HCR) techniques for English alphabets
Momentum contrast for unsupervised visual representation learning
He, K.; Fan, H.; Wu, Y.; Xie, S.; and Girshick, R. 2020 · 2020
Later among the works it cites.
Scatter: selective context attentional scene text recognizer
Litman, R.; Anschel, O.; Tsiper, S.; Litman, R.; Mazor, S.; and Manmatha, R. 2020 · 2020
Later among the works it cites.
Handwritten optical character recognition (OCR): A comprehensive systematic literature review (SLR)
Memon, J.; Sami, M.; Khan, R. A.; and Uddin, M. 2020 · 2020
Later among the works it cites.
DE-GAN: a conditional generative adversarial network for document enhancement
Souibgui, M. A.; and Kessentini, Y. 2020 · 2020
Later among the works it cites.
A conditional GAN based approach for distorted camera captured documents recovery
Souibgui, M. A.; Kessentini, Y.; and Fornés, A. 2020 · 2020
Later among the works it cites.
Sequence-to-sequence contrastive learning for text recognition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sonkusare, M.; and Sahu, N. 2016 · 2016
Cited alongside, same era.
Colorful image colorization
Zhang, R.; Isola, P.; and Efros, A. A. 2016 · 2016
Cited alongside, same era.
Focusing attention: Towards accurate text recognition in natural images
Cheng, Z.; Bai, F.; Xu, Y.; Zheng, G.; Pu, S.; and Zhou, S. 2017 · 2017
Cited alongside, same era.
ICDAR2017 competition on document image binarization (DIBCO 2017)
Pratikakis, I.; Zagoris, K.; Barlas, G.; and Gatos, B. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2018 · 2018
Cited alongside, same era.
Unsupervised representation learning by predicting image rotations
Gidaris, S.; Singh, P.; and Komodakis, N. 2018 · 2018
Cited alongside, same era.
Aberdam, A.; Litman, R.; Tsiper, S.; Anschel, O.; Slossberg, R.; Mazor, S.; Manmatha, R.; and Perona, P. 2021 · 2021
Later among the works it cites.
Beit: Bert pre-training of image transformers
Bao, H.; Dong, L.; and Wei, F. 2021 · 2021
Later among the works it cites.
Vectorization and rasterization: Self-supervised learning for sketch and handwriting
Bhunia, A. K.; Chowdhury, P. N.; Yang, Y.; Hospedales, T. M.; Xiang, T.; and Song, Y.-Z. 2021 · 2021
Later among the works it cites.
Emerging properties in self-supervised vision transformers
Caron, M.; Touvron, H.; Misra, I.; Jégou, H.; Mairal, J.; Bojanowski, P.; and Joulin, A. 2021 · 2021
Later among the works it cites.
Text recognition in the wild: A survey
Chen, X.; Jin, L.; Zhu, Y.; Luo, C.; and Wang, T. 2021 · 2021
Later among the works it cites.
An empirical study of training self-supervised vision transformers
Chen, X.; Xie, S.; and He, K. 2021 · 2021
Later among the works it cites.
PeCo: Perceptual Codebook for BERT Pre-training of Vision Transformers
Dong, X.; Bao, J.; Zhang, T.; Chen, D.; Zhang, W.; Yuan, L.; Chen, D.; Wen, F.; and Yu, N. 2021 · 2021
Later among the works it cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; Uszkoreit, J.; and Houlsby, N. 2021 · 2021
Later among the works it cites.
Masked autoencoders are scalable vision learners
He, K.; Chen, X.; Xie, S.; Li, Y.; Dollár, P.; and Girshick, R. 2021 · 2021
Later among the works it cites.
Complex image processing with less data—Document image binarization by integrating multiple pre-trained U-Net modules
Kang, S.; Iwana, B. K.; and Uchida, S. 2021 · 2021
Later among the works it cites.
Scene text detection and recognition: The deep learning era
Long, S.; He, X.; and Yao, C. 2021 · 2021
Later among the works it cites.
Souibgui, M. A.; Fornés, A.; Kessentini, Y.; and Megyesi, B. 2021 · 2021
Later among the works it cites.
Barlow twins: Self-supervised learning via redundancy reduction
Zbontar, J.; Jing, L.; Misra, I.; LeCun, Y.; and Deny, S. 2021 · 2021
Later among the works it cites.
Enhance to read better: A Multi-Task Adversarial Network for Handwritten Document Image Enhancement
Jemni, S. K.; Souibgui, M. A.; Kessentini, Y.; and Fornés, A. 2022 · 2022
Closest in time.
Semantically contrastive learning for low-light image enhancement
Liang, D.; Li, L.; Wei, M.; Yang, S.; Zhang, L.; Yang, W.; Du, Y.; and Zhou, H. 2022 · 2022
Closest in time.
Perceiving Stroke-Semantic Context: Hierarchical Contrastive Learning for Robust Scene Text Recognition
Liu, H.; Wang, B.; Bao, Z.; Xue, M.; Kang, S.; Jiang, D.; Liu, Y.; and Ren, B. 2022 · 2022
Closest in time.
DocEnTr: An End-to-End Document Image Enhancement Transformer
Souibgui, M. A.; Biswas, S.; Jemni, S. K.; Kessentini, Y.; Fornés, A.; Lladós, J.; and Pal, U. 2022 · 2022
Closest in time.
Context-based Contrastive Learning for Scene Text Recognition
Zhang, X.; Zhu, B.; Yao, X.; Sun, Q.; Li, R.; and Yu, B. 2022 · 2022
Closest in time.