Fetching the paper…
Reading the bibliography…
The history of text can be traced back over thousands of years.
Learning spatial-semantic context with fully convolutional recurrent network for online handwritten Chinese text recognition
Zecheng Xie, Zenghui Sun, Lianwen Jin, Hao Ni, and Terry Lyons. 2017 · 1917
Earlier work this paper cites.
Text Recognition in Images Based on Transformer with Hierarchical Attention. In Proceedings of ICIP . 1945–1949
Yiwei Zhu, Shilin Wang, Zheng Huang, and Kai Chen. 2019 · 1949
Earlier work this paper cites.
Deep matching prior network: Toward tighter multi-oriented text detection. In Proceedings of CVPR . 1962–1969
Yuliang Liu and Lianwen Jin. 2017 · 1969
Earlier work this paper cites.
A novel adaptive morphological approach for degraded character image segmentation
Shigueo Nomura, Keiji Yamanaka, Osamu Katai, Hiroshi Kawakami, and Takayuki Shiose. 2005 · 1975
Earlier work this paper cites.
A computational approach to edge detection
John Canny. 1986 · 1986
Earlier work this paper cites.
Thin-Plate Splines and the Decompositions of Deformations
Fred L Bookstein Principal Warps. 1989 · 1989
Earlier work this paper cites.
Recognition of raised characters for automatic classification of rubber tires
Young Kug Ham, Min Seok Kang, Hong Kyu Chung, Rae-Hong Park, and Gwi Tae Park. 1995 · 1995
Earlier work this paper cites.
A survey of methods and strategies in character segmentation
Richard G Casey and Eric Lecolinet. 1996 · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, Patrick Haffner, et al · 1998
Earlier work this paper cites.
Twenty years of document image analysis in PAMI
George Nagy. 2000 · 2000
Earlier work this paper cites.
Automatic caption localization in compressed video
Yu Zhong, Hongjiang Zhang, and Anil K Jain. 2000 · 2000
Earlier work this paper cites.
Conditional Random Fields: Probabilistic Models for Segmenting and Labeling Sequence Data. In Proceedings of ICML . 282–289
John D. Lafferty, Andrew McCallum, and Fernando C. N. Pereira. 2001 · 2001
Earlier work this paper cites.
Limits on super-resolution and how to break them
Simon Baker and Takeo Kanade. 2002 · 2002
Earlier work this paper cites.
Vision for mobile robot navigation: A survey
Guilherme N DeSouza and Avinash C Kak. 2002 · 2002
Earlier work this paper cites.
Localizing and segmenting text in images and videos
Rainer Lienhart and Axel Wernicke. 2002 · 2002
Earlier work this paper cites.
Lexicon-driven segmentation and recognition of handwritten character strings for Japanese address reading
Cheng-Lin Liu, Masashi Koga, and Hiromichi Fujisawa. 2002 · 2002
Earlier work this paper cites.
Weighted finite-state transducers in speech recognition
Mehryar Mohri, Fernando Pereira, and Michael Riley. 2002 · 2002
Earlier work this paper cites.
Natural language processing
Gobinda G Chowdhury. 2003 · 2003
Earlier work this paper cites.
ICDAR 2003 robust reading competitions. In Proceedings of ICDAR . 682–687
Simon M Lucas, Alex Panaretos, Luis Sosa, Anthony Tang, Shirley Wong, and Robert Young. 2003 · 2003
Earlier work this paper cites.
A robust text detection algorithm in images and video frames. In Proceedings of Joint Conf. Inf., Commun. Signal Process. Pac. Rim Conf. Multimedia . IEEE, 802–806
Qixiang Ye, Wen Gao, Weiqiang Wang, and Wei Zeng. 2003 · 2003
Earlier work this paper cites.
Automatic detection and recognition of signs from natural scenes
Xilin Chen, Jie Yang, Jing Zhang, and Alex Waibel. 2004 · 2004
Earlier work this paper cites.
Histograms of oriented gradients for human detection. In Proceedings of CVPR . 886–893
Navneet Dalal and Bill Triggs. 2005 · 2005
Earlier work this paper cites.
Improved text-detection methods for a camera-based text reading system for blind persons. In Proceedings of ICDAR . 257–261
Nobuo Ezaki, Kimiyasu Kiyota, Bui Truong Minh, Marius Bulacu, and Lambert Schomaker. 2005 · 2005
Earlier work this paper cites.
ICDAR 2005 text locating competition results. In Proceedings of ICDAR . 80–84
Simon M Lucas. 2005 · 2005
Earlier work this paper cites.
Fast and robust text detection in images and video frames
Qixiang Ye, Qingming Huang, Wen Gao, and Debin Zhao. 2005 · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks. In Proceedings of ICML . 369–376
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
CINDI robot: an intelligent Web crawler based on multi-level inspection. In Eleventh International Database Engineering and Applications Symposium (IDEAS) . 93–101
Rui Chen, Bipin C Desai, and Cong Zhou. 2007 · 2007
Earlier work this paper cites.
Moving object detection in spatial domain using background removal techniques-state-of-art
Shireen Y Elhabian, Khaled M El-Sayed, and Sumaya H Ahmed. 2008 · 2008
Earlier work this paper cites.
A novel connectionist system for unconstrained handwriting recognition
Alex Graves, Marcus Liwicki, Santiago Fernández, Roman Bertolami, Horst Bunke, and Jürgen Schmidhuber. 2008 · 2008
Earlier work this paper cites.
A new approach for overlay text detection and extraction from complex video scene
Wonjun Kim and Changick Kim. 2008 · 2008
Earlier work this paper cites.
An adaptive text detection approach in images and video frames. In Proceedings of IJCNN . 72–77
Minhua Li and Chunheng Wang. 2008 · 2008
Earlier work this paper cites.
A camera phone based currency reader for the visually impaired. In Proceedings of ACM SIGACCESS International Conference on Computers and Accessibility . 305–306
Xu Liu. 2008 · 2008
Earlier work this paper cites.
Recaptcha: Human-based character recognition via web security measures
Luis Von Ahn, Benjamin Maurer, Colin McMillen, David Abraham, and Manuel Blum. 2008 · 2008
Earlier work this paper cites.
A gradient difference based technique for video text detection. In Proceedings of ICDAR . 156–160
Palaiahnakote Shivakumara, Trung Quy Phan, and Chew Lim Tan. 2009 · 2009
Earlier work this paper cites.
Combined script and page orientation estimation using the Tesseract OCR engine. In Proceedings of the International Workshop on Multilingual OCR . 6
Ranjith Unnikrishnan and Ray Smith. 2009 · 2009
Earlier work this paper cites.
Detecting text in natural scenes with stroke width transform. In Proceedings of CVPR . IEEE, 2963–2970
Boris Epshtein, Eyal Ofek, and Yonatan Wexler. 2010 · 2010
Earlier work this paper cites.
Scene text extraction with edge constraint and text collinearity. In Proceedings of ICPR . 3983–3986
SeongHun Lee, Min Su Cho, Kyomin Jung, and Jin Hyung Kim. 2010 · 2010
Earlier work this paper cites.
A method for text localization and recognition in real-world images. In Proceedings of ACCV . 770–783
Lukas Neumann and Jiri Matas. 2010 · 2010
Earlier work this paper cites.
Accurate video text detection through classification of low and high contrast images
Palaiahnakote Shivakumara, Weihua Huang, Trung Quy Phan, and Chew Lim Tan. 2010 · 2010
Earlier work this paper cites.
Word spotting in the wild. In Proceedings of ECCV . 591–604
Kai Wang and Serge Belongie. 2010 · 2010
Earlier work this paper cites.
Two-phase kernel estimation for robust motion deblurring. In Proceedings of ECCV . 157–170
Li Xu and Jiaya Jia. 2010 · 2010
Earlier work this paper cites.
Text from corners: a novel approach to detect text and caption in videos
Xu Zhao, Kai-Hsiang Lin, Yun Fu, Yuxiao Hu, Yuncai Liu, and Thomas S Huang. 2010 · 2010
Earlier work this paper cites.
Robustly extracting captions in videos based on stroke-like edges and spatio-temporal analysis
Xiaoqian Liu and Weiqiang Wang. 2011 · 2011
Earlier work this paper cites.
Foreign language abbreviation translation in an instant messaging system
Fang Lu, Corey S McCaffrey, and Elaine I Kuo. 2011 · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning. In Proceedings of NIPS
Yuval Netzer, Tao Wang, Adam Coates, Alessandro Bissacco, Bo Wu, and Andrew Y Ng. 2011 · 2011
Earlier work this paper cites.
A Hybrid Approach to Detect and Localize Texts in Natural Scene Images
Yi-Feng Pan, Xinwen Hou, and Cheng-Lin Liu. 2011 · 2011
Earlier work this paper cites.
ICDAR 2011 robust reading competition challenge 2: Reading text in scene images. In Proceedings of ICDAR . 1491–1496
Asif Shahab, Faisal Shafait, and Andreas Dengel. 2011 · 2011
Earlier work this paper cites.
A new gradient based character segmentation method for video text recognition. In Proceedings of ICDAR . 126–130
Palaiahnakote Shivakumara, Souvik Bhowmick, Bolan Su, Chew Lim Tan, and Umapada Pal. 2011 · 2011
Earlier work this paper cites.
Mobile visual search on printed documents using text and low bit-rate features. In Proceedings of ICIP . 2601–2604
Sam S Tsai, Huizhong Chen, David Chen, Georg Schroth, Radek Grzeszczuk, and Bernd Girod. 2011 · 2011
Earlier work this paper cites.
End-to-end scene text recognition. In Proceedings of ICCV . 1457–1464
Kai Wang, Boris Babenko, and Serge Belongie. 2011 · 2011
Earlier work this paper cites.
Text string detection from natural scenes by structure-based partition and grouping
Chucai Yi and YingLi Tian. 2011 · 2011
Earlier work this paper cites.
Large scale distributed deep networks. In Proceedings of NIPS . 1223–1231
Jeffrey Dean, Greg Corrado, Rajat Monga, Kai Chen, Matthieu Devin, Mark Mao, Marc’aurelio Ranzato, Andrew Senior, Paul Tucker, Ke Yang, et al · 2012
Earlier work this paper cites.
Supervised sequence labelling
Alex Graves. 2012 · 2012
Earlier work this paper cites.
Image Text Detection Using a Bandlet-Based Edge Detector and Stroke Width Transform.. In Proceedings of BMVC . 1–12
Ali Mosleh, Nizar Bouguila, and A Ben Hamza. 2012 · 2012
Earlier work this paper cites.
Real-time scene text localization and recognition. In Proceedings of CVPR . 3538–3545
Lukáš Neumann and Jiří Matas. 2012 · 2012
Earlier work this paper cites.
Convolutional neural networks applied to house numbers digit classification. In Proceedings of ICPR . 3288–3291
Pierre Sermanet, Soumith Chintala, and Yann LeCun. 2012 · 2012
Earlier work this paper cites.
End-to-end text recognition with convolutional neural networks. In Proceedings of ICPR . 3304–3308
Tao Wang, David J Wu, Adam Coates, and Andrew Y Ng. 2012 · 2012
Earlier work this paper cites.
Detecting texts of arbitrary orientations in natural images. In Proceedings of CVPR . 1083–1090
Cong Yao, Xiang Bai, Wenyu Liu, Yi Ma, and Zhuowen Tu. 2012 · 2012
Earlier work this paper cites.
Photoocr: Reading text in uncontrolled conditions. In Proceedings of ICCV . 785–792
Alessandro Bissacco, Mark Cummins, Yuval Netzer, and Hartmut Neven. 2013 · 2013
Earlier work this paper cites.
Whole is greater than sum of parts: Recognizing scene text words. In Proceedings of ICDAR . 398–402
Vibhor Goel, Anand Mishra, Karteek Alahari, and CV Jawahar. 2013 · 2013
Earlier work this paper cites.
Maxout networks. In Proceedings of ICML . 1319–1327
Ian J Goodfellow, David Warde-Farley, Mehdi Mirza, Aaron Courville, and Yoshua Bengio. 2013 · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks. In Proceedings of ICASSP . 6645–6649
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton. 2013 · 2013
Earlier work this paper cites.
ICDAR 2013 robust reading competition. In Proceedings of ICDAR . 1484–1493
Dimosthenis Karatzas, Faisal Shafait, Seiichi Uchida, Masakazu Iwamura, Lluis Gomez i Bigorda, Sergi Robles Mestre, Joan Mas, David Fernandez Mota, Jon Almazan Almazan, and Lluis Pere De Las Heras. 2013 · 2013
Earlier work this paper cites.
Scene Text Detection via Connected Component Clustering and Nontext Filtering
Hyung Il Koo and Duck Hoon Kim. 2013 · 2013
Earlier work this paper cites.
Recognizing text with perspective distortion in natural scenes. In Proceedings of ICCV . 569–576
Trung Quy Phan, Palaiahnakote Shivakumara, Shangxuan Tian, and Chew Lim Tan. 2013 · 2013
Earlier work this paper cites.
Scene text recognition using part-based tree-structured character detection. In Proceedings of CVPR . 2961–2968
Cunzhao Shi, Chunheng Wang, Baihua Xiao, Yang Zhang, Song Gao, and Zhong Zhang. 2013 · 2013
Earlier work this paper cites.
Rotation-invariant features for multi-oriented text detection in natural images
Cong Yao, Xin Zhang, Xiang Bai, Wenyu Liu, Yi Ma, and Zhuowen Tu. 2013 · 2013
Cited alongside, same era.
Text extraction from natural scene image: A survey
Honggang Zhang, Kaili Zhao, Yi-Zhe Song, and Jun Guo. 2013 · 2013
Cited alongside, same era.
Word spotting and recognition with embedded attributes
Jon Almazán, Albert Gordo, Alicia Fornés, and Ernest Valveny. 2014 · 2014
Cited alongside, same era.
End-to-End Text Recognition with Hybrid HMM Maxout Models. In Proceedings of ICLR: Workshop
Ouais Alsharif and Joelle Pineau. 2014 · 2014
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation. In Proceedings of EMNLP . 1724–1734
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Employing Semantic Context for Sparse Information Extraction Assessment
Peipei Li, Haixun Wang, Hongsong Li, and Xindong Wu. 2018 · 2018
Later among the works it cites.
Textboxes++: A single-shot oriented scene text detector
Minghui Liao, Baoguang Shi, and Xiang Bai. 2018 · 2018
Later among the works it cites.
Toward abstractive summarization using semantic representations
Fei Liu, Jeffrey Flanigan, Sam Thomson, Norman Sadeh, and Noah A Smith. 2018b · 2018
Later among the works it cites.
Scene text detection and recognition: The deep learning era
Shangbang Long, Xin He, and Cong Ya. 2018 · 2018
Later among the works it cites.
Mask textspotter: An end-to-end trainable neural network for spotting text with arbitrary shapes. In Proceedings of ECCV . 67–83
Pengyuan Lyu, Minghui Liao, Cong Yao, Wenhao Wu, and Xiang Bai. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Generative adversarial nets. In Proceedings of NIPS . 2672–2680
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Towards end-to-end speech recognition with recurrent neural networks. In Proceedings of ICML . 1764–1772
Alex Graves and Navdeep Jaitly. 2014 · 2014
Cited alongside, same era.
A robust arbitrary text detection system for natural scene images
Anhar Risnumawan, Palaiahankote Shivakumara, Chee Seng Chan, and Chew Lim Tan. 2014 · 2014
Cited alongside, same era.
Accurate scene text recognition based on recurrent neural network. In Proceedings of ACCV . 35–48
Bolan Su and Shijian Lu. 2014 · 2014
Cited alongside, same era.
Text localization and recognition in images and video
Seiichi Uchida. 2014 · 2014
Cited alongside, same era.
A unified framework for multioriented text detection and recognition
Cong Yao, Xiang Bai, and Wenyu Liu. 2014a · 2014
Cited alongside, same era.
Text detection and recognition in imagery: A survey
Qixiang Ye and David Doermann. 2014 · 2014
Cited alongside, same era.
Scene text detection using superpixel-based stroke feature transform and deep learning based region classification
Youbao Tang and Xiangqian Wu. 2018 · 2018
Later among the works it cites.
A Unified Framework for Tracking Based Text Detection and Recognition from Web Videos
Shu Tian, Xu-Cheng Yin, Ya Su, and Hong-Wei Hao. 2018 · 2018
Later among the works it cites.
Scene classification with recurrent attention of VHR remote sensing images
Qi Wang, Shaoteng Liu, Jocelyn Chanussot, and Xuelong Li. 2018a · 2018
Later among the works it cites.
A new CNN-based method for multi-directional car license plate detection
Lele Xie, Tasweer Ahmad, Lianwen Jin, Yuliang Liu, and Sheng Zhang. 2018 · 2018
Later among the works it cites.
A gene–phenotype relationship extraction pipeline from the biomedical literature using a representation learning approach
Wenhui Xing, Junsheng Qi, Xiaohui Yuan, Lin Li, Xiaoyu Zhang, Yuhua Fu, Shengwu Xiong, Lun Hu, and Jing Peng. 2018 · 2018
Later among the works it cites.
A fast uyghur text detector for complex background images
Chenggang Yan, Hongtao Xie, Jianjun Chen, Zhengjun Zha, Xinhong Hao, Yongdong Zhang, and Qionghai Dai. 2018 · 2018
Later among the works it cites.
Tai-Ling Yuan, Zhe Zhu, Kun Xu, Cheng-Jun Li, and Shi-Min Hu. 2018 · 2018
Later among the works it cites.
PrivacyCheck: Automatic Summarization of Privacy Policies Using Data Mining
Razieh Nokhbeh Zaeem, Rachel L German, and K Suzanne Barber. 2018 · 2018
Later among the works it cites.
Verisimilar image synthesis for accurate detection and recognition of texts in scenes. In Proceedings of ECCV . 249–266
Fangneng Zhan, Shijian Lu, and Chuhui Xue. 2018 · 2018
Later among the works it cites.
Feature enhancement network: A refined scene text detector. In Proceedings of AAAI . 2612–2619
Sheng Zhang, Yuliang Liu, Lianwen Jin, and Canjie Luo. 2018 · 2018
Later among the works it cites.
Scene text visual question answering. In Proceedings of ICCV . 4291–4301
Ali Furkan Biten, Ruben Tito, Andres Mafla, Lluis Gomez, Marçal Rusinol, Ernest Valveny, CV Jawahar, and Dimosthenis Karatzas. 2019 · 2019
Later among the works it cites.
Patch Aggregator for Scene Text Script Identification. In Proceedings of ICDAR . 1077–1083
Changxu Cheng, Qiuhui Huang, Xiang Bai, Bin Feng, and Wenyu Liu. 2019 · 2019
Later among the works it cites.
Semi-supervised learning for neural machine translation
Yong Cheng. 2019 · 2019
Later among the works it cites.
Total-Text: toward orientation robustness in scene text detection
Chee-Kheng Ch’ng, Chee Seng Chan, and Cheng-Lin Liu. 2019 · 2019
Later among the works it cites.
ICDAR2019 Robust Reading Challenge on Arbitrary-Shaped Text (RRC-ArT). In Proceedings of ICDAR . 1571–1576
Chee-Kheng Chng, Yuliang Liu, Yipeng Sun, Chun Chet Ng, Canjie Luo, Zihan Ni, ChuanMing Fang, Shuaitao Zhang, Junyu Han, Errui Ding, et al · 2019
Later among the works it cites.
A Comparative Study of Attention-based Encoder-Decoder Approaches to Natural Scene Text Recognition. In Proceedings of ICDAR . 916–921
Fuze Cong, Wenping Hu, Huo Qiang, and Li Guo. 2019 · 2019
Later among the works it cites.
Deep Multi-Scale Context Aware Feature Aggregation for Curved Scene Text Detection
Pengwen Dai, Hua Zhang, and Xiaochun Cao. 2019 · 2019
Later among the works it cites.
End-to-End Information Extraction by Character-Level Embedding and Multi-Stage Attentional U-Net. In Proceedings of BMVC . 96
Tuan Anh Nguyen Dang and Dat Nguyen Thanh. 2019 · 2019
Later among the works it cites.
Learning to draw text in natural images with conditional adversarial networks. In Proceedings of IJCAI . 715–722
Shancheng Fang, Hongtao Xie, Jianjun Chen, Jianlong Tan, and Yongdong Zhang. 2019 · 2019
Later among the works it cites.
Focal CTC Loss for Chinese Optical Character Recognition on Unbalanced Datasets
Xinjie Feng, Hongxun Yao, and Shengping Zhang. 2019b · 2019
Later among the works it cites.
Reading scene text with fully convolutional sequence modeling
Yunze Gao, Yingying Chen, Jinqiao Wang, Ming Tang, and Hanqing Lu. 2019 · 2019
Later among the works it cites.
VD-SAN: Visual-Densely Semantic Attention Network for Image Caption Generation
Xinwei He, Yang Yang, Baoguang Shi, and Xiang Bai. 2019 · 2019
Later among the works it cites.
Express Delivery System based on Fingerprint Identification. In Proceedings of ITNEC . IEEE, 363–367
Hu Huang, Ya Zhong, Shiying Yin, Junlin Xiang, Lijun He, Yu Lv, and Peng Huang. 2019 · 2019
Later among the works it cites.
Show, attend and read: A simple and strong baseline for irregular text recognition. In Proceedings of AAAI . 8610–8617
Hui Li, Peng Wang, Chunhua Shen, and Guyu Zhang. 2019 · 2019
Later among the works it cites.
Mask textspotter: An end-to-end trainable neural network for spotting text with arbitrary shapes
Minghui Liao, Pengyuan Lyu, Minghang He, Cong Yao, Wenhao Wu, and Xiang Bai. 2019a · 2019
Later among the works it cites.
ICDAR 2019 Robust Reading Challenge on Reading Chinese Text on Signboard. In Proceedings of ICDAR . 1577–1581
Xi Liu, Rui Zhang, Yongsheng Zhou, Qianyi Jiang, Qi Song, Nan Li, Kai Zhou, Lei Wang, Dong Wang, Minghui Liao, et al · 2019
Later among the works it cites.
Curved scene text detection via transverse and longitudinal sequence connection
Yuliang Liu, Lianwen Jin, Shuaitao Zhang, Canjie Luo, and Sheng Zhang. 2019c · 2019
Later among the works it cites.
MORAN: A Multi-Object Rectified Attention Network for Scene Text Recognition
Canjie Luo, Lianwen Jin, and Zenghui Sun. 2019 · 2019
Later among the works it cites.
ICDAR2019 Robust Reading Challenge on Multi-lingual Scene Text Detection and Recognition–RRC-MLT-2019. In Proceedings of ICDAR . 1582–1587
Nibal Nayef, Yash Patel, Michal Busta, Pinaki Nath Chowdhury, Dimosthenis Karatzas, Wafa Khlif, Jiri Matas, Umapada Pal, Jean-Christophe Burie, Cheng-lin Liu, et al · 2019
Later among the works it cites.
A Novel Joint Character Categorization and Localization Approach for Character-Level Scene Text Recognition. In Proceedings of ICDAR: Workshops . 83–90
Xianbiao Qi, Yihao Chen, Rong Xiao, Chun-Guang Li, Qin Zou, and Shuguang Cui. 2019 · 2019
Later among the works it cites.
Towards Unconstrained End-to-End Text Spotting. In Proceedings of ICCV . 4704–4714
Siyang Qin, Alessandro Bissacco, Michalis Raptis, Yasuhisa Fujii, and Ying Xiao. 2019 · 2019
Later among the works it cites.
Method and a device for tracking characters that appear on a plurality of images of a video stream of a text
Alain Rouh and Jean Beaudet. 2019 · 2019
Later among the works it cites.
NRTR: A No-Recurrence Sequence-to-Sequence Model For Scene Text Recognition. In Proceedings of ICDAR . 781–786
Fenfen Sheng, Zhineng Chen, and Bo Xu. 2019 · 2019
Later among the works it cites.
Towards VQA models that can read. In Proceedings of CVPR . 8317–8326
Amanpreet Singh, Vivek Natarajan, Meet Shah, Yu Jiang, Xinlei Chen, Dhruv Batra, Devi Parikh, and Marcus Rohrbach. 2019 · 2019
Later among the works it cites.
ICDAR 2019 Competition on Large-scale Street View Text with Partial Labeling–RRC-LSVT. In Proceedings of ICDAR . 1557–1562
Yipeng Sun, Zihan Ni, Chee-Kheng Chng, Yuliang Liu, Canjie Luo, Chun Chet Ng, Junyu Han, Errui Ding, Jingtuo Liu, Dimosthenis Karatzas, et al · 2019
Later among the works it cites.
2D-CTC for Scene Text Recognition
Zhaoyi Wan, Fengming Xie, Yibo Liu, Xiang Bai, and Cong Yao. 2019 · 2019
Later among the works it cites.
A Simple and Robust Convolutional-Attention Network for Irregular Text Recognition
Peng Wang, Lu Yang, Hui Li, Yuyan Deng, Chunhua Shen, and Yanning Zhang. 2019e · 2019
Later among the works it cites.
TextSR: Content-Aware Text Super-Resolution Guided by Recognition
Wenjia Wang, Enze Xie, Peize Sun, Wenhai Wang, Lixun Tian, Chunhua Shen, and Ping Luo. 2019d · 2019
Later among the works it cites.
Editing Text in the Wild. In Proceedings of ACM International Conference on Multimedia . 1500–1508
Liang Wu, Chengquan Zhang, Jiaming Liu, Junyu Han, Jingtuo Liu, Errui Ding, and Xiang Bai. 2019 · 2019
Later among the works it cites.
Convolutional Attention Networks for Scene Text Recognition
Hongtao Xie, Shancheng Fang, Zheng-Jun Zha, Yating Yang, Yan Li, and Yongdong Zhang. 2019a · 2019
Later among the works it cites.
Convolutional Character Networks. In Proceedings of ICCV . 9125–9135
Linjie Xing, Zhi Tian, Weilin Huang, and Matthew R. Scott. 2019 · 2019
Later among the works it cites.
TextField: learning a deep direction field for irregular scene text detection
Yongchao Xu, Yukang Wang, Wei Zhou, Yongpan Wang, Zhibo Yang, and Xiang Bai. 2019 · 2019
Later among the works it cites.
Fully Convolutional Sequence Recognition Network for Water Meter Number Reading
Fan Yang, Lianwen Jin, Songxuan Lai, Xue Gao, and Zhaohai Li. 2019b · 2019
Later among the works it cites.
Video text localization based on Adaboost
Fang Yin, Rui Wu, Xiaoyang Yu, and Guanglu Sun. 2019 · 2019
Later among the works it cites.
Spatial fusion gan for image synthesis. In Proceedings of CVPR . 3653–3662
Fangneng Zhan, Hongyuan Zhu, and Shijian Lu. 2019 · 2019
Later among the works it cites.
Sequence-To-Sequence Domain Adaptation Network for Robust Text Image Recognition. In Proceedings of CVPR . 2740–2749
Yaping Zhang, Shuai Nie, Wenju Liu, Xing Xu, Dongxiang Zhang, and Heng Tao Shen. 2019 · 2019
Later among the works it cites.
Adaptive Embedding Gate for Attention-Based Scene Text Recognition
Xiaoxue Chen, Tianwei Wang, Yuanzhi Zhu, Lianwen Jin, and Canjie Luo. 2020 · 2020
Closest in time.
GTC: Guided Training of CTC Towards Efficient and Accurate Scene Text Recognition. In Proceedings of AAAI
Wenyang Hu, Xiaocong Cai, Jun Hou, Shuai Yi, and Zhiping Lin. 2020 · 2020
Closest in time.
EPAN: Effective parts attention network for scene text recognition
Yunlong Huang, Zenghui Sun, Lianwen Jin, and Canjie Luo. 2020 · 2020
Closest in time.
SCATTER: Selective Context Attentional Scene Text Recognizer. In Proceedings of CVPR
Ron Litman, Oron Anschel, Shahar Tsiper, Roee Litman, Shai Mazor, and R. Manmatha. 2020 · 2020
Closest in time.
Arbitrarily Shaped Scene Text Detection with a Mask Tightness Text Detector
Yuliang Liu, Lianwen Jin, and Chuanming Fang. 2020b · 2020
Closest in time.
UnrealText: Synthesizing Realistic Scene Text Images from the Unreal World. In Proceedings of CVPR
Shangbang Long and Cong Yao. 2020 · 2020
Closest in time.
Separating Content from Style Using Adversarial Learning for Recognizing Text in the Wild
Canjie Luo, Qingxiang Lin, Yuliang Liu, Jin Lianwen, and Shen Chunhua. 2020 · 2020
Closest in time.
TextScanner: Reading Characters in Order for Robust Scene Text Recognition. In Proceedings of AAAI
Zhaoyi Wan, Mingling He, Haoran Chen, Xiang Bai, and Cong Yao. 2020 · 2020
Closest in time.
R-Net: A Relationship Network for Efficient and Accurate Scene Text Detection
Yuxin Wang, Hongtao Xie, Zheng-Jun Zha, Youliang Tian, Zilong Fu, and Yongdong Zhang. 2020c · 2020
Closest in time.
Towards Accurate Scene Text Recognition with Semantic Reasoning Networks. In Proceedings of CVPR
Deli Yu, Xuan Li, Chengquan Zhang, Junyu Han, Jingtuo Liu, and Errui Ding. 2020 · 2020
Closest in time.
Spatial transformer networks. In Proceedings of NIPS . 2017–2025
Max Jaderberg, Karen Simonyan, Andrew Zisserman, et al · 2025
Closest in time.
ASTER: An Attentional Scene Text Recognizer with Flexible Rectification
Baoguang Shi, Mingkun Yang, Xinggang Wang, Pengyuan Lyu, Cong Yao, and Xiang Bai. 2019 · 2048
Closest in time.
Shape-dna: effective character restoration and enhancement for Arabic text documents. In Proceedings of ICPR . 2053–2056
Gulcin Caner and Ismail Haritaoglu. 2010 · 2056
Closest in time.
Embodied question answering. In Proceedings of CVPR: Workshops . 2054–2063
Abhishek Das, Samyak Datta, Georgia Gkioxari, Stefan Lee, Devi Parikh, and Dhruv Batra. 2018 · 2063
Closest in time.
ESIR: End-to-end scene text recognition via iterative image rectification. In Proceedings of CVPR . 2059–2068
Fangneng Zhan and Shijian Lu. 2019 · 2068
Closest in time.