Fetching the paper…
Reading the bibliography…
Existing scene text spotting (i.e., end-to-end text detection and recognition) methods rely on costly bounding box annotations (e.g., text-line, word-level, or character-level bounding boxes).
Deep matching prior network: Toward tighter multi-oriented text detection. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 1962–1969
Yuliang Liu and Lianwen Jin. 2017 · 1969
Earlier work this paper cites.
The pascal visual object classes (VOC) challenge
Mark Everingham, Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman. 2010 · 2010
Earlier work this paper cites.
ICDAR 2011 robust reading competition-challenge 1: reading text in born-digital images (web and email). In Proc. Int. Conf. Doc. Anal. and Recognit. 1485–1490
Dimosthenis Karatzas, S Robles Mestre, Joan Mas, Farshad Nourbakhsh, and P Pratim Roy. 2011 · 2011
Earlier work this paper cites.
Detecting texts of arbitrary orientations in natural images. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. IEEE, 1083–1090
Cong Yao, Xiang Bai, Wenyu Liu, Yi Ma, and Zhuowen Tu. 2012 · 2012
Earlier work this paper cites.
ICDAR 2013 robust reading competition. In Proc. Int. Conf. Doc. Anal. and Recognit. IEEE, 1484–1493
Dimosthenis Karatzas, Faisal Shafait, Seiichi Uchida, Masakazu Iwamura, Lluis Gomez i Bigorda, Sergi Robles Mestre, Joan Mas, David Fernandez Mota, Jon Almazan Almazan, and Lluis Pere De Las Heras. 2013 · 2013
Earlier work this paper cites.
ICDAR 2015 competition on robust reading. In Proc. Int. Conf. Doc. Anal. and Recognit. IEEE, 1156–1160
Dimosthenis Karatzas, Lluis Gomez-Bigorda, Anguelos Nicolaou, Suman Ghosh, Andrew Bagdanov, Masakazu Iwamura, Jiri Matas, Lukas Neumann, Vijay Ramaseshan Chandrasekhar, Shijian Lu, et al · 2015
Earlier work this paper cites.
Synthetic data for text localisation in natural images. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 2315–2324
Ankush Gupta, Andrea Vedaldi, and Andrew Zisserman. 2016 · 2016
Earlier work this paper cites.
Reading text in the wild with convolutional neural networks
Max Jaderberg, Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman. 2016 · 2016
Earlier work this paper cites.
Detecting text in natural image with connectionist text proposal network. In Proc. Eur. Conf. Comp. Vis. Springer, 56–72
Zhi Tian, Weilin Huang, Tong He, Pan He, and Yu Qiao. 2016 · 2016
Earlier work this paper cites.
Deep textspotter: An end-to-end trainable scene text localization and recognition framework. In Proc. IEEE Int. Conf. Comp. Vis. 2204–2212
Michal Busta, Lukas Neumann, and Jiri Matas. 2017 · 2017
Earlier work this paper cites.
Total-Text: A comprehensive dataset for scene text detection and recognition. In Proc. Int. Conf. Doc. Anal. and Recognit. , Vol. 1. IEEE, 935–942
Chee Kheng Ch’ng and Chee Seng Chan. 2017 · 2017
Earlier work this paper cites.
WordSup: Exploiting word annotations for character based text detection. In Proc. IEEE Int. Conf. Comp. Vis. 4940–4949
Han Hu, Chengquan Zhang, Yuxuan Luo, Yuzhuo Wang, Junyu Han, and Errui Ding. 2017 · 2017
Earlier work this paper cites.
Towards end-to-end text spotting with convolutional recurrent neural networks. In Proc. IEEE Int. Conf. Comp. Vis. 5238–5246
Hui Li, Peng Wang, and Chunhua Shen. 2017 · 2017
Earlier work this paper cites.
TextBoxes: A fast text detector with a single deep neural network. In Proc. AAAI Conf. Artificial Intell. 4161–4167
Minghui Liao, Baoguang Shi, Xiang Bai, Xinggang Wang, and Wenyu Liu. 2017 · 2017
Earlier work this paper cites.
ICDAR 2017 robust reading challenge on multi-lingual scene text detection and script identification-RRC-MLT. In Proc. Int. Conf. Doc. Anal. and Recognit. , Vol. 1. IEEE, 1454–1459
Nibal Nayef, Fei Yin, Imen Bizid, Hyunsoo Choi, Yuan Feng, Dimosthenis Karatzas, Zhenbo Luo, Umapada Pal, Christophe Rigaud, Joseph Chazalon, et al · 2017
Cited alongside, same era.
WeText: Scene text detection under weak supervision. In Proc. IEEE Int. Conf. Comp. Vis. 1492–1500
Shangxuan Tian, Shijian Lu, and Chongshou Li. 2017 · 2017
Cited alongside, same era.
Attention is All you Need. In Proc. Advances in Neural Inf. Process. Syst. , Vol. 30
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
EAST: An efficient and accurate scene text detector. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 5551–5560
Xinyu Zhou, Cong Yao, He Wen, Yuzhi Wang, Shuchang Zhou, Weiran He, and Jiajun Liang. 2017 · 2017
Cited alongside, same era.
SEE: Towards semi-supervised end-to-end scene text recognition. In Proc. AAAI Conf. Artificial Intell. 6674–6681
Convolutional Character Networks. In Proc. IEEE Int. Conf. Comp. Vis. 9126–9136
Xing Linjie, Tian Zhi, Huang Weilin, and R. Scott Matthew. 2019 · 2019
Later among the works it cites.
Curved scene text detection via transverse and longitudinal sequence connection
Yuliang Liu, Lianwen Jin, Shuaitao Zhang, Canjie Luo, and Sheng Zhang. 2019 · 2019
Later among the works it cites.
ICDAR 2019 Robust Reading Challenge on Multi-lingual Scene Text Detection and Recognition–RRC-MLT-2019. In Proc. Int. Conf. Doc. Anal. and Recognit. 1582–1587
Nibal Nayef, Yash Patel, Michal Busta, Pinaki Nath Chowdhury, Dimosthenis Karatzas, Wafa Khlif, Jiri Matas, Umapada Pal, Jean-Christophe Burie, Cheng-lin Liu, et al · 2019
Later among the works it cites.
Towards unconstrained end-to-end text spotting. In Proc. IEEE Int. Conf. Comp. Vis. 4704–4714
Siyang Qin, Alessandro Bissacco, Michalis Raptis, Yasuhisa Fujii, and Ying Xiao. 2019 · 2019
Later among the works it cites.
Augment your batch: Improving generalization through instance repetition. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 8129–8138
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Christian Bartz, Haojin Yang, and Christoph Meinel. 2018 · 2018
Cited alongside, same era.
An end-to-end textspotter with explicit alignment and attention. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 5020–5029
Tong He, Zhi Tian, Weilin Huang, Chunhua Shen, Yu Qiao, and Changming Sun. 2018 · 2018
Cited alongside, same era.
FOTS: Fast oriented text spotting with a unified network. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 5676–5685
Xuebo Liu, Ding Liang, Shi Yan, Dagui Chen, Yu Qiao, and Junjie Yan. 2018 · 2018
Cited alongside, same era.
TextSnake: A flexible representation for detecting text of arbitrary shapes. In Proc. Eur. Conf. Comp. Vis. 20–36
Shangbang Long, Jiaqiang Ruan, Wenjie Zhang, Xin He, Wenhao Wu, and Cong Yao. 2018 · 2018
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2018 · 2018
Cited alongside, same era.
Mask TextSpotter: An end-to-end trainable neural network for spotting text with arbitrary shapes. In Proc. Eur. Conf. Comp. Vis. 67–83
Pengyuan Lyu, Minghui Liao, Cong Yao, Wenhao Wu, and Xiang Bai. 2018 · 2018
Cited alongside, same era.
Character region awareness for text detection. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 9365–9374
Youngmin Baek, Bado Lee, Dongyoon Han, Sangdoo Yun, and Hwalsuk Lee. 2019 · 2019
Cited alongside, same era.
ICDAR2019 robust reading challenge on arbitrary-shaped text-RRC-ArT. In Proc. Int. Conf. Doc. Anal. and Recognit. 1571–1576
Chee Kheng Chng, Yuliang Liu, Yipeng Sun, Chun Chet Ng, Canjie Luo, Zihan Ni, ChuanMing Fang, Shuaitao Zhang, Junyu Han, Errui Ding, et al · 2019
Cited alongside, same era.
Elad Hoffer, Tal Ben-Nun, Itay Hubara, Niv Giladi, Torsten Hoefler, and Daniel Soudry. 2020 · 2020
Later among the works it cites.
Mask TextSpotter v3: Segmentation Proposal Network for Robust Scene Text Spotting. In Proc. Eur. Conf. Comp. Vis. 706–722
Minghui Liao, Guan Pang, Jing Huang, Tal Hassner, and Xiang Bai. 2020 · 2020
Later among the works it cites.
ABCNet: Real-time Scene Text Spotting with Adaptive Bezier-Curve Network
Yuliang Liu, Hao Chen, Chunhua Shen, Tong He, Lianwen Jin, and Liangwei Wang. 2020 · 2020
Later among the works it cites.
On layer normalization in the Transformer architecture. In Proc. Int. Conf. Mach. Learn. 10524–10533
Ruibin Xiong, Yunchang Yang, Di He, Kai Zheng, Shuxin Zheng, Chen Xing, Huishuai Zhang, Yanyan Lan, Liwei Wang, and Tieyan Liu. 2020 · 2020
Later among the works it cites.
ABCNet v2: Adaptive Bezier-Curve Network for Real-time End-to-end Text Spotting
Yuliang Liu, Chunhua Shen, Lianwen Jin, Tong He, Peng Chen, Chongyu Liu, and Hao Chen. 2021 · 2021
Closest in time.
MANGO: A Mask Attention Guided One-Stage Scene Text Spotter. In Proc. AAAI Conf. Artificial Intell. 2467–2476
Liang Qiao, Ying Chen, Zhanzhan Cheng, Yunlu Xu, Yi Niu, Shiliang Pu, and Fei Wu. 2021 · 2021
Closest in time.
PAN++: Towards Efficient and Accurate End-to-End Spotting of Arbitrarily-Shaped Text
Wenhai Wang, Enze Xie, Xiang Li, Xuebo Liu, Ding Liang, Yang Zhibo, Tong Lu, and Chunhua Shen. 2021a · 2021
Closest in time.
Fourier contour embedding for arbitrary-shaped text detection. In Proc. IEEE Conf. Comp. Vis. Patt. Recogn. 3123–3131
Yiqin Zhu, Jianyong Chen, Lingyu Liang, Zhanghui Kuang, Lianwen Jin, and Wayne Zhang. 2021 · 2021
Closest in time.
Pix2Seq: A language modeling framework for object detection. In Proc. Int. Conf. Learn. Represent
Ting Chen, Saurabh Saxena, Lala Li, David J Fleet, and Geoffrey Hinton. 2022 · 2022
Closest in time.