Fetching the paper…
Reading the bibliography…
We present MMOCR-an open-source toolbox which provides a comprehensive pipeline for text detection and recognition, as well as their downstream tasks such as named entity recognition and key information extraction.
An Overview of the Tesseract OCR Engine. In ICDAR . 629–633
Ray Smith. 2007 · 2007
Earlier work this paper cites.
Fast R-CNN. In ICCV . 1440–1448
Ross Girshick. 2015 · 2015
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation. In CVPR . 3431–3440
Jonathan Long, Evan Shelhamer, and Trevor Darrell. 2015 · 2015
Earlier work this paper cites.
Faster R-CNN: Towards real-time object detection with region proposal networks. In NIPS . 91–99
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun. 2015 · 2015
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition. In ICLR
Karen Simonyan and Andrew Zisserman. 2015 · 2015
Earlier work this paper cites.
Going deeper with convolutions. In CVPR . 1–9
C Szegedy, W Liu, Y Jia, and P Sermanet. 2015 · 2015
Earlier work this paper cites.
Named Entity Recognition with Bidirectional LSTM-CNNs
Jason P.C. Chiu and Eric Nichols. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition. In CVPR . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Neural architectures for named entity recognition
Guillaume Lample, Miguel Ballesteros, Sandeep Subramanian, Kazuya Kawakami, and Chris Dyer. 2016 · 2016
Earlier work this paper cites.
An End-to-End Trainable Neural Network for Image-Based Sequence Recognition and Its Application to Scene Text Recognition
Baoguang Shi, Xiang Bai, and Cong Yao. 2016a · 2016
Earlier work this paper cites.
Mask R-CNN. In ICCV . 2961–2969
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick. 2017 · 2017
Earlier work this paper cites.
CloudScan - A configuration-free invoice analysis system using recurrent neural networks. In ICDAR . 406–413
Rasmus Berg Palm, Ole Winther, and Florian Laws. 2017 · 2017
Earlier work this paper cites.
YOLO9000: Better, Faster, Stronger. In CVPR . 6517–6525
Joseph Redmon and Ali Farhadi. 2017 · 2017
Earlier work this paper cites.
Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme. In ACL . 1227–1236
Suncong Zheng, Feng Wang, Hongyun Bao, Yuexing Hao, Peng Zhou, and Bo Xu. 2017 · 2017
Cited alongside, same era.
EAST: An Efficient and Accurate Scene Text Detector
Xinyu Zhou, Cong Yao, He Wen, Yuzhi Wang, Shuchang Zhou, Weiran He, and Jiajun Liang. 2017 · 2017
Cited alongside, same era.
Chargrid: Towards Understanding 2D Documents. In EMNLP . 4459–4469
Anoop Raveendra Katti Faddoul, Christian Reisswig Cordula Guder, Sebastian Brarda, Steffen Bickel, Johannes Höhne, and Jean Baptiste. 2018 · 2018
Cited alongside, same era.
TextSnake: A Flexible Representation for Detecting Text of Arbitrary Shapes. In ECCV . 19–35
Shangbang Long, Jiaqiang Ruan, Wenjie Zhang, Xin He, Wenhao Wu, and Cong Yao. 2018 · 2018
Cited alongside, same era.
YOLOv3: An Incremental Improvement
Joseph Redmon and Ali Farhadi. 2018 · 2018
Cited alongside, same era.
Aggregation cross-entropy for sequence recognition. In CVPR . 6538–6547
Zecheng Xie, Yaoxiong Huang, Yuanzhi Zhu, Lianwen Jin, Yuliang Liu, and Lele Xie. 2019 · 2019
Later among the works it cites.
Symmetry-constrained rectification network for scene text recognition. In ICCV . 9146–9155
Mingkun Yang, Yushuo Guan, Minghui Liao, Xin He, Kaigui Bian, Song Bai, Cong Yao, and Xiang Bai. 2019 · 2019
Later among the works it cites.
Real-Time Scene Text Detection with Differentiable Binarization. In AAAI . 11474–11481
Minghui Liao, Zhaoyi Wan, Cong Yao, Kai Chen, and Xiang Bai. 2020 · 2020
Later among the works it cites.
CLUENER2020: Fine-grained Named Entity Recognition Dataset and Benchmark for Chinese
Liang Xu, Yu Tong, Qianqian Dong, Yixuan Liao, Cong Yu, Yin Tian, Weitang Liu, Lu Li, Caiquan Liu, and Xuanwei Zhang. 2020 · 2020
Later among the works it cites.
RobustScanner: Dynamically Enhancing Positional Clues for Robust Text Recognition. In ECCV . 135–151
Xiaoyu Yue, Zhanghui Kuang, Chenhao Lin, Hongbin Sun, and Wayne Zhang. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Boosting up Scene Text Detectors with Guided CNN. In BMVC
Xiaoyu Yue, Zhanghui Kuang, Zhaoyang Zhang, Zhenfang Chen, Pan He, Yu Qiao, and Wei Zhang. 2018 · 2018
Cited alongside, same era.
Character region awareness for text detection. In CVPR . 9365–9374
Youngmin Baek, Bado Lee, Dongyoon Han, Sangdoo Yun, and Hwalsuk Lee. 2019 · 2019
Cited alongside, same era.
Rosetta: Large scale system for text detection and recognition in images
Fedor Borisyuk, Albert Gordo, and Viswanath Sivakumar. 2019 · 2019
Cited alongside, same era.
Geometry normalization networks for accurate scene text detection. In ICCV . 9136–9145
Jiaqi Duan, Youjiang Xu, Zhanghui Kuang, Xiaoyu Yue, Hongbin Sun, Yue Guan, and Wayne Zhang. 2019 · 2019
Cited alongside, same era.
Show, Attend and Read: A Simple and Strong Baseline for Irregular Text Recognition
Hui Li, Peng Wang, Chunhua Shen, and Guyu Zhang. 2019 · 2019
Cited alongside, same era.
Scene text recognition from two-dimensional perspective
Minghui Liao, Jian Zhang, Zhaoyi Wan, Fengming Xie, Jiajun Liang, Pengyuan Lyu, Cong Yao, and Xiang Bai. 2019 · 2019
Cited alongside, same era.
Pyramid Mask Text Detector
Jingchao Liu, Xuebo Liu, Jie Sheng, Ding Liang, Xin Li, and Qingjie Liu. 2019 · 2019
Cited alongside, same era.
Deep Relational Reasoning Graph Network for Arbitrary Shape Text Detection. In CVPR . 9696–9705
Shi-Xue Zhang, Xiaobin Zhu, Jie-Bo Hou, Chang Liu, Chun Yang, Hongfa Wang, and Xu-Cheng Yin. 2020 · 2020
Later among the works it cites.
Deep Dual-resolution Networks for Real-time and Accurate Semantic Segmentation of Road Scenes
Yuanduo Hong, Huihui Pan, Weichao Sun, and Yisong Jia. 2021 · 2021
Closest in time.
Spatial Dual-Modality Graph Reasoning for Key Information Extraction
Hongbin Sun, Zhanghui Kuang, Xiaoyu Yue, Chenhao Lin, and Wayne Zhang. 2021 · 2021
Closest in time.
PGNet: Real-time Arbitrarily-Shaped Text Spotting with Point Gathering Network. In AAAI . 2782–2790
Pengfei Wang, Chengquan Zhang, Fei Qi, Shanshan Liu, Xiaoqiang Zhang, Pengyuan Lyu, Junyu Han, Jingtuo Liu, Errui Ding, and Guangming Shi. 2021 · 2021
Closest in time.
SegOCR: Simple Baseline. In Unpublished Manuscript
Xiaoyu Yue, Zhanghui Kuang, and Wayne Zhang. 2021 · 2021
Closest in time.
Fourier Contour Embedding for Arbitrary-Shaped Text Detection. In CVPR
Yiqin Zhu, Jianyong Chen, Lingyu Liang, Zhuanghui Kuang, Lianwen Jin, and Wayne Zhang. 2021 · 2021
Closest in time.
ASTER : An Attentional Scene Text Recognizer with Flexible Rectification
Baoguang Shi, Mingkun Yang, Xinggang Wang, Pengyuan Lyu, Cong Yao, and Xiang Bai. 2018 · 2048
Closest in time.