Fetching the paper…
Reading the bibliography…
Billions of public domain documents remain trapped in hard copy or lack an accurate digitization.
Mmdetection: Open mmlab detection toolbox and benchmark
Kai Chen, Jiaqi Wang, Jiangmiao Pang, Yuhang Cao, Yu Xiong, Xiaoxiao Li, Shuyang Sun, Wansen Feng, Ziwei Liu, Jiarui Xu, et al. 2019 · 1906
Earlier work this paper cites.
Teikoku Ginko Kaisha Yoroku
Teikoku Koshinjo. 1957 · 1957
Earlier work this paper cites.
The state and fate of linguistic diversity and inclusion in the nlp world
Pratik Joshi, Sebastin Santy, Amar Budhiraja, Kalika Bali, and Monojit Choudhury. 2020 · 2004
Earlier work this paper cites.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
Baoguang Shi, Xiang Bai, and Cong Yao. 2016 · 2016
Earlier work this paper cites.
Cascade r-cnn: Delving into high quality object detection
Zhaowei Cai and Nuno Vasconcelos. 2018 · 2018
Earlier work this paper cites.
Searching for mobilenetv3
Andrew Howard, Mark Sandler, Grace Chu, Liang-Chieh Chen, Bo Chen, Mingxing Tan, Weijun Wang, Yukun Zhu, Ruoming Pang, Vijay Vasudevan, et al. 2019 · 2019
Earlier work this paper cites.
Newspaper copyrights, notices, and renewals
John Mark Ockerbloom. 2019 · 2019
Earlier work this paper cites.
Pytorch image models
Ross Wightman. 2019 · 2019
Earlier work this paper cites.
Detectron2
Yuxin Wu, Alexander Kirillov, Francisco Massa, Wan-Yen Lo, and Ross Girshick. 2019 · 2019
Earlier work this paper cites.
Experiment tracking with weights and biases
Lukas Biewald. 2020 · 2020
Earlier work this paper cites.
YOLOv5 by Ultralytics
Glenn Jocher. 2020 · 2020
Cited alongside, same era.
A large dataset of historical japanese documents with complex layouts
Zejiang Shen, Kaixuan Zhang, and Melissa Dell. 2020 · 2020
Cited alongside, same era.
Assessing the impact of ocr quality on downstream nlp tasks
Daniel van Strien., Kaspar Beelen., Mariona Coll Ardanuy., Kasra Hosseini., Barbara McGillivray., and Giovanni Colavizza. 2020 · 2020
Cited alongside, same era.
Xcit: Cross-covariance image transformers
Alaaeldin Ali, Hugo Touvron, Mathilde Caron, Piotr Bojanowski, Matthijs Douze, Armand Joulin, Ivan Laptev, Natalia Neverova, Gabriel Synnaeve, Jakob Verbeek, et al. 2021 · 2021
Cited alongside, same era.
Handwriting transformers
Ankan Kumar Bhunia, Salman Khan, Hisham Cholakkal, Rao Muhammad Anwer, Fahad Shahbaz Khan, and Mubarak Shah. 2021 · 2021
Cited alongside, same era.
Swin transformer: Hierarchical vision transformer using shifted windows
Exploring plain vision transformer backbones for object detection
Yanghao Li, Hanzi Mao, Ross Girshick, and Kaiming He. 2022 · 2022
Later among the works it cites.
Chronicling America: Historic American Newspapers
Library of Congress. 2022 · 2022
Later among the works it cites.
A convnet for the 2020s
Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie. 2022 · 2022
Later among the works it cites.
Edgenext: efficiently amalgamated cnn-transformer architecture for mobile vision applications
Muhammad Maaz, Abdelrahman Shaker, Hisham Cholakkal, Salman Khan, Syed Waqas Zamir, Rao Muhammad Anwer, and Fahad Shahbaz Khan. 2022 · 2022
Later among the works it cites.
PaddleOCR
PaddlePaddle. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. 2021 · 2021
Cited alongside, same era.
Neural ocr post-hoc correction of historical corpora
Lijun Lyu, Maria Koutraki, Martin Krickl, and Besnik Fetahu. 2021 · 2021
Cited alongside, same era.
Survey of post-ocr processing approaches
Thi Tuyet Hai Nguyen, Adam Jatowt, Mickael Coustaty, and Antoine Doucet. 2021 · 2021
Cited alongside, same era.
Onnx runtime
ONNX. 2021 · 2021
Cited alongside, same era.
Svtr: Scene text recognition with a single visual model
Yongkun Du, Zhineng Chen, Caiyan Jia, Xiaoting Yin, Tianlun Zheng, Chenxia Li, Yuning Du, and Yu-Gang Jiang. 2022 · 2022
Cited alongside, same era.
Trocr github repository
Minghao Li, Tengchao Lv, Lei Cui, Yijuan Lu, Dinei Florencio, Cha Zhang, Zhoujun Li, and Furu Wei. 2021a
Cited in the paper.
Trocr: Transformer-based optical character recognition with pre-trained models
Minghao Li, Tengchao Lv, Lei Cui, Yijuan Lu, Dinei Florencio, Cha Zhang, Zhoujun Li, and Furu Wei. 2021b
Cited in the paper.
Jacob Carlson, Tom Bryan, and Melissa Dell. 2023 · 2023
Closest in time.
American stories: A large-scale structured text dataset of historical u.s. newspapers
Melissa Dell, Jacob Carlson, Tom Bryan, Emily Silcock, Abhishek Arora, Zejiang Shen, Luca D’Amico-Wong, Quan Le, Pablo Querubin, and Leander Heldring. 2023 · 2023
Closest in time.
Tesseract: Open source ocr engine
J Ooms. 2023 · 2023
Closest in time.
Yolo v8 github repository
Ultalytics. 2023 · 2023
Closest in time.