Fetching the paper…
Reading the bibliography…
Thousands of users consult digital archives daily, but the information they can access is unrepresentative of the diversity of documentary history.
Mmdetection: Open mmlab detection toolbox and benchmark
Kai Chen, Jiaqi Wang, Jiangmiao Pang, Yuhang Cao, Yu Xiong, Xiaoxiao Li, Shuyang Sun, Wansen Feng, Ziwei Liu, Jiarui Xu, et al. 2019 · 1906
Earlier work this paper cites.
Jinji koshinroku
Jinji Koshinjo. 1939 · 1939
Earlier work this paper cites.
Nihon shokuinrokj
Jinji Koshinjo. 1954 · 1954
Earlier work this paper cites.
Teikoku Ginko Kaisha Yoroku
Teikoku Koshinjo. 1957 · 1957
Earlier work this paper cites.
The state and fate of linguistic diversity and inclusion in the nlp world
Pratik Joshi, Sebastin Santy, Amar Budhiraja, Kalika Bali, and Monojit Choudhury. 2020 · 2004
Earlier work this paper cites.
Supervised contrastive learning
Prannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna, Yonglong Tian, Phillip Isola, Aaron Maschinot, Ce Liu, and Dilip Krishnan. 2020 · 2004
Earlier work this paper cites.
Bootstrap your own latent: A new approach to self-supervised learning
Jean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec, Pierre H Richemond, Elena Buchatskaya, Carl Doersch, Bernardo Avila Pires, Zhaohan Daniel Guo, Mohammad Gheshlaghi Azar, et al. 2020 · 2006
Earlier work this paper cites.
Ocr post correction for endangered language texts
Shruti Rijhwani, Antonios Anastasopoulos, and Graham Neubig. 2020 · 2011
Earlier work this paper cites.
Grpoly-db: An old greek polytonic document image database
Basilis Gatos, Nikolaos Stamatopoulos, Georgios Louloudis, Giorgos Sfikas, George Retsinas, Vassilis Papavassiliou, Fotini Sunistira, and Vassilis Katsouros. 2015 · 2015
Earlier work this paper cites.
Computational methods for uncovering reprinted texts in antebellum newspapers
David A Smith, Ryan Cordell, and Abby Mullen. 2015 · 2015
Earlier work this paper cites.
An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition
Baoguang Shi, Xiang Bai, and Cong Yao. 2016 · 2016
Earlier work this paper cites.
Impact of ocr errors on the use of digital libraries: Towards a better access to information
Guillaume Chiron, Antoine Doucet, Mickaël Coustaty, Muriel Visani, and Jean-Philippe Moreux. 2017 · 2017
Earlier work this paper cites.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick. 2017 · 2017
Earlier work this paper cites.
Cascade r-cnn: Delving into high quality object detection
Zhaowei Cai and Nuno Vasconcelos. 2018 · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Earlier work this paper cites.
Cascade r-cnn: High quality object detection and instance segmentation
Zhaowei Cai and Nuno Vasconcelos. 2019 · 2019
Earlier work this paper cites.
Searching for mobilenetv3
Andrew Howard, Mark Sandler, Grace Chu, Liang-Chieh Chen, Bo Chen, Mingxing Tan, Weijun Wang, Yukun Zhu, Ruoming Pang, Vijay Vasudevan, et al. 2019 · 2019
Earlier work this paper cites.
Icdar2019 competition on scanned receipt ocr and information extraction
Zheng Huang, Kai Chen, Jianhua He, Xiang Bai, Dimosthenis Karatzas, Shijian Lu, and CV Jawahar. 2019 · 2019
Cited alongside, same era.
Billion-scale similarity search with gpus
Jeff Johnson, Matthijs Douze, and Hervé Jégou. 2019 · 2019
Cited alongside, same era.
Pytorch image models
Ross Wightman. 2019 · 2019
Cited alongside, same era.
Detectron2
Yuxin Wu, Alexander Kirillov, Francisco Massa, Wan-Yen Lo, and Ross Girshick. 2019 · 2019
Cited alongside, same era.
Convolutional character networks
Linjie Xing, Zhi Tian, Weilin Huang, and Matthew R Scott. 2019 · 2019
Cited alongside, same era.
YOLOv5 by Ultralytics
Glenn Jocher. 2020 · 2020
Cited alongside, same era.
An empirical study of training self-supervised vision transformers
Xinlei Chen, Saining Xie, and Kaiming He. 2021 · 2021
Later among the works it cites.
Training vision transformers for image retrieval
Alaaeldin El-Nouby, Natalia Neverova, Ivan Laptev, and Hervé Jégou. 2021 · 2021
Later among the works it cites.
A survey on recent approaches for natural language processing in low-resource scenarios
Michael A. Hedderich, Lukas Lange, Heike Adel, Jannik Strötgen, and Dietrich Klakow. 2021 · 2021
Later among the works it cites.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. 2021 · 2021
Later among the works it cites.
Neural ocr post-hoc correction of historical corpora
Lijun Lyu, Maria Koutraki, Martin Krickl, and Besnik Fetahu. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The newspaper navigator dataset: extracting headlines and visual content from 16 million historic newspaper pages in chronicling america
Benjamin Charles Germain Lee, Jaime Mears, Eileen Jakeway, Meghan Ferriter, Chris Adams, Nathan Yarasavage, Deborah Thomas, Kate Zwaard, and Daniel S Weld. 2020 · 2020
Cited alongside, same era.
Kevin Musgrave, Serge Belongie, and Ser-Nam Lim. 2020 · 2020
Cited alongside, same era.
A large dataset of historical japanese documents with complex layouts
Zejiang Shen, Kaixuan Zhang, and Melissa Dell. 2020 · 2020
Cited alongside, same era.
Revisiting the sibling head in object detector
Guanglu Song, Yu Liu, and Xiaogang Wang. 2020 · 2020
Cited alongside, same era.
Assessing the impact of ocr quality on downstream nlp tasks
Daniel van Strien., Kaspar Beelen., Mariona Coll Ardanuy., Kasra Hosseini., Barbara McGillivray., and Giovanni Colavizza. 2020 · 2020
Cited alongside, same era.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
Survey of post-ocr processing approaches
Thi Tuyet Hai Nguyen, Adam Jatowt, Mickael Coustaty, and Antoine Doucet. 2021 · 2021
Later among the works it cites.
Layoutparser: A unified toolkit for deep learning based document image analysis
Zejiang Shen, Ruochen Zhang, Melissa Dell, Benjamin Charles Germain Lee, Jacob Carlson, and Weining Li. 2021 · 2021
Later among the works it cites.
Svtr: Scene text recognition with a single visual model
Yongkun Du, Zhineng Chen, Caiyan Jia, Xiaoting Yin, Tianlun Zheng, Chenxia Li, Yuning Du, and Yu-Gang Jiang. 2022 · 2022
Later among the works it cites.
Historical newspaper data: A researcher’s guide and toolkit
W Walker Hanlon and Brian Beach. 2022 · 2022
Later among the works it cites.
Exploring plain vision transformer backbones for object detection
Yanghao Li, Hanzi Mao, Ross Girshick, and Kaiming He. 2022 · 2022
Later among the works it cites.
Chronicling America: Historic American Newspapers
Library of Congress. 2022 · 2022
Later among the works it cites.
A convnet for the 2020s
Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie. 2022 · 2022
Later among the works it cites.
A survey of historical document image datasets
Konstantina Nikolaidou, Mathias Seuret, Hamam Mokayed, and Marcus Liwicki. 2022 · 2022
Later among the works it cites.
EfficientOCR: An extensible, open-source package for efficiently digitizing world knowledge
Tom Bryan, Jacob Carlson, Abhishek Arora, and Melissa Dell. 2023 · 2023
Closest in time.
American stories: A large-scale structured text dataset of historical us newspapers
Melissa Dell, Jacob Carlson, Tom Bryan, Emily Silcock, Abhishek Arora, Zejiang Shen, Luca D’Amico-Wong, Quan Le, Pablo Querubin, and Leander Heldring. 2023 · 2023
Closest in time.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al. 2023 · 2023
Closest in time.