Fetching the paper…
Reading the bibliography…
Accurate document layout analysis is a key requirement for high-quality PDF document conversion.
Icdar 2013 table competition
Max Göbel, Tamir Hassan, Ermelinda Oro, and Giorgio Orsi · 2013
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
Ross B. Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik · 2014
Earlier work this paper cites.
Microsoft COCO: common objects in context, 2014
Tsung-Yi Lin, Michael Maire, Serge J. Belongie, Lubomir D. Bourdev, Ross B. Girshick, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick · 2014
Earlier work this paper cites.
Fast R-CNN
Ross B. Girshick · 2015
Earlier work this paper cites.
Information extraction from pdf sources based on rule-based system using integrated formats
Riaz Ahmad, Muhammad Tanvir Afzal, and M. Qadir · 2016
Earlier work this paper cites.
Icdar2017 competition on recognition of documents with complex layouts - rdcl2017
Christian Clausner, Apostolos Antonacopoulos, and Stefan Pletschacher · 2017
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2017
Earlier work this paper cites.
Mask R-CNN
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross B. Girshick · 2017
Earlier work this paper cites.
Corpus conversion service: A machine learning platform to ingest documents at scale
Peter W J Staar, Michele Dolfi, Christoph Auer, and Costas Bekas · 2018
Cited alongside, same era.
ICDAR 2019 Competition on Table Detection and Recognition (cTDaR), April 2019
Hervé Déjean, Jean-Luc Meunier, Liangcai Gao, Yilun Huang, Yu Fang, Florian Kleber, and Eva-Maria Lang · 2019
Cited alongside, same era.
Publaynet: Largest dataset ever for document layout analysis
Xu Zhong, Jianbin Tang, and Antonio Jimeno-Yepes · 2019
Cited alongside, same era.
Efficientdet: Scalable and efficient object detection
Mingxing Tan, Ruoming Pang, and Quoc V. Le · 2019
Cited alongside, same era.
Detectron2, 2019
Yuxin Wu, Alexander Kirillov, Francisco Massa, Wan-Yen Lo, and Ross Girshick · 2019
Cited alongside, same era.
A survey on image data augmentation for deep learning
Layoutlm: Pre-training of text and layout for document image understanding
Yiheng Xu, Minghao Li, Lei Cui, Shaohan Huang, Furu Wei, and Ming Zhou · 2020
Later among the works it cites.
Competition on scientific literature parsing
Antonio Jimeno Yepes, Peter Zhong, and Douglas Burdick · 2021
Later among the works it cites.
ultralytics/yolov5: v6.0 - yolov5n nano models, roboflow integration, tensorflow export, opencv dnn support, October 2021
Glenn Jocher, Alex Stoken, Ayush Chaurasia, Jirka Borovec, NanoCode012, TaoXie, Yonghye Kwon, Kalen Michael, Liu Changyu, Jiacong Fang, Abhiram V, Laughing, tkianai, yxNONG, Piotr Skalski, Adam Hogan, Jebastin Nadar, imyhxy, Lorenzo Mammana, Alex Wang, Cristi Fati, Diego Montes, Jan Hajek, Laurentiu Diaconu, Mai Thanh Minh, Marc, albinxavi, fatih, oleg, and wanghao yang · 2021
Later among the works it cites.
Robust pdf document conversion using recurrent neural networks
Nikolaos Livathinos, Cesar Berrospi, Maksym Lysak, Viktor Kuropiatnyk, Ahmed Nassar, Andre Carvalho, Michele Dolfi, Christoph Auer, Kasper Dinkla, and Peter W. J. Staar · 2021
Later among the works it cites.
Vtlayout: Fusion of visual and text features for document layout analysis, 2021
Shoubin Li, Xuyan Ma, Shuaiqun Pan, Jun Hu, Lin Shi, and Qing Wang · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Connor Shorten and Taghi M. Khoshgoftaar · 2019
Cited alongside, same era.
Docbank: A benchmark dataset for document layout analysis
Minghao Li, Yiheng Xu, Lei Cui, Shaohan Huang, Furu Wei, Zhoujun Li, and Ming Zhou · 2020
Cited alongside, same era.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Cited alongside, same era.
Later among the works it cites.
Vsr: A unified framework for document layout analysis combining vision, semantics and relations, 2021
Peng Zhang, Can Li, Liang Qiao, Zhanzhan Cheng, Shiliang Pu, Yi Niu, and Fei Wu · 2021
Later among the works it cites.
Segmentation for document layout analysis: not dead yet
Logan Markewich, Hao Zhang, Yubin Xing, Navid Lambert-Shirzad, Jiang Zhexin, Roy Lee, Zhi Li, and Seok-Bum Ko · 2022
Closest in time.