Fetching the paper…
Reading the bibliography…
The problem of document structure reconstruction refers to converting digital or scanned documents into corresponding semantic structures.
Recursive X-Y cut using bounding boxes of connected components
Ha, J.; Haralick, R.; and Phillips, I. 1995 · 1995
Earlier work this paper cites.
A fast algorithm for bottom-up document layout analysis
Simon, A.; Pret, J.-C.; and Johnson, A. P. 1997 · 1997
Earlier work this paper cites.
Text Extraction and Document Image Segmentation Using Matched Wavelets and MRF Model
Kumar, S.; Gupta, R.; Khanna, N.; Chaudhury, S.; and Joshi, S. D. 2007 · 2007
Earlier work this paper cites.
Document structure and layout analysis
Namboodiri, A. M.; and Jain, A. K. 2007 · 2007
Earlier work this paper cites.
An Overview of the Tesseract OCR Engine
Smith, R. 2007 · 2007
Earlier work this paper cites.
Metadata extraction from PDF papers for digital library ingest
Marinai, S. 2009 · 2009
Earlier work this paper cites.
Table of contents recognition and extraction for heterogeneous book documents
Wu, Z.; Mitra, P.; and Giles, C. L. 2013 · 2013
Earlier work this paper cites.
On the Properties of Neural Machine Translation: Encoder-Decoder Approaches
Cho, K.; van Merrienboer, B.; Bahdanau, D.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
A hybrid approach to discover semantic hierarchical sections in scholarly documents
Tuarob, S.; Mitra, P.; and Giles, C. L. 2015 · 2015
Earlier work this paper cites.
Ba, J. L.; Kiros, J. R.; and Hinton, G. E. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Tree edit distance: Robust and memory-efficient
Pawlik, M.; and Augsten, N. 2016 · 2016
Cited alongside, same era.
A Deep Learning-Based Formula Detection Method for PDF Documents
Gao, L.; Yi, X.; Liao, Y.; Jiang, Z.; Yan, Z.; and Tang, Z. 2017 · 2017
Cited alongside, same era.
Multi-Scale Multi-Task FCN for Semantic Page Segmentation and Table Detection
He, D.; Cohen, S.; Price, B.; Kifer, D.; and Giles, C. L. 2017a · 2017
Cited alongside, same era.
Mask R-CNN
He, K.; Gkioxari, G.; Dollár, P.; and Girshick, R. B. 2017b · 2017
Cited alongside, same era.
Feature Pyramid Networks for Object Detection
Lin, T.; Dollár, P.; Girshick, R. B.; He, K.; Hariharan, B.; and Belongie, S. J. 2017a · 2017
Cited alongside, same era.
Focal Loss for Dense Object Detection
Lin, T.; Goyal, P.; Girshick, R. B.; He, K.; and Dollár, P. 2017b · 2017
Cited alongside, same era.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Reimers, N.; and Gurevych, I. 2019 · 2019
Later among the works it cites.
PubLayNet: Largest Dataset Ever for Document Layout Analysis
Zhong, X.; Tang, J.; and Jimeno-Yepes, A. 2019 · 2019
Later among the works it cites.
The Financial Document Structure Extraction Shared task (FinToc 2020)
Bentabet, N.-I.; Juge, R.; El Maarouf, I.; Mouilleron, V.; Valsamou-Stanislawski, D.; and El-Haj, M. 2020 · 2020
Later among the works it cites.
DocBank: A Benchmark Dataset for Document Layout Analysis
Li, M.; Xu, Y.; Cui, L.; Huang, S.; Wei, F.; Li, Z.; and Zhou, M. 2020 · 2020
Later among the works it cites.
DocStruct: A Multimodal Method to Extract Hierarchy Structure in Document for General Form Understanding
Wang, Z.; Zhan, M.; Liu, X.; and Liang, D. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Enhancing Table of Contents Extraction by System Aggregation
Nguyen, T.-T.-H.; Doucet, A.; and Coustaty, M. 2017 · 2017
Cited alongside, same era.
Attention is All you Need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, L.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
CNN Based Page Object Detection in Document Images
Yi, X.; Gao, L.; Liao, Y.; Zhang, X.; Liu, R.; and Jiang, Z. 2017 · 2017
Cited alongside, same era.
Cascade R-CNN: Delving Into High Quality Object Detection
Cai, Z.; and Vasconcelos, N. 2018 · 2018
Cited alongside, same era.
Page Object Detection from PDF Document Images by Deep Structured Prediction and Supervised Clustering
Li, X.-H.; Yin, F.; and Liu, C.-L. 2018 · 2018
Cited alongside, same era.
Cui, L.; Xu, Y.; Lv, T.; and Wei, F. 2021 · 2021
Later among the works it cites.
Docparser: Hierarchical document structure parsing from renderings
Rausch, J.; Martinez, O.; Bissig, F.; Zhang, C.; and Feuerriegel, S. 2021 · 2021
Later among the works it cites.
LayoutLMv2: Multi-modal Pre-training for Visually-rich Document Understanding
Xu, Y.; Xu, Y.; Lv, T.; Cui, L.; Wei, F.; Wang, G.; Lu, Y.; Florêncio, D. A. F.; Zhang, C.; Che, W.; Zhang, M.; and Zhou, L. 2021 · 2021
Later among the works it cites.
DocLayNet: A Large Human-Annotated Dataset for Document-Layout Analysis
Pfitzmann, B.; Auer, C.; Dolfi, M.; Nassar, A. S.; and Staar, P. W. J. 2022 · 2022
Later among the works it cites.
Multimodal Pre-training Based on Graph Attention Network for Document Understanding
Zhang, Z.; Ma, J.; Du, J.; Wang, L.; and Zhang, J. 2022 · 2022
Later among the works it cites.