Fetching the paper…
Reading the bibliography…
Document structure analysis (aka document layout analysis) is crucial for understanding the physical layout and logical structure of documents, with applications in information retrieval, document summarization, knowledge extraction, etc.
Y. Y. Tang, S.-W. Lee, C. Y. Suen, Automatic document processing: a survey, Pattern recognition 29 (12) (1996) 1931–1952
1952
Earlier work this paper cites.
H. W. Kuhn, The hungarian method for the assignment problem, Naval research logistics quarterly 2 (1-2) (1955) 83–97
1955
Earlier work this paper cites.
V. I. Levenshtein, et al., Binary codes capable of correcting deletions, insertions, and reversals, in: Soviet physics doklady, Vol. 10, 1966, pp. 707–710
1966
Earlier work this paper cites.
G. Nagy, S. C. Seth, Hierarchical representation of optically scanned documents (1984) 347–349
1984
Earlier work this paper cites.
S. Tsujimoto, H. Asada, Understanding multi-articled documents, in: Proceedings of the International Conference on Pattern Recognition, 1990, pp. 551–556
1990
Earlier work this paper cites.
J. Kreich, A. Luhn, G. Maderlechner, An experimental environment for model based document analysis, in: Proceedings of the International Conference on Document Analysis and Recognition, 1991, pp. 50–58
1991
Earlier work this paper cites.
A. Yamashita, A model based layout understanding method for the document recognition system, in: Proceedings of the International Conference on Document Analysis and Recognition, 1991, pp. 130–140
1991
Earlier work this paper cites.
M. Krishnamoorthy, G. Nagy, S. Seth, M. Viswanathan, Syntactic segmentation and labeling of digitized pages from technical journals, IEEE Transactions on Pattern Analysis and Machine Intelligence 15 (7) (1993) 737–747
1993
Earlier work this paper cites.
A. Conway, Page grammars and page parsing. a syntactic approach to document layout recognition, in: Proceedings of the International Conference on Document Analysis and Recognition, 1993, pp. 761–764
1993
Earlier work this paper cites.
Y. Tateisi, N. Itoh, Using stochastic syntactic analysis for extracting a logical structure from a document image, in: Proceedings of the IAPR International Conference on Pattern Recognition, 1994, pp. 391–394
1994
Earlier work this paper cites.
S. Hochreiter, J. Schmidhuber, Long short-term memory, Neural computation 9 (8) (1997) 1735–1780
1997
Earlier work this paper cites.
S. Mao, A. Rosenfeld, T. Kanungo, Document structure analysis algorithms: a literature survey, in: Proceedings of Document Recognition and Retrieval X, 2003, pp. 197–207
2003
Earlier work this paper cites.
T. M. Breuel, High performance document layout analysis, in: Proceedings of the Symposium on Document Image Understanding Technology, 2003, pp. 209–218
2003
Earlier work this paper cites.
M. Aiello, A. M. Smeulders, Bidimensional relations for reading order detection (2003). URL https://research.rug.nl/en/publications/bidimensional-relations-for-reading-order-detection
2003
Earlier work this paper cites.
J. Meunier, Optimized xy-cut for determining a page reading order, in: Proceedings of the International Conference on Document Analysis and Recognition, 2005, pp. 347–351
2005
Earlier work this paper cites.
M. Ceci, M. Berardi, G. Porcelli, D. Malerba, A data mining approach to reading order detection, in: Proceedings of the International Conference on Document Analysis and Recognition, 2007, pp. 924–928
2007
Earlier work this paper cites.
D. Malerba, M. Ceci, Learning to order: A relational approach, in: Proceedings of the ECML/PKDD International Workshop on Mining Complex Data, Vol. 4944, 2007, pp. 209–223
2007
Earlier work this paper cites.
Z. Wu, P. Mitra, C. L. Giles, Table of contents recognition and extraction for heterogeneous book documents, in: Proceedings of the International Conference on Document Analysis and Recognition, 2013, pp. 1205–1209
2013
Earlier work this paper cites.
R. Girshick, J. Donahue, T. Darrell, J. Malik, Rich feature hierarchies for accurate object detection and semantic segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2014, pp. 580–587
2014
Earlier work this paper cites.
R. Girshick, Fast r-cnn, in: Proceedings of the International Conference on Computer Vision, 2015, pp. 1440–1448
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, J. Sun, Faster r-cnn: Towards real-time object detection with region proposal networks, in: Proceedings of the Advances in Neural Information Processing Systems, 2015, pp. 91–99
2015
Earlier work this paper cites.
J. Long, E. Shelhamer, T. Darrell, Fully convolutional networks for semantic segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2015, pp. 3431–3440
2015
Earlier work this paper cites.
S. Ferilli, A. Pazienza, An abstract argumentation-based strategy for reading order detection, in: Proceedings of the AI*IA Workshop on Intelligent Techniques, Vol. 1509, 2015
2015
Earlier work this paper cites.
J. L. Ba, J. R. Kiros, G. E. Hinton, Layer normalization, arXiv preprint arXiv:1607.06450 (2016)
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, J. Sun, Deep residual learning for image recognition, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2016, pp. 770–778
2016
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, R. Girshick, Mask r-cnn, in: Proceedings of the International Conference on Computer Vision, 2017, pp. 2961–2969
2017
Earlier work this paper cites.
X. Yang, E. Yumer, P. Asente, M. Kraley, D. Kifer, C. Lee Giles, Learning to extract semantic structure from documents using multimodal fully convolutional neural networks, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2017, pp. 5315–5324
2017
Earlier work this paper cites.
L. Gao, X. Yi, Z. Jiang, L. Hao, Z. Tang, ICDAR2017 competition on page object detection, in: Proceedings of the International Conference on Document Analysis and Recognition, 2017, pp. 1417–1422
2017
Earlier work this paper cites.
X. Yi, L. Gao, Y. Liao, X. Zhang, R. Liu, Z. Jiang, Cnn based page object detection in document images, in: Proceedings of the International Conference on Document Analysis and Recognition, Vol. 1, 2017, pp. 230–235
2017
Cited alongside, same era.
D. A. B. Oliveira, M. P. Viana, Fast cnn-based document layout analysis, in: Proceedings of the International Conference on Computer Vision Workshops, 2017, pp. 1173–1180
2017
Cited alongside, same era.
D. He, S. Cohen, B. Price, D. Kifer, C. L. Giles, Multi-scale multi-task fcn for semantic page segmentation and table detection, in: Proceedings of the International Conference on Document Analysis and Recognition, Vol. 1, 2017, pp. 254–261
2017
Cited alongside, same era.
T. Nguyen, A. Doucet, M. Coustaty, Enhancing table of contents extraction by system aggregation, in: Proceedings of the International Conference on Document Analysis and Recognition, 2017, pp. 242–247
2017
Cited alongside, same era.
2021
Later among the works it cites.
M. Minouei, M. R. Soheili, D. Stricker, Document layout analysis with an enhanced object detector, in: Proceedings of the International Conference on Pattern Recognition and Image Analysis, 2021, pp. 1–5
2021
Later among the works it cites.
2022
Later among the works it cites.
J. Li, Y. Xu, T. Lv, L. Cui, C. Zhang, F. Wei, Dit: Self-supervised pre-training for document image transformer, in: Proceedings of the ACM International Conference on Multimedia, 2022, pp. 3530–3539
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Zhang, M. Elhoseiny, S. Cohen, W. Chang, A. Elgammal, Relationship proposal networks, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2017, pp. 5678–5686
2017
Cited alongside, same era.
2017
Cited alongside, same era.
N. D. Vo, K. Nguyen, T. V. Nguyen, K. Nguyen, Ensemble of deep object detectors for page object detection, in: Proceedings of the International Conference on Ubiquitous Information Management and Communication, 2018, pp. 1–6
2018
Cited alongside, same era.
Y. Li, Y. Zou, J. Ma, Deeplayout: A semantic segmentation approach to page layout analysis, in: Proceedings of the International Conference on Intelligent Computing Methodologies, 2018, pp. 266–277
2018
Cited alongside, same era.
X. Li, F. Yin, C. Liu, Page object detection from pdf document images by deep structured prediction and supervised clustering, in: Proceedings of the International Conference on Pattern Recognition, 2018, pp. 3627–3632
2018
Cited alongside, same era.
X. Zhong, J. Tang, A. J. Yepes, Publaynet: largest dataset ever for document layout analysis, in: Proceedings of the International Conference on Document Analysis and Recognition, 2019, pp. 1015–1022
2019
Cited alongside, same era.
Z. Cai, N. Vasconcelos, Cascade r-cnn: High quality object detection and instance segmentation, IEEE Transactions on Pattern Analysis and Machine Intelligence 43 (5) (2019) 1483–1498
2019
Cited alongside, same era.
R. Saha, A. Mondal, C. Jawahar, Graphical object detection in document images, in: Proceedings of the International Conference on Document Analysis and Recognition, 2019, pp. 51–58
2019
Cited alongside, same era.
2022
Later among the works it cites.
H. Yang, W. Hsu, Transformer-based approach for document layout understanding, in: Proceedings of the International Conference on Image Processing, 2022, pp. 4043–4047
2022
Later among the works it cites.
C. Shi, C. Xu, H. Bi, Y. Cheng, Y. Li, H. Zhang, Lateral feature enhancement network for page object detection, IEEE Transactions on Instrumentation and Measurement 71 (2022) 1–10
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Huang, T. Lv, L. Cui, Y. Lu, F. Wei, Layoutlmv3: Pre-training for document ai with unified text and image masking, in: Proceedings of the ACM International Conference on Multimedia, 2022, pp. 4083–4091
2022
Later among the works it cites.
Y. Sang, Y. Zeng, R. Liu, F. Yang, Z. Yao, Y. Pan, Exploiting spatial attention and contextual information for document image segmentation, in: Proceedings of the Advances in Knowledge Discovery and Data Mining, 2022, pp. 261–274
2022
Later among the works it cites.
S. Luo, Y. Ding, S. Long, J. Poon, S. C. Han, Doc-gcn: Heterogeneous graph convolutional networks for document layout analysis, in: Proceedings of the International Conference on Computational Linguistics, 2022, pp. 2906–2916
2022
Later among the works it cites.
R. Wang, Y. Fujii, A. C. Popat, Post-ocr paragraph recognition by graph convolutional networks, in: Proceedings of the IEEE Winter Conference on Applications of Computer Vision, 2022, pp. 493–502
2022
Later among the works it cites.
S. Liu, R. Wang, M. Raptis, Y. Fujii, Unified line and paragraph detection by graph convolutional networks, in: Proceedings of the International Workshop on Document Analysis Systems, 2022, pp. 33–47
2022
Later among the works it cites.
S. Long, S. Qin, D. Panteleev, A. Bissacco, Y. Fujii, M. Raptis, Towards end-to-end unified scene text detection and layout analysis, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 1049–1059
2022
Later among the works it cites.
C. Xue, J. Huang, W. Zhang, S. Lu, C. Wang, S. Bai, Contextual text block detection towards scene text understanding, in: Proceedings of the European Conference on Computer Vision, 2022, pp. 374–391
2022
Later among the works it cites.
Z. Gu, C. Meng, K. Wang, J. Lan, W. Wang, M. Gu, L. Zhang, Xylayoutlm: Towards layout-aware multimodal networks for visually-rich document understanding, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 4583–4592
2022
Later among the works it cites.
L. Quirós, E. Vidal, Reading order detection on handwritten documents, Neural Computing and Applications 34 (12) (2022) 9593–9611
2022
Later among the works it cites.
R. Cao, Y. Cao, G. Zhou, P. Luo, Extracting variable-depth logical document hierarchy from long documents: Method, evaluation, and application, Journal of Computer Science and Technology 37 (3) (2022) 699–718
2022
Later among the works it cites.
P. Hu, Z. Zhang, J. Zhang, J. Du, J. Wu, Multimodal tree decoder for table of contents extraction in document images, in: Proceedings of the International Conference on Pattern Recognition, 2022, pp. 1756–1762
2022
Later among the works it cites.
S. Naik, K. A. Hashmi, A. Pagani, M. Liwicki, D. Stricker, M. Z. Afzal, Investigating attention mechanism for page object detection in document images, Applied Sciences 12 (15) (2022) 7486
2022
Later among the works it cites.
H. Bi, C. Xu, C. Shi, G. Liu, Y. Li, H. Zhang, J. Qu, Srrv: A novel document object detector based on spatial-related relation and vision, IEEE Transactions on Multimedia 25 (2022) 3788–3798
2022
Later among the works it cites.
B. Cheng, I. Misra, A. G. Schwing, A. Kirillov, R. Girdhar, Masked-attention mask transformer for universal image segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 1280–1289
2022
Later among the works it cites.
J. Ma, J. Du, P. Hu, Z. Zhang, J. Zhang, H. Zhu, C. Liu, Hrdoc: Dataset and baseline method toward hierarchical reconstruction of document structures, in: Proceedings of the AAAI Conference on Artificial Intelligence, 2023, pp. 1870–1877
2023
Later among the works it cites.
Z. Zhong, J. Wang, H. Sun, K. Hu, E. Zhang, L. Sun, Q. Huo, A hybrid approach to document layout analysis for heterogeneous document images, in: Proceedings of the International Conference on Document Analysis and Recognition, 2023, pp. 189––206
2023
Later among the works it cites.
R. Wang, Y. Fujii, A. Bissacco, Text reading order in uncontrolled conditions by sparse graph segmentation, in: Proceedings of the International Conference on Document Analysis and Recognition, 2023, pp. 3–21
2023
Later among the works it cites.
H. Zhang, F. Li, S. Liu, L. Zhang, H. Su, J. Zhu, L. M. Ni, H. Shum, DINO: DETR with improved denoising anchor boxes for end-to-end object detection, in: Proceedings of the International Conference on Learning Representations, 2023
2023
Later among the works it cites.
K. Hu, Z. Zhong, L. Sun, Q. Huo, Mathematical formula detection in document images: A new dataset and a new approach, Pattern Recognition 148 (2024) 110212
2024
Closest in time.