Fetching the paper…
Reading the bibliography…
Understanding documents with rich layouts is an essential step towards information extraction.
N. Journet, V. Eglin, J.-Y. Ramel, R. Mullot, Text/graphic labelling of ancient printed documents, in: Proceedings of the ICDAR, 2005, pp. 1010–1014
2005
Earlier work this paper cites.
J. Fang, L. Gao, K. Bai, R. Qiu, X. Tao, Z. Tang, A table detection method for multipage pdf documents via visual seperators and tabular structures, in: ICDAR, 2011
2011
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, C. L. Zitnick, Microsoft coco: Common objects in context, in: Proceedings of the ECCV, 2014
2014
Earlier work this paper cites.
A. Asi, R. Cohen, K. Kedem, J. El-Sana, Simplifying the reading of historical manuscripts, in: Proceedings of the ICDAR, 2015
2015
Earlier work this paper cites.
T. A. Tran, I.-S. Na, S.-H. Kim, Hybrid page segmentation using multilevel homogeneity structure, in: Proceedings of the 9th International Conference on Ubiquitous Information Management and Communication, 2015, pp. 1–6
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, J. Sun, Faster r-cnn: Towards real-time object detection with region proposal networks, in: NIPS, 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, J. Sun, Deep residual learning for image recognition, in: Proceedings of the IEEE Conference on CVPR, 2016, pp. 770–778
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, I. Polosukhin, Attention is all you need, in: NIPS, 2017
2017
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, R. Girshick, Mask r-cnn, in: Proceedings of the ICCV, 2017, pp. 2961–2969
2017
Earlier work this paper cites.
T.-Y. Lin, P. Goyal, R. Girshick, K. He, P. Dollár, Focal loss for dense object detection, in: CVPR, 2017
2017
Earlier work this paper cites.
S. Schreiber, S. Agne, I. Wolf, A. Dengel, S. Ahmed, Deepdesrt: Deep learning for detection and structure recognition of tables in document images, in: ICDAR, 2017
2017
Earlier work this paper cites.
D. He, S. Cohen, B. Price, D. Kifer, C. L. Giles, Multi-scale multi-task fcn for semantic page segmentation and table detection, in: Proceedings of the ICDAR, Vol. 1, 2017, pp. 254–261
2017
Cited alongside, same era.
S. A. Oliveira, B. Seguin, F. Kaplan, dhsegment: A generic deep-learning approach for document segmentation, in: ICFHR, 2018
2018
Cited alongside, same era.
Z. Huang, L. Huang, Y. Gong, C. Huang, X. Wang, Mask scoring r-cnn, in: Proceedings of the IEEE Conference on CVPR, 2019, pp. 6409–6418
2019
Cited alongside, same era.
X. Zhong, J. Tang, A. J. Yepes, Publaynet: largest dataset ever for document layout analysis, in: Proceedings of the ICDAR, 2019, pp. 1015–1022
2019
Cited alongside, same era.
C. Clausner, A. Antonacopoulos, S. Pletschacher, Icdar2019 competition on recognition of documents with complex layouts-rdcl2019, in: Proceedings of the ICDAR, 2019, pp. 1521–1526
L. Wang, C. Wang, Z. Sun, S. Chen, An improved dice loss for pneumothorax segmentation by mining the information of negative areas, IEEE Access, 2020
2020
Later among the works it cites.
Z. Shen, K. Zhang, M. Dell, A large dataset of historical japanese documents with complex layouts, in: Proceedings of the IEEE Conference on CVPRW, 2020, pp. 548–549
2020
Later among the works it cites.
D. Prasad, A. Gadpal, K. Kapadni, M. Visave, K. Sultanpure, Cascadetabnet: An approach for end to end table detection and structure recognition from image-based documents, in: CVPRW, 2020, pp. 572–573
2020
Later among the works it cites.
S. Biswas, P. Riba, J. Lladós, U. Pal, Beyond document object detection: instance-level segmentation of complex layouts, IJDAR, 2021
2021
Later among the works it cites.
S. Appalaraju, B. Jasani, B. U. Kota, Y. Xie, R. Manmatha, Docformer: End-to-end transformer for document understanding, ICCV, 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
K. Li, C. Wigington, C. Tensmeyer, H. Zhao, N. Barmpalios, V. I. Morariu, V. Manjunatha, T. Sun, Y. Fu, Cross-domain document object detection: Benchmark suite and method, in: Proceedings of the IEEE Conference on CVPR, 2020
2020
Cited alongside, same era.
X. Wang, R. Zhang, T. Kong, L. Li, C. Shen, Solov2: Dynamic and fast instance segmentation, NIPS, 2020
2020
Cited alongside, same era.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, S. Zagoruyko, End-to-end object detection with transformers, in: ECCV, 2020
2020
Cited alongside, same era.
Y. Xu, M. Li, L. Cui, S. Huang, F. Wei, M. Zhou, Layoutlm: Pre-training of text and layout for document image understanding, in: ACM SIGKDD, 2020
2020
Cited alongside, same era.
X. Huang, Z. Ge, Z. Jie, O. Yoshie, Nms by representative region: Towards crowded pedestrian detection by proposal pairing, in: CVPR, 2020
2020
Cited alongside, same era.
Cited in the paper.
Cited in the paper.
2021
Later among the works it cites.
R. Guo, D. Niu, L. Qu, Z. Li, Sotr: Segmenting objects with transformers, in: ICCV, 2021
2021
Later among the works it cites.
Z. Shen, R. Zhang, M. Dell, B. C. G. Lee, J. Carlson, W. Li, Layoutparser: A unified toolkit for deep learning based document image analysis, in: ICDAR, 2021, pp. 131–146
2021
Later among the works it cites.
P. Li, J. Gu, J. Kuen, V. I. Morariu, H. Zhao, R. Jain, V. Manjunatha, H. Liu, Selfdoc: Self-supervised document representation learning, in: CVPR, 2021, pp. 5652–5660
2021
Later among the works it cites.
M. Mathew, D. Karatzas, C. Jawahar, Docvqa: A dataset for vqa on document images, in: WACV, 2021
2021
Later among the works it cites.
M. Ju, J. Luo, Z. Wang, H. Luo, Adaptive feature fusion with attention mechanism for multi-scale target detection, Neural Computing and Applications, 2021
2021
Later among the works it cites.