Fetching the paper…
Reading the bibliography…
Pre-training techniques have been verified successfully in a variety of NLP tasks in recent years.
A fast and efficient method for extracting text paragraphs and graphics from unconstrained documents. In Proceedings., 11th IAPR International Conference on Pattern Recognition. Vol. II. Conference B: Pattern Recognition Methodology and Systems . IEEE, 272–276
Frank Lebourgeois, Z Bublinski, and H Emptoz. 1992 · 1992
Earlier work this paper cites.
The document spectrum for page layout analysis
L. O’Gorman. 1993 · 1993
Earlier work this paper cites.
A fast algorithm for bottom-up document layout analysis
Anikó Simon, J-C Pret, and A Peter Johnson. 1997 · 1997
Earlier work this paper cites.
Graph Convolution for Multimodal Information Extraction from Visually Rich Documents. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 2 (Industry Papers) . Association for Computational Linguistics, Minneapolis, Minnesota, 32–39
Xiaojing Liu, Feiyu Gao, Qiong Zhang, and Huasha Zhao. 2019a · 2005
Earlier work this paper cites.
Artificial neural networks for document analysis and recognition
S. Marinai, M. Gori, and G. Soda. 2005 · 2005
Earlier work this paper cites.
Learning nongenerative grammatical models for document analysis. In Tenth IEEE International Conference on Computer Vision (ICCV’05) Volume 1 , Vol. 2. IEEE, 962–969
Michael Shilman, Percy Liang, and Paul Viola. 2005 · 2005
Earlier work this paper cites.
Building a Test Collection for Complex Document Information Processing. In Proceedings of the 29th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval (Seattle, Washington, USA) (SIGIR ’06) . ACM, New York, NY, USA, 665–666
D. Lewis, G. Agam, S. Argamon, O. Frieder, D. Grossman, and J. Heard. 2006 · 2006
Earlier work this paper cites.
Evaluation of SVM, MLP and GMM Classifiers for Layout Analysis of Historical Documents. In 2013 12th International Conference on Document Analysis and Recognition . 1220–1224
H. Wei, M. Baechler, F. Slimane, and R. Ingold. 2013 · 2013
Earlier work this paper cites.
Evaluation of deep convolutional nets for document image classification and retrieval
Adam W. Harley, Alex Ufkes, and Konstantinos G. Derpanis. 2015 · 2015
Earlier work this paper cites.
Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
Shaoqing Ren, Kaiming He, Ross B. Girshick, and Jian Sun. 2015 · 2015
Cited alongside, same era.
A Table Detection Method for PDF Documents Based on Convolutional Neural Networks
Leipeng Hao, Liangcai Gao, Xiaohan Yi, and Zhi Tang. 2016 · 2016
Cited alongside, same era.
Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations
Ranjay Krishna, Yuke Zhu, Oliver Groth, Justin Johnson, Kenji Hata, Joshua Kravitz, Stephanie Chen, Yannis Kalantidis, Li-Jia Li, David A Shamma, Michael Bernstein, and Li Fei-Fei. 2016 · 2016
Cited alongside, same era.
Inception-v4, Inception-ResNet and the Impact of Residual Connections on Learning. In AAAI
Christian Szegedy, Sergey Ioffe, Vincent Vanhoucke, and Alex Alemi. 2016 · 2016
Cited alongside, same era.
Cutting the Error by Half: Investigation of Very Deep CNN and Advanced Training Strategies for Document Image Classification
Muhammad Zeshan Afzal, Andreas Kölsch, Sheraz Ahmed, and Marcus Liwicki. 2017 · 2017
Document Image Classification with Intra-Domain Transfer Learning and Stacked Generalization of Deep Convolutional Neural Networks
Arindam Das, Saikat Roy, and Ujjwal Bhattacharya. 2018 · 2018
Later among the works it cites.
Chargrid: Towards Understanding 2D Documents. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Brussels, Belgium, 4459–4469
Anoop R Katti, Christian Reisswig, Cordula Guder, Sebastian Brarda, Steffen Bickel, Johannes Höhne, and Jean Baptiste Faddoul. 2018 · 2018
Later among the works it cites.
Modular Multimodal Architecture for Document Classification
Tyler Dauphinee, Nikunj Patel, and Mohammad Mehdi Rashidi. 2019 · 2019
Closest in time.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Association for Computational Linguistics, Minneapolis, Minnesota, 4171–4186
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross B. Girshick. 2017 · 2017
Cited alongside, same era.
DeepDeSRT: Deep Learning for Detection and Structure Recognition of Tables in Document Images
Sebastian Schreiber, Stefan Agne, Ivo Wolf, Andreas Dengel, and Sheraz Ahmed. 2017 · 2017
Cited alongside, same era.
Fast CNN-Based Document Layout Analysis
Matheus Palhares Viana and Dário Augusto Borges Oliveira. 2017 · 2017
Cited alongside, same era.
Learning to Extract Semantic Structure from Documents Using Multimodal Fully Convolutional Neural Networks
Xiaowei Yang, Ersin Yumer, Paul Asente, Mike Kraley, Daniel Kifer, and C. Lee Giles. 2017 · 2017
Cited alongside, same era.
Document page decomposition by the bounding-box project. In Proceedings of 3rd International Conference on Document Analysis and Recognition , Vol. 2. IEEE, 1119–1122
Jaekyu Ha, Robert M Haralick, and Ihsin T Phillips. 1995a
Cited in the paper.
Recursive XY cut using bounding boxes of connected components. In Proceedings of 3rd International Conference on Document Analysis and Recognition , Vol. 2. IEEE, 952–955
Jaekyu Ha, Robert M Haralick, and Ihsin T Phillips. 1995b
Cited in the paper.
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Closest in time.
FUNSD: A Dataset for Form Understanding in Noisy Scanned Documents
Guillaume Jaume, Hazim Kemal Ekenel, and Jean-Philippe Thiran. 2019 · 2019
Closest in time.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke S. Zettlemoyer, and Veselin Stoyanov. 2019b · 2019
Closest in time.
Deterministic Routing between Layout Abstractions for Multi-Scale Classification of Visually Rich Documents. In Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, IJCAI-19 . International Joint Conferences on Artificial Intelligence Organization, 3360–3366
Ritesh Sarkhel and Arnab Nandi. 2019 · 2019
Closest in time.
Visual Detection with Context for Document Layout Analysis. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . Association for Computational Linguistics, Hong Kong, China, 3462–3468
Carlos Soto and Shinjae Yoo. 2019 · 2019
Closest in time.
PubLayNet: largest dataset ever for document layout analysis
Xu Zhong, Jianbin Tang, and Antonio Jimeno-Yepes. 2019 · 2019
Closest in time.