Fetching the paper…
Reading the bibliography…
Sequence modeling has demonstrated state-of-the-art performance on natural language and document understanding tasks.
Sofia Serrano and Noah A Smith. 2019 · 1906
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
Bertgrid: Contextualized embedding for 2d document representation and understanding
Timo I Denk and Christian Reisswig. 2019 · 1909
Earlier work this paper cites.
“cloze procedure”: A new tool for measuring readability
Wilson L Taylor. 1953 · 1953
Earlier work this paper cites.
Multivariate statistics: a vector space approach
Morris L Eaton. 1983 · 1983
Earlier work this paper cites.
A framework for computational morphology
David G Kirkpatrick and John D Radke. 1985 · 1985
Earlier work this paper cites.
A fast and efficient method for extracting text paragraphs and graphics from unconstrained documents
Frank Lebourgeois, Zbigniew Bublinski, and Hubert Emptoz. 1992 · 1992
Earlier work this paper cites.
The document spectrum for page layout analysis
Lawrence O’Gorman. 1993 · 1993
Earlier work this paper cites.
Recursive xy cut using bounding boxes of connected components
Jaekyu Ha, Robert M Haralick, and Ihsin T Phillips. 1995 · 1995
Earlier work this paper cites.
A fast algorithm for bottom-up document layout analysis
Anikó Simon, J-C Pret, and A Peter Johnson. 1997 · 1997
Earlier work this paper cites.
Lambert: Layout-aware (language) modeling for information extraction
Łukasz Garncarek, Rafał Powalski, Tomasz Stanisławek, Bartosz Topolski, Piotr Halama, Michał Turski, and Filip Graliński. 2020 · 2002
Earlier work this paper cites.
Kleister: A novel task for information extraction involving long documents with complex layout
Filip Graliński, Tomasz Stanisławek, Anna Wróblewska, Dawid Lipiński, Agnieszka Kaliska, Paulina Rosalska, Bartosz Topolski, and Przemysław Biecek. 2020 · 2003
Earlier work this paper cites.
Artificial neural networks for document analysis and recognition
Simone Marinai, Marco Gori, and Giovanni Soda. 2005 · 2005
Earlier work this paper cites.
Learning nongenerative grammatical models for document analysis
Michael Shilman, Percy Liang, and Paul Viola. 2005 · 2005
Earlier work this paper cites.
Text and layout information extraction from document files of various formats based on the analysis of page description language
Takashi Hirano, Yuichi Okano, Yasuhiro Okada, and Fumio Yoda. 2007 · 2007
Earlier work this paper cites.
A realistic dataset for performance evaluation of document layout analysis
Apostolos Antonacopoulos, David Bridson, Christos Papadopoulos, and Stefan Pletschacher. 2009 · 2009
Earlier work this paper cites.
Design challenges and misconceptions in named entity recognition
Lev Ratinov and Dan Roth. 2009 · 2009
Earlier work this paper cites.
Logical structure recovery in scholarly articles with rich document features
Minh-Thang Luong, Thuy Dung Nguyen, and Min-Yen Kan. 2012 · 2012
Earlier work this paper cites.
Rule-based information extraction is dead! long live rule-based information extraction systems!
Laura Chiticariu, Yunyao Li, and Frederick Reiss. 2013 · 2013
Earlier work this paper cites.
Intellix–end-user trained information extraction for document archiving
Daniel Schuster, Klemens Muthmann, Daniel Esser, Alexander Schill, Michael Berger, Christoph Weidling, Kamil Aliyev, and Andreas Hofmeier. 2013 · 2013
Earlier work this paper cites.
Evaluation of svm, mlp and gmm classifiers for layout analysis of historical documents
Hao Wei, Micheal Baechler, Fouad Slimane, and Rolf Ingold. 2013 · 2013
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Cited alongside, same era.
Neural message passing for quantum chemistry
Justin Gilmer, Samuel S Schoenholz, Patrick F Riley, Oriol Vinyals, and George E Dahl. 2017 · 2017
Cited alongside, same era.
Cloudscan-a configuration-free invoice analysis system using recurrent neural networks
Rasmus Berg Palm, Ole Winther, and Florian Laws. 2017 · 2017
Cited alongside, same era.
Cross-sentence n-ary relation extraction with graph lstms
Nanyun Peng, Hoifung Poon, Chris Quirk, Kristina Toutanova, and Wen-tau Yih. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Visual detection with context for document layout analysis
Carlos Soto and Shinjae Yoo. 2019 · 2019
Later among the works it cites.
A multiscale visualization of attention in the transformer model
Jesse Vig. 2019 · 2019
Later among the works it cites.
Information extraction from text regions with complex tabular structure
Kaixuan Zhang, Zejiang Shen, Jie Zhou, and Melissa Dell. 2019 · 2019
Later among the works it cites.
Cutie: Learning to understand documents with convolutional universal text information extractor
Xiaohui Zhao, Endi Niu, Zhuo Wu, and Xiaoguang Wang. 2019 · 2019
Later among the works it cites.
Form2seq: A framework for higher-order form structure extraction
Milan Aggarwal, Hiresh Gupta, Mausoom Sarkar, and Balaji Krishnamurthy. 2020 · 2020
Later among the works it cites.
Etc: Encoding long and structured data in transformers
Joshua Ainslie, Santiago Ontañón, Chris Alberti, Vaclav Cvicek, Zachary Fisher, Philip Pham, Anirudh Ravula, Sumit Sanghai, Qifan Wang, and Li Yang. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning to extract semantic structure from documents using multimodal fully convolutional neural networks
Xiao Yang, Ersin Yumer, Paul Asente, Mike Kraley, Daniel Kifer, and C Lee Giles. 2017 · 2017
Cited alongside, same era.
Chargrid: Towards understanding 2d documents
Anoop Raveendra Katti, Christian Reisswig, Cordula Guder, Sebastian Brarda, Steffen Bickel, Johannes Höhne, and Jean Baptiste Faddoul. 2018 · 2018
Cited alongside, same era.
Image transformer
Niki Parmar, Ashish Vaswani, Jakob Uszkoreit, Lukasz Kaiser, Noam Shazeer, Alexander Ku, and Dustin Tran. 2018 · 2018
Cited alongside, same era.
Self-attention with relative position representations
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani. 2018 · 2018
Cited alongside, same era.
N-ary relation extraction using graph state lstm
Linfeng Song, Yue Zhang, Zhiguo Wang, and Daniel Gildea. 2018 · 2018
Cited alongside, same era.
Deep visual template-free form parsing
Brian Davis, Bryan Morse, Scott Cohen, Brian Price, and Chris Tensmeyer. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Unilmv2: Pseudo-masked language models for unified language model pre-training
Hangbo Bao, Li Dong, Furu Wei, Wenhui Wang, Nan Yang, Xiaodong Liu, Yu Wang, Jianfeng Gao, Songhao Piao, Ming Zhou, et al. 2020 · 2020
Later among the works it cites.
Representation learning for information extraction from form-like documents
Bodhisattwa Prasad Majumder, Navneet Potti, Sandeep Tata, James Bradley Wendt, Qi Zhao, and Marc Najork. 2020 · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Later among the works it cites.
Layoutlm: Pre-training of text and layout for document image understanding
Yiheng Xu, Minghao Li, Lei Cui, Shaohan Huang, Furu Wei, and Ming Zhou. 2020 · 2020
Later among the works it cites.
Pick: processing key information extraction from documents using improved graph learning-convolutional networks
Wenwen Yu, Ning Lu, Xianbiao Qi, Ping Gong, and Rong Xiao. 2020 · 2020
Later among the works it cites.
Big bird: Transformers for longer sequences
Manzil Zaheer, Guru Guruganesh, Avinava Dubey, Joshua Ainslie, Chris Alberti, Santiago Ontanon, Philip Pham, Anirudh Ravula, Qifan Wang, Li Yang, et al. 2020 · 2020
Later among the works it cites.
Deep relational reasoning graph network for arbitrary shape text detection
Shi-Xue Zhang, Xiaobin Zhu, Jie-Bo Hou, Chang Liu, Chun Yang, Hongfa Wang, and Xu-Cheng Yin. 2020 · 2020
Later among the works it cites.
Docformer: End-to-end transformer for document understanding
Srikar Appalaraju, Bhavan Jasani, Bhargava Urala Kota, Yusheng Xie, and R Manmatha. 2021 · 2021
Later among the works it cites.
Spatial dependency parsing for semi-structured document information extraction
Wonseok Hwang, Jinyeong Yim, Seunghyun Park, Sohee Yang, and Minjoon Seo. 2021 · 2021
Later among the works it cites.
Rope: Reading order equivariant positional encoding for graph-based document information extraction
Chen-Yu Lee, Chun-Liang Li, Chu Wang, Renshen Wang, Yasuhisa Fujii, Siyang Qin, Ashok Popat, and Tomas Pfister. 2021 · 2021
Later among the works it cites.
Going full-tilt boogie on document understanding with text-image-layout transformer
Rafał Powalski, Łukasz Borchmann, Dawid Jurkiewicz, Tomasz Dwojak, Michał Pietruszka, and Gabriela Pałka. 2021 · 2021
Later among the works it cites.
Layoutlmv2: Multi-modal pre-training for visually-rich document understanding
Yang Xu, Yiheng Xu, Tengchao Lv, Lei Cui, Furu Wei, Guoxin Wang, Yijuan Lu, Dinei Florencio, Cha Zhang, Wanxiang Che, et al. 2021 · 2021
Later among the works it cites.
Post-ocr paragraph recognition by graph convolutional networks
Renshen Wang, Yasuhisa Fujii, and Ashok C. Popat. 2022 · 2022
Closest in time.