Fetching the paper…
Reading the bibliography…
Key Information Extraction (KIE) is a challenging multimodal task that aims to extract structured value semantic entities from visually rich documents.
Attention is All you Need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, L.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
EATEN: Entity-Aware Attention for Single Shot Visual Text Extraction
Guo, H.; Qin, X.; Liu, J.; Han, J.; Liu, J.; and Ding, E. 2019 · 2019
Earlier work this paper cites.
ICDAR2019 Competition on Scanned Receipt OCR and Information Extraction
Huang, Z.; Chen, K.; He, J.; Bai, X.; Karatzas, D.; Lu, S.; and Jawahar, C. V. 2019 · 2019
Earlier work this paper cites.
FUNSD: A Dataset for Form Understanding in Noisy Scanned Documents
Jaume, G.; Ekenel, H. K.; and Thiran, J. 2019 · 2019
Earlier work this paper cites.
CORD: a consolidated receipt dataset for post-OCR parsing
Park, S.; Shin, S.; Lee, B.; Lee, J.; Surh, J.; Seo, M.; and Lee, H. 2019 · 2019
Earlier work this paper cites.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
Paszke, A.; Gross, S.; Massa, F.; Lerer, A.; Bradbury, J.; Chanan, G.; Killeen, T.; Lin, Z.; Gimelshein, N.; Antiga, L.; Desmaison, A.; Köpf, A.; Yang, E. Z.; DeVito, Z.; Raison, M.; Tejani, A.; Chilamkurthy, S.; Steiner, B.; Fang, L.; Bai, J.; and Chintala, S. 2019 · 2019
Earlier work this paper cites.
One-shot Text Field labeling using Attention and Belief Propagation for Structure Information Extraction
Cheng, M.; Qiu, M.; Shi, X.; Huang, J.; and Lin, W. 2020 · 2020
Earlier work this paper cites.
Representation Learning for Information Extraction from Form-like Documents
Majumder, B. P.; Potti, N.; Tata, S.; Wendt, J. B.; Zhao, Q.; and Najork, M. 2020 · 2020
Earlier work this paper cites.
Circle Loss: A Unified Perspective of Pair Similarity Optimization
Sun, Y.; Cheng, C.; Zhang, Y.; Zhang, C.; Zheng, L.; Wang, Z.; and Wei, Y. 2020 · 2020
Cited alongside, same era.
Information Extraction from Invoices
Hamdi, A.; Carel, E.; Joseph, A.; Coustaty, M.; and Doucet, A. 2021 · 2021
Cited alongside, same era.
Kleister: Key Information Extraction Datasets Involving Long Documents with Complex Layouts
Stanislawek, T.; Gralinski, F.; Wróblewska, A.; Lipinski, D.; Kaliska, A.; Rosalska, P.; Topolski, B.; and Biecek, P. 2021 · 2021
Cited alongside, same era.
Towards Robust Visual Information Extraction in Real World: New Dataset and Novel Solution
Wang, J.; Liu, C.; Jin, L.; Tang, G.; Zhang, J.; Zhang, S.; Wang, Q.; Wu, Y.; and Cai, M. 2021 · 2021
Cited alongside, same era.
LayoutXLM: Multimodal Pre-training for Multilingual Visually-rich Document Understanding
Xu, Y.; Lv, T.; Cui, L.; Wang, G.; Lu, Y.; Florêncio, D.; Zhang, C.; and Wei, F. 2021 · 2021
Cited alongside, same era.
DocQueryNet: Value Retrieval with Arbitrary Queries for Form-like Documents
Gao, M.; Xue, L.; Ramaiah, C.; Xing, C.; Xu, R.; and Xiong, C. 2022 · 2022
Later among the works it cites.
LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking
Huang, Y.; Lv, T.; Cui, L.; Lu, Y.; and Wei, F. 2022 · 2022
Later among the works it cites.
OCR-Free Document Understanding Transformer
Kim, G.; Hong, T.; Yim, M.; Nam, J.; Park, J.; Yim, J.; Hwang, W.; Yun, S.; Han, D.; and Park, S. 2022 · 2022
Later among the works it cites.
Unified Named Entity Recognition as Word-Word Relation Classification
Li, J.; Fei, H.; Liu, J.; Wu, S.; Zhang, M.; Teng, C.; Ji, D.; and Li, F. 2022 · 2022
Later among the works it cites.
ERNIE-Layout: Layout Knowledge Enhanced Pre-training for Visually-rich Document Understanding
Peng, Q.; Pan, Y.; Wang, W.; Luo, B.; Zhang, Z.; Huang, Z.; Cao, Y.; Yin, W.; Chen, Y.; Zhang, Y.; Feng, S.; Sun, Y.; Tian, H.; Wu, H.; and Wang, H. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Entity Relation Extraction as Dependency Parsing in Visually Rich Documents
Zhang, Y.; Zhang, B.; Wang, R.; Cao, J.; Li, C.; and Bao, Z. 2021 · 2021
Cited alongside, same era.
Query-driven Generative Network for Document Information Extraction in the Wild
Cao, H.; Li, X.; Ma, J.; Jiang, D.; Guo, A.; Hu, Y.; Liu, H.; Liu, Y.; and Ren, B. 2022 · 2022
Cited alongside, same era.
Su, J.; Murtadha, A.; Pan, S.; Hou, J.; Sun, J.; Huang, W.; Wen, B.; and Liu, Y. 2022 · 2022
Later among the works it cites.
A Question-Answering Approach to Key Value Pair Extraction from Form-like Document Images
Hu, K.; Wu, Z.; Zhong, Z.; Lin, W.; Sun, L.; and Huo, Q. 2023 · 2023
Closest in time.