Fetching the paper…
Reading the bibliography…
The output structure of database-like tables, consisting of values structured in horizontal rows and vertical columns identifiable by name, can cover a wide range of NLP tasks.
Injecting knowledge base information into end-to-end joint entity and relation extraction and coreference resolution
Verlinden, S., Zaporojets, K., Deleu, J., Demeester, T., and Develder, C. (2021) · 1957
Earlier work this paper cites.
AxCell: Automatic extraction of results from machine learning papers
Kardas, M., Czapla, P., Stenetorp, P., Ruder, S., Riedel, S., Taylor, R., and Stojnic, R. (2020) · 2004
Earlier work this paper cites.
Ask me anything: Dynamic memory networks for natural language processing
Kumar, A., Irsoy, O., Ondruska, P., Iyyer, M., Bradbury, J., Gulrajani, I., Zhong, V., Paulus, R., and Socher, R. (2016) · 2016
Earlier work this paper cites.
Challenges in data-to-document generation
Wiseman, S., Shieber, S., and Rush, A. (2017) · 2017
Earlier work this paper cites.
The natural language decathlon: Multitask learning as question answering
McCann, B., Keskar, N. S., Xiong, C., and Socher, R. (2018) · 2018
Earlier work this paper cites.
CORD: A consolidated receipt dataset for post-ocr parsing
Park, S., Shin, S., Lee, B., Lee, J., Surh, J., Seo, M., and Lee, H. (2019) · 2019
Earlier work this paper cites.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., Desmaison, A., Kopf, A., Yang, E., DeVito, Z., Raison, M., Tejani, A., Chilamkurthy, S., Steiner, B., Fang, L., Bai, J., and Chintala, S. (2019) · 2019
Earlier work this paper cites.
Insertion transformer: Flexible sequence generation via insertion operations
Stern, M., Chan, W., Kiros, J. R., and Uszkoreit, J. (2019) · 2019
Earlier work this paper cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Z., Dai, Z., Yang, Y., Carbonell, J., Salakhutdinov, R. R., and Le, Q. V. (2019) · 2019
Cited alongside, same era.
From dataset recycling to multi-property extraction and beyond
Dwojak, T., Pietruszka, M., Borchmann, Ł., Chłędowski, J., and Graliński, F. (2020) · 2020
Cited alongside, same era.
UnifiedQA: Crossing format boundaries with a single QA system
Khashabi, D., Min, S., Khot, T., Sabharwal, A., Tafjord, O., Clark, P., and Hajishirzi, H. (2020) · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J. (2020) · 2020
Cited alongside, same era.
End-to-end extraction of structured information from business documents with pointer-generator networks
Sage, C., Aussem, A., Eglin, V., Elghazel, H., and Espinas, J. (2020) · 2020
Cited alongside, same era.
LAMBERT: Layout-aware language modeling for information extraction
Garncarek, Ł., Powalski, R., Stanisławek, T., Topolski, B., Halama, P., Turski, M., and Graliński, F. (2021) · 2021
Later among the works it cites.
Text2Event: Controllable sequence-to-structure generation for end-to-end event extraction
Lu, Y., Lin, H., Xu, J., Han, X., Tang, J., Li, A., Sun, L., Liao, M., and Chen, S. (2021) · 2021
Later among the works it cites.
Going full-TILT boogie on document understanding with text-image-layout transformer
Powalski, R., Borchmann, Ł., Jurkiewicz, D., Dwojak, T., Pietruszka, M., and Pałka, G. (2021) · 2021
Later among the works it cites.
Doc2dict: Information extraction as text generation
Townsend, B., Ito-Fisher, E., Zhang, L., and May, M. (2021) · 2021
Later among the works it cites.
Text-to-table: A new way of information extraction
Wu, X., Zhang, J., and Li, H. (2021) · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Image-based table recognition: Data, model, and evaluation
Zhong, X., ShafieiBavani, E., and Jimeno Yepes, A. (2020) · 2020
Cited alongside, same era.
DUE: End-to-end document understanding benchmark
Borchmann, Ł., Pietruszka, M., Stanislawek, T., Jurkiewicz, D., Turski, M., Szyndler, K., and Graliński, F. (2021) · 2021
Cited alongside, same era.
Evaluating large language models trained on code
Chen, M., Tworek, J., Jun, H., Yuan, Q., de Oliveira Pinto, H. P., Kaplan, J., Edwards, H., Burda, Y., Joseph, N., Brockman, G., Ray, A., Puri, R., Krueger, G., Petrov, M., Khlaaf, H., Sastry, G., Mishkin, P., Chan, B., Gray, S., Ryder, N., Pavlov, M., Power, A., Kaiser, L., Bavarian, M., Winter, C., Tillet, P., Such, F. P., Cummings, D., Plappert, M., Chantzis, F., Barnes, E., Herbert-Voss, A., Guss, W. H., Nichol, A., Paino, A., Tezak, N., Tang, J., Babuschkin, I., Balaji, S., Jain, S., Saunders, W., Hesse, C., Carr, A. N., Leike, J., Achiam, J., Misra, V., Morikawa, E., Radford, A., Knight, M., Brundage, M., Murati, M., Mayer, K., Welinder, P., McGrew, B., Amodei, D., McCandlish, S., Sutskever, I., and Zaremba, W. (2021) · 2021
Cited alongside, same era.
Later among the works it cites.
LayoutLMv2: Multi-modal pre-training for visually-rich document understanding
Xu, Y., Xu, Y., Lv, T., Cui, L., Wei, F., Wang, G., Lu, Y., Florencio, D., Zhang, C., Che, W., Zhang, M., and Zhou, L. (2021) · 2021
Later among the works it cites.
DWIE: An entity-centric dataset for multi-task document-level information extraction
Zaporojets, K., Deleu, J., Develder, C., and Demeester, T. (2021) · 2021
Later among the works it cites.
Representations for question answering from documents with tables and text
Zayats, V., Toutanova, K., and Ostendorf, M. (2021) · 2021
Later among the works it cites.