Fetching the paper…
Reading the bibliography…
Typically, information extraction (IE) requires a pipeline approach: first, a sequence labeling model is trained on manually annotated documents to extract relevant spans; then, when a new document arrives, a model predicts spans which are then post-processed and standardized to convert the information into a database entry.
Robust lexical features for improved neural network named-entity recognition
Abbas Ghaddar and Phillippe Langlais. 2018 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Effective modeling of encoder-decoder architecture for joint entity and relation extraction
Tapas Nayak and H. T. Ng. 2020 · 1911
Earlier work this paper cites.
The ATIS spoken language systems pilot corpus
Charles T. Hemphill, John J. Godfrey, and George R. Doddington. 1990 · 1990
Earlier work this paper cites.
Rapid adaptation of bert for information extraction on domain-specific business documents
Ruixue Zhang, W. Yang, Luyun Lin, Zhengkai Tu, Yuqing Xie, Zihang Fu, Yuhao Xie, Luchen Tan, Kun Xiong, and Jimmy Lin. 2020 · 2002
Earlier work this paper cites.
A simple language model for task-oriented dialogue
Ehsan Hosseini-Asl, B. McCann, Chien-Sheng Wu, Semih Yavuz, and R. Socher. 2020 · 2005
Earlier work this paper cites.
Soloist: Few-shot task-oriented dialog with a single pre-trained auto-regressive model
Baolin Peng, C. Li, Jin chao Li, Shahin Shayandeh, L. Liden, and Jianfeng Gao. 2020 · 2005
Earlier work this paper cites.
Linformer: Self-attention with linear complexity
Sinong Wang, Belinda Z. Li, Madian Khabsa, Han Fang, and Hao Ma. 2020 · 2006
Earlier work this paper cites.
Leveraging passage retrieval with generative models for open domain question answering
Gautier Izacard and E. Grave. 2020 · 2007
Earlier work this paper cites.
Active learning with real annotation costs
B. Settles, M. Craven, and Lewis A. Friedland. 2008 · 2008
Earlier work this paper cites.
Natural language processing (almost) from scratch
Ronan Collobert, J. Weston, L. Bottou, Michael Karlen, K. Kavukcuoglu, and P. Kuksa. 2011 · 2011
Earlier work this paper cites.
Seqgensql - a robust sequence generation model for structured query language
Ning Li, Bethany Keller, Mark Butler, and Daniel Matthew Cer. 2020 · 2011
Earlier work this paper cites.
CoNLL-2012 shared task: Modeling multilingual unrestricted coreference in OntoNotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Olga Uryupina, and Yuchen Zhang. 2012 · 2012
Earlier work this paper cites.
Semantic parsing as machine translation
Jacob Andreas, Andreas Vlachos, and Stephen Clark. 2013 · 2013
Cited alongside, same era.
Towards robust linguistic analysis using OntoNotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Hwee Tou Ng, Anders Björkelund, Olga Uryupina, Yuchen Zhang, and Zhi Zhong. 2013 · 2013
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. 2014 · 2014
Cited alongside, same era.
Bidirectional lstm-crf models for sequence tagging
Zhiheng Huang, W. Xu, and Kai Yu. 2015 · 2015
Cited alongside, same era.
From noisy questions to minecraft texts: Annotation challenges in extreme syntax scenario
Héctor Martínez Alonso, Djamé Seddah, and Benoît Sagot. 2016 · 2016
Cited alongside, same era.
Transformer-xl: Attentive language models beyond a fixed-length context
Zihang Dai, Z. Yang, Yiming Yang, J. Carbonell, Quoc V. Le, and R. Salakhutdinov. 2019 · 2019
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Attend, copy, parse end-to-end information extraction from documents
R. Palm, Florian Laws, and O. Winther. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
A. Radford, Jeffrey Wu, R. Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Kleister: A novel task for information extraction involving long documents with complex layout
Filip Graliński, Tomasz Stanisławek, Anna Wróblewska, Dawid Lipiński, Agnieszka Kaliska, Paulina Rosalska, Bartosz Topolski, and Przemysław Biecek. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Chen, B. Xu, C. Zhang, and Carlos Guestrin. 2016 · 2016
Cited alongside, same era.
Named entity recognition with bidirectional LSTM-CNNs
Jason P.C. Chiu and Eric Nichols. 2016 · 2016
Cited alongside, same era.
Neural architectures for named entity recognition
Guillaume Lample, Miguel Ballesteros, Sandeep Subramanian, Kazuya Kawakami, and Chris Dyer. 2016 · 2016
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, L. Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Universal language model fine-tuning for text classification
J. Howard and Sebastian Ruder. 2018 · 2018
Cited alongside, same era.
The natural language decathlon: Multitask learning as question answering
B. McCann, N. Keskar, Caiming Xiong, and R. Socher. 2018 · 2018
Cited alongside, same era.
Extracting relational facts by an end-to-end neural model with copy mechanism
Xiangrong Zeng, Daojian Zeng, Shizhu He, Kang Liu, and Jun Zhao. 2018 · 2018
Cited alongside, same era.
Representation learning for information extraction from form-like documents
Bodhisattwa Majumder, Navneet Potti, Sandeep Tata, James B. Wendt, Qi Zhao, and Marc Najork. 2020 · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, M. Matena, Yanqi Zhou, W. Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
End-to-end extraction of structured information from business documents with pointer-generator networks
Clément Sage, A. Aussem, V. Eglin, H. Elghazel, and Jérémy Espinas. 2020 · 2020
Later among the works it cites.
Project deepform: Extract information from documents
Jonathan Stray Stacey Svetlichnaya. 2020 · 2020
Later among the works it cites.
Layoutlm: Pre-training of text and layout for document image understanding
Yiheng Xu, Minghao Li, Lei Cui, Shaohan Huang, Furu Wei, and M. Zhou. 2020 · 2020
Later among the works it cites.
Seq2sql: Generating structured queries from natural language using reinforcement learning
Victor Zhong, Caiming Xiong, and Richard Socher. 2017 · 2020
Later among the works it cites.
Structured prediction as translation between augmented natural languages
Giovanni Paolini, Ben Athiwaratkun, Jason Krone, Jie Ma, Alessandro Achille, RISHITA ANUBHAI, Cicero Nogueira dos Santos, Bing Xiang, and Stefano Soatto. 2021 · 2021
Closest in time.