Fetching the paper…
Reading the bibliography…
In this paper, we propose KnowCoder, a Large Language Model (LLM) to conduct Universal Information Extraction (UIE) via code generation.
GCDT: A global context enhanced deep transition architecture for sequence labeling
Yijin Liu, Fandong Meng, Jinchao Zhang, Jinan Xu, Yufeng Chen, and Jie Zhou. 2019 · 1906
Earlier work this paper cites.
Megatron-lm: Training multi-billion parameter language models using model parallelism
Mohammad Shoeybi, Mostofa Patwary, Raul Puri, Patrick LeGresley, Jared Casper, and Bryan Catanzaro. 2019 · 1909
Earlier work this paper cites.
Cross-lingual name tagging and linking for 282 languages
Xiaoman Pan, Boliang Zhang, Jonathan May, Joel Nothman, Kevin Knight, and Heng Ji. 2017a · 1958
Earlier work this paper cites.
Cross-lingual name tagging and linking for 282 languages
Xiaoman Pan, Boliang Zhang, Jonathan May, Joel Nothman, Kevin Knight, and Heng Ji. 2017b · 1958
Earlier work this paper cites.
Genia corpus—a semantically annotated corpus for bio-textmining
Jin-Dong Kim, Tomoko Ohta, Yuka Tateisi, and Jun’ichi Tsujii. 2003 · 2003
Earlier work this paper cites.
Introduction to the conll-2003 shared task: Language-independent named entity recognition
Erik F. Tjong Kim Sang and Fien De Meulder. 2003 · 2003
Earlier work this paper cites.
Ace 2004 multilingual training corpus
Alexis Mitchell, Stephanie Strassel, Shudong Huang, and Ramez Zakhary. 2005 · 2004
Earlier work this paper cites.
A linear programming formulation for global inference in natural language tasks
Dan Roth and Wen tau Yih. 2004 · 2004
Earlier work this paper cites.
ACE 2005 Multilingual Training Corpus
C. Walker and Linguistic Data Consortium. 2005 · 2005
Earlier work this paper cites.
Semeval-2010 task 8: Multi-way classification of semantic relations between pairs of nominals
Iris Hendrickx, Su Nam Kim, Zornitsa Kozareva, Preslav Nakov, Diarmuid Ó Séaghdha, Sebastian Padó, Marco Pennacchiotti, Lorenza Romano, and Stan Szpakowicz. 2010 · 2010
Earlier work this paper cites.
Modeling relations and their mentions without labeled text
Sebastian Riedel, Limin Yao, and Andrew McCallum. 2010 · 2010
Earlier work this paper cites.
Development of a benchmark corpus to support the automatic extraction of drug-related adverse effects from medical case reports
Harsha Gurulingappa, Abdul Mateen Rajput, Angus Roberts, Juliane Fluck, Martin Hofmann-Apitius, and Luca Toldo. 2012 · 2012
Earlier work this paper cites.
Ontonotes release 5.0 ldc2013t19
Ralph Weischedel, Martha Palmer, Mitchell Marcus, Eduard Hovy, Sameer Pradhan, Lance Ramshaw, Nianwen Xue, Ann Taylor, Jeff Kaufman, Michelle Franchini, et al. 2013 · 2013
Earlier work this paper cites.
Ncbi disease corpus: A resource for disease name recognition and concept normalization
Rezarta Islamaj Dogan, Robert Leaman, and Zhiyong Lu. 2014 · 2014
Earlier work this paper cites.
openbiocorpora anatem
openbiocorpora. 2015 · 2015
Earlier work this paper cites.
Relation classification via recurrent neural network
Dongxu Zhang and Dong Wang. 2015 · 2015
Earlier work this paper cites.
Broad Twitter corpus: A diverse named entity recognition resource
Leon Derczynski, Kalina Bontcheva, and Ian Roberts. 2016 · 2016
Cited alongside, same era.
Biocreative v cdr task corpus: a resource for chemical disease relation extraction
Jiao Li, Yueping Sun, Robin J. Johnson, Daniela Sciaky, Chih-Hsuan Wei, Robert Leaman, Allan Peter Davis, Carolyn J. Mattingly, Thomas C. Wiegers, and Zhiyong Lu. 2016 · 2016
Cited alongside, same era.
Automatically labeled data generation for large scale event extraction
Yubo Chen, Shulin Liu, Xiang Zhang, Kang Liu, and Jun Zhao. 2017 · 2017
Cited alongside, same era.
Results of the wnut2017 shared task on novel and emerging entity recognition
Leon Derczynski, Eric Nichols, Marieke Van Erp, and Nut Limsopatham. 2017 · 2017
Cited alongside, same era.
Improving distantly supervised relation extraction using word and entity based attention
Sharmistha Jat, Siddhesh Khandelwal, and Partha Pratim Talukdar. 2018 · 2018
Cited alongside, same era.
Unified structure generation for universal information extraction
Yaojie Lu, Qing Liu, Dai Dai, Xinyan Xiao, Hongyu Lin, Xianpei Han, Le Sun, and Hua Wu. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
MultiNERD: A multilingual, multi-genre and fine-grained dataset for named entity recognition (and disambiguation)
Simone Tedeschi and Roberto Navigli. 2022 · 2022
Later among the works it cites.
Code4struct: Code generation for few-shot structured prediction from natural language
Xingyao Wang, Sha Li, and Heng Ji. 2022 · 2022
Later among the works it cites.
Promptner: Prompting for named entity recognition
Dhananjay Ashok and Zachary C Lipton. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2018 · 2018
Cited alongside, same era.
Yi Luan, Luheng He, Mari Ostendorf, and Hannaneh Hajishirzi. 2018 · 2018
Cited alongside, same era.
Biomedical named entity recognition at scale
Veysel Kocaman and David Talby. 2020 · 2020
Cited alongside, same era.
Crossner: Evaluating cross-domain named entity recognition
Zihan Liu, Yan Xu, Tiezheng Yu, Wenliang Dai, Ziwei Ji, Samuel Cahyawijaya, Andrea Madotto, and Pascale Fung. 2020 · 2020
Cited alongside, same era.
Knowledge graph based synthetic corpus generation for knowledge-enhanced language model pre-training
Oshin Agarwal, Heming Ge, Siamak Shakeri, and Rami Al-Rfou. 2021 · 2021
Cited alongside, same era.
Knowledge Graphs
Aidan Hogan, Eva Blomqvist, Michael Cochez, Claudia d’Amato, Gerard de Melo, Claudio Gutiérrez, Sabrina Kirrane, José Emilio Labra Gayo, Roberto Navigli, Sebastian Neumaier, Axel-Cyrille Ngonga Ngomo, Axel Polleres, Sabbir M. Rashid, Anisa Rula, Lukas Schmelzeisen, Juan F. Sequeda, Steffen Staab, and Antoine Zimmermann. 2021 · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E Gonzalez, et al. 2023 · 2023
Later among the works it cites.
Instructie: A chinese instruction-based information extraction dataset
Honghao Gui, Jintian Zhang, Hongbin Ye, and Ningyu Zhang. 2023 · 2023
Later among the works it cites.
Retrieval-augmented code generation for universal information extraction
Yucan Guo, Zixuan Li, Xiaolong Jin, Yantao Liu, Yutao Zeng, Wenxuan Liu, Xiang Li, Pan Yang, Long Bai, Jiafeng Guo, et al. 2023 · 2023
Later among the works it cites.
Codeie: Large code generation models are better few-shot information extractors
Peng Li, Tianxiang Sun, Qiong Tang, Hang Yan, Yuanbin Wu, Xuanjing Huang, and Xipeng Qiu. 2023 · 2023
Later among the works it cites.
Universal information extraction as unified semantic matching
Jie Lou, Yaojie Lu, Dai Dai, Wei Jia, Hongyu Lin, Xianpei Han, Le Sun, and Hua Wu. 2023 · 2023
Later among the works it cites.
Gollie: Annotation guidelines improve zero-shot information-extraction
Oscar Sainz, Iker García-Ferrero, Rodrigo Agerri, Oier Lopez de Lacalle, German Rigau, and Eneko Agirre. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
Large language models for generative information extraction: A survey
Derong Xu, Wei Chen, Wenjun Peng, Chao Zhang, Tong Xu, Xiangyu Zhao, Xian Wu, Yefeng Zheng, and Enhong Chen. 2023 · 2023
Later among the works it cites.
Knowlm technical report
Ningyu Zhang, Jintian Zhang, Xiaohan Wang, Honghao Gui, Kangwei Liu, Yinuo Jiang, Xiang Chen, Shengyu Mao, Shuofei Qiao, Yuqi Zhu, Zhen Bi, Jing Chen, Xiaozhuan Liang, Yixin Ou, Runnan Fang, Zekun Xi, Xin Xu, Lei Li, Peng Wang, Mengru Wang, Yunzhi Yao, Bozhong Tian, Yin Fang, Guozhou Zheng, and Huajun Chen. 2023 · 2023
Later among the works it cites.
Universalner: Targeted distillation from large language models for open named entity recognition
Wenxuan Zhou, Sheng Zhang, Yu Gu, Muhao Chen, and Hoifung Poon. 2023 · 2023
Later among the works it cites.