Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have shown impressive ability for open-domain NLP tasks.
The ATIS spoken language systems pilot corpus
Charles T. Hemphill, John J. Godfrey, and George R. Doddington. 1990 · 1990
Earlier work this paper cites.
Is learning the n-th thing any easier than learning the first?
Sebastian Thrun. 1995 · 1995
Earlier work this paper cites.
Multitask learning
Rich Caruana. 1997 · 1997
Earlier work this paper cites.
A novel use of statistical parsing to extract information from text
Scott Miller, Heidi Fox, Lance Ramshaw, and Ralph Weischedel. 2000 · 2000
Earlier work this paper cites.
Cluener2020: Fine-grained named entity recognition dataset and benchmark for chinese
Liang Xu, Yu Tong, Qianqian Dong, Cong Yu, Yin Tian, Weitang Liu, Lu Li, and Xuanwei Zhang. 2020b · 2001
Earlier work this paper cites.
Learning question classifiers
Xin Li and Dan Roth. 2002 · 2002
Earlier work this paper cites.
Introduction to the conll-2003 shared task: Language-independent named entity recognition
Erik F. Tjong Kim Sang and Fien De Meulder. 2003 · 2003
Earlier work this paper cites.
Cluecorpus2020: A large-scale chinese corpus for pre-training language model
Liang Xu, Xuanwei Zhang, and Qianqian Dong. 2020c · 2003
Earlier work this paper cites.
Introduction to the bio-entity recognition task at JNLPBA
Nigel Collier and Jin-Dong Kim. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
ACE 2005 Multilingual Training Corpus
Walker, Christopher, Strassel, Stephanie, Medero, Julie, and Maeda, Kazuaki. 2006 · 2005
Earlier work this paper cites.
The third international Chinese language processing bakeoff: Word segmentation and named entity recognition
Gina-Anne Levow. 2006 · 2006
Earlier work this paper cites.
Dynamic conditional random fields: Factorized probabilistic models for labeling and segmenting sequence data
Charles Sutton, Andrew McCallum, and Khashayar Rohanimanesh. 2007 · 2007
Earlier work this paper cites.
A unified architecture for natural language processing: Deep neural networks with multitask learning
Ronan Collobert and Jason Weston. 2008 · 2008
Earlier work this paper cites.
Overview of biocreative ii gene mention recognition
Larry Smith, Lorraine Tanabe, Rie Ando, Cheng-Ju Kuo, I-Fang Chung, Chun-Nan Hsu, Yu-Shi Lin, Roman Klinger, Christoph Friedrich, Kuzman Ganchev, Manabu Torii, Hongfang Liu, Barry Haddow, Craig Struble, Richard Povinelli, Andreas Vlachos, William Baumgartner Jr, Lawrence Hunter, Bob Carpenter, and W. Wilbur. 2008 · 2008
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. 2011 · 2011
Earlier work this paper cites.
Proceedings of the Fifteenth Conference on Computational Natural Language Learning: Shared Task . Association for Computational Linguistics, Portland, Oregon, USA
Sameer Pradhan, editor. 2011 · 2011
Earlier work this paper cites.
Hidden factors and hidden topics: understanding rating dimensions with review text
Julian McAuley and Jure Leskovec. 2013 · 2013
Earlier work this paper cites.
Towards robust linguistic analysis using OntoNotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Hwee Tou Ng, Anders Björkelund, Olga Uryupina, Yuchen Zhang, and Zhi Zhong. 2013 · 2013
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D. Manning, Andrew Ng, and Christopher Potts. 2013 · 2013
Earlier work this paper cites.
Ncbi disease corpus
Rezarta Islamaj Dogan, Robert Leaman, and Zhiyong Lu. 2014 · 2014
Earlier work this paper cites.
Anatomical entity recognition with a hierarchical framework augmented by external resources
Yan Xu, Ji Hua, Zhaoheng Ni, Qinlang Chen, Yubo Fan, Sophia Ananiadou, Eric I-Chao Chang, and Junichi Tsujii. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning. 2015 · 2015
Earlier work this paper cites.
The chemdner corpus of chemicals and drugs and its annotation principles
Martin Krallinger, Obdulia Rabal, Florian Leitner, et al. 2015 · 2015
Earlier work this paper cites.
Named entity recognition for Chinese social media with jointly trained embeddings
Nanyun Peng and Mark Dredze. 2015 · 2015
Earlier work this paper cites.
Short text clustering via convolutional neural networks
Jiaming Xu, Peng Wang, Guanhua Tian, Bo Xu, Jun Zhao, Fangyuan Wang, and Hongwei Hao. 2015 · 2015
Earlier work this paper cites.
Biocreative v cdr task corpus: a resource for chemical disease relation extraction
Jiao Li, Yueping Sun, Robin Johnson, Daniela Sciaky, Chih-Hsuan Wei, Robert Leaman, Allan Peter Davis, Carolyn Mattingly, Thomas Wiegers, and Zhiyong lu. 2016a · 2016
Earlier work this paper cites.
Recurrent neural network for text classification with multi-task learning
Pengfei Liu, Xipeng Qiu, and Xuanjing Huang. 2016 · 2016
Earlier work this paper cites.
THUCTC: An Efficient Chinese Text Classifier
Sun Maosong, Li Jingyang, Guo Zhipeng, Zhao Yu, Zheng Yabin, Si Xiance, and Liu Zhiyuan. 2016 · 2016
Earlier work this paper cites.
Results of the WNUT16 named entity recognition shared task
Benjamin Strauss, Bethany Toma, Alan Ritter, Marie-Catherine de Marneffe, and Wei Xu. 2016 · 2016
Earlier work this paper cites.
Results of the wnut2017 shared task on novel and emerging entity recognition
Leon Derczynski, Eric Nichols, Marieke Van Erp, and Nut Limsopatham. 2017 · 2017
Earlier work this paper cites.
MultiWOZ - a large-scale multi-domain Wizard-of-Oz dataset for task-oriented dialogue modelling
Paweł Budzianowski, Tsung-Hsien Wen, Bo-Hsiang Tseng, Iñigo Casanueva, Stefan Ultes, Osman Ramadan, and Milica Gašić. 2018 · 2018
Earlier work this paper cites.
Ultra-fine entity typing
Eunsol Choi, Omer Levy, Yejin Choi, and Luke Zettlemoyer. 2018 · 2018
Earlier work this paper cites.
Alice Coucke, Alaa Saade, Adrien Ball, Théodore Bluche, Alexandre Caulier, David Leroy, Clément Doumouro, Thibault Gisselbrecht, Francesco Caltagirone, Thibaut Lavril, Maël Primet, and Joseph Dureau. 2018 · 2018
Earlier work this paper cites.
SemEval-2018 task 7: Semantic relation extraction and classification in scientific papers
Kata Gábor, Davide Buscaldi, Anne-Kathrin Schumann, Behrang QasemiZadeh, Haïfa Zargayouna, and Thierry Charnois. 2018 · 2018
Earlier work this paper cites.
DuReader: a Chinese machine reading comprehension dataset from real-world applications
Wei He, Kai Liu, Jing Liu, Yajuan Lyu, Shiqi Zhao, Xinyan Xiao, Yuan Liu, Yizhong Wang, Hua Wu, Qiaoqiao She, Xuan Liu, Tian Wu, and Haifeng Wang. 2018 · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? a new dataset for open book question answering
Todor Mihaylov, Peter Clark, Tushar Khot, and Ashish Sabharwal. 2018 · 2018
Earlier work this paper cites.
Glue: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
Cail2018: A large-scale legal dataset for judgment prediction
Chaojun Xiao, Haoxi Zhong, Zhipeng Guo, Cunchao Tu, Zhiyuan Liu, Maosong Sun, Yansong Feng, Xianpei Han, Zhen Hu, Heng Wang, and Jianfeng Xu. 2018 · 2018
Earlier work this paper cites.
WikiQA: A challenge dataset for open-domain question answering
Yi Yang, Wen-tau Yih, and Christopher Meek. 2015 · 2018
Earlier work this paper cites.
Chinese NER using lattice LSTM
Yue Zhang and Jie Yang. 2018 · 2018
Cited alongside, same era.
Bipar: A bilingual parallel dataset for multilingual and cross-lingual reading comprehension on novels
Yimin Jing, Deyi Xiong, and Zhen Yan. 2019 · 2019
Cited alongside, same era.
An evaluation dataset for intent classification and out-of-scope prediction
Stefan Larson, Anish Mahendran, Joseph J. Peper, Christopher Clarke, Andrew Lee, Parker Hill, Jonathan K. Kummerfeld, Kevin Leach, Michael A. Laurenzano, Lingjia Tang, and Jason Mars. 2019 · 2019
Cited alongside, same era.
Duie: A large-scale chinese dataset for information extraction
Shuangjie Li, Wei He, Yabing Shi, Wenbin Jiang, Haijin Liang, Ye Jiang, Yang Zhang, Yajuan Lyu, and Yong Zhu. 2019 · 2019
Cited alongside, same era.
Multi-task deep neural networks for natural language understanding
Xiaodong Liu, Pengcheng He, Weizhu Chen, and Jianfeng Gao. 2019 · 2019
Cited alongside, same era.
GLM: General language model pretraining with autoregressive blank infilling
Zhengxiao Du, Yujie Qian, Xiao Liu, Ming Ding, Jiezhong Qiu, Zhilin Yang, and Jie Tang. 2022 · 2022
Later among the works it cites.
Proceedings of BigScience Episode #5 – Workshop on Challenges & Perspectives in Creating Large Language Models . Association for Computational Linguistics, virtual+Dublin
Angela Fan, Suzana Ilic, Thomas Wolf, and Matthias Gallé, editors. 2022 · 2022
Later among the works it cites.
Development of a benchmark corpus to support entity recognition in job descriptions
Thomas AF Green, Diana Maynard, and Chenghua Lin. 2022 · 2022
Later among the works it cites.
Training compute-optimal large language models
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, Elena Buchatskaya, Trevor Cai, Eliza Rutherford, Diego de Las Casas, Lisa Anne Hendricks, Johannes Welbl, Aidan Clark, Tom Hennigan, Eric Noland, Katie Millican, George van den Driessche, Bogdan Damoc, Aurelia Guy, Simon Osindero, Karen Simonyan, Erich Elsen, Jack W. Rae, Oriol Vinyals, and Laurent Sifre. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R Thomas McCoy, Ellie Pavlick, and Tal Linzen. 2020 · 2019
Cited alongside, same era.
Ipre: a dataset for inter-personal relationship extraction
Haitao Wang, Zhengqiu He, Jin Ma, Wenliang Chen, and Min Zhang. 2019 · 2019
Cited alongside, same era.
CATSLU: The 1st chinese audio-textual spoken language understanding challenge
Su Zhu, Zijian Zhao, Tiejun Zhao, Chengqing Zong, and Kai Yu. 2019 · 2019
Cited alongside, same era.
Political advertising dataset: the use case of the Polish 2020 presidential elections
Lukasz Augustyniak, Krzysztof Rajda, Tomasz Kajdanowicz, and Michał Bernaczyk. 2020 · 2020
Cited alongside, same era.
SLURP: A Spoken Language Understanding Resource Package
Emanuele Bastianelli, Andrea Vanzo, Pawel Swietojanski, and Verena Rieser. 2020 · 2020
Cited alongside, same era.
Subjqa: A dataset for subjectivity and review comprehension
Johannes Bjerva, Nikita Bhutani, Behzad Golshan, Wang-Chiew Tan, and Isabelle Augenstein. 2020 · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, et al. 2020 · 2020
Cited alongside, same era.
LoRA: Low-rank adaptation of large language models
Edward J Hu, yelong shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Later among the works it cites.
MultiCoNER: A large-scale multilingual dataset for complex named entity recognition
Shervin Malmasi, Anjie Fang, Besnik Fetahu, Sudipta Kar, and Oleg Rokhlenko. 2022 · 2022
Later among the works it cites.
Chatgpt: Optimizing language models for dialogue
TB OpenAI. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Later among the works it cites.
In-BoXBART: Get instructions into biomedical multi-task learning
Mihir Parmar, Swaroop Mishra, Mirali Purohit, Man Luo, Murad Mohammad, and Chitta Baral. 2022 · 2022
Later among the works it cites.
Multitask prompted training enables zero-shot task generalization
Victor Sanh, Albert Webson, Colin Raffel, et al. 2022 · 2022
Later among the works it cites.
MultiNERD: A multilingual, multi-genre and fine-grained dataset for named entity recognition (and disambiguation)
Simone Tedeschi and Roberto Navigli. 2022 · 2022
Later among the works it cites.
DeepStruct: Pretraining of language models for structure prediction
Chenguang Wang, Xiao Liu, Zui Chen, Haoyun Hong, Jie Tang, and Dawn Song. 2022a · 2022
Later among the works it cites.
Super-NaturalInstructions: Generalization via declarative instructions on 1600+ NLP tasks
Yizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi, et al. 2022c · 2022
Later among the works it cites.
LEVEN: A large-scale Chinese legal event detection dataset
Feng Yao, Chaojun Xiao, Xiaozhi Wang, Zhiyuan Liu, Lei Hou, Cunchao Tu, Juanzi Li, Yun Liu, Weixing Shen, and Maosong Sun. 2022 · 2022
Later among the works it cites.
Rohan Anil, Andrew M. Dai, Orhan Firat, et al. 2023 · 2023
Closest in time.
Promptner: Prompting for named entity recognition
Dhananjay Ashok and Zachary C. Lipton. 2023 · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, Harsha Nori, Hamid Palangi, Marco Tulio Ribeiro, and Yi Zhang. 2023 · 2023
Closest in time.
Maybe only 0.5needed: A preliminary exploration of low training data instruction tuning
Hao Chen, Yiming Zhang, Qi Zhang, Hantao Yang, Xiaomeng Hu, Xuetao Ma, Yifan Yanggong, and Junbo Zhao. 2023 · 2023
Closest in time.
SemEval-2023 task 2: Fine-grained multilingual named entity recognition (MultiCoNER 2)
Besnik Fetahu, Sudipta Kar, Zhiyu Chen, Oleg Rokhlenko, and Shervin Malmasi. 2023 · 2023
Closest in time.
Chatgpt outperforms crowd-workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Mael Kubli. 2023 · 2023
Closest in time.
AutoGPT
Significant Gravitas. 2023 · 2023
Closest in time.
Instruction tuned models are quick learners
Himanshu Gupta, Saurabh Arjun Sawant, Swaroop Mishra, Mutsumi Nakamura, Arindam Mitra, Santosh Mashetty, and Chitta Baral. 2023 · 2023
Closest in time.
Data-efficient finetuning using cross-task nearest neighbors
Hamish Ivison, Noah A. Smith, Hannaneh Hajishirzi, and Pradeep Dasigi. 2023 · 2023
Closest in time.
Opt-iml: Scaling language model instruction meta learning through the lens of generalization
Srinivasan Iyer, Xi Victoria Lin, Ramakanth Pasunuru, Todor Mihaylov, Daniel Simig, Ping Yu, Kurt Shuster, Tianlu Wang, Qing Liu, Punit Singh Koura, Xian Li, Brian O’Horo, Gabriel Pereyra, Jeff Wang, Christopher Dewan, Asli Celikyilmaz, Luke Zettlemoyer, and Ves Stoyanov. 2023 · 2023
Closest in time.
Exploring the benefits of training expert language models over instruction tuning
Joel Jang, Seungone Kim, Seonghyeon Ye, Doyoung Kim, Lajanugen Logeswaran, Moontae Lee, Kyungjae Lee, and Minjoon Seo. 2023 · 2023
Closest in time.
Amr parsing with instruction fine-tuned pre-trained language models
Young-Suk Lee, Ramón Fernandez Astudillo, Radu Florian, Tahira Naseem, and Salim Roukos. 2023 · 2023
Closest in time.
A comprehensive evaluation of chatgpt’s zero-shot text-to-sql capability
Aiwei Liu, Xuming Hu, Lijie Wen, and Philip S. Yu. 2023 · 2023
Closest in time.
The flan collection: Designing data and methods for effective instruction tuning
Shayne Longpre, Le Hou, Tu Vu, Albert Webson, Hyung Won Chung, Yi Tay, Denny Zhou, Quoc V. Le, Barret Zoph, Jason Wei, and Adam Roberts. 2023 · 2023
Closest in time.
Pivoine: Instruction tuning for open-world information extraction
Keming Lu, Xiaoman Pan, Kaiqiang Song, Hongming Zhang, Dong Yu, and Jianshu Chen. 2023 · 2023
Closest in time.
Crosslingual generalization through multitask finetuning
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, Adam Roberts, Stella Biderman, Teven Le Scao, M Saiful Bari, Sheng Shen, Zheng-Xin Yong, Hailey Schoelkopf, Xiangru Tang, Dragomir Radev, Alham Fikri Aji, Khalid Almubarak, Samuel Albanie, Zaid Alyafeai, Albert Webson, Edward Raff, and Colin Raffel. 2023 · 2023
Closest in time.
Is chatgpt a general-purpose natural language processing task solver?
Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen, Michihiro Yasunaga, and Diyi Yang. 2023 · 2023
Closest in time.
Bloom: A 176b-parameter open-access multilingual language model
Teven Le Scao, Angela Fan, Christopher Akiki, et al. 2023 · 2023
Closest in time.
Revisiting relation extraction in the era of large language models
Somin Wadhwa, Silvio Amir, and Byron Wallace. 2023 · 2023
Closest in time.
Zero-shot information extraction via chatting with chatgpt
Xiang Wei, Xingyu Cui, Ning Cheng, Xiaobin Wang, Xin Zhang, Shen Huang, Pengjun Xie, Jinan Xu, Yufeng Chen, Meishan Zhang, Yong Jiang, and Wenjuan Han. 2023 · 2023
Closest in time.
GLM-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, Weng Lam Tam, Zixuan Ma, Yufei Xue, Jidong Zhai, Wenguang Chen, Zhiyuan Liu, Peng Zhang, Yuxiao Dong, and Jie Tang. 2023 · 2023
Closest in time.
Deepke-llm: A large language model based knowledge extraction toolkit
Ningyu Zhang, Jintian Zhang, Xiaohan Wang, et al. 2023 · 2023
Closest in time.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, Yifan Du, Chen Yang, Yushuo Chen, Zhipeng Chen, Jinhao Jiang, Ruiyang Ren, Yifan Li, Xinyu Tang, Zikang Liu, Peiyu Liu, Jian-Yun Nie, and Ji-Rong Wen. 2023 · 2023
Closest in time.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric. P Xing, Hao Zhang, Joseph E. Gonzalez, and Ion Stoica. 2023 · 2023
Closest in time.
Can chatgpt reproduce human-generated labels? a study of social computing tasks
Yiming Zhu, Peixian Zhang, Ehsan-Ul Haq, Pan Hui, and Gareth Tyson. 2023 · 2023
Closest in time.