Fetching the paper…
Reading the bibliography…
In this study, we aim to reduce generation latency for Named Entity Recognition (NER) with Large Language Models (LLMs).
The genia corpus: An annotated research abstract corpus in molecular biology domain
Tomoko Ohta, Yuka Tateisi, Jin-Dong Kim, Hideki Mima, and Junichi Tsujii · 2002
Earlier work this paper cites.
Introduction to the conll-2003 shared task
Erik F. Tjong Kim Sang and Fien De Meulder · 2003
Earlier work this paper cites.
The third international Chinese language processing bakeoff: Word segmentation and named entity recognition
Gina-Anne Levow · 2006
Earlier work this paper cites.
Researching english as a lingua franca in asia: The asian corpus of english (ace) project
Andy Kirkpatrick · 2010
Earlier work this paper cites.
Asgard: A portable architecture for multilingual dialogue systems
Jingjing Liu, Panupong Pasupat, Scott Cyphers, and Jim Glass · 2013
Earlier work this paper cites.
Towards robust linguistic analysis using OntoNotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Hwee Tou Ng, Anders Björkelund, Olga Uryupina, Yuchen Zhang, and Zhi Zhong · 2013
Earlier work this paper cites.
Named entity recognition for chinese social media with jointly trained embeddings
Nanyun Peng and Mark Dredze · 2015
Earlier work this paper cites.
SGDR: Stochastic gradient descent with warm restarts
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Chinese NER using lattice LSTM
Yue Zhang and Jie Yang · 2018
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2018
Earlier work this paper cites.
Depth-adaptive transformer
Maha Elbayad, Jiatao Gu, Edouard Grave, and Michael Auli · 2019
Earlier work this paper cites.
Better modeling of incomplete annotations for named entity recognition
Zhanming Jie, Pengjun Xie, Wei Lu, Ruixue Ding, and Linlin Li · 2019
Earlier work this paper cites.
A neural multi-digraph model for chinese ner with gazetteers
Ruixue Ding, Pengjun Xie, Xiaoyan Zhang, Wei Lu, Linlin Li, and Luo Si · 2019
Earlier work this paper cites.
Icdar2019 competition on scanned receipt ocr and information extraction
Zheng Huang, Kai Chen, Jianhua He, Xiang Bai, Dimosthenis Karatzas, Shijian Lu, and CV Jawahar · 2019
Earlier work this paper cites.
Simplify the usage of lexicon in Chinese NER
Ruotian Ma, Minlong Peng, Qi Zhang, Zhongyu Wei, and Xuanjing Huang · 2020
Earlier work this paper cites.
Structured prediction as translation between augmented natural languages
Giovanni Paolini, Ben Athiwaratkun, Jason Krone, Jie Ma, Alessandro Achille, RISHITA ANUBHAI, Cicero Nogueira dos Santos, Bing Xiang, and Stefano Soatto · 2020
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
The pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, et al · 2020
Earlier work this paper cites.
Lexicon enhanced Chinese sequence labeling using BERT adapter
Wei Liu, Xiyan Fu, Yue Zhang, and Wenming Xiao · 2021
Earlier work this paper cites.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant · 2021
Cited alongside, same era.
Spatten: Efficient sparse attention architecture with cascade token and head pruning
Hanrui Wang, Zhekai Zhang, and Song Han · 2021
Cited alongside, same era.
Crosslingual generalization through multitask finetuning
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, Adam Roberts, Stella Biderman, Teven Le Scao, M Saiful Bari, Sheng Shen, Zheng-Xin Yong, Hailey Schoelkopf, et al · 2022
Cited alongside, same era.
A rationale-centric framework for human-in-the-loop machine learning
Jinghui Lu, Linyi Yang, Brian Namee, and Yue Zhang · 2022
Cited alongside, same era.
GPT3.int8(): 8-bit matrix multiplication for transformers at scale
Tim Dettmers, Mike Lewis, Younes Belkada, and Luke Zettlemoyer · 2022
Cited alongside, same era.
Gollie: Annotation guidelines improve zero-shot information-extraction
Oscar Sainz, Iker García-Ferrero, Rodrigo Agerri, Oier Lopez de Lacalle, German Rigau, and Eneko Agirre · 2023
Later among the works it cites.
Promptner: Prompting for named entity recognition
Dhananjay Ashok and Zachary C Lipton · 2023
Later among the works it cites.
Learning in-context learning for named entity recognition
Jiawei Chen, Yaojie Lu, Hongyu Lin, Jie Lou, Wei Jia, Dai Dai, Hua Wu, Boxi Cao, Xianpei Han, and Le Sun · 2023
Later among the works it cites.
Dan Friedman, Alexander Wettig, and Danqi Chen · 2023
Later among the works it cites.
OPTQ: Accurate quantization for generative pre-trained transformers
Elias Frantar, Saleh Ashkboos, Torsten Hoefler, and Dan Alistarh · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Confident adaptive language modeling
Tal Schuster, Adam Fisch, Jai Gupta, Mostafa Dehghani, Dara Bahri, Vinh Tran, Yi Tay, and Donald Metzler · 2022
Cited alongside, same era.
DeepStruct: Pretraining of language models for structure prediction
Chenguang Wang, Xiao Liu, Zui Chen, Haoyun Hong, Jie Tang, and Dawn Song · 2022
Cited alongside, same era.
Unified named entity recognition as word-word relation classification
Jingye Li, Hao Fei, Jiang Liu, Shengqiong Wu, Meishan Zhang, Chong Teng, Donghong Ji, and Fei Li · 2022
Cited alongside, same era.
Rethinking the role of demonstrations: What makes in-context learning work?
Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe, Mike Lewis, Hannaneh Hajishirzi, and Luke Zettlemoyer · 2022
Cited alongside, same era.
Unified low-resource sequence labeling by sample-aware dynamic sparse finetuning
Sarkar Snigdha Sarathi Das, Haoran Zhang, Peng Shi, Wenpeng Yin, and Rui Zhang · 2023
Cited alongside, same era.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al · 2023
Cited alongside, same era.
What makes pre-trained language models better zero-shot learners?
Jinghui Lu, Dongsheng Zhu, Weidong Han, Rui Zhao, Brian Mac Namee, and Fei Tan · 2023
Cited alongside, same era.
Later among the works it cites.
Accelerating transformer inference for translation via parallel decoding
Andrea Santilli, Silvio Severino, Emilian Postolache, Valentino Maiorca, Michele Mancusi, Riccardo Marin, and Emanuele Rodola · 2023
Later among the works it cites.
Smoothquant: Accurate and efficient post-training quantization for large language models
Guangxuan Xiao, Ji Lin, Mickael Seznec, Hao Wu, Julien Demouth, and Song Han · 2023
Later among the works it cites.
Copy is all you need
Tian Lan, Deng Cai, Yan Wang, Heyan Huang, and Xian-Ling Mao · 2023
Later among the works it cites.
Nag-ner: a unified non-autoregressive generation framework for various ner tasks
Xinpeng Zhang, Ming Tan, Jingfan Zhang, and Wei Zhu · 2023
Later among the works it cites.
A local information perception enhancement–based method for chinese ner
Miao Zhang and Ling Lu · 2023
Later among the works it cites.
Gliner: Generalist model for named entity recognition using bidirectional transformer
Urchade Zaratiana, Nadi Tomeh, Pierre Holat, and Thierry Charnois · 2023
Later among the works it cites.
Rwkv: Reinventing rnns for the transformer era
Bo Peng, Eric Alcaide, Quentin Anthony, Alon Albalak, Samuel Arcadinho, Huanqi Cao, Xin Cheng, Michael Chung, Matteo Grella, Kranthi Kiran GV, et al · 2023
Later among the works it cites.
Goat: Fine-tuned llama outperforms gpt-4 on arithmetic tasks
Tiedong Liu and Bryan Kian Hsiang Low · 2023
Later among the works it cites.
Towards understanding chain-of-thought prompting: An empirical study of what matters
Boshi Wang, Sewon Min, Xiang Deng, Jiaming Shen, You Wu, Luke Zettlemoyer, and Huan Sun · 2023
Later among the works it cites.
Jinghui Lu, Haiyang Yu, Yanjie Wang, Yongjie Ye, Jingqun Tang, Ziwei Yang, Binghong Wu, Qi Liu, Hao Feng, Han Wang, et al · 2024
Closest in time.
Rethinking negative instances for generative named entity recognition
Yuyang Ding, Juntao Li, Pinzheng Wang, Zecheng Tang, Bowen Yan, and Min Zhang · 2024
Closest in time.
Universalner: Targeted distillation from large language models for open named entity recognition
Wenxuan Zhou, Sheng Zhang, Yu Gu, Muhao Chen, and Hoifung Poon · 2024
Closest in time.
Medusa: Simple llm inference acceleration framework with multiple decoding heads
Tianle Cai, Yuhong Li, Zhengyang Geng, Hongwu Peng, Jason D Lee, Deming Chen, and Tri Dao · 2024
Closest in time.