A stack-propagation framework with token-level intent detection for spoken language understanding
Libo Qin, Wanxiang Che, Yangming Li, Haoyang Wen, and T. Liu. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Later among the works it cites.
Multi-view sequence-to-sequence models with conversational structure for abstractive dialogue summarization
Jiaao Chen and Diyi Yang. 2020 · 2020
Later among the works it cites.
Keybert: Minimal keyword extraction with bert
Maarten Grootendorst. 2020 · 2020
Later among the works it cites.
How can we know what language models know?
Zhengbao Jiang, Frank F. Xu, Jun Araki, and Graham Neubig. 2020 · 2020
Later among the works it cites.
How domain terminology affects meeting summarization performance
Jia Jin Koay, Alexander Roustai, Xiaojin Dai, Dillon Burns, Alec Kerrigan, and Fei Liu. 2020 · 2020
Later among the works it cites.
Data augmentation using pre-trained transformer models
Varun Kumar, Ashutosh Choudhary, and Eunah Cho. 2020 · 2020
Later among the works it cites.
BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Later among the works it cites.
Birds have four legs?! NumerSense: Probing Numerical Commonsense Knowledge of Pre-Trained Language Models
Bill Yuchen Lin, Seyeon Lee, Rahul Khanna, and Xiang Ren. 2020 · 2020
Later among the works it cites.
Stanza: A python natural language processing toolkit for many human languages
Peng Qi, Yuhao Zhang, Yuhui Zhang, Jason Bolton, and Christopher D. Manning. 2020 · 2020
Later among the works it cites.
Modeling content importance for summarization with pre-trained language models
Liqiang Xiao, Lu Wang, Hao He, and Yaohui Jin. 2020 · 2020
Later among the works it cites.
Improving abstractive dialogue summarization with graph structures and topic words
Lulu Zhao, Weiran Xu, and Jun Guo. 2020 · 2020
Later among the works it cites.
Ask2transformers: Zero-shot domain labelling with pre-trained language models
Oscar Sainz and German Rigau. 2021 · 2021
Closest in time.