Fetching the paper…
Reading the bibliography…
Language models have steadily increased in size over the past few years.
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, et al (2020) Language models are few-shot learners. Advances in neural information processing systems 33:1877–1901
1901
Earlier work this paper cites.
1905
Earlier work this paper cites.
2001
Earlier work this paper cites.
Papineni K, Roukos S, Ward T, Zhu WJ (2002) Bleu: a method for automatic evaluation of machine translation. In: Proceedings of the 40th annual meeting of the Association for Computational Linguistics, pp 311–318
2002
Earlier work this paper cites.
Manning CD (2008) Introduction to information retrieval. Syngress Publishing,
2008
Earlier work this paper cites.
Robertson S, Zaragoza H (2009) The probabilistic relevance framework: Bm25 and beyond. Found Trends Inf Retr 3(4):333–389, DOI
2009
Earlier work this paper cites.
Zhao T, Zhao R, Eskenazi M (2017) Learning discourse-level diversity for neural dialog models using conditional variational autoencoders. arXiv preprint arXiv:170310960
2017
Earlier work this paper cites.
Trinh TH, Le QV (2018) A simple method for commonsense reasoning. DOI
2018
Earlier work this paper cites.
Beltagy I, Lo K, Cohan A (2019) Scibert: A pretrained language model for scientific text. arXiv preprint arXiv:190310676
2019
Earlier work this paper cites.
Gupta P, Mehri S, Zhao T, Pavel A, Eskenazi M, Bigham J (2019) Investigating evaluation of open-domain dialogue systems with human generated multiple references. In: Proceedings of the 20th Annual SIGdial Meeting on Discourse and Dialogue, Association for Computational Linguistics, Stockholm, Sweden, pp 379–391, DOI
2019
Earlier work this paper cites.
Liu Y, Ott M, Goyal N, Du J, Joshi M, Chen D, Levy O, Lewis M, Zettlemoyer L, Stoyanov V (2019) Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:190711692
2019
Earlier work this paper cites.
Radford A, Wu J, Child R, Luan D, Amodei D, Sutskever I, et al (2019) Language models are unsupervised multitask learners. OpenAI blog 1(8):9
2019
Earlier work this paper cites.
Zhang Y, Sun S, Galley M, Chen YC, Brockett C, Gao X, Gao J, Liu J, Dolan B (2019) Dialogpt: Large-scale generative pre-training for conversational response generation. arXiv preprint arXiv:191100536
2019
Earlier work this paper cites.
Baumgartner J, Zannettou S, Keegan B, Squire M, Blackburn J (2020) The pushshift reddit dataset. In: Proceedings of the international AAAI conference on web and social media, vol 14, pp 830–839
2020
Earlier work this paper cites.
Huang L, Ye Z, Qin J, Lin L, Liang X (2020) GRADE: Automatic graph-enhanced coherence metric for evaluating open-domain dialogue systems. In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP), Association for Computational Linguistics, Online, pp 9230–9240, DOI
2020
Earlier work this paper cites.
Jiang Z, Xu FF, Araki J, Neubig G (2020) How can we know what language models know? Transactions of the Association for Computational Linguistics 8:423–438
2020
Cited alongside, same era.
Mehri S, Eskenazi M (2020a) Unsupervised evaluation of interactive dialog with DialoGPT. In: Proceedings of the 21th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Association for Computational Linguistics, 1st virtual meeting, pp 225–235, URL
2020
Cited alongside, same era.
Mehri S, Eskenazi M (2020b) USR: An unsupervised and reference free evaluation metric for dialog generation. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, Association for Computational Linguistics, Online, pp 681–707, DOI
2020
Cited alongside, same era.
Phy V, Zhao Y, Aizawa A (2020) Deconstruct to reconstruct a configurable evaluation metric for open-domain dialogue systems. In: Proceedings of the 28th International Conference on Computational Linguistics, International Committee on Computational Linguistics, Barcelona, Spain (Online), pp 4164–4178, DOI
Wei J, Bosma M, Zhao VY, Guu K, Yu AW, Lester B, Du N, Dai AM, Le QV (2021) Finetuned language models are zero-shot learners. arXiv preprint arXiv:210901652
2021
Later among the works it cites.
Zhang C, Chen Y, D’Haro LF, Zhang Y, Friedrichs T, Lee G, Li H (2021a) DynaEval: Unifying turn and dialogue level evaluation. In: Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), Association for Computational Linguistics, Online, pp 5676–5689, DOI
2021
Later among the works it cites.
Zhao Z, Wallace E, Feng S, Klein D, Singh S (2021) Calibrate before use: Improving few-shot performance of language models. In: International Conference on Machine Learning, PMLR, pp 12,697–12,706
2021
Later among the works it cites.
BigScience Workshop (2022) Bloom (revision 4ab0472). DOI
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
Raffel C, Shazeer N, Roberts A, Lee K, Narang S, Matena M, Zhou Y, Li W, Liu PJ (2020) Exploring the limits of transfer learning with a unified text-to-text transformer. Journal of Machine Learning Research 21(140):1–67, URL
2020
Cited alongside, same era.
Sai AB, Mohankumar AK, Arora S, Khapra MM (2020) Improving dialog evaluation with a multi-reference adversarial dataset and large scale pretraining. Transactions of the Association for Computational Linguistics 8:810–827
2020
Cited alongside, same era.
Zhao T, Lala D, Kawahara T (2020) Designing precise and robust dialogue response evaluators. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, Association for Computational Linguistics, Online, pp 26–33, DOI
2020
Cited alongside, same era.
Agarwal O, Yang Y, Wallace BC, Nenkova A (2021) Interpretability analysis for named entity recognition to understand system predictions and how they can improve. Computational Linguistics 47(1):117–140
2021
Cited alongside, same era.
Chen Z, Sadoc J, D’Haro LF, Banchs R, Rudnicky A (2021) Automatic evaluation and moderation of open-domain dialogue systems. arXiv preprint arXiv:211102110
2021
Cited alongside, same era.
Gupta P, Tsvetkov Y, Bigham J (2021) Synthesizing adversarial negative responses for robust response ranking and evaluation. In: Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021, Association for Computational Linguistics, Online, pp 3867–3883, DOI
2021
Cited alongside, same era.
Liu J, Shen D, Zhang Y, Dolan B, Carin L, Chen W (2021) What makes good in-context examples for gpt-
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Chowdhery A, Narang S, Devlin J, Bosma M, Mishra G, Roberts A, Barham P, Chung HW, Sutton C, Gehrmann S, et al (2022) Palm: Scaling language modeling with pathways. arXiv preprint arXiv:220402311
2022
Later among the works it cites.
Chung HW, Hou L, Longpre S, Zoph B, Tay Y, Fedus W, Li E, Wang X, Dehghani M, Brahma S, et al (2022) Scaling instruction-finetuned language models. arXiv preprint arXiv:221011416
2022
Later among the works it cites.
Gupta P, Jiao C, Yeh YT, Mehri S, Eskenazi M, Bigham JP (2022) Improving zero and few-shot generalization in dialogue through instruction tuning. arXiv preprint arXiv:220512673
2022
Later among the works it cites.
Ouyang L, Wu J, Jiang X, Almeida D, Wainwright CL, Mishkin P, Zhang C, Agarwal S, Slama K, Ray A, et al (2022) Training language models to follow instructions with human feedback. arXiv preprint arXiv:220302155
2022
Later among the works it cites.
Sap M, LeBras R, Fried D, Choi Y (2022) Neural theory-of-mind? on the limits of social intelligence in large lms. arXiv preprint arXiv:221013312
2022
Later among the works it cites.
Smith S, Patwary M, Norick B, LeGresley P, Rajbhandari S, Casper J, Liu Z, Prabhumoye S, Zerveas G, Korthikanti V, et al (2022) Using deepspeed and megatron to train megatron-turing nlg 530b, a large-scale generative language model. arXiv preprint arXiv:220111990
2022
Later among the works it cites.
2022
Later among the works it cites.
Thoppilan R, De Freitas D, Hall J, Shazeer N, Kulshreshtha A, Cheng HT, Jin A, Bos T, Baker L, Du Y, et al (2022) Lamda: Language models for dialog applications. arXiv preprint arXiv:220108239
2022
Later among the works it cites.
Wei J, Tay Y, Bommasani R, Raffel C, Zoph B, Borgeaud S, Yogatama D, Bosma M, Zhou D, Metzler D, et al (2022) Emergent abilities of large language models. arXiv preprint arXiv:220607682
2022
Later among the works it cites.
Zhang S, Roller S, Goyal N, Artetxe M, Chen M, Chen S, Dewan C, Diab M, Li X, Lin XV, et al (2022) Opt: Open pre-trained transformer language models. arXiv preprint arXiv:220501068
2022
Later among the works it cites.