Fetching the paper…
Reading the bibliography…
As a natural language assistant, ChatGPT is capable of performing various tasks, including but not limited to article generation, code completion, and data analysis.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Albert: A lite bert for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut · 2019
Earlier work this paper cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le · 2019
Earlier work this paper cites.
Unified language model pre-training for natural language understanding and generation
Li Dong, Nan Yang, Wenhui Wang, Furu Wei, Xiaodong Liu, Yu Wang, Jianfeng Gao, Ming Zhou, and Hsiao-Wuen Hon · 2019
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, et al · 2019
Earlier work this paper cites.
Moverscore: Text generation evaluating with contextualized embeddings and earth mover distance
Wei Zhao, Maxime Peyrard, Fei Liu, et al · 2019
Earlier work this paper cites.
Language models are few-shot learners, July 2020
Tom B. Brown, Benjamin Mann, Nick Ryder, et al · 2020
Earlier work this paper cites.
Electra: Pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V Le, and Christopher D Manning · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu · 2020
Earlier work this paper cites.
Fine-tuning language models from human preferences, January 2020
Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu, et al · 2020
Earlier work this paper cites.
Comet: A neural framework for mt evaluation
Ricardo Rei, Craig Stewart, Ana C Farinha, and Alon Lavie · 2020
Earlier work this paper cites.
Bleurt: Learning robust metrics for text generation
Thibault Sellam, Dipanjan Das, and Ankur P Parikh · 2020
Earlier work this paper cites.
Finetuned language models are zero-shot learners
Jason Wei, Maarten Bosma, Vincent Y. Zhao, et al · 2021
Earlier work this paper cites.
Multitask prompted training enables zero-shot task generalization
Victor Sanh, Albert Webson, Colin Raffel, et al · 2021
Cited alongside, same era.
Evaluating large language models trained on code, July 2021
Mark Chen, Jerry Tworek, Heewoo Jun, et al · 2021
Cited alongside, same era.
Bartscore: Evaluating generated text as text generation
Weizhe Yuan, Graham Neubig, and Pengfei Liu · 2021
Cited alongside, same era.
GPT-J-6B: A 6 Billion Parameter Autoregressive Language Model
Ben Wang and Aran Komatsuzaki · 2021
Cited alongside, same era.
Opt: Open pre-trained transformer language models, June 2022
Susan Zhang, Stephen Roller, Naman Goyal, et al · 2022
Cited alongside, same era.
Palm: Scaling language modeling with pathways, October 2022
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, et al · 2022
Improving alignment of dialogue agents via targeted human judgements
Amelia Glaese, Nat McAleese, Maja Trębacz, John Aslanides, Vlad Firoiu, Timo Ewalds, Maribeth Rauh, Laura Weidinger, Martin Chadwick, Phoebe Thacker, et al · 2022
Later among the works it cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, et al · 2022
Later among the works it cites.
Learning to summarize from human feedback, February 2022
Nisan Stiennon, Long Ouyang, Jeff Wu, et al · 2022
Later among the works it cites.
Webgpt: Browser-assisted question-answering with human feedback, June 2022
Reiichiro Nakano, Jacob Hilton, Suchir Balaji, et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Bloom: A 176b-parameter open-access multilingual language model, December 2022
BigScience Workshop, Teven Le Scao, Angela Fan, et al · 2022
Cited alongside, same era.
Gpt-neox-20b: An open-source autoregressive language model, April 2022
Sid Black, Stella Biderman, Eric Hallahan, et al · 2022
Cited alongside, same era.
Training compute-optimal large language models, March 2022
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, et al · 2022
Cited alongside, same era.
Crosslingual generalization through multitask finetuning, November 2022
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, et al · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models, October 2022
Hyung Won Chung, Le Hou, Shayne Longpre, et al · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback, March 2022
Long Ouyang, Jeff Wu, Xu Jiang, et al · 2022
Cited alongside, same era.
Krishna Pillutla, Lang Liu, John Thickstun, et al · 2022
Later among the works it cites.
Is chatgpt a general-purpose natural language processing task solver?, February 2023
Chengwei Qin, Aston Zhang, Zhuosheng Zhang, et al · 2023
Closest in time.
Codegen: An open large language model for code with multi-turn program synthesis, February 2023
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, et al · 2023
Closest in time.
Large language models are state-of-the-art evaluators of translation quality, February 2023
Tom Kocmi and Christian Federmann · 2023
Closest in time.
Is chatgpt a good nlg evaluator? a preliminary study, March 2023
Jiaan Wang, Yunlong Liang, Fandong Meng, et al · 2023
Closest in time.
Pretraining language models with human preferences, February 2023
Tomasz Korbak, Kejian Shi, Angelica Chen, et al · 2023
Closest in time.
Chataug: Leveraging chatgpt for text data augmentation, February 2023
Haixing Dai, Zhengliang Liu, Wenxiong Liao, et al · 2023
Closest in time.
Chatgpt: Beginning of an end of manual annotation? use case of automatic genre identification, March 2023
Taja Kuzman, Nikola Ljubešić, and Igor Mozetič · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, et al · 2023
Closest in time.