Fetching the paper…
Reading the bibliography…
Recent advances in natural language processing (NLP) have led to the development of large language models (LLMs) such as ChatGPT.
Language models are few-shot learners
Brown T., Mann B., Ryder N., Subbiah M., Kaplan J. D., Dhariwal P., Neelakantan A., Shyam P., Sastry G., Askell A. et al · 1901
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Liu Y., Ott M., Goyal N., Du J., Joshi M., Chen D., Levy O., Lewis M., Zettlemoyer L. & Stoyanov V · 1907
Earlier work this paper cites.
Release strategies and the social impacts of language models
Solaiman I., Brundage M., Clark J., Askell A., Herbert-Voss A., Wu J., Radford A., Krueger G., Kim J. W., Kreps S. et al · 1908
Earlier work this paper cites.
Megatron-LM: Training multi-billion parameter language models using model parallelism
Shoeybi M., Patwary M., Puri R., LeGresley P., Casper J. & Catanzaro B · 1909
Earlier work this paper cites.
Www’18 open challenge: Financial opinion mining and question answering
Maia M., Handschuh S., Freitas A., Davis B., McDermott R., Zarrouk M. & Balahur A · 1942
Earlier work this paper cites.
Building a treebank for French
Abeillé A., Clément L. & Kinyon A · 2000
Earlier work this paper cites.
Attacking neural text detectors
Wolff M. & Wolff S · 2002
Earlier work this paper cites.
Meddialog: a large-scale medical dialogue dataset
Chen S., Ju Z., Dong X., Fang H., Wang S., Yang Y., Zeng J., Zhang R., Zhang R., Zhou M., Zhu P. & Xie P · 2004
Earlier work this paper cites.
The radicalization risks of gpt-3 and advanced neural language models
McGuffie K. & Newhouse A · 2009
Earlier work this paper cites.
Wikiqa: A challenge dataset for open-domain question answering
Yang Y., Yih S. W.-t. & Meek C · 2015
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Christiano P. F., Leike J., Brown T., Martic M., Legg S. & Amodei D · 2017
Earlier work this paper cites.
Improving language understanding by generative pre-training
Radford A., Narasimhan K., Salimans T. & Sutskever I · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin J., Chang M.-W., Lee K. & Toutanova K · 2019
Earlier work this paper cites.
ELI5: long form question answering
Fan A., Jernite Y., Perez E., Grangier D., Weston J. & Auli M · 2019
Cited alongside, same era.
Nlp augmentation
Ma E · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Radford A., Wu J., Child R., Luan D., Amodei D. & Sutskever I · 2019
Cited alongside, same era.
Defending against neural fake news
Zellers R., Holtzman A., Rashkin H., Bisk Y., Farhadi A., Roesner F. & Choi Y · 2019
Cited alongside, same era.
ELECTRA: Pre-training text encoders as discriminators rather than generators
Clark K., Luong M.-T., Le Q. V. & Manning C. D · 2020
Cited alongside, same era.
Unsupervised cross-lingual representation learning at scale
Conneau A., Khandelwal K., Goyal N., Chaudhary V., Wenzek G., Guzmán F., Grave E., Ott M., Zettlemoyer L. & Stoyanov V · 2020
Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity
Fedus W., Zoph B. & Shazeer N · 2021
Later among the works it cites.
Debertav3: Improving deberta using electra-style pre-training with gradient-disentangled embedding sharing
He P., Gao J. & Chen W · 2021
Later among the works it cites.
Machine translated text detection through text similarity with round-trip translation
Nguyen-Son H.-Q., Thao T., Hidano S., Gupta I. & Kiyomoto S · 2021
Later among the works it cites.
Scaling language models: Methods, analysis & insights from training gopher
Rae J. W., Borgeaud S., Cai T., Millican K., Hoffmann J., Song F., Aslanides J., Henderson S., Ring R., Young S. et al · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
CamemBERT: a tasty French language model
Martin L., Muller B., Ortiz Suárez P. J., Dupont Y., Romary L., de la Clergerie É., Seddah D. & Sagot B · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel C., Shazeer N., Roberts A., Lee K., Narang S., Matena M., Zhou Y., Li W. & Liu P. J · 2020
Cited alongside, same era.
Learning to summarize with human feedback
Stiennon N., Ouyang L., Wu J., Ziegler D., Lowe R., Voss C., Radford A., Amodei D. & Christiano P. F · 2020
Cited alongside, same era.
Authorship attribution for neural text generation
Uchendu A., Le T., Shu K. & Lee D · 2020
Cited alongside, same era.
On the dangers of stochastic parrots: Can language models be too big?
Bender E. M., Gebru T., McMillan-Major A. & Shmitchell S · 2021
Cited alongside, same era.
MFAQ: a multilingual FAQ dataset
De Bruyn M., Lotfi E., Buhmann J. & Daelemans W · 2021
Cited alongside, same era.
Weidinger L., Mellor J., Rauh M., Griffin C., Uesato J., Huang P.-S., Cheng M., Glaese M., Balle B., Kasirzadeh A. et al · 2021
Later among the works it cites.
Palm: Scaling language modeling with pathways
Chowdhery A., Narang S., Devlin J., Bosma M., Mishra G., Roberts A., Barham P., Chung H. W., Sutton C., Gehrmann S. et al · 2022
Later among the works it cites.
Training compute-optimal large language models
Hoffmann J., Borgeaud S., Mensch A., Buchatskaya E., Cai T., Rutherford E., Casas D. d. L., Hendricks L. A., Welbl J., Clark A. et al · 2022
Later among the works it cites.
Automatic detection of entity-manipulated text using factual knowledge
Jawahar G., Abdul-Mageed M. & Lakshmanan L · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Ouyang L., Wu J., Jiang X., Almeida D., Wainwright C., Mishkin P., Zhang C., Agarwal S., Slama K., Gray A., Schulman J., Hilton J., Kelton F., Miller L., Simens M., Askell A., Welinder P., Christiano P., Leike J. & Lowe R · 2022
Later among the works it cites.
Data-efficient french language modeling with camemberta
Antoun W., Sagot B. & Seddah D · 2023
Closest in time.
How close is chatgpt to human experts? comparison corpus, evaluation, and detection
Guo B., Zhang X., Wang Z., Jiang M., Nie J., Ding Y., Yue J. & Wu Y · 2023
Closest in time.
Detectgpt: Zero-shot machine-generated text detection using probability curvature
Mitchell E., Lee Y., Khazatsky A., Manning C. D. & Finn C · 2023
Closest in time.
Can ai-generated text be reliably detected?
Sadasivan V. S., Kumar A., Balasubramanian S., Wang W. & Feizi S · 2023
Closest in time.