“Language Models are Few-Shot Learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“Towards deep learning models resistant to adversarial attacks”
Original
Aleksander Madry et al · 2017
Earlier work this paper cites.
“Attention is all you need”
Ashish Vaswani et al · 2017
Earlier work this paper cites.
“Certified adversarial robustness via randomized smoothing”
Jeremy Cohen, Elan Rosenfeld and Zico Kolter · 2019
Earlier work this paper cites.
“Adversarial examples: Attacks and defenses for deep learning”
Xiaoyong Yuan, Pan He, Qile Zhu and Xiaolin Li · 2019
Earlier work this paper cites.
“Globally-robust neural networks”
Klas Leino, Zifan Wang and Matt Fredrikson · 2021
Earlier work this paper cites.
“What’s in the box? an analysis of undesirable content in the Common Crawl corpus”
Alexandra Luccioni and Joseph Viviano · 2021
Earlier work this paper cites.
“Constitutional ai: Harmlessness from ai feedback”
Original
Yuntao Bai et al · 2022
Earlier work this paper cites.
“A survey for in-context learning”
Original
Qingxiu Dong et al · 2022
Earlier work this paper cites.
“Improving alignment of dialogue agents via targeted human judgements”
Original
Amelia Glaese et al · 2022
Earlier work this paper cites.
“Large language models are zero-shot reasoners”
Takeshi Kojima et al · 2022
Earlier work this paper cites.
“The Threat of Offensive AI to Organizations”
Yisroel Mirsky et al · 2022
Earlier work this paper cites.