Show your work: Scratchpads for intermediate computation with language models
Original
Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari, Henryk Michalewski, Jacob Austin, David Bieber, David Dohan, Aitor Lewkowycz, Maarten Bosma, David Luan, et al · 2021
Later among the works it cites.
Multitask prompted training enables zero-shot task generalization
Original
Victor Sanh, Albert Webson, Colin Raffel, Stephen H Bach, Lintang Sutawika, Zaid Alyafeai, Antoine Chaffin, Arnaud Stiegler, Teven Le Scao, Arun Raja, et al · 2021
Later among the works it cites.
Self-diagnosis and self-debiasing: A proposal for reducing corpus-based bias in NLP
Timo Schick, Sahana Udupa, and Hinrich Schütze · 2021
Later among the works it cites.
Embedding heterogeneous networks into hyperbolic space without meta-path
Lili Wang, Chongyang Gao, Chenghan Huang, Ruibo Liu, Weicheng Ma, and Soroush Vosoughi · 2021
Later among the works it cites.
Ethical and social risks of harm from language models
Original
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al · 2021
Later among the works it cites.
Calibrate before use: Improving few-shot performance of language models
Zihao Zhao, Eric Wallace, Shi Feng, Dan Klein, and Sameer Singh · 2021
Later among the works it cites.
Constitutional ai: Harmlessness from ai feedback
Original
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, et al · 2022
Later among the works it cites.
People found a really easy way to make Siri curse
BusinessInsider · 2022
Later among the works it cites.
Understanding iterative revision from human-written text
Wanyu Du, Vipul Raheja, Dhruv Kumar, Zae Myung Kim, Melissa Lopez, and Dongyeop Kang · 2022
Later among the works it cites.
Training compute-optimal large language models
Original
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, Elena Buchatskaya, Trevor Cai, Eliza Rutherford, Diego de Las Casas, Lisa Anne Hendricks, Johannes Welbl, Aidan Clark, et al · 2022
Later among the works it cites.
Microsoft’s virtual assistant ’will get mad’ if you ’say things that are particularly a–holeish’
Insider · 2022
Later among the works it cites.
Can language models learn from explanations in context?
Original
Andrew K Lampinen, Ishita Dasgupta, Stephanie CY Chan, Kory Matthewson, Michael Henry Tessler, Antonia Creswell, James L McClelland, Jane X Wang, and Felix Hill · 2022
Later among the works it cites.
Coauthor: Designing a human-ai collaborative writing dataset for exploring language model capabilities
Original
Mina Lee, Percy Liang, and Qian Yang · 2022
Later among the works it cites.
Non-parallel text style transfer with self-parallel supervision
Ruibo Liu, Chongyang Gao, Chenyan Jia, Guangxuan Xu, and Soroush Vosoughi · 2022
Later among the works it cites.
Knowledge infused decoding
Ruibo Liu, Guoqing Zheng, Shashank Gupta, Radhika Gaonkar, Chongyang Gao, Soroush Vosoughi, Milad Shokouhi, and Ahmed Hassan Awadallah · 2022
Later among the works it cites.
Quantifying and alleviating political bias in language models
Ruibo Liu, Chenyan Jia, Jason Wei, Guangxuan Xu, and Soroush Vosoughi · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Original
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Red teaming language models with language models
Original
Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving · 2022
Later among the works it cites.
Learning to model editing processes
Original
Machel Reid and Graham Neubig · 2022
Later among the works it cites.
Commonsenseqa 2.0: Exposing the limits of ai through gamification
Original
Alon Talmor, Ori Yoran, Ronan Le Bras, Chandra Bhagavatula, Yoav Goldberg, Yejin Choi, and Jonathan Berant · 2022
Later among the works it cites.
Benchmarking generalization via in-context instructions on 1,600+ language tasks
Original
Yizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi, Yeganeh Kordi, Amirreza Mirzaei, Anjana Arunkumar, Arjun Ashok, Arut Selvan Dhanasekaran, Atharva Naik, David Stap, et al · 2022
Later among the works it cites.
Chain of thought prompting elicits reasoning in large language models
Original
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou · 2022
Later among the works it cites.
The moral integrity corpus: A benchmark for ethical dialogue systems
Original
Caleb Ziems, Jane A Yu, Yi-Chia Wang, Alon Halevy, and Diyi Yang · 2022
Later among the works it cites.
Movie-DiC: a movie dialogue corpus for research and development
Rafael E. Banchs · 2040
Closest in time.