Fetching the paper…
Reading the bibliography…
Large language models that exhibit instruction-following behaviour represent one of the biggest recent upheavals in conversational interfaces, a trend in large part fuelled by the release of OpenAI's ChatGPT, a proprietary large language model for text generation fine-tuned through reinforcement learning from human feedback (LLM+RLHF).
Fine-Tuning Language Models from Human Preferences
Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B. Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving. 2020 · 1909
Earlier work this paper cites.
Tools for conviviality
Ivan Illich. 1973 · 1973
Earlier work this paper cites.
The protein data bank: A computer-based archival file for macromolecular structures
Frances C. Bernstein, Thomas F. Koetzle, Graheme J. B. Williams, Edgar F. Meyer, Michael D. Brice, John R. Rodgers, Olga Kennard, Takehiko Shimanouchi, and Mitsuo Tasumi. 1977 · 1977
Earlier work this paper cites.
TAMER: Training an Agent Manually via Evaluative Reinforcement. In 2008 7th IEEE International Conference on Development and Learning . 292–297
W. Bradley Knox and Peter Stone. 2008 · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In 2009 IEEE Conference on Computer Vision and Pattern Recognition . 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
On Bullshit
Harry G. Frankfurt. 2009 · 2009
Earlier work this paper cites.
Extracting Training Data from Large Language Models
Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, Alina Oprea, and Colin Raffel. 2021 · 2012
Earlier work this paper cites.
None But Ourselves Can Free Our Minds: Critical Computational Literacy as a Pedagogy of Resistance
Clifford H. Lee and Elisabeth Soep. 2016 · 2016
Earlier work this paper cites.
How open science helps researchers succeed
Erin C. McKiernan, Philip E. Bourne, C. Titus Brown, Stuart Buck, Amye Kenall, Jennifer Lin, Damon McDougall, Brian A. Nosek, Karthik Ram, Courtney K. Soderberg, Jeffrey R. Spies, Kaitlin Thaney, Andrew Updegrove, Kara H. Woo, and Tal Yarkoni. 2016 · 2016
Earlier work this paper cites.
ImageNet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2017 · 2017
Earlier work this paper cites.
On Reproducible AI: Towards Reproducible Research, Open Science, and Digital Scholarship in AI Publications
Odd Erik Gundersen, Yolanda Gil, and David W. Aha. 2018 · 2018
Earlier work this paper cites.
State of the Art: Reproducibility in Artificial Intelligence
Odd Erik Gundersen and Sigbjørn Kjensmo. 2018 · 2018
Earlier work this paper cites.
Deep TAMER: Interactive Agent Shaping in High-Dimensional State Spaces
Garrett Warnell, Nicholas Waytowich, Vernon Lawhern, and Peter Stone. 2018 · 2018
Earlier work this paper cites.
Open Science, Open Data, and Open Scholarship: European Policies to Make Science Fit for the Twenty-First Century
Jean-Claude Burgelman, Corina Pascu, Katarzyna Szkuta, Rene Von Schomberg, Athanasios Karalopoulos, Konstantinos Repanas, and Michel Schouppe. 2019 · 2019
Earlier work this paper cites.
Model Cards for Model Reporting. In Proceedings of the Conference on Fairness, Accountability, and Transparency (FAT* ’19) . Association for Computing Machinery, New York, NY, USA, 220–229
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru. 2019 · 2019
Earlier work this paper cites.
The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, Shawn Presser, and Connor Leahy. 2020 · 2020
Earlier work this paper cites.
Transparency and reproducibility in artificial intelligence
Benjamin Haibe-Kains, George Alexandru Adam, Ahmed Hosny, Farnoosh Khodakarami, Levi Waldron, Bo Wang, Chris McIntosh, Anna Goldenberg, Anshul Kundaje, Casey S. Greene, Tamara Broderick, Michael M. Hoffman, Jeffrey T. Leek, Keegan Korthauer, Wolfgang Huber, Alvis Brazma, Joelle Pineau, Robert Tibshirani, Trevor Hastie, John P. A. Ioannidis, John Quackenbush, and Hugo J. W. L. Aerts. 2020 · 2020
Earlier work this paper cites.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?. In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency . ACM, Virtual Event Canada, 610–623
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Earlier work this paper cites.
Algorithmic injustice: a relational ethics approach
Abeba Birhane. 2021 · 2021
Earlier work this paper cites.
Towards Decolonising Computational Sciences
Abeba Birhane and Olivia Guest. 2021 · 2021
Earlier work this paper cites.
Large image datasets: A pyrrhic win for computer vision?. In 2021 IEEE Winter Conference on Applications of Computer Vision (WACV) . 1536–1546
Abeba Birhane and Vinay Uday Prabhu. 2021 · 2021
Cited alongside, same era.
Multimodal datasets: misogyny, pornography, and malignant stereotypes
Abeba Birhane, Vinay Uday Prabhu, and Emmanuel Kahembwe. 2021 · 2021
Cited alongside, same era.
Towards Accountability for Machine Learning Datasets: Practices from Software Engineering and Infrastructure. In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’21) . Association for Computing Machinery, New York, NY, USA, 560–575
Ben Hutchinson, Andrew Smart, Alex Hanna, Emily Denton, Christina Greer, Oddur Kjartansson, Parker Barnes, and Margaret Mitchell. 2021 · 2021
Cited alongside, same era.
Not Directly Stated, Not Explicitly Stored: Conversational Agents and the Privacy Threat of Implicit Information. In Adjunct Proceedings of the 29th ACM Conference on User Modeling, Adaptation and Personalization (UMAP ’21) . Association for Computing Machinery, New York, NY, USA, 388–391
Collaboration challenges in building ML-enabled systems: communication, documentation, engineering, and process. In Proceedings of the 44th International Conference on Software Engineering (ICSE ’22) . Association for Computing Machinery, New York, NY, USA, 413–425
Nadia Nahar, Shurui Zhou, Grace Lewis, and Christian Kästner. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Later among the works it cites.
EleutherAI: Going Beyond "Open Science" to "Science in the Open"
Jason Phang, Herbie Bradley, Leo Gao, Louis Castricato, and Stella Biderman. 2022 · 2022
Later among the works it cites.
A systematic evaluation of large language models of code. In Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming (MAPS 2022) . Association for Computing Machinery, New York, NY, USA, 1–10
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Martha Larson, Nelleke Oostdijk, and Frederik Zuiderveen Borgesius. 2021 · 2021
Cited alongside, same era.
Reusable Templates and Guides For Documenting Datasets and Models for Natural Language Processing and Generation: A Case Study of the HuggingFace and GEM Data and Model Cards. In Proceedings of the 1st Workshop on Natural Language Generation, Evaluation, and Metrics (GEM 2021) . Association for Computational Linguistics, Online, 121–135
Angelina McMillan-Major, Salomey Osei, Juan Diego Rodriguez, Pawan Sasanka Ammanamanchi, Sebastian Gehrmann, and Yacine Jernite. 2021 · 2021
Cited alongside, same era.
Data and its (dis)contents: A survey of dataset development and use in machine learning research
Amandalynne Paullada, Inioluwa Deborah Raji, Emily M. Bender, Emily Denton, and Alex Hanna. 2021 · 2021
Cited alongside, same era.
Changing the World by Changing the Data. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Online, 2182–2194
Anna Rogers. 2021 · 2021
Cited alongside, same era.
“Everyone wants to do the model work, not the data work”: Data Cascades in High-Stakes AI. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . ACM, Yokohama Japan, 1–15
Nithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong, Praveen Paritosh, and Lora M Aroyo. 2021 · 2021
Cited alongside, same era.
Highly accurate protein structure prediction for the human proteome
Kathryn Tunyasuvunakool, Jonas Adler, Zachary Wu, Tim Green, Michal Zielinski, Augustin Žídek, Alex Bridgland, Andrew Cowie, Clemens Meyer, Agata Laydon, Sameer Velankar, Gerard J. Kleywegt, Alex Bateman, Richard Evans, Alexander Pritzel, Michael Figurnov, Olaf Ronneberger, Russ Bates, Simon A. A. Kohl, Anna Potapenko, Andrew J. Ballard, Bernardino Romera-Paredes, Stanislav Nikolov, Rishub Jain, Ellen Clancy, David Reiman, Stig Petersen, Andrew W. Senior, Koray Kavukcuoglu, Ewan Birney, Pushmeet Kohli, John Jumper, and Demis Hassabis. 2021 · 2021
Cited alongside, same era.
Emerging Trends: SOTA-Chasing
Kenneth Ward Church and Valia Kordoni. 2022 · 2022
Cited alongside, same era.
Dynamic Planning in Open-Ended Dialogue using Reinforcement Learning
Deborah Cohen, Moonkyung Ryu, Yinlam Chow, Orgad Keller, Ido Greenberg, Avinatan Hassidim, Michael Fink, Yossi Matias, Idan Szpektor, Craig Boutilier, and Gal Elidan. 2022 · 2022
Cited alongside, same era.
Interactive Model Cards: A Human-Centered Approach to Model Documentation. In 2022 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’22) . Association for Computing Machinery, New York, NY, USA, 427–439
Anamaria Crisan, Margaret Drouhard, Jesse Vig, and Nazneen Rajani. 2022 · 2022
Cited alongside, same era.
Frank F. Xu, Uri Alon, Graham Neubig, and Vincent Josua Hellendoorn. 2022 · 2022
Later among the works it cites.
The growing influence of industry in AI research
Nur Ahmed, Muntasir Wahed, and Neil C. Thompson. 2023 · 2023
Closest in time.
Ten years after ImageNet: a 360° perspective on artificial intelligence
Sanjay Chawla, Preslav Nakov, Ahmed Ali, Wendy Hall, Issa Khalil, Xiaosong Ma, Husrev Taha Sencar, Ingmar Weber, Michael Wooldridge, and Ting Yu. 2023 · 2023
Closest in time.
Statement from the listed authors of Stochastic Parrots on the "AI pause" letter
Timnit Gebru, Emily M. Bender, Angelina McMillan-Major, and Margaret Mitchell. 2023 · 2023
Closest in time.
Trustworthy AI: From Principles to Practices
Bo Li, Peng Qi, Bo Liu, Shuai Di, Jingen Liu, Jiquan Pei, Jinfeng Yi, and Bowen Zhou. 2023 · 2023
Closest in time.
Data Statements: From Technical Concept to Community Practice
Angelina McMillan-Major, Emily M. Bender, and Batya Friedman. 2023 · 2023
Closest in time.
Augmented Language Models: a Survey
Grégoire Mialon, Roberto Dessì, Maria Lomeli, Christoforos Nalmpantis, Ram Pasunuru, Roberta Raileanu, Baptiste Rozière, Timo Schick, Jane Dwivedi-Yu, Asli Celikyilmaz, Edouard Grave, Yann LeCun, and Thomas Scialom. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
OpenAI Might Invite Legal Trouble
Mohit Pandey. 2023 · 2023
Closest in time.
Inside the secret list of websites that make AI like ChatGPT sound smart
Kevin Schaul, Szu Yu Chen, and Nitasha Tiku. 2023 · 2023
Closest in time.
The Gradient of Generative AI Release: Methods and Considerations. In Proceedings of the 2023 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’23) . Association for Computing Machinery, New York, NY, USA, 111–122
Irene Solaiman. 2023 · 2023
Closest in time.
Why open-source generative AI models are an ethical way forward for science
Arthur Spirling. 2023 · 2023
Closest in time.
LLaMA: Open and Efficient Foundation Language Models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. 2023 · 2023
Closest in time.
Self-Instruct: Aligning Language Models with Self-Generated Instructions
Yizhong Wang, Yeganeh Kordi, Swaroop Mishra, Alisa Liu, Noah A. Smith, Daniel Khashabi, and Hannaneh Hajishirzi. 2023 · 2023
Closest in time.
BLOOM: A 176B-Parameter Open-Access Multilingual Language Model
BigScience Workshop. 2023 · 2023
Closest in time.
Baize: An Open-Source Chat Model with Parameter-Efficient Tuning on Self-Chat Data
Canwen Xu, Daya Guo, Nan Duan, and Julian McAuley. 2023 · 2023
Closest in time.