Fetching the paper…
Reading the bibliography…
The recent emergence and adoption of Machine Learning technology, and specifically of Large Language Models, has drawn attention to the need for systematic and transparent management of language data.
Towards Standardization of Data Licenses: The Montreal Data License
Misha Benjamin, Paul Gagnon, Negar Rostamzadeh, Chris Pal, Yoshua Bengio, and Alex Shee. 2019 · 1903
Earlier work this paper cites.
Show your work: Improved reporting of experimental results
Jesse Dodge, Suchin Gururangan, Dallas Card, Roy Schwartz, and Noah A. Smith. 2019 · 1909
Earlier work this paper cites.
Universal declaration of human rights . Vol. 3381
United Nations. General Assembly. 1949 · 1949
Earlier work this paper cites.
Black Skin, White Masks
Frantz Fanon. 1952 · 1952
Earlier work this paper cites.
International covenant on civil and political rights
UN General Assembly. 1966a · 1966
Earlier work this paper cites.
Language Behavior in a Black Urban Community
Claudia Mitchell-Kernan. 1971 · 1971
Earlier work this paper cites.
The development of attitudes toward dialect in Italian children
Cristiana Cremona and Elizabeth Bates. 1977 · 1977
Earlier work this paper cites.
Purity and danger: an analysis of the concepts of pollution and taboo
Mary Douglas. 1978 · 1978
Earlier work this paper cites.
Chicano Spanish: Cross Hispanic Language Attitudes toward Specific Lexical Items
Robert LeRoy Giron. 1982 · 1982
Earlier work this paper cites.
Situated Knowledges: The Science Question in Feminism and the Privilege of Partial Perspective
Donna Harawy. 1988 · 1988
Earlier work this paper cites.
Principles of Linguistic Change, Volume 1: Internal Factors
William Labov. 1994 · 1994
Earlier work this paper cites.
Our global neighbourhood: the report of the Commission on Global Governance
The Commission on Global Governance. 1995 · 1995
Earlier work this paper cites.
Corporations and Human Rights: A Theory of Legal Responsibility
Steven Ratner. 2001 · 2001
Earlier work this paper cites.
Masakhane - Machine Translation For Africa
Iroro Orife, Julia Kreutzer, Blessing Sibanda, Daniel Whitenack, Kathleen Siminyu, Laura Martinus, Jamiil Toure Ali, Jade Z. Abbott, Vukosi Marivate, Salomon Kabongo, Musie Meressa, Espoir Murhabazi, Orevaoghene Ahia, Elan Van Biljon, Arshath Ramkilowan, Adewale Akinfaderin, Alp Öktem, Wole Akin, Ghollah Kioko, Kevin Degila, Herman Kamper, Bonaventure Dossou, Chris Emezue, Kelechi Ogueji, and Abdallah Bashir. 2020 · 2003
Earlier work this paper cites.
ICT Tools to Support Public Participation in Water Resources Governance & Planning: Experiences from the Design and Testing of a Multi-Media Platform
Ângela Guimarães Pereira, Jean Daniel Rinaudo, Paul Jeffrey, J E M Blasques, Serafin Corral Quintana, Nathalie Courtois, Silvio Funtowicz, and V. Petit. 2003 · 2003
Earlier work this paper cites.
Power in global governance . Vol. 98
Michael Barnett and Raymond Duvall. 2004 · 2004
Earlier work this paper cites.
Don’t look now, but we’ve created a bureaucracy: the nature and roles of policies and rules in wikipedia. In Proceedings of the SIGCHI conference on human factors in computing systems . 1101–1110
Brian Butler, Elisabeth Joyce, and Jacqueline Pike. 2008 · 2008
Earlier work this paper cites.
Middle-class African Americans: Reactions and attitudes toward African American English
Jacquelyn Rahman. 2008 · 2008
Earlier work this paper cites.
Decentralization in Wikipedia governance
Andrea Forte, Vanesa Larco, and Amy Bruckman. 2009 · 2009
Earlier work this paper cites.
Semantic Variation: Meaning in society and in sociolinguistics . Vol. 2
Ruqaiya Hasan. 2009 · 2009
Earlier work this paper cites.
META-GOVERNANCE: VALUES, NORMS AND PRINCIPLES, AND THE MAKING OF HARD CHOICES
Jan Peter Kooiman and Svein Jentoft. 2009 · 2009
Earlier work this paper cites.
The work of sustaining order in Wikipedia: The banning of a vandal. In Proceedings of the 2010 ACM conference on Computer supported cooperative work . 117–126
R Stuart Geiger and David Ribes. 2010 · 2010
Earlier work this paper cites.
The regime complex for climate change
Robert O Keohane and David G Victor. 2011 · 2011
Earlier work this paper cites.
An integrative framework for collaborative governance
Kirk Emerson, Tina Nabatchi, and Stephen Balogh. 2012 · 2012
Earlier work this paper cites.
A framework for assessing power in collaborative governance processes
Jill M Purdy. 2012 · 2012
Earlier work this paper cites.
The N Word: Its History and Use in the African American Community
Jacquelyn Rahman. 2012 · 2012
Earlier work this paper cites.
Best practices and admissibility of forensic author identification
Carole E Chaski. 2013 · 2013
Earlier work this paper cites.
The institutional collective action framework
Richard C Feiock. 2013 · 2013
Earlier work this paper cites.
Work-to-rule: the emergence of algorithmic governance in Wikipedia. In Proceedings of the 6th International Conference on Communities and Technologies . 80–89
Claudia Müller-Birn, Leonhard Dobusch, and James D Herbsleb. 2013 · 2013
Earlier work this paper cites.
Regime complexes: A buzz, a boom, or a boost for global governance
Amandine Orsini, Jean-Frédéric Morin, and Oran Young. 2013 · 2013
Earlier work this paper cites.
Introduction: The institutional fragmentation of global environmental governance: Causes, consequences, and responses
Fariborz Zelli and Harro Van Asselt. 2013 · 2013
Earlier work this paper cites.
Human rights for more than one voice: rethinking political space beyond the global/local divide
Rebecca Adami. 2014 · 2014
Earlier work this paper cites.
2 The Disintegration of Property
Thomas C Grey. 2014 · 2014
Earlier work this paper cites.
Technology ecosystem governance
Jonathan Wareham, Paul B Fox, and Josep Lluís Cano Giner. 2014 · 2014
Earlier work this paper cites.
A Sociolinguistic Approach to Vernacular Varieties: Stigmas and Prejudices in the Case of the West Country Dialect
Cristina García-Bermejo Gallego. 2015 · 2015
Earlier work this paper cites.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books. In Proceedings of the IEEE international conference on computer vision . 19–27
Yukun Zhu, Ryan Kiros, Rich Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Earlier work this paper cites.
Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word Embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Zou, Venkatesh Saligrama, and Adam Kalai. 2016 · 2016
Earlier work this paper cites.
The multiple roles of sustainability indicators in informational governance: Between intended use and unanticipated influence
Markku Lehtonen, Léa Sébastien, and Thomas Bauler. 2016 · 2016
Earlier work this paper cites.
Language and linguistics on trial: Hearing Rachel Jeantel (and other vernacular speakers) in the courtroom and beyond
John Rickford and Sharese King. 2016 · 2016
Earlier work this paper cites.
Public artworks and the freedom of panorama controversy: a case of Wikimedia influence
Mélanie Dulong De Rosnay and Pierre-Carl Langlais. 2017 · 2017
Earlier work this paper cites.
UCI Machine Learning Repository
Dheeru Dua and Casey Graff. 2017 · 2017
Earlier work this paper cites.
An army of me: Sockpuppets in online discussion communities. In Proceedings of the 26th International Conference on World Wide Web . 857–866
Srijan Kumar, Justin Cheng, Jure Leskovec, and VS Subrahmanian. 2017 · 2017
Earlier work this paper cites.
On the coloniality of human rights
Nelson Maldonado-Torres. 2017 · 2017
Earlier work this paper cites.
Regulating data as property: a new construct for moving forward
Jeffrey Ritter and Anna Mayer. 2017 · 2017
Earlier work this paper cites.
Human rights literacy: A quest for meaning
Cornelia Roux and Petro Du Preez. 2013 · 2017
Earlier work this paper cites.
Meaningful Information and the Right to Explanation
Andrew D Selbst and Julia Powles. 2017 · 2017
Earlier work this paper cites.
Artificial Intelligence’s Fair Use Crisis
Benjamin LW Sobel. 2017 · 2017
Earlier work this paper cites.
Men Also Like Shopping: Reducing Gender Bias Amplification using Corpus-level Constraints. In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Copenhagen, Denmark, 2979–2989
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2017 · 2017
Earlier work this paper cites.
REGULATION (EU) 2018/1725 OF THE EUROPEAN PARLIAMENT AND OF THE COUNCIL
2018 · 2018
Earlier work this paper cites.
Data Statements for Natural Language Processing: Toward Mitigating System Bias and Enabling Better Science
Emily M. Bender and Batya Friedman. 2018 · 2018
Cited alongside, same era.
2018 reform of EU data protection rules
European Commission. 2018 · 2018
Cited alongside, same era.
“Participant” Perceptions of Twitter Research Ethics
Casey Fiesler and Nicholas Proferes. 2018 · 2018
Cited alongside, same era.
Implementation of an Open Science Policy in the context of management of CLARIN language resources. In CLARIN Annual Conference 2018
A. Kelli, Krister Lindén, Kadri Vider, P. Labropoulou, E. Ketzan, Pawel Kamocki, and P. Stranák. 2018 · 2018
Cited alongside, same era.
Deep Contextualized Word Representations. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2018, New Orleans, Louisiana, USA, June 1-6, 2018, Volume 1 (Long Papers) , Marilyn A. Walker, Heng Ji, and Amanda Stent (Eds.). Association for Computational Linguistics, 2227–2237
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?. In FAccT ’21: 2021 ACM Conference on Fairness, Accountability, and Transparency, Virtual Event / Toronto, Canada, March 3-10, 2021 , Madeleine Clare Elish, William Isaac, and Richard S. Zemel (Eds.). ACM, 610–623
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Later among the works it cites.
Unreliable Guidelines: Reliable sources and marginalized communities in French, English, and Spanish Wikipedias
Amber Berson, Monika Sengul-Jones, and Melissa Tamani. 2021 · 2021
Later among the works it cites.
Introducing the NeurIPS 2021 Paper Checklist
Alina Beygelzimer, Yann Dauphin, Percy Liang, and Jennifer Wortman Vaughan. 2021 · 2021
Later among the works it cites.
Algorithmic injustice: a relational ethics approach
Abeba Birhane. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Cited alongside, same era.
Improving Language Understanding by Generative Pre-Training
Alec Radford and Karthik Narasimhan. 2018 · 2018
Cited alongside, same era.
Technology governance and the innovation process
David E Winickoff and Sebastian M Pfotenhauer. 2018 · 2018
Cited alongside, same era.
Intellectual Property and Human Rights 2.0
Peter K Yu. 2018 · 2018
Cited alongside, same era.
Openness, inclusion and self-affirmation: Indigenous knowledge in open knowledge projects
Nathalie Casemajor, Christian Coocoo, and Karine Gentelet. 2019 · 2019
Cited alongside, same era.
Plug and play language models: A simple approach to controlled text generation
Sumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung, Eric Frank, Piero Molino, Jason Yosinski, and Rosanne Liu. 2019 · 2019
Cited alongside, same era.
Racial Bias in Hate Speech and Abusive Language Detection Datasets. In Proceedings of the Third Workshop on Abusive Language Online . Association for Computational Linguistics, 25–35
Thomas Davidson, Debasmita Bhattacharya, and Ingmar Weber. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2019, Minneapolis, MN, USA, June 2-7, 2019, Volume 1 (Long and Short Papers) , Jill Burstein, Christy Doran, and Thamar Solorio (Eds.). Association for Computational Linguistics, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Abeba Birhane, Pratyusha Kalluri, Dallas Card, William Agnew, Ravit Dotan, and Michelle Bao. 2021a · 2021
Later among the works it cites.
Multimodal datasets: misogyny, pornography, and malignant stereotypes
Abeba Birhane, Vinay Uday Prabhu, and Emmanuel Kahembwe. 2021b · 2021
Later among the works it cites.
Multimodal datasets: misogyny, pornography, and malignant stereotypes
Abeba Birhane, Vinay Uday Prabhu, and Emmanuel Kahembwe. 2021c · 2021
Later among the works it cites.
Systematic Inequalities in Language Technology Performance across the World’s Langu ages
Damián E. Blasi, Antonios Anastasopoulos, and Graham Neubig. 2021 · 2021
Later among the works it cites.
Toward Gender-Inclusive Coreference Resolution: An Analysis of Gender and Bias Throughout the Machine Learning Lifecycle
Yang Trista Cao and Hal Daumé III. 2021 · 2021
Later among the works it cites.
Responsible NLP Research Checklist
Marine Carpuat, Marie-Catherine de Marneffe, and Ivan Vladimir Meza Ruiz. 2021 · 2021
Later among the works it cites.
Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets
Isaac Caswell, Julia Kreutzer, Lisa Wang, Ahsan Wahab, Daan van Esch, Nasanbayar Ulzii-Orshikh, Allahsera Tapo, Nishant Subramani, Artem Sokolov, Claytone Sikasote, Monang Setyawan, Supheakmungkol Sarin, Sokhar Samb, Benoît Sagot, Clara Rivera, Annette Rios, Isabel Papadimitriou, Salomey Osei, Pedro Javier Ortiz Suárez, Iroro Orife, Kelechi Ogueji, Rubungo Andre Niyongabo, Toan Q. Nguyen, Mathias Müller, André Müller, Shamsuddeen Hassan Muhammad, Nanda Muhammad, Ayanda Mnyakeni, Jamshidbek Mirzakhalov, Tapiwanashe Matangira, Colin Leong, Nze Lawson, Sneha Kudugunta, Yacine Jernite, Mathias Jenny, Orhan Firat, Bonaventure F. P. Dossou, Sakhile Dlamini, Nisansa de Silva, Sakine Çabuk Ballı, Stella Biderman, Alessia Battisti, Ahmed Baruwa, Ankur Bapna, Pallavi Baljekar, Israel Abebe Azime, Ayodele Awokoya, Duygu Ataman, Orevaoghene Ahia, Oghenefego Ahia, Sweta Agrawal, and Mofetoluwa Adeyemi. 2021 · 2021
Later among the works it cites.
The Problem of Zombie Datasets: A Framework For Deprecating Datasets
Frances Corry, Hamsini Sridharan, Alexandra Luccioni, Mike Ananny, Jason Schultz, and Kate Crawford. 2021 · 2021
Later among the works it cites.
On the genealogy of machine learning datasets: A critical history of ImageNet
Emily Denton, Alex Hanna, Razvan Amironesei, Andrew Smart, and Hilary Nicole. 2021 · 2021
Later among the works it cites.
Documenting the English Colossal Clean Crawled Corpus
Jesse Dodge, Maarten Sap, Ana Marasovic, William Agnew, Gabriel Ilharco, Dirk Groeneveld, and Matt Gardner. 2021 · 2021
Later among the works it cites.
Disparate Limbo: How Administrative Law Erased Antidiscrimination
David Freeman Engstrom, Daniel E Ho, and Cristina Isabel Ceballos. 2021 · 2021
Later among the works it cites.
Proposal for a Regulation Laying down Harmonised Rules on Artificial Intelligence
European Commission. 2021 · 2021
Later among the works it cites.
A Survey of Race, Racism, and Anti-Racism in NLP
Anjalie Field, Su Lin Blodgett, Zeerak Waseem, and Yulia Tsvetkov. 2021 · 2021
Later among the works it cites.
The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, Shawn Presser, and Connor Leahy. 2021 · 2021
Later among the works it cites.
The gem benchmark: Natural language generation, its evaluation and metrics
Sebastian Gehrmann, Tosin Adewumi, Karmanya Aggarwal, Pawan Sasanka Ammanamanchi, Aremu Anuoluwapo, Antoine Bosselut, Khyathi Raghavi Chandu, Miruna Clinciu, Dipanjan Das, Kaustubh D Dhole, et al · 2021
Later among the works it cites.
Towards Accountability for Machine Learning Datasets: Practices from Software Engineering and Infrastructure. In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency (Virtual Event, Canada) (FAccT ’21) . Association for Computing Machinery, New York, NY, USA, 560–575
Ben Hutchinson, Andrew Smart, Alex Hanna, Emily Denton, Christina Greer, Oddur Kjartansson, Parker Barnes, and Margaret Mitchell. 2021 · 2021
Later among the works it cites.
Quentin Lhoest, Albert Villanova del Moral, Yacine Jernite, Abhishek Thakur, Patrick von Platen, Suraj Patil, Julien Chaumond, Mariama Drame, Julien Plu, Lewis Tunstall, Joe Davison, Mario Šaško, Gunjan Chhablani, Bhavitvya Malik, Simon Brandeis, Teven Le Scao, Victor Sanh, Canwen Xu, Nicolas Patry, Angelina McMillan-Major, Philipp Schmid, Sylvain Gugger, Clément Delangue, Théo Matussière, Lysandre Debut, Stas Bekman, Pierric Cistac, Thibault Goehringer, Victor Mustar, François Lagunas, Alexander Rush, and Thomas Wolf. 2021 · 2021
Later among the works it cites.
Are We Learning Yet? A Meta Review of Evaluation Failures Across Machine Learning
Thomas Liao, Rohan Taori, Inioluwa Deborah Raji, and Ludwig Schmidt. 2021 · 2021
Later among the works it cites.
What’s in the Box? An Analysis of Undesirable Content in the Common Crawl Corpus
Alexandra Sasha Luccioni and Joseph D Viviano. 2021 · 2021
Later among the works it cites.
Reusable Templates and Guides For Documenting Datasets and Models for Natural Language Processing and Generation: A Case Study of the HuggingFace and GEM Data and Model Cards. In Proceedings of the 1st Workshop on Natural Language Generation, Evaluation, and Metrics (GEM 2021) . Association for Computational Linguistics, Online, 121–135
Angelina McMillan-Major, Salomey Osei, Juan Diego Rodriguez, Pawan Sasanka Ammanamanchi, Sebastian Gehrmann, and Yacine Jernite. 2021 · 2021
Later among the works it cites.
No Intruder, no Validity: Evaluation Criteria for Privacy-Preserving Text Anonymization
Maximilian Mozes and Bennett Kleinberg. 2021 · 2021
Later among the works it cites.
Data and its (dis)contents: A survey of dataset development and use in machine learning research
Amandalynne Paullada, Inioluwa Deborah Raji, Emily M. Bender, Emily Denton, and Alex Hanna. 2021 · 2021
Later among the works it cites.
EU Law Analysis: The Ola & Uber Judgments: For the First Time a Court Recognises a GDPR Right to an Explanation for Algorithmic Decision-Making
Steve Peers. 2021 · 2021
Later among the works it cites.
Mitigating dataset harms requires stewardship: Lessons from 1000 papers
Kenny Peng, Arunesh Mathur, and A. Narayanan. 2021 · 2021
Later among the works it cites.
Improving reproducibility in machine learning research: a report from the NeurIPS 2019 reproducibility program
Joelle Pineau, Philippe Vincent-Lamarre, Koustuv Sinha, Vincent Larivière, Alina Beygelzimer, Florence d’Alché Buc, Emily Fox, and Hugo Larochelle. 2021 · 2021
Later among the works it cites.
A Human Rights Approach to Responsible AI
Vinod Prabhakaran, Iason Gabriel, Timnit Gebru, and Margaret Mitchell. 2021 · 2021
Later among the works it cites.
Learning Transferable Visual Models From Natural Language Supervision. In ICML
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021 · 2021
Later among the works it cites.
AI and the Everything in the Whole Wide World Benchmark
Inioluwa Deborah Raji, Emily M. Bender, Amandalynne Paullada, Emily Denton, and Alex Hanna. 2021 · 2021
Later among the works it cites.
Changing the World by Changing the Data. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Association for Computational Linguistics, Online, 2182–2194
Anna Rogers. 2021 · 2021
Later among the works it cites.
Just What do You Think You’re Doing, Dave?’ A Checklist for Responsible Data Use in NLP
Anna Rogers, Tim Baldwin, and Kobi Leins. 2021 · 2021
Later among the works it cites.
“Everyone wants to do the model work, not the data work”: Data Cascades in High-Stakes AI. In proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–15
Nithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong, Praveen Paritosh, and Lora M Aroyo. 2021 · 2021
Later among the works it cites.
Do datasets have politics? Disciplinary values in computer vision dataset development
Morgan Klaus Scheuerman, Alex Hanna, and Emily Denton. 2021 · 2021
Later among the works it cites.
LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Christoph Schuhmann, Richard Vencu, Romain Beaumont, Robert Kaczmarczyk, Clayton Mullis, Aarush Katta, Theo Coombes, Jenia Jitsev, and Aran Komatsuzaki. 2021 · 2021
Later among the works it cites.
Disembodied Machine Learning: On the Illusion of Objectivity in NLP
Zeerak Talat, Smarika Lulz, Joachim Bingel, and Isabelle Augenstein. 2021 · 2021
Later among the works it cites.
Adversarial glue: A multi-task benchmark for robustness evaluation of language models
Boxin Wang, Chejian Xu, Shuohang Wang, Zhe Gan, Yu Cheng, Jianfeng Gao, Ahmed Hassan Awadallah, and Bo Li. 2021 · 2021
Later among the works it cites.
Reconciling legal and technical approaches to algorithmic bias
Alice Xiang. 2021 · 2021
Later among the works it cites.
Bot-Adversarial Dialogue for Safe Conversational Agents. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Online, 2950–2968
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan. 2021 · 2021
Later among the works it cites.
Stella Biderman, Kieran Bicheno, and Leo Gao. 2022 · 2022
Closest in time.
Help:Cloud Services Introduction — Wikitech,
Wikitech contributors. 2021a · 2022
Closest in time.
Wikipedia:Five pillars
Wikipedia contributors. 2021b · 2022
Closest in time.
Training Compute-Optimal Large Language Models
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, Elena Buchatskaya, Trevor Cai, Eliza Rutherford, Diego de Las Casas, Lisa Anne Hendricks, Johannes Welbl, Aidan Clark, Tom Hennigan, Eric Noland, Katie Millican, George van den Driessche, Bogdan Damoc, Aurelia Guy, Simon Osindero, Karen Simonyan, Erich Elsen, Jack W. Rae, Oriol Vinyals, and L. Sifre. 2022 · 2022
Closest in time.
Documenting Geographically and Contextually Diverse Data Sources: The BigScience Catalogue of Language Data and Resources
Angelina McMillan-Major, Zaid Alyafeai, Stella Biderman, Kimbo Chen, Francesco De Toni, Gerard Dupont, Hady Elsahar, Chris Emezue, Alham Fikri Aji, Suzana Ilic, Nurulaqilla Khamis, Colin Leong, Maraim Masoud, Aitor Soroa, Pedro Ortiz Suarez, Zeerak Talat, Daniel van Strien, and Yacine Jernite. 2022 · 2022
Closest in time.
Linguistically-Inclusive Natural Language Processing
Samson Tan. 2022 · 2022
Closest in time.
International covenant on economic, social and cultural rights
UN General Assembly. 1966b · 2057
Closest in time.