Fetching the paper…
Reading the bibliography…
Coloniality, the continuation of colonial harms beyond "official" colonization, has pervasive effects across society and scientific fields.
Language Models are Few-shot [l]earners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Racial Bias in Hate Speech and Abusive Language Detection Datasets
Thomas Davidson, Debasmita Bhattacharya, and Ingmar Weber. 2019 · 1905
Earlier work this paper cites.
Energy and Policy Considerations for Deep Learning in NLP
Emma Strubell, Ananya Ganesh, and Andrew McCallum. 2019 · 1906
Earlier work this paper cites.
Cross-lingual Name Tagging and Linking for 282 Languages
Xiaoman Pan, Boliang Zhang, Jonathan May, Joel Nothman, Kevin Knight, and Heng Ji. 2017 · 1958
Earlier work this paper cites.
A Dying Colonialism
Frantz Fanon. 1959 · 1959
Earlier work this paper cites.
Equitability Indices: Dependence on the Species Count
Andrew L Sheldon. 1969 · 1969
Earlier work this paper cites.
El Instituto Lingüístico de Verano, Instrumento del Imperialismo
TA DelValls. 1978 · 1978
Earlier work this paper cites.
Orientalism
Edward Said. 1978 · 1978
Earlier work this paper cites.
Brown v. Board of Education and the Interest-convergence Dilemma
Derrick A Bell Jr. 1980 · 1980
Earlier work this paper cites.
Do Artifacts Have Politics?
Langdon Winner. 1980 · 1980
Earlier work this paper cites.
The Social Life of Things: Commodities in Cultural Perspective
Arjun Appadurai. 1988 · 1988
Earlier work this paper cites.
The ATIS Spoken Language Systems Pilot Corpus
Charles T. Hemphill, John J. Godfrey, and George R. Doddington. 1990 · 1990
Earlier work this paper cites.
MUC-3 Evaluation Metrics
Nancy Chinchor. 1991 · 1991
Earlier work this paper cites.
Decolonising the Mind: The Politics of Language in African Literature
Ngugi Wa Thiong’o. 1992 · 1992
Earlier work this paper cites.
Inequality Reexamined
Amartya Sen. 1995 · 1995
Earlier work this paper cites.
Overview of Results of the MUC-6 Evaluation
Beth M. Sundheim. 1995 · 1995
Earlier work this paper cites.
On Oakland’s Ebonics
Jacquelyne Johnson Jackson. 1997 · 1997
Earlier work this paper cites.
Every Time I Fire a Linguist, my Performance Goes Up, and Other Myths of the Statistical natural Language Processing Revolution
Julia Hirschberg. 1998 · 1998
Earlier work this paper cites.
The Mobile Media Actor-Network in Urban India
Neha Kumar and Nimmi Rangaswamy. 2013 · 1998
Earlier work this paper cites.
Sorting Things Out
Geoffrey Bowker and Susan Leigh Star. 1999 · 1999
Earlier work this paper cites.
Foundations of Statistical Natural Language Processing
Christopher Manning and Hinrich Schutze. 1999 · 1999
Earlier work this paper cites.
Mock Ebonics: Linguistic Racism in Parodies of Ebonics on the Internet
Maggie Ronkin and Helen E Karn. 1999 · 1999
Earlier work this paper cites.
Decolonizing Methodologies: Research and Indigenous Peoples
Linda Tuhiwai Smith. 1999 · 1999
Earlier work this paper cites.
Regional Variations in the Phonological Characteristics of African American Vernacular English
Linette N Hinton and Karen E Pollock. 2000 · 2000
Earlier work this paper cites.
Coloniality of Power and Eurocentrism in Latin America
Anibal Quijano. 2000 · 2000
Earlier work this paper cites.
Colonial Linguistics
Joseph Errington. 2001 · 2001
Earlier work this paper cites.
Second-level Digital Divide: Mapping Differences in People’s Online Skills
Eszter Hargittai. 2001 · 2001
Earlier work this paper cites.
Development as Freedom
Amartya Sen. 2001 · 2001
Earlier work this paper cites.
Semiotics and the Social Analysis of Material Things
Webb Keane. 2003 · 2003
Earlier work this paper cites.
Masakhane–Machine Translation for Africa
Iroro Orife, Julia Kreutzer, Blessing Sibanda, Daniel Whitenack, Kathleen Siminyu, Laura Martinus, Jamiil Toure Ali, Jade Abbott, Vukosi Marivate, Salomon Kabongo, et al. 2020 · 2003
Earlier work this paper cites.
Internal Colonialism: An American Theory of Race
Ramón A. Gutiérrez. 2004 · 2004
Earlier work this paper cites.
Reassembling the Social: an Introduction to Actor-Network-Theory
Bruno Latour. 2005 · 2005
Earlier work this paper cites.
Measuring Linguistic Diversity on the Internet
John Paolillo, Daniel Pimienta, Daniel Prado, et al. 2005 · 2005
Earlier work this paper cites.
A Questão Da Gambiarra
Rodrigo Naumann Boufleur. 2006 · 2006
Earlier work this paper cites.
Indigenous, Ethnic and Cultural Articulations of New Media
Ramesh Srinivasan. 2006 · 2006
Earlier work this paper cites.
Using Actor-Network Theory to Analyze e-Government Implementation in Developing Countries
Carolyne Stanforth. 2006 · 2006
Earlier work this paper cites.
Corpus Development and Publication
Stephanie Strassel and Andrew W Cole. 2006 · 2006
Earlier work this paper cites.
Designing for Appropriation
Alan Dix. 2007 · 2007
Earlier work this paper cites.
On the Coloniality of Being: Contributions to the Development of a Concept
Nelson Maldonado-Torres. 2007 · 2007
Earlier work this paper cites.
Principles of Evaluation in Natural Language Processing
Patrick Paroubek, Stéphane Chaudiron, and Lynette Hirschman. 2007 · 2007
Earlier work this paper cites.
Statistical Machine Translation of Australian Aboriginal Languages: Morphological Analysis with Languages of Differing Morphological Richness
Simon Zwarts and Mark Dras. 2007 · 2007
Earlier work this paper cites.
Neo-colonialism and Research Collaboration in Central Africa
Nelius Boshoff. 2009 · 2009
Earlier work this paper cites.
Arabic Natural Language Processing: Challenges and Solutions
Ali Farghaly and Khaled Shaalan. 2009 · 2009
Earlier work this paper cites.
ICT4WHAT?-Using the Choice Framework to Operationalise the Capability Approach to Development
Dorothea Kleine. 2009 · 2009
Earlier work this paper cites.
GROBID: Combining Automatic Bibliographic Data Recognition and Term Extraction for Scholarship Publications
Patrice Lopez. 2009 · 2009
Earlier work this paper cites.
Twelve Years of Measuring Linguistic Diversity in the Internet: Balance and Perspectives
Daniel Pimienta, Daniel Prado, and Álvaro Blanco. 2009 · 2009
Earlier work this paper cites.
Theano: A CPU and GPU Math Compiler in Python
James Bergstra, Olivier Breuleux, Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, Guillaume Desjardins, Joseph Turian, David Warde-Farley, and Yoshua Bengio. 2010 · 2010
Earlier work this paper cites.
Materials and Medicine: Trade, Conquest and Therapeutics in the Eighteenth Century
Pratik Chakrabarti. 2010 · 2010
Earlier work this paper cites.
Investigating African-American Vernacular English in Transformer-Based Text Generation
Sophie Groenwold, Lily Ou, Aesha Parekh, Samhita Honnavalli, Sharon Levy, Diba Mirza, and William Yang Wang. 2020a · 2010
Earlier work this paper cites.
Inequality, Income Shares and Poverty: The Practical Meaning of Gini Coefficients
Malte Luebker. 2010 · 2010
Earlier work this paper cites.
On Achieving and Evaluating Language-independence in NLP
Emily M Bender. 2011 · 2011
Earlier work this paper cites.
Sustainability: Design for the Pluriverse
Arturo Escobar. 2011 · 2011
Earlier work this paper cites.
Towards a Computational History of the ACL: 1980-2008
Ashton Anderson, Dan Jurafsky, and Daniel A. McFarland. 2012 · 2012
Earlier work this paper cites.
Entangled: An Archaeology of the Relationships between Humans and Things
Ian Hodder. 2012 · 2012
Earlier work this paper cites.
Language Presence in the Real World and Cyberspace
Daniel Prado. 2012 · 2012
Earlier work this paper cites.
Decolonization is not a Metaphor
Eve Tuck and K Wayne Yang. 2012 · 2012
Earlier work this paper cites.
Digital Language Death
András Kornai. 2013 · 2013
Earlier work this paper cites.
Understanding Jugaad: ICTD and the Tensions of Appropriation, Innovation and Utility
Nimmi Rangaswamy and Melissa Densmore. 2013 · 2013
Earlier work this paper cites.
Arabizi Detection and Conversion to Arabic
Kareem Darwish. 2014 · 2014
Earlier work this paper cites.
Uneven Geographies of User-generated Information: Patterns of Increasing Informational Poverty
Mark Graham, Bernie Hogan, Ralph K Straumann, and Ahmed Medhat. 2014 · 2014
Earlier work this paper cites.
Introducing the Index
Nature. 2014 · 2014
Earlier work this paper cites.
The Digital Divide Shifts to Differences in Usage
Alexander JAM Van Deursen and Jan AGM Van Dijk. 2014 · 2014
Earlier work this paper cites.
A Massively Parallel Corpus: The Bible in 100 Languages
Christos Christodouloupoulos and Mark Steedman. 2015 · 2015
Earlier work this paper cites.
The Origins of African American Vernacular English
Donald Winford. 2015 · 2015
Earlier work this paper cites.
Mobile Technology Appropriation in a Distant Mirror: Baroquization, Areolization, and Cannibalism
François Bar, Matthew S Weber, and Francis Pisani. 2016 · 2016
Earlier work this paper cites.
The Social Impact of Natural Language Processing
Dirk Hovy and Shannon L Spruit. 2016 · 2016
Earlier work this paper cites.
Learning a POS Tagger for AAVE-like Language
Anna Jørgensen, Dirk Hovy, Anders Søgaard, et al. 2016 · 2016
Earlier work this paper cites.
How Translation Alters Sentiment
Saif M Mohammad, Mohammad Salameh, and Svetlana Kiritchenko. 2016 · 2016
Cited alongside, same era.
Voicing the Other: Mock AAVE on Social Media
Hanna L Smokoski. 2016 · 2016
Cited alongside, same era.
An Indigenous Feminist’s Take on the Ontological Turn:‘Ontology’is Just Another Word for Colonialism
Zoe Todd. 2016 · 2016
Cited alongside, same era.
Su Lin Blodgett and Brendan O’Connor. 2017 · 2017
Cited alongside, same era.
What can Machine Learning Do? Workforce Implications
Erik Brynjolfsson and Tom Mitchell. 2017 · 2017
Cited alongside, same era.
Writer Profiling without the Writer’s Text
David Jurgens, Yulia Tsvetkov, and Dan Jurafsky. 2017b · 2017
Automating Linguistics
Jacqueline Léon. 2021 · 2021
Later among the works it cites.
MTOP: A Comprehensive Multilingual Task-Oriented Semantic Parsing Benchmark
Haoran Li, Abhinav Arora, Shuohui Chen, Anchit Gupta, Sonal Gupta, and Yashar Mehdad. 2021 · 2021
Later among the works it cites.
Pollution is Colonialism
Max Liboiron. 2021 · 2021
Later among the works it cites.
Findings of the AmericasNLP 2021 Shared Task on Open Machine Translation for Indigenous Languages of the Americas
Manuel Mager, Arturo Oncevay, Abteen Ebrahimi, John Ortega, Annette Rios, Angela Fan, Ximena Gutierrez-Vasques, Luis Chiruzzo, Gustavo Giménez-Lugo, Ricardo Ramos, Ivan Vladimir Meza Ruiz, Rolando Coto-Solano, Alexis Palmer, Elisabeth Mager-Hois, Vishrav Chaudhary, Graham Neubig, Ngoc Thang Vu, and Katharina Kann. 2021 · 2021
Later among the works it cites.
Measuring Biological Diversity
Anne E Magurran. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The Chinese Typewriter: A History
Thomas S Mullaney. 2017 · 2017
Cited alongside, same era.
Data Statements for Natural Language Processing: Toward Mitigating System Bias and Enabling Better Science
Emily M. Bender and Batya Friedman. 2018 · 2018
Cited alongside, same era.
XNLI: Evaluating Cross-lingual Sentence Representations
Alexis Conneau, Guillaume Lample, Ruty Rinott, Adina Williams, Samuel R Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Cited alongside, same era.
ACM Code of Ethics and Professional Conduct
DW Gotterbarn, Bo Brinkman, Catherine Flick, Michael S Kirkpatrick, Keith Miller, Kate Vazansky, and Marty J Wolf. 2018 · 2018
Cited alongside, same era.
Challenges of Language Technologies for the Indigenous Languages of the Americas
Manuel Mager, Ximena Gutierrez-Vasques, Gerardo Sierra, and Ivan Meza-Ruiz. 2018 · 2018
Cited alongside, same era.
Universal Dependencies 2.2
Joakim Nivre, Mitchell Abrams, Željko Agić, Lars Ahrenberg, Lene Antonsen, Maria Jesus Aranzabe, Gashaw Arutie, Masayuki Asahara, Luma Ateyah, Mohammed Attia, et al. 2018 · 2018
Cited alongside, same era.
Zion Mengesha, Courtney Heldreth, Michal Lahav, Juliana Sublewski, and Elyse Tuennerman. 2021 · 2021
Later among the works it cites.
Carbon Emissions and Large Neural Network Training
David Patterson, Joseph Gonzalez, Quoc Le, Chen Liang, Lluis-Miquel Munguia, Daniel Rothchild, David So, Maud Texier, and Jeff Dean. 2021 · 2021
Later among the works it cites.
AI and the Everything in the Whole Wide World Benchmark
Inioluwa Deborah Raji, Emily M Bender, Amandalynne Paullada, Emily Denton, and Alex Hanna. 2021 · 2021
Later among the works it cites.
XTREME-R: Towards More Challenging and Nuanced Multilingual Evaluation
Sebastian Ruder, Noah Constant, Jan Botha, Aditya Siddhant, Orhan Firat, Jinlan Fu, Pengfei Liu, Junjie Hu, Dan Garrette, Graham Neubig, et al. 2021 · 2021
Later among the works it cites.
“everyone wants to do the model work, not the data work”: Data Cascades in High-Stakes AI
Nithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong, Praveen Paritosh, and Lora M Aroyo. 2021 · 2021
Later among the works it cites.
Beyond Fair Pay: Ethical Implications of NLP Crowdsourcing
Boaz Shmueli, Jan Fell, Soumya Ray, and Lun-Wei Ku. 2021 · 2021
Later among the works it cites.
mT5: A Massively Multilingual Pre-trained Text-to-Text Transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2021
Later among the works it cites.
One Country, 700+ Languages: NLP Challenges for Underrepresented Languages and Dialects in Indonesia
Alham Fikri Aji, Genta Indra Winata, Fajri Koto, Samuel Cahyawijaya, Ade Romadhony, Rahmad Mahendra, Kemal Kurniawan, David Moeljadi, Radityo Eko Prasojo, Timothy Baldwin, Jey Han Lau, and Sebastian Ruder. 2022 · 2022
Later among the works it cites.
JamPatoisNLI: A Jamaican Patois Natural Language Inference Dataset
Ruth-Ann Armstrong, John Hewitt, and Christopher Manning. 2022 · 2022
Later among the works it cites.
Local Languages, Third Spaces, and Other High-resource Scenarios
Steven Bird. 2022 · 2022
Later among the works it cites.
The Values Encoded in Machine Learning Research
Abeba Birhane, Pratyusha Kalluri, Dallas Card, William Agnew, Ravit Dotan, and Michelle Bao. 2022b · 2022
Later among the works it cites.
Towards a Deep Multi-layered Dialectal Language Analysis: A Case Study of African-American English
Jamell Dacon. 2022 · 2022
Later among the works it cites.
AmericasNLI: Evaluating Zero-shot Natural Language Understanding of Pretrained Multilingual Models in Truly Low-resource Languages
Abteen Ebrahimi, Manuel Mager, Arturo Oncevay, Vishrav Chaudhary, Luis Chiruzzo, Angela Fan, John Ortega, Ricardo Ramos, Annette Rios, Ivan Vladimir Meza Ruiz, Gustavo Giménez-Lugo, Elisabeth Mager, Graham Neubig, Alexis Palmer, Rolando Coto-Solano, Thang Vu, and Katharina Kann. 2022 · 2022
Later among the works it cites.
Dataset Geography: Mapping Language Data to Language Users
Fahim Faisal, Yinkai Wang, and Antonios Anastasopoulos. 2022 · 2022
Later among the works it cites.
Jack FitzGerald, Christopher Hench, Charith Peris, Scott Mackie, Kay Rottmann, Ana Sanchez, Aaron Nash, Liam Urbach, Vishesh Kakarala, Richa Singh, Swetha Ranganath, Laurie Crist, Misha Britan, Wouter Leeuwis, Gokhan Tur, and Prem Natarajan. 2022 · 2022
Later among the works it cites.
Exploring the Role of Grammar and Word Choice in Bias Toward African American English (AAE) in Hate Speech Classification
Camille Harris, Matan Halevy, Ayanna Howard, Amy Bruckman, and Diyi Yang. 2022 · 2022
Later among the works it cites.
Training Compute-Optimal Large Language Models
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, Elena Buchatskaya, Trevor Cai, Eliza Rutherford, Diego de Las Casas, Lisa Anne Hendricks, Johannes Welbl, Aidan Clark, Tom Hennigan, Eric Noland, Katie Millican, George van den Driessche, Bogdan Damoc, Aurelia Guy, Simon Osindero, Karen Simonyan, Erich Elsen, Jack W. Rae, Oriol Vinyals, and Laurent Sifre. 2022 · 2022
Later among the works it cites.
The Ghost in the Machine has an American Accent: Value Conflict in GPT-3
Rebecca L Johnson, Giada Pistilli, Natalia Menédez-González, Leslye Denisse Dias Duran, Enrico Panai, Julija Kalpokiene, and Donald Jay Bertulfo. 2022 · 2022
Later among the works it cites.
Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets
Julia Kreutzer, Isaac Caswell, Lisa Wang, Ahsan Wahab, Daan van Esch, Nasanbayar Ulzii-Orshikh, Allahsera Tapo, Nishant Subramani, Artem Sokolov, Claytone Sikasote, Monang Setyawan, Supheakmungkol Sarin, Sokhar Samb, Benoît Sagot, Clara Rivera, Annette Rios, Isabel Papadimitriou, Salomey Osei, Pedro Ortiz Suarez, Iroro Orife, Kelechi Ogueji, Andre Niyongabo Rubungo, Toan Q. Nguyen, Mathias Müller, André Müller, Shamsuddeen Hassan Muhammad, Nanda Muhammad, Ayanda Mnyakeni, Jamshidbek Mirzakhalov, Tapiwanashe Matangira, Colin Leong, Nze Lawson, Sneha Kudugunta, Yacine Jernite, Mathias Jenny, Orhan Firat, Bonaventure F. P. Dossou, Sakhile Dlamini, Nisansa de Silva, Sakine Çabuk Ballı, Stella Biderman, Alessia Battisti, Ahmed Baruwa, Ankur Bapna, Pallavi Baljekar, Israel Abebe Azime, Ayodele Awokoya, Duygu Ataman, Orevaoghene Ahia, Oghenefego Ahia, Sweta Agrawal, and Mofetoluwa Adeyemi. 2022 · 2022
Later among the works it cites.
What a Creole Wants, What a Creole Needs
Heather Lent, Kelechi Ogueji, Miryam de Lhoneux, Orevaoghene Ahia, and Anders Søgaard. 2022 · 2022
Later among the works it cites.
Tackling Hate Speech in Low-resource Languages with Context Experts
Daniel Nkemelu, Harshil Shah, Michael Best, and Irfan Essa. 2022 · 2022
Later among the works it cites.
Training Language Models to Follow Instructions with Human Feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
Square One Bias in NLP: Towards a Multi-Dimensional Exploration of the Research Manifold
Sebastian Ruder, Ivan Vulić, and Anders Søgaard. 2022 · 2022
Later among the works it cites.
Geographic Citation Gaps in NLP Research
Mukund Rungta, Janvijay Singh, Saif M Mohammad, and Diyi Yang. 2022 · 2022
Later among the works it cites.
Primum Non Nocere: Before working with Indigenous data, the ACL must confront ongoing colonialism
Lane Schwartz. 2022 · 2022
Later among the works it cites.
Compute Trends Across Three Eras of Machine Learning
Jaime Sevilla, Lennart Heim, Anson Ho, Tamay Besiroglu, Marius Hobbhahn, and Pablo Villalobos. 2022 · 2022
Later among the works it cites.
When is ML Data Good?: Valuing in Public Health Datafication
Divy Hasmukhbhai Thakkar, Azra Ismail, Pratyush Kumar, Alex Hanna, Nithya Sambasivan, and Neha Kumar. 2022 · 2022
Later among the works it cites.
Taxonomy of Risks Posed by Language Models
Laura Weidinger, Jonathan Uesato, Maribeth Rauh, Conor Griffin, Po-Sen Huang, John Mellor, Amelia Glaese, Myra Cheng, Borja Balle, Atoosa Kasirzadeh, et al. 2022 · 2022
Later among the works it cites.
VALUE: Understanding Dialect Disparity in NLU
Caleb Ziems, Jiaao Chen, Camille Harris, Jessica Anderson, and Diyi Yang. 2022 · 2022
Later among the works it cites.
Do All Languages Cost the Same? Tokenization in the Era of Commercial Language Models
Orevaoghene Ahia, Sachin Kumar, Hila Gonen, Jungo Kasai, David R. Mortensen, Noah A. Smith, and Yulia Tsvetkov. 2023 · 2023
Closest in time.
Varepsilon kú Mask: Integrating Yorùbá Cultural Greetings into Machine Translation
Idris Akinade, Jesujoba Alabi, David Adelani, Clement Odoje, and Dietrich Klakow. 2023 · 2023
Closest in time.
Probing Pre-Trained Language Models for Cross-Cultural Differences in Values
Arnav Arora, Lucie-aimée Kaffee, and Isabelle Augenstein. 2023 · 2023
Closest in time.
Social Commonsense for Explanation and Cultural Bias Discovery
Lisa Bauer, Hanna Tischer, and Mohit Bansal. 2023 · 2023
Closest in time.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
BIG bench authors. 2023 · 2023
Closest in time.
Decolonizing Data, One Language at a Time
Claudia Magallanes Blanco, Sabelo Mhlambi, Nanjala Nyabola, Nick Couldry, Toussaint Nothias, and Kathleen Siminyu. 2023 · 2023
Closest in time.
Evaluation for Change
Rishi Bommasani. 2023 · 2023
Closest in time.
Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models
Myra Cheng, Esin Durmus, and Dan Jurafsky. 2023 · 2023
Closest in time.
UniMax: Fairer and More Effective Language Sampling for Large-Scale Multilingual Pretraining
Hyung Won Chung, Xavier Garcia, Adam Roberts, Yi Tay, Orhan Firat, Sharan Narang, and Noah Constant. 2023 · 2023
Closest in time.
Towards Measuring the Representation of Subjective Global Opinions in Language Models
Esin Durmus, Karina Nyugen, Thomas I. Liao, Nicholas Schiefer, Amanda Askell, Anton Bakhtin, Carol Chen, Zac Hatfield-Dodds, Danny Hernandez, Nicholas Joseph, Liane Lovitt, Sam McCandlish, Orowa Sikder, Alex Tamkin, Janel Thamkul, Jared Kaplan, Jack Clark, and Deep Ganguli. 2023 · 2023
Closest in time.
Findings of the AmericasNLP 2023 Shared Task on Machine Translation into Indigenous Languages
Abteen Ebrahimi, Manuel Mager, Shruti Rijhwani, Enora Rice, Arturo Oncevay, Claudia Baltazar, María Cortés, Cynthia Montaño, John E. Ortega, Rolando Coto-solano, Hilaria Cruz, Alexis Palmer, and Katharina Kann. 2023 · 2023
Closest in time.
Yanai Elazar, Akshita Bhagia, Ian Magnusson, Abhilasha Ravichander, Dustin Schwenk, Alane Suhr, Pete Walsh, Dirk Groeneveld, Luca Soldaini, Sameer Singh, Hanna Hajishirzi, Noah A. Smith, and Jesse Dodge. 2023 · 2023
Closest in time.
GPTs are GPTs: An Early Look at the Labor Market Impact Potential of Large Language Models
Tyna Eloundou, Sam Manning, Pamela Mishkin, and Daniel Rock. 2023 · 2023
Closest in time.
"Honestly, I Think TikTok Has a Vendetta Against Black Creators": Understanding Black Content Creator Experiences on TikTok
Camille Harris, Amber Gayle Johnson, Sadie Palmer, Diyi Yang, and Amy Bruckman. 2023 · 2023
Closest in time.
Hate Speech Classifiers are Culturally Insensitive
Nayeon Lee, Chani Jung, and Alice Oh. 2023 · 2023
Closest in time.
The Data Provenance Initiative: A Large Scale Audit of Dataset Licensing & Attribution in AI
Shayne Longpre, Robert Mahari, Anthony Chen, Naana Obeng-Marnu, Damien Sileo, William Brannon, Niklas Muennighoff, Nathan Khazam, Jad Kabbara, Kartik Perisetla, Xinyi Wu, Enrico Shippole, Kurt Bollacker, Tongshuang Wu, Luis Villa, Sandy Pentland, Deb Roy, and Sara Hooker. 2023 · 2023
Closest in time.
Ethical Considerations for Machine Translation of Indigenous Languages: Giving a Voice to the Speakers
Manuel Mager, Elisabeth Mager, Katharina Kann, and Ngoc Thang Vu. 2023 · 2023
Closest in time.
Naamapadam: A Large-Scale Named Entity Annotated Data for Indic Languages
Arnav Mhaske, Harshit Kedia, Sumanth Doddapaneni, Mitesh M. Khapra, Pratyush Kumar, Rudra Murthy, and Anoop Kunchukuttan. 2023 · 2023
Closest in time.
Flickr Africa: Examining Geo-Diversity in Large-Scale, Human-Centric Visual Data
Keziah Naggita, Julienne LaChance, and Alice Xiang. 2023 · 2023
Closest in time.
Having Beer after Prayer? Measuring Cultural Bias in Large Language Models
Tarek Naous, Michael J Ryan, and Wei Xu. 2023 · 2023
Closest in time.
Mini But Mighty: Efficient Multilingual Pretraining with Linguistically-informed Data Selection
Tolúlọpẹ́ Ògúnrẹ̀mí, Dan Jurafsky, and Christopher Manning. 2023 · 2023
Closest in time.
Decolonizing NLP for “Low-resource Languages”: Applying Abeba Birhane’s Relational Ethics
Wilhelmina Onyothi Ògúnrẹ̀mí, Tolúlọpẹ́and Nekoto and Saron Samuel. 2023 · 2023
Closest in time.
Multilingual BERT has an Accent: Evaluating English Influences on Fluency in Multilingual Models
Isabel Papadimitriou, Kezia Lopez, and Dan Jurafsky. 2023 · 2023
Closest in time.
The ACL OCL Corpus: Advancing Open Science in Computational Linguistics
Shaurya Rohatgi, Yanxia Qin, Benjamin Aw, Niranjana Unnithan, and Min-Yen Kan. 2023 · 2023
Closest in time.
First Tragedy, then Parse: History Repeats Itself in the New Era of Large Language Models
Naomi Saphra, Eve Fleisig, Kyunghyun Cho, and Adam Lopez. 2023 · 2023
Closest in time.
On second thought, let’s not think step by step! bias and toxicity in zero-shot reasoning
Omar Shaikh, Hongxin Zhang, William Held, Michael Bernstein, and Diyi Yang. 2023 · 2023
Closest in time.
Llama 2: Open Foundation and Fine-tuned Chat Models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Closest in time.
Is ChatGPT a Good Teacher Coach? Measuring Zero-Shot Performance For Scoring and Providing Actionable Insights on Classroom Instruction
Rose Wang and Dorottya Demszky. 2023 · 2023
Closest in time.
The BabyLM Challenge: Sample-efficient Pretraining on a Developmentally Plausible Corpus
Alex Warstadt, Leshem Choshen, Aaron Mueller, Adina Williams, Ethan Wilcox, and Chengxu Zhuang. 2023 · 2023
Closest in time.
Low-Resource Languages Jailbreak GPT-4
Zheng-Xin Yong, Cristina Menghini, and Stephen H. Bach. 2023 · 2023
Closest in time.
Synthetic Lies: Understanding AI-generated Misinformation and Evaluating Algorithmic and Human Solutions
Jiawei Zhou, Yixuan Zhang, Qianni Luo, Andrea G Parker, and Munmun De Choudhury. 2023 · 2023
Closest in time.
Multi-VALUE: A framework for cross-dialectal English NLP
Caleb Ziems, William Held, Jingfeng Yang, Jwala Dhamala, Rahul Gupta, and Diyi Yang. 2023 · 2023
Closest in time.
Does ‘well-being’ Translate on Twitter?
Laura Smith, Salvatore Giorgi, Rishi Solanki, Johannes Eichstaedt, H. Andrew Schwartz, Muhammad Abdul-Mageed, Anneke Buffone, and Lyle Ungar. 2016 · 2047
Closest in time.