Fetching the paper…
Reading the bibliography…
We evaluate how well LLMs understand African American Language (AAL) in comparison to their performance on White Mainstream English (WME), the encouraged "standard" form of English taught in American classrooms.
Variable use of african american english across two language sampling contexts
Julie A. Washington, Holly K. Craig, and Amy J. Kushmaul. 1998 · 1998
Earlier work this paper cites.
Mock ebonics: Linguistic racism in parodies of ebonics on the internet
Maggie Ronkin and Helen E. Karn. 1999 · 1999
Earlier work this paper cites.
Regional variations in the phonological characteristics of african american vernacular english
Linette N. Hinton and Karen E. Pollock. 2000 · 2000
Earlier work this paper cites.
Sociocultural and historical contexts of African American English
Sonja L Lanehart. 2001 · 2001
Earlier work this paper cites.
7. “nuthin’ but a g thang”: Grammar and language ideology in hip hop identity
Marcyliena H. Morgan. 2001 · 2001
Earlier work this paper cites.
African American English: A Linguistic Introduction
Lisa J. Green. 2009 · 2009
Earlier work this paper cites.
Articulate while Black: Barack Obama, language, and race in the US
H Samy Alim and Geneva Smitherman. 2012 · 2012
Earlier work this paper cites.
The n word: Its history and use in the african american community
Jacquelyn Rahman. 2012 · 2012
Earlier work this paper cites.
755SWB (Speaking while Black): Linguistic Profiling and Discrimination Based on Speech as a Surrogate for Race against Speakers of African American Vernacular English
John Baugh. 2015 · 2015
Earlier work this paper cites.
Practice recommendations for pain assessment by self-report with african american older adults
Staja “Star” Booker, Chris Pasero, and Keela A. Herr. 2015 · 2015
Earlier work this paper cites.
Demographic dialectal variation in social media: A case study of African-American English
Su Lin Blodgett, Lisa Green, and Brendan O’Connor. 2016 · 2016
Earlier work this paper cites.
Learning a POS tagger for AAVE-like language
Anna Jørgensen, Dirk Hovy, and Anders Søgaard. 2016 · 2016
Earlier work this paper cites.
Racial disparity in natural language processing: A case study of social media african-american english
Su Lin Blodgett and Brendan O’Connor. 2017 · 2017
Earlier work this paper cites.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M. Bender and Batya Friedman. 2018 · 2018
Earlier work this paper cites.
Twitter Universal Dependency parsing for African-American and mainstream American English
Su Lin Blodgett, Johnny Wei, and Brendan O’Connor. 2018 · 2018
Earlier work this paper cites.
Examining gender and race bias in two hundred sentiment analysis systems
Svetlana Kiritchenko and Saif Mohammad. 2018 · 2018
Earlier work this paper cites.
Newsroom employees are less diverse than u.s. workers overall
Pew Research Center. 2018 · 2018
Earlier work this paper cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
Bias and fairness in natural language processing
Kai-Wei Chang, Vinodkumar Prabhakaran, and Vicente Ordonez. 2019 · 2019
Earlier work this paper cites.
Unified language model pre-training for natural language understanding and generation
Li Dong, Nan Yang, Wenhui Wang, Furu Wei, Xiaodong Liu, Yu Wang, Jianfeng Gao, Ming Zhou, and Hsiao-Wuen Hon. 2019 · 2019
Earlier work this paper cites.
Toward understanding the n-words
Jessica A Grieser. 2019 · 2019
Earlier work this paper cites.
Machine translation for arabic dialects (survey)
Salima Harrat, Karima Meftouh, and Kamel Smaili. 2019 · 2019
Earlier work this paper cites.
Dissecting racial bias in an algorithm used to manage the health of populations
Ziad Obermeyer, Brian Powers, Christine Vogeli, and Sendhil Mullainathan. 2019 · 2019
Earlier work this paper cites.
The woman worked as a babysitter: On biases in language generation
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng. 2019 · 2019
Cited alongside, same era.
Has nigga been reappropriated as a term of endearment? (a qualitative and quantitative analysis)
Hiram Smith. 2019 · 2019
Cited alongside, same era.
Linguistic justice: Black language, literacy, identity, and pedagogy
April Baker-Bell. 2020 · 2020
Cited alongside, same era.
Language (technology) is power: A critical survey of “bias” in NLP
Su Lin Blodgett, Solon Barocas, Hal Daumé III, and Hanna Wallach. 2020 · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Cited alongside, same era.
National-level disparities in internet access among low-income and black and hispanic youth: Current population survey
M Margaret Dolcini, Jesse A Canchola, Joseph A Catania, Marissa M Song Mayeda, Erin L Dietz, Coral Cotto-Negrón, and Vasudha Narayanan. 2021 · 2021
Later among the works it cites.
Sources of variation in the speech of african americans: Perspectives from sociophonetics
Charlie Farrington, Sharese King, and Mary Kohn. 2021 · 2021
Later among the works it cites.
The corpus of regional african american language
Tyler Kendall and Charlie Farrington. 2021 · 2021
Later among the works it cites.
Subjective bias in abstractive summarization
Lei Li, Wei Liu, Marina Litvak, Natalia Vanetik, Jiacheng Pei, Yinan Liu, and Siya Qi. 2021 · 2021
Later among the works it cites.
“i don’t think these devices are very culturally sensitive.”—impact of automated speech recognition errors on african americans
Zion Mengesha, Courtney Heldreth, Michal Lahav, Juliana Sublewski, and Elyse Tuennerman. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
It’s easier to translate out of English than into it: Measuring neural translation difficulty by cross-mutual information
Emanuele Bugliarello, Sabrina J. Mielke, Antonios Anastasopoulos, Ryan Cotterell, and Naoaki Okazaki. 2020 · 2020
Cited alongside, same era.
Investigating African-American Vernacular English in transformer-based text generation
Sophie Groenwold, Lily Ou, Aesha Parekh, Samhita Honnavalli, Sharon Levy, Diba Mirza, and William Yang Wang. 2020 · 2020
Cited alongside, same era.
Racial disparities in automated speech recognition
Allison Koenecke, Andrew Nam, Emily Lake, Joe Nudell, Minnie Quartey, Zion Mengesha, Connor Toups, John R. Rickford, Dan Jurafsky, and Sharad Goel. 2020 · 2020
Cited alongside, same era.
BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Cited alongside, same era.
Understanding Racial Disparities in Automatic Speech Recognition: The Case of Habitual “be”
Joshua L. Martin and Kevin Tang. 2020 · 2020
Cited alongside, same era.
Artie bias corpus: An open dataset for detecting demographic bias in speech applications
Josh Meyer, Lindy Rauchenstein, Joshua D. Eisenberg, and Nicholas Howell. 2020 · 2020
Cited alongside, same era.
Contextual analysis of social media: The promise and challenge of eliciting context in social media posts with natural language processing
Desmond U. Patton, William R. Frey, Kyle A. McGregor, Fei-Tzin Lee, Kathleen McKeown, and Emanuel Moss. 2020 · 2020
Cited alongside, same era.
Challenges in automated debiasing for toxic language detection
Xuhui Zhou, Maarten Sap, Swabha Swayamdipta, Yejin Choi, and Noah Smith. 2021 · 2021
Later among the works it cites.
Challenges in measuring bias via open-ended language generation
Afra Feyza Akyürek, Muhammed Yusuf Kocyigit, Sejin Paik, and Derry Tanti Wijaya. 2022 · 2022
Later among the works it cites.
Measuring annotator agreement generally across complex structured, multi-object, and free-text annotation tasks
Alexander Braylan, Omar Alonso, and Matthew Lease. 2022 · 2022
Later among the works it cites.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Alex Castro-Ros, Marie Pellat, Kevin Robinson, Dasha Valter, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Zhao, Yanping Huang, Andrew Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei. 2022 · 2022
Later among the works it cites.
Demystifying prompts in language models via perplexity estimation
Hila Gonen, Srini Iyer, Terra Blevins, Noah A. Smith, and Luke Zettlemoyer. 2022 · 2022
Later among the works it cites.
The Black side of the river: Race, language, and belonging in Washington, DC
Jessica A Grieser. 2022 · 2022
Later among the works it cites.
A medical chatbot using machine learning and natural language understanding
I-Ching Hsu and Jiun-De Yu. 2022 · 2022
Later among the works it cites.
Automatic Dialect Density Estimation for African American English
Alexander Johnson, Kevin Everson, Vijay Ravi, Anissa Gladney, Mari Ostendorf, and Abeer Alwan. 2022 · 2022
Later among the works it cites.
Algorithmic bias: review, synthesis, and future research directions
Nima Kordzadeh and Maryam Ghasemaghaei. 2022 · 2022
Later among the works it cites.
Corpus-guided contrast sets for morphosyntactic feature detection in low-resource English varieties
Tessa Masis, Anissa Neal, Lisa Green, and Brendan O’Connor. 2022 · 2022
Later among the works it cites.
Disambiguation of morpho-syntactic features of African American English – the case of habitual be
Harrison Santiago, Joshua Martin, Sarah Moeller, and Kevin Tang. 2022 · 2022
Later among the works it cites.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Later among the works it cites.
Unintended bias in language model-driven conversational recommendation
Tianshu Shen, Jiaru Li, Mohamed Reda Bouadjenek, Zheda Mai, and Scott Sanner. 2022 · 2022
Later among the works it cites.
VALUE: Understanding dialect disparity in NLU
Caleb Ziems, Jiaao Chen, Camille Harris, Jessica Anderson, and Diyi Yang. 2022 · 2022
Later among the works it cites.
Critical race theory: An introduction , volume 87
Richard Delgado and Jean Stefancic. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
U.s. journalists’ beats vary widely by gender and other factors
Pew Research Center. 2023 · 2023
Closest in time.