Fetching the paper…
Reading the bibliography…
This paper aims to help structure the risk landscape associated with large-scale Language Models (LMs).
The Curious Case of Neural Text Degeneration
A. Holtzman, J. Buys, L. Du, M. Forbes, and Y. Choi · 1904
Earlier work this paper cites.
Gender Bias in Contextualized Word Embeddings
J. Zhao, T. Wang, M. Yatskar, R. Cotterell, V. Ordonez, and K.-W. Chang · 1904
Earlier work this paper cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
A. Wang, Y. Pruksachatkun, N. Nangia, A. Singh, J. Michael, F. Hill, O. Levy, and S. R. Bowman · 1905
Earlier work this paper cites.
Defending Against Neural Fake News
R. Zellers, A. Holtzman, H. Rashkin, Y. Bisk, A. Farhadi, F. Roesner, and Y. Choi · 1905
Earlier work this paper cites.
Measuring Bias in Contextualized Word Representations
K. Kurita, N. Vyas, A. Pareek, A. W. Black, and Y. Tsvetkov · 1906
Earlier work this paper cites.
Energy and Policy Considerations for Deep Learning in NLP
E. Strubell, A. Ganesh, and A. McCallum · 1906
Earlier work this paper cites.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov · 1907
Earlier work this paper cites.
A Survey on Bias and Fairness in Machine Learning
N. Mehrabi, F. Morstatter, N. Saxena, K. Lerman, and A. Galstyan · 1908
Earlier work this paper cites.
Release Strategies and the Social Impacts of Language Models
I. Solaiman, M. Brundage, J. Clark, A. Askell, A. Herbert-Voss, J. Wu, A. Radford, G. Krueger, J. W. Kim, S. Kreps, M. McCain, A. Newhouse, J. Blazakis, K. McGuffie, and J. Wang · 1908
Earlier work this paper cites.
Toward Gender-Inclusive Coreference Resolution
Y. T. Cao and H. Daumé III · 1910
Earlier work this paper cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 1910
Earlier work this paper cites.
Reducing Sentiment Bias in Language Models via Counterfactual Evaluation
P.-S. Huang, H. Zhang, R. Jiang, R. Stanforth, J. Welbl, J. Rae, V. Maini, D. Yogatama, and P. Kohli · 1911
Earlier work this paper cites.
Negated and Misprimed Probes for Pretrained Language Models: Birds Can Talk, But Cannot Fly
N. Kassner and H. Schütze · 1911
Earlier work this paper cites.
Predictive Biases in Natural Language Processing Models: A Conceptual Framework and Overview
D. Shah, H. A. Schwartz, and D. Hovy · 1912
Earlier work this paper cites.
Essai d’une recherche statistique sur le texte du roman “eugène oněgin”, illustrant la liaison des épreuves en chaîne
A. Markov · 1913
Earlier work this paper cites.
Predicting age groups of Twitter users based on language and metadata features
A. A. Morgan-Lopez, A. E. Kim, R. F. Chew, and P. Ruddle · 1932
Earlier work this paper cites.
Automatic personality assessment through social media language
G. Park, H. A. Schwartz, J. C. Eichstaedt, M. L. Kern, M. Kosinski, D. J. Stillwell, L. H. Ungar, and M. E. P. Seligman · 1939
Earlier work this paper cites.
Deep neural networks are more accurate than humans at detecting sexual orientation from facial images
Y. Wang and M. Kosinski · 1939
Earlier work this paper cites.
On Achieving and Evaluating Language-Independence in NLP
E. M. Bender · 1945
Earlier work this paper cites.
A mathematical theory of communication
C. E. Shannon · 1948
Earlier work this paper cites.
Continuous speech recognition by statistical methods
F. Jelinek · 1976
Earlier work this paper cites.
Do Artifacts Have Politics?
L. Winner · 1980
Earlier work this paper cites.
Secrecy and Openness in Science: Ethical Considerations
S. Bok · 1982
Earlier work this paper cites.
Feminism and Methodology: Social Science Issues
S. G. Harding · 1987
Earlier work this paper cites.
Situated knowledges: The science question in feminism and the privilege of partial perspective
D. Haraway · 1988
Earlier work this paper cites.
Stochastic modeling for automatic speech understanding
J. K. Baker · 1990
Earlier work this paper cites.
Scepticism
C. Hookway · 1990
Earlier work this paper cites.
The role of language in the persistence of stereotypes
A. Maass and L. Arcuri · 1992
Earlier work this paper cites.
Machine Learning Bias, Statistical Bias, and Statistical Variance of Decision Tree Algorithms
T. Dietterich and E. B. Kong · 1995
Earlier work this paper cites.
English with an Accent: Language, Ideology and Discrimination in the United States
R. Lippi · 1997
Earlier work this paper cites.
Sorting Things Out: Classification and Its Consequences
G. C. Bowker and S. L. Star · 1999
Earlier work this paper cites.
Infant-like Social Interactions between a Robot and a Human Caregiver
C. Breazeal and B. Scassellati · 2000
Earlier work this paper cites.
I. D. Raji, A. Smart, R. N. White, M. Mitchell, T. Gebru, B. Hutchinson, J. Smith-Loud, D. Theron, and P. Barnes · 2001
Earlier work this paper cites.
Toward an Afrocentric Feminist Epistemology
P. Hill Collins and N. Denzin · 2003
Earlier work this paper cites.
WES: Agent-based User Interaction Simulation on Real Infrastructure
J. Ahlgren, M. E. Berezin, K. Bojarczuk, E. Dulskyte, I. Dvortsova, J. George, N. Gucevska, M. Harman, R. Lämmel, E. Meijer, S. Sapora, and J. Spahr-Summers · 2004
Earlier work this paper cites.
The Haraway Reader
D. J. Haraway · 2004
Earlier work this paper cites.
The State and Fate of Linguistic Diversity and Inclusion in the NLP World
P. Joshi, S. Santy, A. Budhiraja, K. Bali, and M. Choudhury · 2004
Earlier work this paper cites.
Epistemic Relativism
S. Luper · 2004
Earlier work this paper cites.
StereoSet: Measuring stereotypical bias in pretrained language models
M. Nadeem, A. Bethke, and S. Reddy · 2004
Earlier work this paper cites.
Language modelling’s generative model: is it rational?
K. Sparck Jones · 2004
Earlier work this paper cites.
Language (Technology) is Power: A Critical Survey of "Bias" in NLP
S. L. Blodgett, S. Barocas, H. Daumé III, and H. Wallach · 2005
Earlier work this paper cites.
Language Models are Few-Shot Learners
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei · 2005
Earlier work this paper cites.
Intersectional Bias in Hate Speech and Abusive Language Datasets
J. Y. Kim, C. Ortiz, S. Nam, S. Santiago, and V. Datta · 2005
Earlier work this paper cites.
D. Martin Jr., V. Prabhakaran, J. Kuhlberg, A. Smart, and W. S. Isaac · 2005
Earlier work this paper cites.
The Law and Ethics of Trade Secrets: A Case Study
K. M. Saunders · 2005
Earlier work this paper cites.
Calibrating Noise to Sensitivity in Private Data Analysis
C. Dwork, F. McSherry, K. Nissim, and A. Smith · 2006
Earlier work this paper cites.
DeBERTa: Decoding-enhanced BERT with Disentangled Attention
P. He, X. Liu, J. Gao, and W. Chen · 2006
Earlier work this paper cites.
Secrecy and National Security Investigations
N. A. Sales · 2006
Earlier work this paper cites.
Bringing the People Back In: Contesting Benchmark Machine Learning Datasets
E. Denton, A. Hanna, R. Amironesei, A. Smart, H. Nicole, and M. K. Scheuerman · 2007
Earlier work this paper cites.
Participation is not a Design Fix for Machine Learning
M. Sloane, E. Moss, O. Awomolo, and L. Forlano · 2007
Earlier work this paper cites.
Race and epistemologies of ignorance
S. Sullivan and N. Tuana, editors · 2007
Earlier work this paper cites.
“Just Roll Your Mouse Over Me”: Designing Virtual Women for Customer Service on the Web
S. Zdenek · 2007
Earlier work this paper cites.
Neural net language models, January 2008
Y. Bengio · 2008
Earlier work this paper cites.
Discovering and Categorising Language Biases in Reddit
X. Ferrer, T. van Nuenen, J. M. Such, and N. Criado · 2008
Earlier work this paper cites.
Aligning AI With Shared Human Values
D. Hendrycks, C. Burns, S. Basart, A. Critch, J. Li, D. Song, and J. Steinhardt · 2008
Earlier work this paper cites.
Question and Answer Test-Train Overlap in Open-Domain Question Answering Datasets
P. Lewis, P. Stenetorp, and S. Riedel · 2008
Earlier work this paper cites.
RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models
S. Gehman, S. Gururangan, M. Sap, Y. Choi, and N. A. Smith · 2009
Earlier work this paper cites.
The Radicalization Risks of GPT-3 and Advanced Neural Language Models
K. McGuffie and A. Newhouse · 2009
Earlier work this paper cites.
Facilitating benign deceit in mediated communication
W. Moncur, J. Masthoff, and E. Reiter · 2009
Earlier work this paper cites.
Training Production Language Models without Memorizing User Data
S. Ramaswamy, O. Thakkar, R. Mathews, G. Andrew, H. B. McMahan, and F. Beaufays · 2009
Earlier work this paper cites.
Fair Hate Speech Detection through Evaluation of Social Group Counterfactuals
A. M. Davani, A. Omrani, B. Kennedy, M. Atari, X. Ren, and M. Dehghani · 2010
Earlier work this paper cites.
CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Masked Language Models
N. Nangia, C. Vania, R. Bhalerao, and S. R. Bowman · 2010
Earlier work this paper cites.
Information hazards: A typology of potential harms from knowledge
N. Bostrom et al · 2011
Earlier work this paper cites.
Automatic Detection of Machine Generated Text: A Critical Survey
G. Jawahar, M. Abdul-Mageed, and L. V. S. Lakshmanan · 2011
Earlier work this paper cites.
Loose tweets: an analysis of privacy leaks on twitter
H. Mao, X. Shuai, and A. Kapadia · 2011
Earlier work this paper cites.
Conversational Agents and Natural Language Interaction: Techniques and Effective Practices
D. Perez-Marin and I. Pascual-Nieto · 2011
Earlier work this paper cites.
Our Twitter Profiles, Our Selves: Predicting Personality with Twitter
D. Quercia, M. Kosinski, D. Stillwell, and J. Crowcroft · 2011
Earlier work this paper cites.
Thinking Inside the Box: Controlling and Using an Oracle AI
S. Armstrong, A. Sandberg, and N. Bostrom · 2012
Earlier work this paper cites.
Extracting Training Data from Large Language Models
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, U. Erlingsson, A. Oprea, and C. Raffel · 2012
Earlier work this paper cites.
Discipline and punish: the birth of the prison
M. Foucault and A. Sheridan · 2012
Earlier work this paper cites.
A Distributional Approach to Controlled Text Generation
M. Khalifa, H. Elsahar, and M. Dymetman · 2012
Earlier work this paper cites.
Anthropomorphism of computers: Is it mindful or mindless?
Y. Kim and S. S. Sundar · 2012
Earlier work this paper cites.
UNKs Everywhere: Adapting Multilingual Language Models to New Scripts
J. Pfeiffer, I. Vulić, I. Gurevych, and S. Ruder · 2012
Earlier work this paper cites.
The Effects of Culturally Congruent Educational Technologies on Student Achievement
S. Finkelstein, E. Yarzebinski, C. Vaughn, A. Ogan, and J. Cassell · 2013
Earlier work this paper cites.
China ’employs 2 million to police internet’
K. Hunt and C. Xu · 2013
Earlier work this paper cites.
Rising Income Inequality: Technology, or Trade and Financial Globalization?
F. Jaumotte, S. Lall, and C. Papageorgiou · 2013
Earlier work this paper cites.
Private traits and attributes are predictable from digital records of human behavior
M. Kosinski, D. Stillwell, and T. Graepel · 2013
Earlier work this paper cites.
Duplicate and fake publications in the scientific literature: how many SCIgen papers in computer science?
C. Labbé and D. Labbé · 2013
Earlier work this paper cites.
Providing Adaptive Health Updates Across the Personal Social Network
W. Moncur, J. Masthoff, E. Reiter, Y. Freer, and H. Nguyen · 2013
Earlier work this paper cites.
"How Old Do You Think I Am?" A Study of Language and Age in Twitter
D. Nguyen, R. Gravel, D. Trieschnigg, and T. Meder · 2013
Earlier work this paper cites.
Patent-Busting: The Public Patent Foundation, Gene Patents and the Seed Wars
M. Rimmer · 2013
Earlier work this paper cites.
Developing a framework for responsible innovation
J. Stilgoe, R. Owen, and P. Macnaghten · 2013
Earlier work this paper cites.
Superintelligence: paths, dangers, strategies
N. Bostrom · 2014
Earlier work this paper cites.
Echo Chamber or Public Sphere? Predicting Political Orientation and Measuring Political Homophily in Twitter Using Big Data
E. Colleoni, A. Rozza, and A. Arvidsson · 2014
Earlier work this paper cites.
Speech and language processing
D. Jurafsky and J. H. Martin · 2014
Earlier work this paper cites.
Predicting political preference of Twitter users
A. Makazhanov, D. Rafiei, and M. Waqar · 2014
Earlier work this paper cites.
The Racial Formation of Chatbots
M. Marino · 2014
Earlier work this paper cites.
Reclaiming Queer: Activist & Academic Rhetorics of Resistance
E. Rand · 2014
Earlier work this paper cites.
Publishers withdraw more than 120 gibberish papers
R. Van Noorden · 2014
Earlier work this paper cites.
Dark Matters
S. Browne · 2015
Earlier work this paper cites.
Unspeaking on Facebook? Testing network effects on self-censorship of political expressions in social network sites
K. H. Kwon, S.-I. Moon, and M. A. Stefanone · 2015
Earlier work this paper cites.
Down the (White) Rabbit Hole: The Extreme Right and Online Recommender Systems
D. O’Callaghan, D. Greene, M. Conway, J. Carthy, and P. Cunningham · 2015
Earlier work this paper cites.
Computer-based personality judgments are more accurate than those made by humans
W. Youyou, M. Kosinski, and D. Stillwell · 2015
Earlier work this paper cites.
Anthropomorphism: Opportunities and Challenges in Human–Robot Interaction
J. Złotowski, D. Proudfoot, K. Yogeeswaran, and C. Bartneck · 2015
Earlier work this paper cites.
Deep Learning with Differential Privacy
M. Abadi, A. Chu, I. Goodfellow, H. B. McMahan, I. Mironov, K. Talwar, and L. Zhang · 2016
Earlier work this paper cites.
Machine Bias
J. Angwin, J. Larson, S. Mattu, and L. Kirchner · 2016
Earlier work this paper cites.
Big Data’s Disparate Impact
S. Barocas and A. D. Selbst · 2016
Earlier work this paper cites.
Demographic Dialectal Variation in Social Media: A Case Study of African-American English
S. L. Blodgett, L. Green, and B. O’Connor · 2016
Earlier work this paper cites.
Doxing: a conceptual analysis
D. M. Douglas · 2016
Earlier work this paper cites.
DeepMind AI Reduces Google Data Centre Cooling Bill by 40%, July 2016
R. Evans and J. Gao · 2016
Earlier work this paper cites.
Equality of Opportunity in Supervised Learning
M. Hardt, E. Price, and N. Srebro · 2016
Cited alongside, same era.
Tay, Microsoft’s AI chatbot, gets a crash course in racism from Twitter
E. Hunt · 2016
Cited alongside, same era.
Smartphone-Based Conversational Agents and Responses to Questions About Mental Health, Interpersonal Violence, and Physical Health
A. S. Miner, A. Milstein, S. Schueller, R. Hegde, C. Mangurian, and E. Linos · 2016
Cited alongside, same era.
Online censors are a barrier to sex education, 2016
P. Oosterhoff · 2016
Cited alongside, same era.
The Black Box Society
F. Pasquale · 2016
Cited alongside, same era.
Indiscriminate mass surveillance and the public sphere
T. Stahl · 2016
Giving GPT-3 a Turing Test, July 2020
K. Lacker · 2020
Later among the works it cites.
Gender stereotypes are reflected in the distributional structure of 25 languages
M. Lewis and G. Lupyan · 2020
Later among the works it cites.
Racial mirroring effects on human-agent interaction in psychotherapeutic conversations
Y. Liao and J. He · 2020
Later among the works it cites.
Trends in U.S. income and wealth inequality
J. Menasce Horowitz, R. Igielnik, and R. Kochhar · 2020
Later among the works it cites.
Recommender systems and their ethical challenges
S. Milano, M. Taddeo, and L. Floridi · 2020
Later among the works it cites.
Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence
S. Mohamed, M.-T. Png, and W. Isaac · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Physiognomy’s New Clothes, May 2017
B. Agüera y Arcas, M. Mitchell, and A. Todorov · 2017
Cited alongside, same era.
S. L. Blodgett and B. O’Connor · 2017
Cited alongside, same era.
Semantics derived automatically from language corpora contain human-like biases
A. Caliskan, J. J. Bryson, and A. Narayanan · 2017
Cited alongside, same era.
Towards A Rigorous Science of Interpretable Machine Learning
F. Doshi-Velez and B. Kim · 2017
Cited alongside, same era.
Online Harassment 2017
M. Duggan · 2017
Cited alongside, same era.
Fake news infiltrates financial markets
C. Flood · 2017
Cited alongside, same era.
Interpreters and Translators: Occupational Outlook Handbook
B. of Labor Statistics · 2020
Later among the works it cites.
Misinformation in action: Fake news exposure is linked to lower trust in media, higher trust in government when your side is in power
K. Ognyanova, D. Lazer, R. E. Robertson, and C. Wilson · 2020
Later among the works it cites.
Social Media and Democracy: The State of the Field, Prospects for Reform
N. Persily and J. A. Tucker · 2020
Later among the works it cites.
Researchers made an OpenAI GPT-3 medical chatbot as an experiment. It told a mock patient to kill themselves
K. Quach · 2020
Later among the works it cites.
Handle with Care: Lessons for Data Science from Black Female Scholars
I. D. Raji · 2020
Later among the works it cites.
Could NLG systems injure or even kill people?, October 2020
E. Reiter · 2020
Later among the works it cites.
Turing-NLG: A 17-billion-parameter language model by Microsoft, February 2020
C. Rosset · 2020
Later among the works it cites.
Why You Should Do NLP Beyond English, August 2020
S. Ruder · 2020
Later among the works it cites.
Teaching GPT-3 to Identify Nonsense, July 2020
A. Sabeti · 2020
Later among the works it cites.
The ethics of nudging: An overview
A. T. Schmidt and B. Engelen · 2020
Later among the works it cites.
Bots Are Destroying Political Discourse As We Know It
B. Schneier · 2020
Later among the works it cites.
How AI is getting better at detecting hate speech, November 2020
M. Schroepfer · 2020
Later among the works it cites.
Green AI
R. Schwartz, J. Dodge, N. A. Smith, and O. Etzioni · 2020
Later among the works it cites.
Tackling threats to informed decision-making in democratic societies
E. Seger, S. Avin, G. Pearson, M. Briers, S. Ó Heigeartaigh, and H. Bacon · 2020
Later among the works it cites.
Who’s driving innovation
J. Stilgoe · 2020
Later among the works it cites.
Fiction by Neil Gaiman and Terry Pratchett by GPT-3, July 2020
summerstay on Reddit · 2020
Later among the works it cites.
Verse by Verse, 2020
VersebyVerse · 2020
Later among the works it cites.
Does GPT-2 Know Your Phone Number?, December 2020
D. Wallace, F. Tramer, M. Jagielski, and A. Herbert-Voss · 2020
Later among the works it cites.
Persistent Anti-Muslim Bias in Large Language Models
A. Abid, M. Farooqi, and J. Zou · 2021
Closest in time.
MasakhaNER: Named Entity Recognition for African Languages
D. I. Adelani, J. Abbott, G. Neubig, D. D’souza, J. Kreutzer, C. Lignos, C. Palen-Michel, H. Buzaaba, S. Rijhwani, S. Ruder, S. Mayhew, I. A. Azime, S. Muhammad, C. C. Emezue, J. Nakatumba-Nabende, P. Ogayo, A. Aremu, C. Gitau, D. Mbaye, J. Alabi, S. M. Yimam, T. Gwadabe, I. Ezeani, R. A. Niyongabo, J. Mukiibi, V. Otiende, I. Orife, D. David, S. Ngom, T. Adewumi, P. Rayson, M. Adeyemi, G. Muriuki, E. Anebi, C. Chukwuneke, N. Odu, E. P. Wairagala, S. Oyerinde, C. Siro, T. S. Bateesa, T. Oloyede, Y. Wambui, V. Akinode, D. Nabagereka, M. Katusiime, A. Awokoya, M. MBOUP, D. Gebreyohannes, H. Tilaye, K. Nwaike, D. Wolde, A. Faye, B. Sibanda, O. Ahia, B. F. P. Dossou, K. Ogueji, T. I. DIOP, A. Diallo, A. Akinfaderin, T. Marengereke, and S. Osei · 2021
Closest in time.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell · 2021
Closest in time.
Stereotyping Norwegian Salmon: An Inventory of Pitfalls in Fairness Benchmark Datasets
S. L. Blodgett, G. Lopez, A. Olteanu, R. Sim, and H. Wallach · 2021
Closest in time.
On the Opportunities and Risks of Foundation Models
R. Bommasani, D. A. Hudson, E. Adeli, R. Altman, S. Arora, S. von Arx, M. S. Bernstein, J. Bohg, A. Bosselut, E. Brunskill, E. Brynjolfsson, S. Buch, D. Card, R. Castellon, N. Chatterji, A. Chen, K. Creel, J. Q. Davis, D. Demszky, C. Donahue, M. Doumbouya, E. Durmus, S. Ermon, J. Etchemendy, K. Ethayarajh, L. Fei-Fei, C. Finn, T. Gale, L. Gillespie, K. Goel, N. Goodman, S. Grossman, N. Guha, T. Hashimoto, P. Henderson, J. Hewitt, D. E. Ho, J. Hong, K. Hsu, J. Huang, T. Icard, S. Jain, D. Jurafsky, P. Kalluri, S. Karamcheti, G. Keeling, F. Khani, O. Khattab, P. W. Koh, M. Krass, R. Krishna, R. Kuditipudi, A. Kumar, F. Ladhak, M. Lee, T. Lee, J. Leskovec, I. Levent, X. L. Li, X. Li, T. Ma, A. Malik, C. D. Manning, S. Mirchandani, E. Mitchell, Z. Munyikwa, S. Nair, A. Narayan, D. Narayanan, B. Newman, A. Nie, J. C. Niebles, H. Nilforoshan, J. Nyarko, G. Ogut, L. Orr, I. Papadimitriou, J. S. Park, C. Piech, E. Portelance, C. Potts, A. Raghunathan, R. Reich, H. Ren, F. Rong, Y. Roohani, C. Ruiz, J. Ryan, C. Ré, D. Sadigh, S. Sagawa, K. Santhanam, A. Shih, K. Srinivasan, A. Tamkin, R. Taori, A. W. Thomas, F. Tramèr, R. E. Wang, W. Wang, B. Wu, J. Wu, Y. Wu, S. M. Xie, M. Yasunaga, J. You, M. Zaharia, M. Zhang, T. Zhang, X. Zhang, Y. Zhang, L. Zheng, K. Zhou, and P. Liang · 2021
Closest in time.
Truth, Lies, and Truth, Lies, and Automation: How Language Models Could Change DisinformationAutomation: How Language Models Could Change Disinformation
B. Buchanan, A. Lohn, M. Musser, and S. Katerina · 2021
Closest in time.
Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets
I. Caswell, J. Kreutzer, L. Wang, A. Wahab, D. van Esch, N. Ulzii-Orshikh, A. Tapo, N. Subramani, A. Sokolov, C. Sikasote, M. Setyawan, S. Sarin, S. Samb, B. Sagot, C. Rivera, A. Rios, I. Papadimitriou, S. Osei, P. J. O. Suárez, I. Orife, K. Ogueji, R. A. Niyongabo, T. Q. Nguyen, M. Müller, A. Müller, S. H. Muhammad, N. Muhammad, A. Mnyakeni, J. Mirzakhalov, T. Matangira, C. Leong, N. Lawson, S. Kudugunta, Y. Jernite, M. Jenny, O. Firat, B. F. P. Dossou, S. Dlamini, N. de Silva, S. c. Ballı, S. Biderman, A. Battisti, A. Baruwa, A. Bapna, P. Baljekar, I. A. Azime, A. Awokoya, D. Ataman, O. Ahia, O. Ahia, S. Agrawal, and M. Adeyemi · 2021
Closest in time.
GitHub Copilot · Your AI pair programmer, 2021
CopilotonGitHub · 2021
Closest in time.
Atlas of AI
K. Crawford · 2021
Closest in time.
GPT-3: What’s it good for?
R. Dale · 2021
Closest in time.
Anticipating Safety Issues in E2E Conversational AI: Framework and Tooling
E. Dinan, G. Abercrombie, A. S. Bergman, S. Spruit, D. Hovy, Y.-L. Boureau, and V. Rieser · 2021
Closest in time.
Korean app-maker Scatter Lab fined for using private data to create homophobic and lewd chatbot
L. Dobberstein · 2021
Closest in time.
Documenting Large Webtext Corpora: A Case Study on the Colossal Clean Crawled Corpus
J. Dodge, M. Sap, A. Marasović, W. Agnew, G. Ilharco, D. Groeneveld, M. Mitchell, and M. Gardner · 2021
Closest in time.
Chinese AI lab challenges Google, OpenAI with a model of 1.75 trillion parameters
C. Du · 2021
Closest in time.
Disentangling polarisation and civic empowerment in the digital age : The role of filter bubbles and echo chambers in the rise of populism
W. H. Dutton and C. T. Robertson · 2021
Closest in time.
Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity
W. Fedus, B. Zoph, and N. Shazeer · 2021
Closest in time.
The Challenge of Value Alignment: from Fairer Algorithms to AI Safety
I. Gabriel and V. Ghazavi · 2021
Closest in time.
What happened to jobs at high risk of automation?
A. Georgieff and A. Milanez · 2021
Closest in time.
GPT-3 Contains Disturbing Bias Against Muslims, January 2021
D. Gershgorn · 2021
Closest in time.
Bias Mitigated Learning from Differentially Private Synthetic Data: A Cautionary Tale
S. Ghalebikesabi, H. Wilde, J. Jewson, A. Doucet, S. Vollmer, and C. Holmes · 2021
Closest in time.
Black Feminist Musings on Algorithmic Oppression
L. M. Hampton · 2021
Closest in time.
Epistemic values in feature importance methods: Lessons from feminist epistemology
L. Hancox-Li and I. E. Kumar · 2021
Closest in time.
How AI Is Learning to Identify Toxic Online Content
L. H. Hanu, J. Thewlis, and S. Haco · 2021
Closest in time.
The Importance of Modeling Social Factors of Language: Theory and Practice
D. Hovy and D. Yang · 2021
Closest in time.
Towards Accountability for Machine Learning Datasets: Practices from Software Engineering and Infrastructure
B. Hutchinson, A. Smart, A. Hanna, E. Denton, C. Greer, O. Kjartansson, P. Barnes, and M. Mitchell · 2021
Closest in time.
How U.S. Companies & Partisans Hack Democracy to Undermine Your Voice
L. James · 2021
Closest in time.
Unintended Bias and Identity Terms, October 2021
Jigsaw · 2021
Closest in time.
Reasons, Values, Stakeholders: A Philosophical Framework for Explainable Artificial Intelligence
A. Kasirzadeh · 2021
Closest in time.
Z. Kenton, T. Everitt, L. Weidinger, I. Gabriel, V. Mikulik, and G. Irving · 2021
Closest in time.
Chatbot Gone Awry Starts Conversations About AI Ethics in South Korea
D. Kim · 2021
Closest in time.
Offensive, aggressive, and hate speech analysis: From data-centric to human-centered approach
J. Kocoń, A. Figas, M. Gruza, D. Puchalska, T. Kajdanowicz, and P. Kazienko · 2021
Closest in time.
Algorithmic bias: review, synthesis, and future research directions
N. Kordzadeh and M. Ghasemaghaei · 2021
Closest in time.
Pitfalls of Static Language Modelling
A. Lazaridou, A. Kuncoro, E. Gribovskaya, D. Agrawal, A. Liska, T. Terzi, M. Gimenez, C. d. M. d’Autume, S. Ruder, D. Yogatama, K. Cao, T. Kocisky, S. Young, and P. Blunsom · 2021
Closest in time.
TeraPipe: Token-Level Pipeline Parallelism for Training Large-Scale Language Models
Z. Li, S. Zhuang, S. Guo, D. Zhuo, H. Zhang, D. Song, and I. Stoica · 2021
Closest in time.
TruthfulQA: Measuring How Models Mimic Human Falsehoods
S. Lin, J. Hilton, and O. Evans · 2021
Closest in time.
Explainable AI: A Review of Machine Learning Interpretability Methods
P. Linardatos, V. Papastefanopoulos, and S. Kotsiantis · 2021
Closest in time.
What’s in the Box? A Preliminary Analysis of Undesirable Content in the Common Crawl Corpus
A. S. Luccioni and J. D. Viviano · 2021
Closest in time.
Gender and Representation Bias in GPT-3 Generated Stories
L. Lucy and D. Bamman · 2021
Closest in time.
Can Conversing with a Computer Increase Turnout? Mobilization Using Chatbot Communication
C. B. Mann · 2021
Closest in time.
On the importance of ethnographic methods in AI research
V. Marda and S. Narayan · 2021
Closest in time.
Understanding Human Impressions of Artificial Intelligence
K. McKee, X. Bai, and S. Fiske · 2021
Closest in time.
A Survey on Bias and Fairness in Machine Learning
N. Mehrabi, F. Morstatter, N. Saxena, K. Lerman, and A. Galstyan · 2021
Closest in time.
DeepMind’s Lila Ibrahim: ‘It’s hard not to go through imposter syndrome’
M. Murgia · 2021
Closest in time.
How to Recognize AI Snake Oil, January 2021
A. Narayanan · 2021
Closest in time.
Synthetic Data for Deep Learning , volume 174 of Springer Optimization and Its Applications
S. I. Nikolenko · 2021
Closest in time.
Workshop on NLP for Positive Impact at ACL-IJCNLP, 2021
NLP for Positive Impact 2021 · 2021
Closest in time.
HONEST: Measuring Hurtful Sentence Completion in Language Models
D. Nozza, F. Bianchi, and D. Hovy · 2021
Closest in time.
Carbon Emissions and Large Neural Network Training
D. Patterson, J. Gonzalez, Q. Le, C. Liang, L.-M. Munguia, D. Rothchild, D. So, M. Texier, and J. Dean · 2021
Closest in time.
Perspective API | Developers, 2021
PerspectiveAPI · 2021
Closest in time.
GPT-3 Powers the Next Generation of Apps, March 2021
A. Pilipiszyn · 2021
Closest in time.
Scaling language models: Methods, analysis & insights from training Gopher
J. Rae, S. Borgeaud, T. Cai, K. Millican, J. Hoffmann, F. Song, J. Aslanides, S. Henderson, R. Ring, S. Young, E. Rutherford, T. Hennigan, J. Menick, A. Cassirer, R. Powell, G. van den Driessche, L. A. Hendricks, M. Rauh, P.-S. Huang, A. Glaese, J. Welbl, S. Dathathri, S. Huang, J. Uesato, J. Mellor, I. Higgins, A. Creswell, N. McAleese, A. Wu, E. Elsen, S. Jayakumar, E. Buchatskaya, D. Budden, E. Sutherland, K. Simonyan, M. Paganini, L. Sifre, L. Martens, X. L. Li, A. Kuncoro, A. Nematzadeh, E. Gribovskaya, D. Donato, A. Lazaridou, A. Mensch, J.-B. Lespiau, M. Tsimpoukelli, N. Grigorev, D. Fritz, T. Sottiaux, M. Pajarskas, T. Pohlen, Z. Gong, D. Toyama, C. de Masson d’Autume, Y. Li, T. Terzi, I. Babuschkin, A. Clark, D. de Las Casas, A. Guy, J. Bradbury, M. Johnson, L. Weidinger, I. Gabriel, W. Isaac, E. Lockhart, S. Osindero, L. Rimell, C. Dyer, O. Vinyals, K. Ayoub, J. Stanway, L. Bennett, D. Hassabis, K. Kavukcuoglu, and G. Irving · 2021
Closest in time.
Generating Fake Cyber Threat Intelligence Using Transformer-Based Models
P. Ranade, A. Piplai, S. Mittal, A. Joshi, and T. Finin · 2021
Closest in time.
Re-imagining Algorithmic Fairness in India and Beyond
N. Sambasivan, E. Arnesen, B. Hutchinson, T. Doshi, and V. Prabhakaran · 2021
Closest in time.
Societal Biases in Language Generation: Progress and Challenges
E. Sheng, K.-W. Chang, P. Natarajan, and N. Peng · 2021
Closest in time.
Process for Adapting Language Models to Society (PALMS) with Values-Targeted Datasets
I. Solaiman and C. Dennison · 2021
Closest in time.
Political Discourse Analysis: A Case Study of Code Mixing and Code Switching in Political Speeches
D. Sravani, L. Kameswari, and R. Mamidi · 2021
Closest in time.
The Stanford Natural Language Processing Group, 2021
StanfordNaturalProcessingGroup · 2021
Closest in time.
ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation
Y. Sun, S. Wang, S. Feng, S. Ding, C. Pang, J. Shang, J. Liu, X. Chen, Y. Zhao, Y. Lu, W. Liu, Z. Wu, W. Gong, J. Liang, Z. Shang, P. Sun, W. Liu, X. Ouyang, D. Yu, H. Tian, H. Wu, and H. Wang · 2021
Closest in time.
Understanding the Capabilities, Limitations, and Societal Impact of Large Language Models
A. Tamkin, M. Brundage, J. Clark, and D. Ganguli · 2021
Closest in time.
Fairness for Unobserved Characteristics: Insights from Technological Impacts on Queer Communities
N. Tomasev, K. R. McKee, J. Kay, and S. Mohamed · 2021
Closest in time.
The Right to Explanation
K. Vredenburgh · 2021
Closest in time.
Should CC-Licensed Content be Used to Train AI? It Depends., March 2021
B. Vézina and S. Hinchcliff Pearson · 2021
Closest in time.
Directional Bias Amplification
A. Wang and O. Russakovsky · 2021
Closest in time.
Towards Zero-Label Language Learning
Z. Wang, A. W. Yu, O. Firat, and Y. Cao · 2021
Closest in time.
Challenges in Detoxifying Language Models
J. Welbl, A. Glaese, J. Uesato, S. Dathathri, J. Mellor, L. A. Hendricks, K. Anderson, P. Kohli, B. Coppin, and P.-S. Huang · 2021
Closest in time.
Language Models are Few-shot Multilingual Learners
G. I. Winata, A. Madotto, Z. Lin, R. Liu, J. Yosinski, and P. Fung · 2021
Closest in time.
Detoxifying Language Models Risks Marginalizing Minority Voices
A. Xu, E. Pathak, E. Wallace, S. Gururangan, M. Sap, and D. Klein · 2021
Closest in time.
Censorship of Online Encyclopedias: Implications for NLP Models
E. Yang and M. E. Roberts · 2021
Closest in time.
K. Yee, U. Tantipongpipat, and S. Mishra · 2021
Closest in time.
A systematic review: The YouTube recommender system and pathways to problematic content
M. Yesilada and S. Lewandowsky · 2021
Closest in time.
Calibrate Before Use: Improving Few-Shot Performance of Language Models
T. Z. Zhao, E. Wallace, S. Feng, D. Klein, and S. Singh · 2021
Closest in time.
Trends in the diffusion of misinformation on social media
H. Allcott, M. Gentzkow, and C. Yu · 2053
Closest in time.
Algorithmic content moderation: Technical and political challenges in the automation of platform governance
R. Gorwa, R. Binns, and C. Katzenbach · 2053
Closest in time.
Data centre water consumption
D. Mytton · 2059
Closest in time.
‘I’d Blush if I Could’: Digital Assistants, Disembodied Cyborgs and the Problem of Gender
H. Bergen · 2069
Closest in time.
The Social Impact of Natural Language Processing
D. Hovy and S. L. Spruit · 2096
Closest in time.