Fetching the paper…
Reading the bibliography…
As the capabilities of generative language models continue to advance, the implications of biases ingrained within these models have garnered increasing attention from researchers, practitioners, and the broader public.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al., Language models are few-shot learners, Advances in neural information processing systems 33 (2020) 1877–1901
1901
Earlier work this paper cites.
C. Geertz, The interpretation of cultures, Vol. 5043, Basic books, 1973
1973
Earlier work this paper cites.
G. Hofstede, Culture’s consequences: International differences in work-related values, Vol. 5, sage, 1984
1984
Earlier work this paper cites.
P. Bourdieu, Language and symbolic power, Harvard University Press, 1991
1991
Earlier work this paper cites.
Y. Bengio, R. Ducharme, P. Vincent, A neural probabilistic language model, Advances in neural information processing systems 13 (2000)
2000
Earlier work this paper cites.
N. Fairclough, Language and power, Pearson Education, 2001
2001
Earlier work this paper cites.
S. S. Mufwene, The Ecology of Language Evolution. Cambridge Approaches to Language Contact., ERIC, 2001
2001
Earlier work this paper cites.
R. Inglehart, Christian Welzel Modernization, Cultural Change, and Democracy The Human Development Sequence, Cambridge: Cambridge university press, 2005
2005
Earlier work this paper cites.
G. Lakoff, M. Johnson, Metaphors we live by, University of Chicago press, 2008
2008
Earlier work this paper cites.
H. Jenkins, M. Deuze, Convergence culture (2008)
2008
Earlier work this paper cites.
W. Wallach, C. Allen, I. Smit, Machine morality: bottom-up and top-down approaches for modelling human moral faculties, AI & Society 22 (2008) 565–582
2008
Earlier work this paper cites.
H. Cramer, V. Evers, S. Ramlal, M. Van Someren, L. Rutledge, N. Stash, L. Aroyo, B. Wielinga, The effects of transparency on trust in and acceptance of a content-based art recommender, User Modeling and User-adapted interaction 18 (2008) 455–496
2008
Earlier work this paper cites.
J. H. Hill, The everyday language of white racism, John Wiley & Sons, 2009
2009
Earlier work this paper cites.
R. Munro, S. Bethard, V. Kuperman, V. T. Lai, R. Melnick, C. Potts, T. Schnoebelen, H. Tily, Crowdsourcing and language studies: the new generation of linguistic data, in: NAACL Workshop on Creating Speech and Language Data With Amazon’s Mechanical Turk, Association for Computational Linguistics, 2010, pp. 122–130
2010
Earlier work this paper cites.
M. Castells, The rise of the network society, John wiley & sons, 2011
2011
Earlier work this paper cites.
A. Torralba, A. A. Efros, Unbiased look at dataset bias, in: CVPR 2011, IEEE, 2011, pp. 1521–1528
2011
Earlier work this paper cites.
B. L. Whorf, Language, thought, and reality: Selected writings, MIT press, 2012
2012
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, J. Dean, Distributed representations of words and phrases and their compositionality, Advances in neural information processing systems 26 (2013)
2013
Earlier work this paper cites.
M. Foucault, Archaeology of knowledge, Routledge, 2013
2013
Earlier work this paper cites.
I. Sutskever, O. Vinyals, Q. V. Le, Sequence to sequence learning with neural networks, Advances in neural information processing systems 27 (2014)
2014
Earlier work this paper cites.
D. K. Citron, F. Pasquale, The scored society: Due process for automated predictions, Wash. L. Rev. 89 (2014) 1
2014
Earlier work this paper cites.
L. Taylor, R. Schroeder, Is bigger better? the emergence of big data as a tool for international development policy, GeoJournal 80 (2015) 503–518
2015
Earlier work this paper cites.
M. K. Lee, D. Kusbit, E. Metsky, L. Dabbish, Working with machines: The impact of algorithmic and data-driven management on human workers, in: Proceedings of the 33rd annual ACM conference on human factors in computing systems, 2015, pp. 1603–1612
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
T. Bolukbasi, K.-W. Chang, J. Y. Zou, V. Saligrama, A. T. Kalai, Man is to computer programmer as woman is to homemaker? debiasing word embeddings, Advances in neural information processing systems 29 (2016)
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
D. Hovy, S. L. Spruit, The social impact of natural language processing, in: Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), 2016, pp. 591–598
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Etzioni, O. Etzioni, Keeping AI legal, Vand. J. Ent. & Tech. L. 19 (2016) 133
2016
Earlier work this paper cites.
B. D. Mittelstadt, P. Allo, M. Taddeo, S. Wachter, L. Floridi, The ethics of algorithms: Mapping the debate, Big Data & Society 3 (2) (2016) 2053951716679679
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, I. Polosukhin, Attention is all you need, Advances in neural information processing systems 30 (2017)
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
H. Chen, X. Liu, D. Yin, J. Tang, A survey on dialogue systems: Recent advances and new frontiers, Acm SIGKDD Explorations Newsletter 19 (2) (2017) 25–35
2017
Earlier work this paper cites.
A. Caliskan, J. J. Bryson, A. Narayanan, Semantics derived automatically from language corpora contain human-like biases, Science 356 (6334) (2017) 183–186
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. Johnson, M. Schuster, Q. V. Le, M. Krikun, Y. Wu, Z. Chen, N. Thorat, F. Viégas, M. Wattenberg, G. Corrado, et al., Google’s multilingual neural machine translation system: Enabling zero-shot translation, Transactions of the Association for Computational Linguistics 5 (2017) 339–351
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. B. Zafar, I. Valera, M. Gomez Rodriguez, K. P. Gummadi, Fairness beyond disparate treatment & disparate impact: Learning classification without disparate mistreatment, in: Proceedings of the 26th international conference on world wide web, 2017, pp. 1171–1180
2017
Earlier work this paper cites.
S. Wachter, B. Mittelstadt, L. Floridi, Transparent, explainable, and accountable AI for robotics, Science robotics 2 (6) (2017) eaan6080
2017
Earlier work this paper cites.
C. O’neil, Weapons of math destruction: How big data increases inequality and threatens democracy, Crown, 2017
2017
Earlier work this paper cites.
B. Goodman, S. Flaxman, European union regulations on algorithmic decision-making and a “right to explanation”, AI magazine 38 (3) (2017) 50–57
2017
Earlier work this paper cites.
K. Shahriari, M. Shahriari, Ieee standard review—ethically aligned design: A vision for prioritizing human wellbeing with artificial intelligence and autonomous systems, in: 2017 IEEE Canada International Humanitarian Technology Conference (IHTC), IEEE, 2017, pp. 197–201
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Radford, K. Narasimhan, T. Salimans, I. Sutskever, et al., Improving language understanding by generative pre-training (2018)
2018
Earlier work this paper cites.
T. Young, D. Hazarika, S. Poria, E. Cambria, Recent trends in deep learning based natural language processing, IEEE Computational intelligenCe magazine 13 (3) (2018) 55–75
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Karakanta, J. Dehdari, J. van Genabith, Neural machine translation for low-resource languages without parallel corpora, Machine Translation 32 (2018) 167–189
2018
Earlier work this paper cites.
C. Christianson, J. Duncan, B. Onyshkevych, Overview of the DARPA LORELEI program, Machine Translation 32 (2018) 3–9
2018
Earlier work this paper cites.
J. Buolamwini, T. Gebru, Gender shades: Intersectional accuracy disparities in commercial gender classification, in: Conference on fairness, accountability and transparency, PMLR, 2018, pp. 77–91
2018
Earlier work this paper cites.
E. M. Bender, B. Friedman, Data statements for natural language processing: Toward mitigating system bias and enabling better science, Transactions of the Association for Computational Linguistics 6 (2018) 587–604
2018
Earlier work this paper cites.
R. Binns, Fairness in machine learning: Lessons from political philosophy, in: Conference on fairness, accountability and transparency, PMLR, 2018, pp. 149–159
2018
Earlier work this paper cites.
N. Garg, L. Schiebinger, D. Jurafsky, J. Zou, Word embeddings quantify 100 years of gender and ethnic stereotypes, Proceedings of the National Academy of Sciences 115 (16) (2018) E3635–E3644
2018
Earlier work this paper cites.
L. Dixon, J. Li, J. Sorensen, N. Thain, L. Vasserman, Measuring and mitigating unintended bias in text classification, in: Proceedings of the 2018 AAAI/ACM Conference on AI, Ethics, and Society, 2018, pp. 67–73
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
H. C. Triandis, Individualism and collectivism, Routledge, 2018
2018
Cited alongside, same era.
M. Bogen, A. Rieke, Help wanted: An examination of hiring algorithms, equity, and bias (2018)
2018
Cited alongside, same era.
T. Gillespie, Custodians of the Internet: Platforms, content moderation, and the hidden decisions that shape social media, Yale University Press, 2018
2018
Cited alongside, same era.
M. A. Gianfrancesco, S. Tamang, J. Yazdany, G. Schmajuk, Potential biases in machine learning algorithms using electronic health record data, JAMA internal medicine 178 (11) (2018) 1544–1547
2018
Cited alongside, same era.
M. C. Elish, D. Boyd, Situating methods in the magic of Big Data and AI, Communication monographs 85 (1) (2018) 57–80
2018
Cited alongside, same era.
N. A. Smith, Contextual word representations: putting words into computers, Communications of the ACM 63 (6) (2020) 66–74
2020
Later among the works it cites.
Z. Jiang, F. F. Xu, J. Araki, G. Neubig, How can we know what language models know?, Transactions of the Association for Computational Linguistics 8 (2020) 423–438
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Guidotti, A. Monreale, S. Ruggieri, F. Turini, F. Giannotti, D. Pedreschi, A survey of methods for explaining black box models, ACM computing surveys (CSUR) 51 (5) (2018) 1–42
2018
Cited alongside, same era.
J. Heer, The partnership on AI, AI Matters 4 (3) (2018) 25–26
2018
Cited alongside, same era.
S. Pichai, Ai at Google: our principles, The Keyword 7 (2018) 1–3
2018
Cited alongside, same era.
D. Reisman, J. Schultz, K. Crawford, M. Whittaker, Algorithmic impact assessments: A practical framework for public agency, AI Now (2018)
2018
Cited alongside, same era.
C. Cath, S. Wachter, B. Mittelstadt, M. Taddeo, L. Floridi, Artificial intelligence and the ‘good society’: the US, EU, and UK approach, Science and engineering ethics 24 (2018) 505–528
2018
Cited alongside, same era.
B. H. Zhang, B. Lemoine, M. Mitchell, Mitigating unwanted biases with adversarial learning, in: Proceedings of the 2018 AAAI/ACM Conference on AI, Ethics, and Society, 2018, pp. 335–340
2018
Cited alongside, same era.
V. Eubanks, Automating inequality: How high-tech tools profile, police, and punish the poor, St. Martin’s Press, 2018
2018
Cited alongside, same era.
2020
Later among the works it cites.
S. Larsson, F. Heintz, Transparency in artificial intelligence, Internet Policy Review 9 (2) (2020)
2020
Later among the works it cites.
I. D. Raji, A. Smart, R. N. White, M. Mitchell, T. Gebru, B. Hutchinson, J. Smith-Loud, D. Theron, P. Barnes, Closing the AI accountability gap: Defining an end-to-end framework for internal algorithmic auditing, in: Proceedings of the 2020 conference on fairness, accountability, and transparency, 2020, pp. 33–44
2020
Later among the works it cites.
M. R. Morris, Ai and accessibility, Communications of the ACM 63 (6) (2020) 35–37
2020
Later among the works it cites.
R. Schwartz, J. Dodge, N. A. Smith, O. Etzioni, Green AI, Communications of the ACM 63 (12) (2020) 54–63
2020
Later among the works it cites.
M. Raghavan, S. Barocas, J. Kleinberg, K. Levy, Mitigating bias in algorithmic hiring: Evaluating claims and practices, in: Proceedings of the 2020 conference on fairness, accountability, and transparency, 2020, pp. 469–481
2020
Later among the works it cites.
I. V. Pasquetto, B. Swire-Thompson, M. A. Amazeen, F. Benevenuto, N. M. Brashier, R. M. Bond, L. C. Bozarth, C. Budak, U. K. Ecker, L. K. Fazio, et al., Tackling misinformation: What researchers could do with social media data, The Harvard Kennedy School Misinformation Review (2020)
2020
Later among the works it cites.
J. K. Paulus, D. M. Kent, Predictably unequal: understanding and addressing concerns that algorithmic clinical prediction may increase health disparities, NPJ digital medicine 3 (1) (2020) 99
2020
Later among the works it cites.
K. Yeung, Recommendation of the council on artificial intelligence (OECD), International legal materials 59 (1) (2020) 27–34
2020
Later among the works it cites.
S. Yan, H.-t. Kao, E. Ferrara, Fair class balancing: Enhancing model fairness without observing sensitive attributes, in: Proceedings of the 29th ACM International Conference on Information & Knowledge Management, 2020, pp. 1715–1724
2020
Later among the works it cites.
S. Costanza-Chock, Design justice: Community-led practices to build the worlds we need, The MIT Press, 2020
2020
Later among the works it cites.
J. Stoyanovich, B. Howe, H. Jagadish, Responsible data management, Proceedings of the VLDB Endowment 13 (12) (2020)
2020
Later among the works it cites.
E. M. Bender, T. Gebru, A. McMillan-Major, S. Shmitchell, On the dangers of stochastic parrots: Can language models be too big?, in: Proceedings of the 2021 ACM conference on fairness, accountability, and transparency, 2021, pp. 610–623
2021
Later among the works it cites.
D. Hovy, S. Prabhumoye, Five sources of bias in natural language processing, Language and Linguistics Compass 15 (8) (2021) e12432
2021
Later among the works it cites.
H. R. Kirk, Y. Jun, F. Volpin, H. Iqbal, E. Benussi, F. Dreyer, A. Shtedritski, Y. Asano, Bias out-of-the-box: An empirical analysis of intersectional occupational biases in popular generative language models, Advances in neural information processing systems 34 (2021) 2611–2624
2021
Later among the works it cites.
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. B. Brown, D. Song, U. Erlingsson, et al., Extracting training data from large language models., in: USENIX Security Symposium, Vol. 6, 2021
2021
Later among the works it cites.
T. Gebru, J. Morgenstern, B. Vecchione, J. W. Vaughan, H. Wallach, H. D. Iii, K. Crawford, Datasheets for datasets, Communications of the ACM 64 (12) (2021) 86–92
2021
Later among the works it cites.
U. Ehsan, Q. V. Liao, M. Muller, M. O. Riedl, J. D. Weisz, Expanding explainability: Towards social transparency in AI systems, in: Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems, 2021, pp. 1–19
2021
Later among the works it cites.
H. Smith, Clinical AI: opacity, accountability, responsibility and liability, AI & SOCIETY 36 (2) (2021) 535–545
2021
Later among the works it cites.
2021
Later among the works it cites.
N. Mehrabi, F. Morstatter, N. Saxena, K. Lerman, A. Galstyan, A survey on bias and fairness in machine learning, ACM Computing Surveys (CSUR) 54 (6) (2021) 1–35
2021
Later among the works it cites.
Y. K. Dwivedi, L. Hughes, E. Ismagilova, G. Aarts, C. Coombs, T. Crick, Y. Duan, R. Dwivedi, J. Edwards, A. Eirug, et al., Artificial intelligence (ai): Multidisciplinary perspectives on emerging challenges, opportunities, and agenda for research, practice and policy, International Journal of Information Management 57 (2021) 101994
2021
Later among the works it cites.
M. Felderer, R. Ramler, Quality assurance for ai-based systems: Overview and challenges (introduction to interactive session), in: Software Quality: Future Perspectives on Software Engineering Quality: 13th International Conference, SWQD 2021, Vienna, Austria, January 19–21, 2021, Proceedings 13, Springer, 2021, pp. 33–42
2021
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
K.-C. Yang, E. Ferrara, F. Menczer, Botometer 101: Social bot practicum for computational social scientists, Journal of Computational Social Science 5 (2) (2022) 1511–1528
2022
Later among the works it cites.
2022
Later among the works it cites.
E. Ferrara, Twitter spam and false accounts prevalence, detection, and characterization: A survey, First Monday 27 (12) (2022)
2022
Later among the works it cites.
S. Biderman, K. Bicheno, L. Gao, Datasheet for the pile, arXiv preprint arXiv:2201.07311 (2022)
2022
Later among the works it cites.
2022
Later among the works it cites.
T. Dettmers, M. Lewis, Y. Belkada, L. Zettlemoyer, LLM.int8 (): 8-bit matrix multiplication for transformers at scale, Advances in Neural Information Processing Systems 35 (2022) 30318–30332
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
OpenAI, Gpt-4 technical report (2023)
2023
Closest in time.
S. Ranathunga, E.-S. A. Lee, M. Prifti Skenduli, R. Shekhar, M. Alam, R. Kaur, Neural machine translation for low-resource languages: A survey, ACM Computing Surveys 55 (11) (2023) 1–37
2023
Closest in time.
E. Ferrara, Social bot detection in the age of ChatGPT: Challenges and opportunities, First Monday 28 (6) (2023)
2023
Closest in time.
A. Gilson, C. W. Safranek, T. Huang, V. Socrates, L. Chi, R. A. Taylor, D. Chartash, et al., How does ChatGPT perform on the united states medical licensing examination? the implications of large language models for medical education and knowledge assessment, JMIR Medical Education 9 (1) (2023) e45312
2023
Closest in time.
J. H. Choi, K. E. Hickman, A. Monahan, D. Schwarcz, ChatGPT goes to law school, Available at SSRN (2023)
2023
Closest in time.
2023
Closest in time.
OpenAI, GPT-4 System Card (2023)
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
R. W. McGee, Is ChatGPT biased against conservatives? an empirical study, An Empirical Study (February 15, 2023) (2023)
2023
Closest in time.
2023
Closest in time.
V. Kumar, H. Koorehdavoudi, M. Moshtaghi, A. Misra, A. Chadha, E. Ferrara, Controlled text generation with hidden representation transformations, in: Findings of the Association for Computational Linguistics: ACL 2023, Association for Computational Linguistics, Toronto, Canada, 2023, pp. 9440–9455
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
F. Ezzeddine, O. Ayoub, S. Giordano, G. Nogara, I. Sbeity, E. Ferrara, L. Luceri, Exposing influence campaigns in the age of llms: a behavioral-based ai approach to detecting state-sponsored trolls, EPJ Data Science 12 (1) (2023) 46
2023
Closest in time.
2023
Closest in time.
Y. H. Ezzeldin, S. Yan, C. He, E. Ferrara, S. Avestimehr, Fairfed: Enabling group fairness in federated learning, in: Proceedings of the 37th AAAI Conference on Artificial Intelligence, 2023
2023
Closest in time.