Mitigating unwanted biases with adversarial learning
Brian Hu Zhang, Blake Lemoine, and Margaret Mitchell. 2018 · 2018
Earlier work this paper cites.
Stereotypical bias removal for hate speech detection task using knowledge-based generalizations
Pinkesh Badjatiya, Manish Gupta, and Vasudeva Varma. 2019 · 2019
Earlier work this paper cites.
Bias in bios: A case study of semantic representation bias in a high-stakes setting
Maria De-Arteaga, Alexey Romanov, Hanna Wallach, Jennifer Chayes, Christian Borgs, Alexandra Chouldechova, Sahin Geyik, Krishnaram Kenthapadi, and Adam Tauman Kalai. 2019 · 2019
Earlier work this paper cites.
Lipstick on a pig: Debiasing methods cover up systematic gender biases in word embeddings but do not remove them
Hila Gonen and Yoav Goldberg. 2019 · 2019
Earlier work this paper cites.
Debiasing vandalism detection models at wikidata
Stefan Heindorf, Yan Scholten, Gregor Engels, and Martin Potthast. 2019 · 2019
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like moral choices
Sophie Jentzsch, Patrick Schramowski, Constantin Rothkopf, and Kristian Kersting. 2019 · 2019
Earlier work this paper cites.
Eraser: A benchmark to evaluate rationalized nlp models
Jay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric Lehman, Caiming Xiong, Richard Socher, and Byron C Wallace. 2020 · 2020
Earlier work this paper cites.