Fetching the paper…
Reading the bibliography…
Bias elimination and recent probing studies attempt to remove specific information from embedding spaces.
Algorithms in Combinatorial Geometry
Herbert Edelsbrunner. 1987 · 1987
Earlier work this paper cites.
Harms of gender exclusivity and challenges in non-binary representation in language technologies
Sunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian, Jeff Phillips, and Kai-Wei Chang. 2021b · 1994
Earlier work this paper cites.
V-measure: A conditional entropy-based external cluster evaluation measure
Andrew Rosenberg and Julia Hirschberg. 2007 · 2007
Earlier work this paper cites.
Scikit-learn: Machine learning in python
Fabian Pedregosa, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, et al. 2011 · 2011
Earlier work this paper cites.
Simlex-999: Evaluating semantic models with (genuine) similarity estimation
Felix Hill, Roi Reichart, and Anna Korhonen. 2015 · 2015
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Earlier work this paper cites.
Equality of opportunity in supervised learning
Moritz Hardt, Eric Price, and Nati Srebro. 2016 · 2016
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Aylin Caliskan, Joanna J Bryson, and Arvind Narayanan. 2017 · 2017
Earlier work this paper cites.
Bag of tricks for efficient text classification
Armand Joulin, Edouard Grave, Piotr Bojanowski, and Tomas Mikolov. 2017 · 2017
Earlier work this paper cites.
Learning adversarially fair and transferable representations
David Madras, Elliot Creager, Toniann Pitassi, and Richard Zemel. 2018 · 2018
Earlier work this paper cites.
Firearms and tigers are dangerous, kitchen knives and zebras are not: Testing whether word embeddings can tell
Pia Sommerauer and Antske Fokkens. 2018 · 2018
Earlier work this paper cites.
Mitigating unwanted biases with adversarial learning
Brian Hu Zhang, Blake Lemoine, and Margaret Mitchell. 2018 · 2018
Earlier work this paper cites.
Gender bias in coreference resolution: Evaluation and debiasing methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2018a · 2018
Cited alongside, same era.
Learning gender-neutral word embeddings
Jieyu Zhao, Yichao Zhou, Zeyu Li, Wei Wang, and Kai-Wei Chang. 2018b · 2018
Cited alongside, same era.
Bias in bios: A case study of semantic representation bias in a high-stakes setting
Maria De-Arteaga, Alexey Romanov, Hanna Wallach, Jennifer Chayes, Christian Borgs, Alexandra Chouldechova, Sahin Geyik, Krishnaram Kenthapadi, and Adam Tauman Kalai. 2019 · 2019
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Understanding undesirable word embedding associations
Kawin Ethayarajh, David Duvenaud, and Graeme Hirst. 2019 · 2019
Cited alongside, same era.
Amnesic probing: Behavioral explanation with amnesic counterfactuals
Yanai Elazar, Shauli Ravfogel, Alon Jacovi, and Yoav Goldberg. 2021 · 2021
Later among the works it cites.
Obstructing classification via projection
Pantea Haghighatkhah, Wouter Meulemans, Bettina Speckmann, Jérôme Urhausen, and Kevin Verbeek. 2021 · 2021
Later among the works it cites.
The rediscovery hypothesis: Language models need to meet linguistics
Vassilina Nikoulina, Maxat Tezekbayev, Nuradil Kozhakhmet, Madina Babazhanova, Matthias Gallé, and Zhenisbek Assylbekov. 2021 · 2021
Later among the works it cites.
How conservative are language models? adapting to the introduction of gender-neutral pronouns
Stephanie Brandl, Ruixiang Cui, and Anders Søgaard. 2022 · 2022
Closest in time.
Can transformer be too compositional? analysing idiom processing in neural machine translation
Verna Dankers, Christopher Lucas, and Ivan Titov. 2022 · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lipstick on a pig: Debiasing methods cover up systematic gender biases in word embeddings but do not remove them
Hila Gonen and Yoav Goldberg. 2019 · 2019
Cited alongside, same era.
What’s in a name? Reducing bias in bios without access to protected attributes
Alexey Romanov, Maria De-Arteaga, Hanna Wallach, Jennifer Chayes, Christian Borgs, Alexandra Chouldechova, Sahin Geyik, Krishnaram Kenthapadi, Anna Rumshisky, and Adam Kalai. 2019 · 2019
Cited alongside, same era.
Controlling the imprint of passivization and negation in contextualized representations
Hande Celikkanat, Sami Virpioja, Jörg Tiedemann, and Marianna Apidianaki. 2020 · 2020
Cited alongside, same era.
Null it out: Guarding protected attributes by iterative nullspace projection
Shauli Ravfogel, Yanai Elazar, Hila Gonen, Michael Twiton, and Yoav Goldberg. 2020 · 2020
Cited alongside, same era.
Exploring the linear subspace hypothesis in gender bias mitigation
Francisco Vargas and Ryan Cotterell. 2020 · 2020
Cited alongside, same era.
Geometric probing of word vectors
Madina Babazhanova, Maxat Tezekbayev, and Zhenisbek Assylbekov. 2021 · 2021
Cited alongside, same era.
OSCaR: Orthogonal subspace correction and rectification of biases in word embeddings
Sunipa Dev, Tao Li, Jeff M Phillips, and Vivek Srikumar. 2021a · 2021
Cited alongside, same era.
Hila Gonen, Shauli Ravfogel, and Yoav Goldberg. 2022 · 2022
Closest in time.
Obstructing classification via projection
Pantea Haghighatkhah, Wouter Meulemans, Bettina Speckmann, Jérôme Urhausen, and Kevin Verbeek. 2022 · 2022
Closest in time.
Unit testing for concepts in neural networks
Charles Lovering and Ellie Pavlick. 2022 · 2022
Closest in time.
Linear adversarial concept erasure
Shauli Ravfogel, Michael Twiton, Yoav Goldberg, and Ryan D Cotterell. 2022 · 2022
Closest in time.
Gold doesn’t always glitter: Spectral removal of linear and nonlinear guarded attribute information
Shun Shao, Yftah Ziser, and Shay B. Cohen. 2022 · 2022
Closest in time.
Probing word syntactic representations in the brain by a feature elimination method
Xiaohan Zhang, Shaonan Wang, Nan Lin, Jiajun Zhang, and Chengqing Zong. 2022 · 2022
Closest in time.