Fetching the paper…
Reading the bibliography…
We describe a simple and effective method (Spectral Attribute removaL; SAL) to remove private or guarded information from neural representations.
Hila Gonen and Yoav Goldberg. 2019 · 1903
Earlier work this paper cites.
Thomas Manzini, Yao Chong Lim, Yulia Tsvetkov, and Alan W Black. 2019 · 1904
Earlier work this paper cites.
Understanding undesirable word embedding associations
Kawin Ethayarajh, David Duvenaud, and Graeme Hirst. 2019 · 1908
Earlier work this paper cites.
Placing search in context: The concept revisited
Lev Finkelstein, Evgeniy Gabrilovich, Yossi Matias, Ehud Rivlin, Zach Solan, Gadi Wolfman, and Eytan Ruppin. 2001 · 2001
Earlier work this paper cites.
Latent semantic analysis
Susan T Dumais. 2004 · 2004
Earlier work this paper cites.
Null it out: Guarding protected attributes by iterative nullspace projection
Shauli Ravfogel, Yanai Elazar, Hila Gonen, Michael Twiton, and Yoav Goldberg. 2020 · 2004
Earlier work this paper cites.
Large-scale learning of word relatedness with constraints
Guy Halawi, Gideon Dror, Evgeniy Gabrilovich, and Yehuda Koren. 2012 · 2012
Earlier work this paper cites.
Censoring representations with an adversary
Harrison Edwards and Amos Storkey. 2015 · 2015
Earlier work this paper cites.
Simlex-999: Evaluating semantic models with (genuine) similarity estimation
Felix Hill, Roi Reichart, and Anna Korhonen. 2015 · 2015
Earlier work this paper cites.
Demographic dialectal variation in social media: A case study of african-american english
Su Lin Blodgett, Lisa Green, and Brendan O’Connor. 2016 · 2016
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Cited alongside, same era.
Equality of opportunity in supervised learning
Moritz Hardt, Eric Price, and Nati Srebro. 2016 · 2016
Cited alongside, same era.
Fasttext. zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski, Matthijs Douze, Hérve Jégou, and Tomas Mikolov. 2016 · 2016
Cited alongside, same era.
Bjarke Felbo, Alan Mislove, Anders Søgaard, Iyad Rahwan, and Sune Lehmann. 2017 · 2017
Cited alongside, same era.
Cleaning the null space: A privacy mechanism for predictors
Ke Xu, Tongyi Cao, Swair Shah, Crystal Maung, and Haim Schweitzer. 2017 · 2017
Cited alongside, same era.
Learning gender-neutral word embeddings
Jieyu Zhao, Yichao Zhou, Zeyu Li, Wei Wang, and Kai-Wei Chang. 2018 · 2018
Later among the works it cites.
Adversarial removal of demographic attributes revisited
Maria Barrett, Yova Kementchedjhieva, Yanai Elazar, Desmond Elliott, and Anders Søgaard. 2019 · 2019
Later among the works it cites.
Bias in bios: A case study of semantic representation bias in a high-stakes setting
Maria De-Arteaga, Alexey Romanov, Hanna Wallach, Jennifer Chayes, Christian Borgs, Alexandra Chouldechova, Sahin Geyik, Krishnaram Kenthapadi, and Adam Tauman Kalai. 2019 · 2019
Later among the works it cites.
Diverse adversaries for mitigating bias in training
Xudong Han, Timothy Baldwin, and Trevor Cohn. 2021 · 2021
Later among the works it cites.
Learning disentangled textual representations via statistical measures of similarity
Pierre Colombo, Guillaume Staerman, Nathan Noiry, and Pablo Piantanida. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Privacy-preserving neural representations of text
Maximin Coavoux, Shashi Narayan, and Shay B Cohen. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Adversarial removal of demographic attributes from text data
Yanai Elazar and Yoav Goldberg. 2018 · 2018
Cited alongside, same era.
Towards robust and privacy-preserving text representations
Yitong Li, Timothy Baldwin, and Trevor Cohn. 2018 · 2018
Cited alongside, same era.
Native language cognate effects on second language lexical choice
Ella Rabinovich, Yulia Tsvetkov, and Shuly Wintner. 2018 · 2018
Cited alongside, same era.
Closest in time.
fairlib: A unified framework for assessing and improving classification fairness
Xudong Han, Aili Shen, Yitong Li, Lea Frermann, Timothy Baldwin, and Trevor Cohn. 2022 · 2022
Closest in time.
Linear adversarial concept erasure
Shauli Ravfogel, Michael Twiton, Yoav Goldberg, and Ryan D Cotterell. 2022 · 2022
Closest in time.
Erasure of unaligned attributes from neural representations
Shun Shao, Yftah Ziser, and Shay B. Cohen. 2023 · 2023
Closest in time.
Domain-adversarial training of neural networks
Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor Lempitsky. 2016 · 2030
Closest in time.