Fetching the paper…
Reading the bibliography…
Machine learning (ML) algorithms can often exhibit discriminatory behavior, negatively affecting certain populations across protected groups.
Generalized procrustes analysis
John C Gower · 1975
Earlier work this paper cites.
Second order properties of error surfaces: Learning time and generalization
Yann LeCun, Ido Kanter, and Sara Solla · 1990
Earlier work this paper cites.
Content and cluster analysis: assessing representational similarity in neural systems
Aarre Laakso and Garrison Cottrell · 2000
Earlier work this paper cites.
Representational similarity analysis-connecting the branches of systems neuroscience
Nikolaus Kriegeskorte, Marieke Mur, and Peter A Bandettini · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al · 2009
Earlier work this paper cites.
Fairness through awareness
Cynthia Dwork, Moritz Hardt, Toniann Pitassi, Omer Reingold, and Richard Zemel · 2012
Earlier work this paper cites.
Data preprocessing techniques for classification without discrimination
Faisal Kamiran and Toon Calders · 2012
Earlier work this paper cites.
Representational geometry: integrating cognition, computation, and the brain
Nikolaus Kriegeskorte and Rogier A Kievit · 2013
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2013
Earlier work this paper cites.
Learning fair representations
Rich Zemel, Yu Wu, Kevin Swersky, Toni Pitassi, and Cynthia Dwork · 2013
Earlier work this paper cites.
Deep learning for healthcare decision making with emrs
Znaonui Liang, Gang Zhang, Jimmy Xiangji Huang, and Qmming Vivian Hu · 2014
Earlier work this paper cites.
Certifying and removing disparate impact
Michael Feldman, Sorelle A Friedler, John Moeller, Carlos Scheidegger, and Suresh Venkatasubramanian · 2015
Earlier work this paper cites.
Convergent learning: Do different neural networks learn the same representations?
Yixuan Li, Jason Yosinski, Jeff Clune, Hod Lipson, and John Hopcroft · 2015
Earlier work this paper cites.
The variational fair autoencoder
Christos Louizos, Kevin Swersky, Yujia Li, Max Welling, and Richard Zemel · 2015
Earlier work this paper cites.
Predicting judicial decisions of the european court of human rights: A natural language processing perspective
Nikolaos Aletras, Dimitrios Tsarapatsanis, Daniel Preoţiuc-Pietro, and Vasileios Lampos · 2016
Earlier work this paper cites.
Machine bias
Julia Angwin, Jeff Larson, Surya Mattu, and Lauren Kirchner · 2016
Earlier work this paper cites.
Commentaries on the Laws of England
William Blackstone · 2016
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai · 2016
Earlier work this paper cites.
Censoring representations with an adversary
Harrison Edwards and Amos J. Storkey · 2016
Earlier work this paper cites.
Cultural shift or linguistic drift? comparing two computational measures of semantic change
William L Hamilton, Jure Leskovec, and Dan Jurafsky · 2016
Earlier work this paper cites.
Equality of opportunity in supervised learning
Moritz Hardt, Eric Price, Eric Price, and Nati Srebro · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Fairness in learning: Classic and contextual bandits
Matthew Joseph, Michael Kearns, Jamie H Morgenstern, and Aaron Roth · 2016
Earlier work this paper cites.
Data and analysis for ’How we analyzed the COMPAS recidivism algorithm’
Jeff Larson, Surya Mattu, Lauren Kirchner, and Julia Angwin · 2016
Earlier work this paper cites.
The limitations of deep learning in adversarial settings
Nicolas Papernot, Patrick McDaniel, Somesh Jha, Matt Fredrikson, Z Berkay Celik, and Ananthram Swami · 2016
Earlier work this paper cites.
Data decisions and theoretical implications when adversarially learning fair representations
Alex Beutel, Jilin Chen, Zhe Zhao, and Ed H Chi · 2017
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Aylin Caliskan, Joanna J Bryson, and Arvind Narayanan · 2017
Earlier work this paper cites.
Dynamics of scene representations in the human brain revealed by magnetoencephalography and deep neural networks
Radoslaw Martin Cichy, Aditya Khosla, Dimitrios Pantazis, and Aude Oliva · 2017
Earlier work this paper cites.
Conscientious classification: A data scientist’s guide to discrimination-aware classification
Brian d’Alessandro, Cathy O’Neil, and Tom LaGatta · 2017
Earlier work this paper cites.
Counterfactual fairness
Matt J Kusner, Joshua Loftus, Chris Russell, and Ricardo Silva · 2017
Earlier work this paper cites.
Weapons of math destruction: How big data increases inequality and threatens democracy
Cathy O’neil · 2017
Earlier work this paper cites.
Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability
Maithra Raghu, Justin Gilmer, Jason Yosinski, and Jascha Sohl-Dickstein · 2017
Earlier work this paper cites.
All you need is beyond a good init: Exploring better solution for training extremely deep convolutional neural networks with orthonormality and modulation
Di Xie, Jiang Xiong, and Shiliang Pu · 2017
Cited alongside, same era.
A reductions approach to fair classification
Alekh Agarwal, Alina Beygelzimer, Miroslav Dudik, John Langford, and Hanna Wallach · 2018
Cited alongside, same era.
Gender shades: Intersectional accuracy disparities in commercial gender classification
Joy Buolamwini and Timnit Gebru · 2018
Cited alongside, same era.
Empirical risk minimization under fairness constraints
Michele Donini, Luca Oneto, Shai Ben-David, John Shawe-Taylor, and Massimiliano Pontil · 2018
Cited alongside, same era.
Word embeddings quantify 100 years of gender and ethnic stereotypes
Nikhil Garg, Londa Schiebinger, Dan Jurafsky, and James Zou · 2018
Cited alongside, same era.
Gender bias in contextualized word embeddings
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Ryan Cotterell, Vicente Ordonez, and Kai-Wei Chang · 2019
Later among the works it cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Later among the works it cites.
Convergent algorithms for (relaxed) minimax fairness
Emily Diana, Wesley Gill, Michael Kearns, Krishnaram Kenthapadi, and Aaron Roth · 2020
Later among the works it cites.
Fact: A diagnostic for group fairness trade-offs
Joon Sik Kim, Jiahao Chen, and Ameet Talwalkar · 2020
Later among the works it cites.
Fairness without demographics through adversarially reweighted learning
Preethi Lahoti, Alex Beutel, Jilin Chen, Kang Lee, Flavien Prost, Nithum Thain, Xuezhi Wang, and Ed H. Chi · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Robert Geirhos, Carlos R. M. Temme, Jonas Rauber, Heiko H. Schütt, Matthias Bethge, and Felix A. Wichmann · 2018
Cited alongside, same era.
Non-discriminatory machine learning through convex fairness criteria
Naman Goel, Mohammad Yaghini, and Boi Faltings · 2018
Cited alongside, same era.
Preventing disparate treatment in sequential decision making
Hoda Heidari and Andreas Krause · 2018
Cited alongside, same era.
Preventing fairness gerrymandering: Auditing and learning for subgroup fairness
Michael Kearns, Seth Neel, Aaron Roth, and Zhiwei Steven Wu · 2018
Cited alongside, same era.
Learning adversarially fair and transferable representations
David Madras, Elliot Creager, Toniann Pitassi, and Richard S. Zemel · 2018
Cited alongside, same era.
Insights on representational similarity in neural networks with canonical correlation
Ari Morcos, Maithra Raghu, and Samy Bengio · 2018
Cited alongside, same era.
Inclusivefacenet: Improving face attribute detection with race and gender diversity
Hee Jung Ryu, Hartwig Adam, and Margaret Mitchell · 2018
Cited alongside, same era.
Later among the works it cites.
Gender bias in neural natural language processing
Kaiji Lu, Piotr Mardziel, Fangjing Wu, Preetam Amancharla, and Anupam Datta · 2020
Later among the works it cites.
Minimax pareto fairness: A multi objective perspective
Natalia Martinez, Martin Bertran, and Guillermo Sapiro · 2020
Later among the works it cites.
Fnnc: Achieving fairness through neural networks
Manisha Padala and Sujit Gujar · 2020
Later among the works it cites.
Anatomy of catastrophic forgetting: Hidden representations and task semantics
Vinay V Ramasesh, Ethan Dyer, and Maithra Raghu · 2020
Later among the works it cites.
Post-hoc methods for debiasing neural networks
Yash Savani, Colin White, and Naveen Sundar Govindarajulu · 2020
Later among the works it cites.
An approach for prediction of loan approval using machine learning algorithm
Mohammad Ahmad Sheikh, Amit Kumar Goel, and Tapas Kumar · 2020
Later among the works it cites.
Measuring robustness to natural distribution shifts in image classification
Rohan Taori, Achal Dave, Vaishaal Shankar, Nicholas Carlini, Benjamin Recht, and Ludwig Schmidt · 2020
Later among the works it cites.
Mitigating bias in face recognition using skewness-aware reinforcement learning
Mei Wang and Weihong Deng · 2020
Later among the works it cites.
Towards fairness in visual recognition: Effective strategies for bias mitigation, 2020
Zeyu Wang, Klint Qinami, Ioannis Christos Karakozis, Kyle Genova, Prem Nair, Kenji Hata, and Olga Russakovsky · 2020
Later among the works it cites.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al · 2021
Later among the works it cites.
Graph-based similarity of neural network representations
Zuohui Chen, Yao Lu, Jinxuan Hu, Wen Yang, Qi Xuan, Zhen Wang, and Xiaoniu Yang · 2021
Later among the works it cites.
Debiasing pre-trained contextualised embeddings
Masahiro Kaneko and Danushka Bollegala · 2021
Later among the works it cites.
Orthogonal over-parameterized training
Weiyang Liu, Rongmei Lin, Zhen Liu, James M Rehg, Liam Paull, Li Xiong, Le Song, and Adrian Weller · 2021
Later among the works it cites.
An empirical survey of the effectiveness of debiasing techniques for pre-trained language models
Nicholas Meade, Elinor Poole-Dayan, and Siva Reddy · 2021
Later among the works it cites.
A survey on bias and fairness in machine learning
Ninareh Mehrabi, Fred Morstatter, Nripsuta Saxena, Kristina Lerman, and Aram Galstyan · 2021
Later among the works it cites.
An ecologically motivated image dataset for deep learning yields better models of human vision
Johannes Mehrer, Courtney J Spoerer, Emer C Jones, Nikolaus Kriegeskorte, and Tim C Kietzmann · 2021
Later among the works it cites.
Do wide and deep networks learn the same things? uncovering how neural network representations vary with width and depth
Thao Nguyen, Maithra Raghu, and Simon Kornblith · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Do vision transformers see like convolutional neural networks?
Maithra Raghu, Thomas Unterthiner, Simon Kornblith, Chiyuan Zhang, and Alexey Dosovitskiy · 2021
Later among the works it cites.
Towards a comprehensive understanding and accurate evaluation of societal biases in pre-trained transformers
Andrew Silva, Pradyumna Tambwekar, and Matthew Gombolay · 2021
Later among the works it cites.
Relative representations enable zero-shot latent space communication
Luca Moschella, Valentino Maiorca, Marco Fumero, Antonio Norelli, Francesco Locatello, and Emanuele Rodola · 2022
Later among the works it cites.
A review on fairness in machine learning
Dana Pessach and Erez Shmueli · 2022
Later among the works it cites.
Graph neural networks in recommender systems: a survey
Shiwen Wu, Fei Sun, Wentao Zhang, Xu Xie, and Bin Cui · 2022
Later among the works it cites.
A survey on fairness in large language models
Yingji Li, Mengnan Du, Rui Song, Xin Wang, and Ying Wang · 2023
Closest in time.
Modeldiff: A framework for comparing learning algorithms
Harshay Shah, Sung Min Park, Andrew Ilyas, and Aleksander Madry · 2023
Closest in time.
Self-supervised learning for recommender systems: A survey
Junliang Yu, Hongzhi Yin, Xin Xia, Tong Chen, Jundong Li, and Zi Huang · 2023
Closest in time.