Fetching the paper…
Reading the bibliography…
As the adoption of LLMs becomes more widespread in software coding ecosystems, a pressing issue has emerged: does the generated code contain social bias and unfairness, such as those related to age, gender, and race? This issue concerns the integrity, fairness, and ethical foundation of software applications that depend on the code generated by these models but are underexplored in the literature.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics . Philadelphia, Pennsylvania, USA: Association for Computational Linguistics, Jul. 2002, pp. 311–318. [Online]. Available: https://aclanthology.org/P02-1040
2002
Earlier work this paper cites.
G. Andreeva, J. Ansell, and J. Crook, “Impact of anti-discrimination laws on credit scoring,” Journal of Financial Services Marketing , vol. 9, pp. 22–33, 2004
2004
Earlier work this paper cites.
C.-Y. Lin, “ROUGE: A package for automatic evaluation of summaries,” in Text Summarization Branches Out . Barcelona, Spain: Association for Computational Linguistics, Jul. 2004, pp. 74–81. [Online]. Available: https://aclanthology.org/W04-1013
2004
Earlier work this paper cites.
N. Ahmad and A. N. Abd Alla, “Smart evaluation for job vacancy application system,” in 2009 Second International Conference on the Applications of Digital Information and Web Technologies . IEEE, 2009, pp. 452–455
2009
Earlier work this paper cites.
A. N. Mukherjee, S. Bhattacharyya, and R. Bera, “Role of information technology in human resource management of sme: A study on the use of applicant tracking system,” IBMRD’s Journal of Management & Research , pp. 1–22, 2014
2014
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
S. A. Friedler, C. Scheidegger, S. Venkatasubramanian, S. Choudhary, E. P. Hamilton, and D. Roth, “A comparative study of fairness-enhancing interventions in machine learning,” in Proceedings of the conference on fairness, accountability, and transparency , 2019, pp. 329–338
2019
Earlier work this paper cites.
M. Kearns, S. Neel, A. Roth, and Z. S. Wu, “An empirical study of rich subgroup fairness for machine learning,” in Proceedings of the conference on fairness, accountability, and transparency , 2019, pp. 100–109
2019
Earlier work this paper cites.
M. Hernandez, D. R. Avery, S. D. Volpone, and C. R. Kaiser, “Bargaining while black: The role of race in salary negotiations.” Journal of Applied Psychology , vol. 104, no. 4, p. 581, 2019
2019
Earlier work this paper cites.
N. Mehrabi, F. Morstatter, N. A. Saxena, K. Lerman, and A. G. Galstyan, “A survey on bias and fairness in machine learning,” ACM Computing Surveys (CSUR) , vol. 54, pp. 1 – 35, 2019. [Online]. Available: https://api.semanticscholar.org/CorpusID:201666566
2019
Earlier work this paper cites.
S. Biswas and H. Rajan, “Do the machine learning models on a crowd sourced platform exhibit bias? an empirical study on model fairness,” in ESEC/FSE’2020: The 28th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering , November 8-November 13, 2020 2020
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” Journal of Machine Learning Research , vol. 21, no. 140, pp. 1–67, 2020. [Online]. Available: http://jmlr.org/papers/v21/20-074.html
2020
Earlier work this paper cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” in Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020, NeurIPS 2020, December 6-12, 2020, virtual , H. Larochelle, M. Ranzato, R. Hadsell, M. Balcan, and H. Lin, Eds., 2020. [Online]. Available: https://proceedings.neurips.cc/paper/2020/hash/1457c0d6bfcb4967418bfb8ac142f64a-Abstract.html
2020
Earlier work this paper cites.
B. Rozière, M. Lachaux, L. Chanussot, and G. Lample, “Unsupervised translation of programming languages,” in Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020, NeurIPS 2020, December 6-12, 2020, virtual , H. Larochelle, M. Ranzato, R. Hadsell, M. Balcan, and H. Lin, Eds., 2020. [Online]. Available: https://proceedings.neurips.cc/paper/2020/hash/ed23fbf18c2cd35f8c7f8de44f85c08d-Abstract.html
2020
Earlier work this paper cites.
L. L. Taylor, J. N. Lahey, M. I. Beck, and J. E. Froyd, “How to do a salary equity study: With an illustrative example from higher education,” Public personnel management , vol. 49, no. 1, pp. 57–82, 2020
2020
Earlier work this paper cites.
M. Nadeem, A. Bethke, and S. Reddy, “Stereoset: Measuring stereotypical bias in pretrained language models,” in Annual Meeting of the Association for Computational Linguistics , 2020
2020
Earlier work this paper cites.
S. Dutta, D. Wei, H. Yueksel, P.-Y. Chen, S. Liu, and K. Varshney, “Is there a trade-off between fairness and accuracy? a perspective using mismatched hypothesis testing,” in International conference on machine learning . PMLR, 2020, pp. 2803–2813
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang, and M. Zhou, “CodeBERT: A pre-trained model for programming and natural languages,” in Findings of the Association for Computational Linguistics: EMNLP 2020 . Online: Association for Computational Linguistics, Nov. 2020, pp. 1536–1547. [Online]. Available: https://aclanthology.org/2020.findings-emnlp.139
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
M. Hasan, T. Muttaqueen, A. A. Ishtiaq, K. S. Mehrab, M. M. A. Haque, T. Hasan, W. U. Ahmad, A. Iqbal, and R. Shahriyar, “Codesc: A large code-description parallel dataset,” in Findings of the Association for Computational Linguistics: ACL/IJCNLP 2021, Online Event, August 1-6, 2021 , ser. Findings of ACL, C. Zong, F. Xia, W. Li, and R. Navigli, Eds., vol. ACL/IJCNLP 2021. Association for Computational Linguistics, 2021, pp. 210–218. [Online]. Available: https://doi.org/10.18653/v1/2021.findings-acl.18
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
F. Ding, M. Hardt, J. Miller, and L. Schmidt, “Retiring adult: New datasets for fair machine learning,” Advances in neural information processing systems , vol. 34, pp. 6478–6490, 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
N. Mehrabi, F. Morstatter, N. Saxena, K. Lerman, and A. Galstyan, “A survey on bias and fairness in machine learning,” ACM computing surveys (CSUR) , vol. 54, no. 6, pp. 1–35, 2021
2021
Earlier work this paper cites.
H. Chang and R. Shokri, “On the privacy risks of algorithmic fairness,” in 2021 IEEE European Symposium on Security and Privacy (EuroS&P) . IEEE, 2021, pp. 292–303
2021
Earlier work this paper cites.
J.-P. Platteau and D. U. Ontiveros, “Cognitive bias in insurance: evidence from a health scheme in india,” World Development , vol. 144, p. 105498, 2021
2021
Earlier work this paper cites.
P. Barlas, K. Kyriakou, O. Guest, S. Kleanthous, and J. Otterbacher, “To" see" is to stereotype: Image tagging algorithms, gender recognition, and the accuracy-fairness trade-off,” Proceedings of the ACM on Human-Computer Interaction , vol. 4, no. CSCW3, pp. 1–31, 2021
2021
Earlier work this paper cites.
A. F. Cooper, E. Abrams, and N. Na, “Emergent unfairness in algorithmic fairness-accuracy trade-off research,” in Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society , 2021, pp. 46–54
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
J.-B. Alayrac, J. Donahue, P. Luc, A. Miech, I. Barr, Y. Hasson, K. Lenc, A. Mensch, K. Millican, M. Reynolds et al. , “Flamingo: a visual language model for few-shot learning,” Advances in Neural Information Processing Systems , vol. 35, pp. 23 716–23 736, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou et al. , “Chain-of-thought prompting elicits reasoning in large language models,” Advances in Neural Information Processing Systems , vol. 35, pp. 24 824–24 837, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Cited alongside, same era.
A. M. Mir, E. Latoskinas, S. Proksch, and G. Gousios, “Type4py: Practical deep similarity learning-based type inference for python,” in 44th IEEE/ACM 44th International Conference on Software Engineering, ICSE 2022, Pittsburgh, PA, USA, May 25-27, 2022 . ACM, 2022, pp. 2241–2252. [Online]. Available: https://doi.org/10.1145/3510003.3510124
2022
Cited alongside, same era.
T. Ahmed and P. T. Devanbu, “Few-shot training llms for project-specific code-summarization,” in 37th IEEE/ACM International Conference on Automated Software Engineering, ASE 2022, Rochester, MI, USA, October 10-14, 2022 . ACM, 2022, pp. 177:1–177:5. [Online]. Available: https://doi.org/10.1145/3551349.3559555
2022
Cited alongside, same era.
C. Lemieux, J. P. Inala, S. K. Lahiri, and S. Sen, “Codamosa: Escaping coverage plateaus in test generation with pre-trained large language models,” in 45th IEEE/ACM International Conference on Software Engineering, ICSE 2023, Melbourne, Australia, May 14-20, 2023 . IEEE, 2023, pp. 919–931. [Online]. Available: https://doi.org/10.1109/ICSE48619.2023.00085
2023
Closest in time.
2023
Closest in time.
W. U. Ahmad, M. G. R. Tushar, S. Chakraborty, and K. Chang, “AVATAR: A parallel corpus for java-python program translation,” in Findings of the Association for Computational Linguistics: ACL 2023, Toronto, Canada, July 9-14, 2023 , A. Rogers, J. L. Boyd-Graber, and N. Okazaki, Eds. Association for Computational Linguistics, 2023, pp. 2268–2281. [Online]. Available: https://doi.org/10.18653/v1/2023.findings-acl.143
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Zan, B. Chen, D. Yang, Z. Lin, M. Kim, B. Guan, Y. Wang, W. Chen, and J. Lou, “CERT: continual pre-training on sketches for library-oriented code generation,” in Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence, IJCAI 2022, Vienna, Austria, 23-29 July 2022 , L. D. Raedt, Ed. ijcai.org, 2022, pp. 2369–2375. [Online]. Available: https://doi.org/10.24963/ijcai.2022/329
2022
Cited alongside, same era.
N. Jain, S. Vaidyanath, A. S. Iyer, N. Natarajan, S. Parthasarathy, S. K. Rajamani, and R. Sharma, “Jigsaw: Large language models meet program synthesis,” in 44th IEEE/ACM 44th International Conference on Software Engineering, ICSE 2022, Pittsburgh, PA, USA, May 25-27, 2022 . ACM, 2022, pp. 1219–1231. [Online]. Available: https://doi.org/10.1145/3510003.3510203
2022
Cited alongside, same era.
2022
Cited alongside, same era.
T. Le Quy, A. Roy, V. Iosifidis, W. Zhang, and E. Ntoutsi, “A survey on datasets for fairness-aware machine learning,” Wiley Interdisciplinary Reviews: Data Mining and Knowledge Discovery , vol. 12, no. 3, p. e1452, 2022
2022
Cited alongside, same era.
P. Besse, E. del Barrio, P. Gordaliza, J.-M. Loubes, and L. Risser, “A survey of bias in machine learning through the prism of statistical parity,” The American Statistician , vol. 76, no. 2, pp. 188–198, 2022
2022
Cited alongside, same era.
F. Xia, T. Guo, X. Bai, A. Shatte, Z. Liu, and J. Tang, “Summer: Bias-aware prediction of graduate employment based on educational big data,” ACM/IMS Transactions on Data Science (TDS) , vol. 2, no. 4, pp. 1–24, 2022
2022
Cited alongside, same era.
A. Papadaki, N. Martinez, M. A. Bertran, G. Sapiro, and M. R. Rodrigues, “Federated fairness without access to demographics,” in Workshop on Federated Learning: Recent Advances and New Challenges (in Conjunction with NeurIPS 2022) , 2022
2022
Cited alongside, same era.
A. Papadaki, N. Martinez, M. Bertran, G. Sapiro, and M. Rodrigues, “Minimax demographic group fairness in federated learning,” in Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency , 2022, pp. 142–159
2022
Cited alongside, same era.
A. Wang, V. V. Ramaswamy, and O. Russakovsky, “Towards intersectionality in machine learning: Including more identities, handling underrepresentation, and performing evaluation,” in Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency , 2022, pp. 336–349
2022
Cited alongside, same era.
J. Wei, G. Durrett, and I. Dillig, “Typet5: Seq2seq type inference using static analysis,” in The Eleventh International Conference on Learning Representations, ICLR 2023, Kigali, Rwanda, May 1-5, 2023 . OpenReview.net, 2023. [Online]. Available: https://openreview.net/pdf?id=4TyNEhI2GdN
2023
Closest in time.
J. Liu, C. S. Xia, Y. Wang, and L. ZHANG, “Is your code generated by chatGPT really correct? rigorous evaluation of large language models for code generation,” in Thirty-seventh Conference on Neural Information Processing Systems , 2023. [Online]. Available: https://openreview.net/forum?id=1qvx610Cu7
2023
Closest in time.
S. Wang, Z. Li, H. Qian, C. Yang, Z. Wang, M. Shang, V. Kumar, S. Tan, B. Ray, P. Bhatia, R. Nallapati, M. K. Ramanathan, D. Roth, and B. Xiang, “Recode: Robustness evaluation of code generation models,” in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2023, Toronto, Canada, July 9-14, 2023 , A. Rogers, J. L. Boyd-Graber, and N. Okazaki, Eds. Association for Computational Linguistics, 2023, pp. 13 818–13 843. [Online]. Available: https://doi.org/10.18653/v1/2023.acl-long.773
2023
Closest in time.
2023
Closest in time.
F. Cassano, J. Gouwar, D. Nguyen, S. Nguyen, L. Phipps-Costin, D. Pinckney, M. Yee, Y. Zi, C. J. Anderson, M. Q. Feldman, A. Guha, M. Greenberg, and A. Jangda, “Multipl-e: A scalable and polyglot approach to benchmarking neural code generation,” IEEE Trans. Software Eng. , vol. 49, no. 7, pp. 3675–3691, 2023. [Online]. Available: https://doi.org/10.1109/TSE.2023.3267446
2023
Closest in time.
B. Athiwaratkun, S. K. Gouda, Z. Wang, X. Li, Y. Tian, M. Tan, W. U. Ahmad, S. Wang, Q. Sun, M. Shang, S. K. Gonugondla, H. Ding, V. Kumar, N. Fulton, A. Farahani, S. Jain, R. Giaquinto, H. Qian, M. K. Ramanathan, and R. Nallapati, “Multi-lingual evaluation of code generation models,” in The Eleventh International Conference on Learning Representations, ICLR 2023, Kigali, Rwanda, May 1-5, 2023 . OpenReview.net, 2023. [Online]. Available: https://openreview.net/pdf?id=Bo7eeXm6An8
2023
Closest in time.
Y. Lai, C. Li, Y. Wang, T. Zhang, R. Zhong, L. Zettlemoyer, W. Yih, D. Fried, S. I. Wang, and T. Yu, “DS-1000: A natural and reliable benchmark for data science code generation,” in International Conference on Machine Learning, ICML 2023, 23-29 July 2023, Honolulu, Hawaii, USA , ser. Proceedings of Machine Learning Research, A. Krause, E. Brunskill, K. Cho, B. Engelhardt, S. Sabato, and J. Scarlett, Eds., vol. 202. PMLR, 2023, pp. 18 319–18 345. [Online]. Available: https://proceedings.mlr.press/v202/lai23b.html
2023
Closest in time.
P. Yin, W. Li, K. Xiao, A. Rao, Y. Wen, K. Shi, J. Howland, P. Bailey, M. Catasta, H. Michalewski, O. Polozov, and C. Sutton, “Natural language to code generation in interactive data science notebooks,” in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2023, Toronto, Canada, July 9-14, 2023 , A. Rogers, J. L. Boyd-Graber, and N. Okazaki, Eds. Association for Computational Linguistics, 2023, pp. 126–173. [Online]. Available: https://doi.org/10.18653/v1/2023.acl-long.9
2023
Closest in time.
2023
Closest in time.
E. Nijkamp, B. Pang, H. Hayashi, L. Tu, H. Wang, Y. Zhou, S. Savarese, and C. Xiong, “Codegen: An open large language model for code with multi-turn program synthesis,” in The Eleventh International Conference on Learning Representations, ICLR 2023, Kigali, Rwanda, May 1-5, 2023 . OpenReview.net, 2023. [Online]. Available: https://openreview.net/pdf?id=iaYcJKpY2B_
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
D. Shrivastava, H. Larochelle, and D. Tarlow, “Repository-level prompt generation for large language models of code,” in International Conference on Machine Learning, ICML 2023, 23-29 July 2023, Honolulu, Hawaii, USA , ser. Proceedings of Machine Learning Research, A. Krause, E. Brunskill, K. Cho, B. Engelhardt, S. Sabato, and J. Scarlett, Eds., vol. 202. PMLR, 2023, pp. 31 693–31 715. [Online]. Available: https://proceedings.mlr.press/v202/shrivastava23a.html
2023
Closest in time.
F. Zhang, B. Chen, Y. Zhang, J. Keung, J. Liu, D. Zan, Y. Mao, J. Lou, and W. Chen, “Repocoder: Repository-level code completion through iterative retrieval and generation,” in Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, EMNLP 2023, Singapore, December 6-10, 2023 , H. Bouamor, J. Pino, and K. Bali, Eds. Association for Computational Linguistics, 2023, pp. 2471–2484. [Online]. Available: https://aclanthology.org/2023.emnlp-main.151
2023
Closest in time.
Z. Chen, J. M. Zhang, F. Sarro, and M. Harman, “A comprehensive empirical study of bias mitigation methods for machine learning classifiers,” ACM Transactions on Software Engineering and Methodology , vol. 32, no. 4, pp. 1–30, 2023
2023
Closest in time.
X. Han, Z. Jiang, H. Jin, Z. Liu, N. Zou, Q. Wang, and X. Hu, “Retiring dp: New distribution-level metrics for demographic parity,” Transactions on Machine Learning Research , 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
J. Ferry, “Addresing interpretability fairness & privacy in machine learning through combinatorial optimization methods,” Ph.D. dissertation, Université Paul Sabatier-Toulouse III, 2023
2023
Closest in time.
J. Ferry, U. Aïvodji, S. Gambs, M.-J. Huguet, and M. Siala, “Exploiting fairness to enhance sensitive attributes reconstruction,” in 2023 IEEE Conference on Secure and Trustworthy Machine Learning (SaTML) . IEEE, 2023, pp. 18–41
2023
Closest in time.
2023
Closest in time.
J. M. Alvarez, K. M. Scott, B. Berendt, and S. Ruggieri, “Domain adaptive decision trees: Implications for accuracy and fairness,” in Proceedings of the 2023 ACM Conference on Fairness, Accountability, and Transparency , 2023, pp. 423–433
2023
Closest in time.
J. Simson, F. Pfisterer, and C. Kern, “Using multiverse analysis to evaluate the influence of model design decisions on algorithmic fairness,” in HHAI 2023: Augmenting Human Intellect . IOS Press, 2023, pp. 382–384
2023
Closest in time.
G. Nguyen, S. Biswas, and H. Rajan, “Fix fairness, don’t ruin accuracy: Performance aware fairness repair using automl,” in ESEC/FSE’2023: The 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering , December 3-9, 2023 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
“Adult income dataset,” www.kaggle.com/datasets/wenruliu/adult-income-dataset , 2023, accessed on August 1, 2023
2023
Closest in time.
“Employee dataset,” www.kaggle.com/datasets/tawfikelmetwally/employee-dataset , 2023, accessed on August 1, 2023
2023
Closest in time.
“Us health insurance dataset,” www.kaggle.com/datasets/teertha/ushealthinsurancedataset , 2023, accessed on August 1, 2023
2023
Closest in time.
Z. Chen, J. Zhang, F. Sarro, and M. Harman, “Fairness improvement with multiple protected attributes: How far are we?” in 46th International Conference on Software Engineering (ICSE 2024) . ACM, 2023
2023
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
D. OBrien, S. Biswas, S. Imtiaz, R. Abdalkareem, E. Shihab, and H. Rajan, “Are prompt engineering and todo comments friends or foes? an evaluation on github copilot,” in ICSE’2024: The 46th International Conference on Software Engineering , April 14-April 20 2024
2024
Closest in time.
B. Bharti, P. Yi, and J. Sulam, “Estimating and controlling for equalized odds via sensitive attribute predictors,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Z. Chen, J. M. Zhang, F. Sarro, and M. Harman, “Fairness improvement with multiple protected attributes: How far are we?” IEEE/ACM, 2024
2024
Closest in time.
2024
Closest in time.
L. Ling, “Evaluating social bias in code generation models,” in Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering , 2024, pp. 695–697
2024
Closest in time.