Fetching the paper…
Reading the bibliography…
As the utilization of large language models (LLMs) has proliferated world-wide, it is crucial for them to have adequate knowledge and fair representation for diverse global cultures.
Extracting cultural commonsense knowledge at scale
Tuan-Phong Nguyen, Simon Razniewski, Aparna Varde, and Gerhard Weikum · 1917
Earlier work this paper cites.
The interpretation of cultures , volume 5019
Clifford Geertz · 1973
Earlier work this paper cites.
Marked and unmarked: A choice between unequals in semiotic structure
Linda R Waugh · 1982
Earlier work this paper cites.
The world values survey
Christian W Haerpfer and Kseniya Kizilova · 2012
Earlier work this paper cites.
Correlation coefficients: appropriate use and interpretation
Patrick Schober, Christa Boer, and Lothar A Schwarte · 2018
Earlier work this paper cites.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, Alexander H Miller, and Sebastian Riedel · 2019
Earlier work this paper cites.
Symbolic knowledge distillation: from general language models to commonsense models
Peter West, Chandra Bhagavatula, Jack Hessel, Jena Hwang, Liwei Jiang, Ronan Le Bras, Ximing Lu, Sean Welleck, and Yejin Choi · 2022
Earlier work this paper cites.
Geomlama: Geo-diverse commonsense probing on multilingual pre-trained language models
Da Yin, Hritik Bansal, Masoud Monajatipoor, Liunian Harold Li, and Kai-Wei Chang · 2022
Earlier work this paper cites.
Marked personas: Using natural language prompts to measure stereotypes in language models
Myra Cheng, Esin Durmus, and Dan Jurafsky · 2023
Earlier work this paper cites.
Redpajama: an open dataset for training large language models, October 2023
Together Computer · 2023
Earlier work this paper cites.
Evaluation of african american language bias in natural language generation
Nicholas Deas, Jessica A. Grieser, Shana Kleiner, Desmond Upton Patton, Elsbeth Turcan, and Kathleen McKeown · 2023
Earlier work this paper cites.
Towards measuring the representation of subjective global opinions in language models
Esin Durmus, Karina Nyugen, Thomas Liao, Nicholas Schiefer, Amanda Askell, Anton Bakhtin, Carol Chen, Zac Hatfield-Dodds, Danny Hernandez, Nicholas Joseph, Liane Lovitt, Sam McCandlish, Orowa Sikder, Alex Tamkin, Janel Thamkul, Jared Kaplan, Jack Clark, and Deep Ganguli · 2023
Earlier work this paper cites.
Eticor: Corpus for analyzing llms for etiquettes
Ashutosh Dwivedi, Pradhyumna Lavania, and Ashutosh Modi · 2023
Cited alongside, same era.
Yanai Elazar, Akshita Bhagia, Ian Magnusson, Abhilasha Ravichander, Dustin Schwenk, Alane Suhr, Pete Walsh, Dirk Groeneveld, Luca Soldaini, Sameer Singh, et al · 2023
Cited alongside, same era.
Culturally aware natural language inference
Jing Huang and Diyi Yang · 2023
Cited alongside, same era.
Amr Keleg and Walid Magdy · 2023
Cited alongside, same era.
Khyati Khandelwal, Manuel Tonneau, Andrew M. Bean, Hannah Rose Kirk, and Scott A. Hale · 2023
Auditing and mitigating cultural bias in llms, 2023
Yan Tao, Olga Viberg, Ryan S. Baker, and Rene F. Kizilcec · 2023
Later among the works it cites.
Copal-id: Indonesian language reasoning with local culture and nuances
Haryo Akbarianto Wibowo, Erland Hilman Fuadi, Made Nindyatama Nityasya, Radityo Eko Prasojo, and Alham Fikri Aji · 2023
Later among the works it cites.
Normbank: A knowledge bank of situational social norms
Caleb Ziems, Jane Dwivedi-Yu, Yi-Chia Wang, Alon Halevy, and Diyi Yang · 2023
Later among the works it cites.
Towards measuring and modeling” culture” in llms: A survey
Muhammad Farid Adilazuarda, Sagnik Mukherjee, Pradhyumna Lavania, Siddhant Singh, Ashutosh Dwivedi, Alham Fikri Aji, Jacki O’Neill, Ashutosh Modi, and Monojit Choudhury · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Intersectional stereotypes in large language models: Dataset and analysis
Weicheng Ma, Brian Chiang, Tong Wu, Lili Wang, and Soroush Vosoughi · 2023
Cited alongside, same era.
Global voices, local biases: Socio-cultural prejudices across languages
Anjishnu Mukherjee, Chahat Raj, Ziwei Zhu, and Antonios Anastasopoulos · 2023
Cited alongside, same era.
Having beer after prayer? measuring cultural bias in large language models
Tarek Naous, Michael J Ryan, and Wei Xu · 2023
Cited alongside, same era.
Extracting cultural commonsense knowledge at scale
Tuan-Phong Nguyen, Simon Razniewski, Aparna Varde, and Gerhard Weikum · 2023
Cited alongside, same era.
Fork: A bite-sized test set for probing culinary cultural biases in commonsense reasoning models
Shramay Palta and Rachel Rudinger · 2023
Cited alongside, same era.
Knowledge of cultural moral norms in large language models
Aida Ramezani and Yang Xu · 2023
Cited alongside, same era.
What do llamas really think? revealing preference biases in language model representations
Raphael Tang, Xinyu Crystina Zhang, Jimmy J. Lin, and Ferhan Ture · 2023
Cited alongside, same era.
Badr AlKhamissi, Muhammad ElNokrashy, Mai AlKhamissi, and Mona Diab · 2024
Closest in time.
Massively multi-cultural knowledge acquisition & lm benchmarking
Yi Ren Fung, Ruining Zhao, Jae Doo, Chenkai Sun, and Heng Ji · 2024
Closest in time.
Olmo: Accelerating the science of language models
Dirk Groeneveld, Iz Beltagy, Pete Walsh, Akshita Bhagia, Rodney Kinney, Oyvind Tafjord, Ananya Harsh Jha, Hamish Ivison, Ian Magnusson, Yizhong Wang, et al · 2024
Closest in time.
Culturellm: Incorporating cultural differences into large language models
Cheng Li, Mengzhou Chen, Jindong Wang, Sunayana Sitaram, and Xing Xie · 2024
Closest in time.
Unintended impacts of llm alignment on global representation
Michael J Ryan, William Held, and Diyi Yang · 2024
Closest in time.
Dolma: An open corpus of three trillion tokens for language model pretraining research
Luca Soldaini, Rodney Kinney, Akshita Bhagia, Dustin Schwenk, David Atkinson, Russell Authur, Ben Bogin, Khyathi Chandu, Jennifer Dumas, Yanai Elazar, et al · 2024
Closest in time.
A roadmap to pluralistic alignment
Taylor Sorensen, Jared Moore, Jillian Fisher, Mitchell Gordon, Niloofar Mireshghallah, Christopher Michael Rytting, Andre Ye, Liwei Jiang, Ximing Lu, Nouha Dziri, et al · 2024
Closest in time.