Fetching the paper…
Reading the bibliography…
Large language models (LLMs) are reported to be partial to certain cultures owing to the training data dominance from the English corpora.
Extracting cultural commonsense knowledge at scale
Tuan-Phong Nguyen, Simon Razniewski, Aparna Varde, and Gerhard Weikum · 1917
Earlier work this paper cites.
Direct experience and attitude-behavior consistency
Russell H Fazio and Mark P Zanna · 1981
Earlier work this paper cites.
Culture’s consequences: International differences in work-related values , volume 5
Geert Hofstede · 1984
Earlier work this paper cites.
Modernity at large: Cultural dimensions of globalization
Arjun Appadurai · 1996
Earlier work this paper cites.
Nltk: The natural language toolkit
Edward Loper and Steven Bird · 2002
Earlier work this paper cites.
Social influence and the collective dynamics of opinion formation
Mehdi Moussaïd, Juliane E Kämmer, Pantelis P Analytis, and Hansjörg Neth · 2013
Earlier work this paper cites.
Measuring the Reliability of Hate Speech Annotations: The Case of the European Refugee Crisis
Björn Ross, Michael Rist, Guillermo Carbonell, Benjamin Cabrera, Nils Kurowsky, and Michael Wojatzki · 2016
Earlier work this paper cites.
Hateful symbols or hateful people? predictive features for hate speech detection on twitter
Zeerak Waseem and Dirk Hovy · 2016
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber · 2017
Earlier work this paper cites.
Offensive comments in the brazilian web: a dataset and baseline results
Rogers P. de Pelle and Viviane P. Moreira · 2017
Earlier work this paper cites.
Bangla-abusive-comment-dataset
aimansnigdha · 2018
Earlier work this paper cites.
Overview of mex-a3t at ibereval 2018: Authorship and aggressiveness analysis in mexican spanish tweets
Miguel Á Álvarez-Carmona, Estefanıa Guzmán-Falcón, Manuel Montes-y Gómez, Hugo Jair Escalante, Luis Villasenor-Pineda, Verónica Reyes-Meza, and Antonio Rico-Sulayes · 2018
Earlier work this paper cites.
Overview of the task on automatic misogyny identification at ibereval 2018
Elisabetta Fersini, Paolo Rosso, Maria Anzovino, et al · 2018
Earlier work this paper cites.
Overview of the germeval 2018 shared task on the identification of offensive language
Michael Wiegand, Melanie Siegel, and Josef Ruppenhofer · 2018
Earlier work this paper cites.
UCI Machine Learning Repository, 2019
Turkish Spam V01 · 2019
Earlier work this paper cites.
Semeval-2019 task 5: Multilingual detection of hate speech against immigrants and women in twitter
Valerio Basile, Cristina Bosco, Elisabetta Fersini, Debora Nozza, Viviana Patti, Francisco Manuel Rangel Pardo, Paolo Rosso, and Manuela Sanguinetti · 2019
Earlier work this paper cites.
Detect camouflaged spam content via stoneskipping: Graph and text joint embedding for chinese character variation representation
Zhuoren Jiang, Zhe Gao, Guoxiu He, Yangyang Kang, Changlong Sun, Qiong Zhang, Luo Si, and Xiaozhong Liu · 2019
Earlier work this paper cites.
Jigsaw-multilingual-toxicity
Kaggle · 2019
Earlier work this paper cites.
Multilingual and multi-aspect hate speech analysis
Nedjma Ousidhoum, Zizheng Lin, Hongming Zhang, Yangqiu Song, and Dit-Yan Yeung · 2019
Earlier work this paper cites.
Detecting and monitoring hate speech in twitter
Juan Carlos Pereira-Kohatsu, Lara Quijano-Sánchez, Federico Liberatore, and Miguel Camacho-Collados · 2019
Earlier work this paper cites.
Developing a multilingual annotated corpus of misogyny and aggression
Shiladitya Bhattacharya, Siddharth Singh, Ritesh Kumar, Akanksha Bansal, Akash Bhagat, Yogesh Dawer, Bornini Lahiri, and Atul Kr. Ojha · 2020
Earlier work this paper cites.
I feel offended, don’t be abusive! implicit/explicit messages in offensive and abusive language
Tommaso Caselli, Valerio Basile, Jelena Mitrović, Inga Kartoziya, and Michael Granitzer · 2020
Earlier work this paper cites.
A corpus of turkish offensive language on social media
Çağrı Çöltekin · 2020
Earlier work this paper cites.
A multi-platform arabic news comment dataset for offensive language detection
Shammur Absar Chowdhury, Hamdy Mubarak, Ahmed Abdelali, Soon-gyo Jung, Bernard J Jansen, and Joni Salminen · 2020
Earlier work this paper cites.
Korean hatespeech dataset
daanVeer · 2020
Earlier work this paper cites.
The pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, et al · 2020
Earlier work this paper cites.
Hasoc2020
HASOC · 2020
Earlier work this paper cites.
F Husain · 2020
Cited alongside, same era.
Joao A Leite, Diego F Silva, Kalina Bontcheva, and Carolina Scarton · 2020
Cited alongside, same era.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Cited alongside, same era.
BEEP! Korean corpus of online news comments for toxic speech detection
Jihyung Moon, Won Ik Cho, and Junbum Lee · 2020
Cited alongside, same era.
Offensive language identification in greek
Zeses Pitenis, Marcos Zampieri, and Tharindu Ranasinghe · 2020
Persianllama: Towards building first persian large language model
Mohammad Amin Abbasi, Arash Ghafouri, Mahdi Firouzmandi, Hassan Naderi, and Behrouz Minaei Bidgoli · 2023
Later among the works it cites.
Mega: Multilingual evaluation of generative ai
Kabir Ahuja, Rishav Hada, Millicent Ochieng, Prachi Jain, Harshita Diddee, Samuel Maina, Tanuja Ganu, Sameer Segal, Maxamed Axmed, Kalika Bali, et al · 2023
Later among the works it cites.
Assessing cross-cultural alignment between chatgpt and human societies: An empirical study. arxiv
Y Cao, L Zhou, S Lee, L Cabello, M Chen, and D Hershcovich · 2023
Later among the works it cites.
Harmonizing global voices: Culturally-aware models for enhanced content moderation
Alex J Chan, José Luis Redondo García, Fabrizio Silvestri, Colm O’Donnel, and Konstantina Palla · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Solid: A large-scale semi-supervised dataset for offensive language identification
Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Marcos Zampieri, and Preslav Nakov · 2020
Cited alongside, same era.
Marcos Zampieri, Preslav Nakov, Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Hamdy Mubarak, Leon Derczynski, Zeses Pitenis, and Çağrı Çöltekin · 2020
Cited alongside, same era.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, et al · 2021
Cited alongside, same era.
Angel Felipe Magnossao de Paula and Ipek Baris Schlicht · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2021
Cited alongside, same era.
5k turkish tweets with incivil content
Kaggle · 2021
Cited alongside, same era.
Detecting abusive instagram comments in turkish using convolutional neural network and machine learning methods
Habibe Karayiğit, Çiğdem İnan Acı, and Ali Akdağlı · 2021
Cited alongside, same era.
Jiaming Ji, Tianyi Qiu, Boyuan Chen, Borong Zhang, Hantao Lou, Kaile Wang, Yawen Duan, Zhonghao He, Jiayi Zhou, Zhaowei Zhang, et al · 2023
Later among the works it cites.
Better to ask in english: Cross-lingual evaluation of large language models for healthcare queries
Yiqiao Jin, Mohit Chandra, Gaurav Verma, Yibo Hu, Munmun De Choudhury, and Srijan Kumar · 2023
Later among the works it cites.
Large language models as superpositions of cultural perspectives
Grgur Kovač, Masataka Sawayama, Rémy Portelas, Cédric Colas, Peter Ford Dominey, and Pierre-Yves Oudeyer · 2023
Later among the works it cites.
Self-alignment with instruction backtranslation
Xian Li, Ping Yu, Chunting Zhou, Timo Schick, Luke Zettlemoyer, Omer Levy, Jason Weston, and Mike Lewis · 2023
Later among the works it cites.
Taiwan llm: Bridging the linguistic divide with a culturally aligned language model
Yen-Ting Lin and Yun-Nung Chen · 2023
Later among the works it cites.
When less is more: Investigating data pruning for pretraining llms at scale
Max Marion, Ahmet Üstün, Luiza Pozzobon, Alex Wang, Marzieh Fadaee, and Sara Hooker · 2023
Later among the works it cites.
Reem I Masoud, Ziquan Liu, Martin Ferianc, Philip Treleaven, and Miguel Rodrigues · 2023
Later among the works it cites.
Having beer after prayer? measuring cultural bias in large language models
Tarek Naous, Michael J Ryan, and Wei Xu · 2023
Later among the works it cites.
Large language models can replicate cross-cultural differences in personality
Paweł Niszczota and Mateusz Janczak · 2023
Later among the works it cites.
Typhoon: Thai large language models
Kunat Pipatanakul, Phatrasek Jirabovonvisut, Potsawee Manakul, Sittipong Sripaisarnmongkol, Ruangsak Patomwong, Pathomporn Chokchainant, and Kasima Tharnpipitchai · 2023
Later among the works it cites.
Ethical reasoning over moral alignment: A case and framework for in-context ethical policies in llms
Abhinav Rao, Aditi Khandelwal, Kumar Tanmay, Utkarsh Agarwal, and Monojit Choudhury · 2023
Later among the works it cites.
Large language model alignment: A survey
Tianhao Shen, Renren Jin, Yufei Huang, Chuang Liu, Weilong Dong, Zishan Guo, Xinwei Wu, Yan Liu, and Deyi Xiong · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models, 2023
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Fabrication and errors in the bibliographic citations generated by chatgpt
William H Walters and Esther Isabelle Wilder · 2023
Later among the works it cites.
Cvalues: Measuring the values of chinese large language models from safety to responsibility
Guohai Xu, Jiayi Liu, Ming Yan, Haotian Xu, Jinghui Si, Zhuoran Zhou, Peng Yi, Xing Gao, Jitao Sang, Rong Zhang, Ji Zhang, Chao Peng, Fei Huang, and Jingren Zhou · 2023
Later among the works it cites.
From instructions to intrinsic human values–a survey of alignment goals for big models
Jing Yao, Xiaoyuan Yi, Xiting Wang, Jindong Wang, and Xing Xie · 2023
Later among the works it cites.
Metamath: Bootstrap your own mathematical questions for large language models
Longhui Yu, Weisen Jiang, Han Shi, Jincheng Yu, Zhengying Liu, Yu Zhang, James T Kwok, Zhenguo Li, Adrian Weller, and Weiyang Liu · 2023
Later among the works it cites.
Promptbench: Towards evaluating the robustness of large language models on adversarial prompts
Kaijie Zhu, Jindong Wang, Jiaheng Zhou, Zichen Wang, Hao Chen, Yidong Wang, Linyi Yang, Wei Ye, Yue Zhang, Neil Zhenqiang Gong, et al · 2023
Later among the works it cites.
Self-play fine-tuning converts weak language models to strong language models
Zixiang Chen, Yihe Deng, Huizhuo Yuan, Kaixuan Ji, and Quanquan Gu · 2024
Closest in time.
Dataset of arabic spam and ham tweets
Sanaa Kaddoura and Safaa Henno · 2024
Closest in time.
Mala-500: Massive language adaptation of large language models, 2024
Peiqin Lin, Shaoxiong Ji, Jörg Tiedemann, André F. T. Martins, and Hinrich Schütze · 2024
Closest in time.
Blend: A benchmark for llms on everyday knowledge in diverse cultures and languages
Junho Myung, Nayeon Lee, Yi Zhou, Jiho Jin, Rifki Afina Putri, Dimosthenis Antypas, Hsuvas Borkakoty, Eunsu Kim, Carla Perez-Almendros, Abinew Ali Ayele, et al · 2024
Closest in time.
Weiyan Shi, Ryan Li, Yutong Zhang, Caleb Ziems, Raya Horesh, Rogério Abreu de Paula, Diyi Yang, et al · 2024
Closest in time.