Fetching the paper…
Reading the bibliography…
Large-scale deployment of large language models (LLMs) in various applications, such as chatbots and virtual assistants, requires LLMs to be culturally sensitive to the user to ensure inclusivity.
Extracting cultural commonsense knowledge at scale
Nguyen, Tuan-Phong, Simon Razniewski, Aparna S. Varde, and Gerhard Weikum. 2023a · 1917
Earlier work this paper cites.
Extracting Cultural Commonsense Knowledge at Scale
Nguyen, Tuan-Phong, Simon Razniewski, Aparna S. Varde, and Gerhard Weikum. 2023b · 1917
Earlier work this paper cites.
How Communication Works , volume 586
Schramm, W. 1954 · 1954
Earlier work this paper cites.
The dynamic elements of culture
Halpern, Ben. 1955 · 1955
Earlier work this paper cites.
handbook i: cognitive domain
Bloom, BS. 1956 · 1956
Earlier work this paper cites.
The concept of culture
White, Leslie A. 1959 · 1959
Earlier work this paper cites.
CrowS-pairs: A challenge dataset for measuring social biases in masked language models
Nangia, Nikita, Clara Vania, Rasika Bhalerao, and Samuel R. Bowman. 2020 · 1967
Earlier work this paper cites.
The influence of culture on visual perception
Segall, Marshall H., Donald T. Campbell, and Melville J. Herskovits. 1967 · 1967
Earlier work this paper cites.
Pictorial depth perception: a developmental study
Jahoda, Gustav and Harry McGurk. 1974 · 1974
Earlier work this paper cites.
Bloom’s taxonomy of educational objectives for the cognitive domain: Philosophical and educational issues
Furst, Edward J. 1981 · 1981
Earlier work this paper cites.
The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale
Kuznetsova, Alina, Hassan Rom, Neil Alldrin, Jasper Uijlings, Ivan Krasin, Jordi Pont-Tuset, Shahab Kamali, Stefan Popov, Matteo Malloci, Alexander Kolesnikov, et al. 2020 · 1981
Earlier work this paper cites.
Marked and unmarked: A choice between unequals in semiotic structure
WAUGH, LINDA R. 1982 · 1982
Earlier work this paper cites.
Gender stereotypes in children’s books: their prevalence and influence on cognitive and affective development
Peterson, Sharyl Bender and Mary Alyce Lach. 1990 · 1990
Earlier work this paper cites.
Cultural alignment in response to strategic organizational change: New considerations for a change framework
Bennett III, Robert H, Paul A Fadil, and Robin T Greenwood. 1994 · 1994
Earlier work this paper cites.
Culture and change: An introduction
Naylor, Larry. 1996 · 1996
Earlier work this paper cites.
Measuring individual differences in implicit cognition: the implicit association test
Greenwald, Anthony G, Debbie E McGhee, and Jordan LK Schwartz. 1998 · 1998
Earlier work this paper cites.
Formality of language: definition, measurement and behavioral determinants
Heylighen, Francis and Jean-Marc Dewaele. 1999 · 1999
Earlier work this paper cites.
Globalization’s cultural consequences
Holton, Robert. 2000 · 2000
Earlier work this paper cites.
A textbook of translation
Newmark, Peter. 2003 · 2003
Earlier work this paper cites.
Culture and point of view
Nisbett, Richard E. and Takahiko Masuda. 2003 · 2003
Earlier work this paper cites.
A neo-boasian conception of cultural boundaries
Bashkow, Ira. 2004 · 2004
Earlier work this paper cites.
Deconstructing travel: Cultural perspectives on tourism
Berger, Arthur Asa. 2004 · 2004
Earlier work this paper cites.
Genericskb: A knowledge base of generic statements
Bhakthavatsalam, Sumithra, Chloe Anastasiades, and Peter Clark. 2020 · 2005
Earlier work this paper cites.
Culture’s recent consequences
Hofstede, Geert. 2005 · 2005
Earlier work this paper cites.
Culture, context, and behavior
Matsumoto, David. 2007 · 2007
Earlier work this paper cites.
Cross-cultural research methods
Ember, Carol R. 2009 · 2009
Earlier work this paper cites.
Creating standardized video recordings of multimodal interactions across cultures
Rehm, Matthias, Elisabeth André, Nikolaus Bee, Birgit Endrass, Michael Wissner, Yukiko Nakano, Afia Akhter Lipi, Toyoaki Nishida, and Hung-Hsuan Huang. 2009 · 2009
Earlier work this paper cites.
The weirdest people in the world?
Henrich, Joseph, Steven J Heine, and Ara Norenzayan. 2010 · 2010
Earlier work this paper cites.
Cultures and organizations: Software of the mind, third edition , 3 edition
Hofstede, Geert, Gert Jan Hofstede, and Michael Minkov. 2010 · 2010
Earlier work this paper cites.
Creative captioning: An ai grand challenge based on the dixit board game
Kunda, Maithilee and Irina Rabkina. 2020 · 2010
Earlier work this paper cites.
Atlas of the World’s Languages in Danger
Moseley, Christopher. 2010 · 2010
Earlier work this paper cites.
Intra-, inter-, and cross-cultural classification of vocal affect
Neiberg, Daniel, Petri Laukka, and Hillary Anger Elfenbein. 2011 · 2011
Earlier work this paper cites.
Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Roemmele, Melissa, Cosmin Adrian Bejan, and Andrew S Gordon. 2011 · 2011
Earlier work this paper cites.
Universals and cultural differences in forming personality trait judgments from faces
Walker, Mirella, Fang Jiang, Thomas Vetter, and Sabine Sczesny. 2011 · 2011
Earlier work this paper cites.
The world values survey
Haerpfer, Christian W and Kseniya Kizilova. 2012 · 2012
Earlier work this paper cites.
An overview of the schwartz theory of basic values
Schwartz, Shalom H. 2012 · 2012
Earlier work this paper cites.
Thinking, fast and slow, Daniel Kahneman, Farrar, Straus & Giroux
Arvai, Joseph. 2013 · 2013
Earlier work this paper cites.
The (un)faithful machine translator
Jones, Ruth and Ann Irvine. 2013 · 2013
Earlier work this paper cites.
Memes in digital culture
Shifman, Limor. 2013 · 2013
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, Tsung-Yi, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Young, Peter, Alice Lai, Micah Hodosh, and Julia Hockenmaier. 2014 · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
Antol, Stanislaw, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C. Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Earlier work this paper cites.
Commonsense reasoning and commonsense knowledge in artificial intelligence
Davis, Ernest and Gary Marcus. 2015 · 2015
Earlier work this paper cites.
Visual affect around the world: A large-scale multilingual visual sentiment ontology
Jou, Brendan, Tao Chen, Nikolaos Pappas, Miriam Redi, Mercan Topkara, and Shih-Fu Chang. 2015 · 2015
Earlier work this paper cites.
Emotion recognition from natural speech – emotional profiles
Sapinski, Tomasz and Dorota Kamińska. 2015 · 2015
Earlier work this paper cites.
Modelling human factors in perceptual multimedia quality: On the role of personality and culture
Scott, Michael ’Adrir’, Sharath Chandra Guntuku, Huan Yang, Weisi Lin, and George Ghinea. 2015 · 2015
Earlier work this paper cites.
Machine-assisted translation of literary text: A case study
Toral, Antonio and Andy Way. 2015 · 2015
Earlier work this paper cites.
The relation between language, culture, and thought
Imai, Mutsumi, Junko Kanero, and Takahiko Masuda. 2016 · 2016
Earlier work this paper cites.
Verbal Communication Styles and Culture
Liu, Meina. 2016 · 2016
Earlier work this paper cites.
Open information extraction systems and downstream applications
Mausam. 2016 · 2016
Earlier work this paper cites.
Human culture and science fiction: A review of the literature, 1980-2016
Menadue, Christopher Benjamin and Karen Diane Cheer. 2017 · 2016
Earlier work this paper cites.
Multilingual visual sentiment concept matching
Pappas, Nikolaos, Miriam Redi, Mercan Topkara, Brendan Jou, Hongyi Liu, Tao Chen, and Shih-Fu Chang. 2016 · 2016
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Caliskan, Aylin, Joanna J. Bryson, and Arvind Narayanan. 2017 · 2017
Earlier work this paper cites.
Modelling the influence of cultural information on vision-based human home activity recognition
Menicatti, Roberto, Barbara Bruno, and Antonio Sgorbissa. 2017 · 2017
Earlier work this paper cites.
Universal Dependencies
Nivre, Joakim, Daniel Zeman, Filip Ginter, and Francis Tyers. 2017 · 2017
Earlier work this paper cites.
Conceptnet 5.5: An open multilingual graph of general knowledge
Speer, Robyn, Joshua Chin, and Catherine Havasi. 2017 · 2017
Earlier work this paper cites.
Examining the millennials’ ethical profile: Assessing demographic variations in their personal value orientations
Weber, James and Michael J Urick. 2017 · 2017
Earlier work this paper cites.
Towards indonesian speech-emotion automatic recognition (i-spear)
Wunarso, Novita Belinda and Yustinus Eko Soelistio. 2017 · 2017
Earlier work this paper cites.
Intercultural Communication on Web sites: a Cross-Cultural Analysis of Web sites from High-Context Cultures and Low-Context Cultures
Würtz, Elizabeth. 2017 · 2017
Earlier work this paper cites.
Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks
Zhang, Han, Tao Xu, and Hongsheng Li. 2017 · 2017
Earlier work this paper cites.
Cosmocult card game: A methodological tool to understand the hybrid and peripheral cultural consumption of young people
Bekesas, Wilson Roberto, Mauro Berimbau, Renato Vercesi Mader, Joana Angelica Pellerano, Viviane Riegel, and Joana Pellerano. 2018 · 2018
Earlier work this paper cites.
Towards a description of chinese morphemic concepts and semantic word-formation
Liu, Yang, Zi Lin, and Sichen Kang. 2018 · 2018
Earlier work this paper cites.
WikiArt emotions: An annotated dataset of emotions evoked by art
Mohammad, Saif and Svetlana Kiritchenko. 2018 · 2018
Earlier work this paper cites.
Survey on emotional body gesture recognition
Noroozi, Fatemeh, Ciprian Adrian Corneanu, Dorota Kamińska, Tomasz Sapiński, Sergio Escalera, and Gholamreza Anbarjafari. 2018 · 2018
Earlier work this paper cites.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
Xu, Tao, Pengchuan Zhang, Qiuyuan Huang, Han Zhang, Zhe Gan, Xiaolei Huang, and Xiaodong He. 2018 · 2018
Earlier work this paper cites.
SWAG: A large-scale adversarial dataset for grounded commonsense inference
Zellers, Rowan, Yonatan Bisk, Roy Schwartz, and Yejin Choi. 2018 · 2018
Earlier work this paper cites.
Knowledge representation for culturally competent personal robots: requirements, design principles, implementation, and assessment
Bruno, Barbara, Carmine Tommaso Recchiuto, Irena Papadopoulos, Alessandro Saffiotti, Christina Koulouglioti, Roberto Menicatti, Fulvio Mastrogiovanni, Renato Zaccaria, and Antonio Sgorbissa. 2019 · 2019
Earlier work this paper cites.
Sewa db: A rich database for audio-visual emotion and sentiment research in the wild
Kossaifi, Jean, Robert Walecki, Yannis Panagakis, Jie Shen, Maximilian Schmitt, Fabien Ringeval, Jing Han, Vedhas Pandit, Antoine Toisoul, Björn Schuller, et al. 2019 · 2019
Earlier work this paper cites.
OK-VQA: A visual question answering benchmark requiring external knowledge
Marino, Kenneth, Mohammad Rastegari, Ali Farhadi, and Roozbeh Mottaghi. 2019 · 2019
Earlier work this paper cites.
Detecting hofstede cultural dimensions
Migon Favaretto, Rodolfo, Soraia Raupp Musse, Angelo Brandelli Costa, Rodolfo Migon Favaretto, Soraia Raupp Musse, and Angelo Brandelli Costa. 2019 · 2019
Earlier work this paper cites.
Generating diverse high-fidelity images with VQ-VAE-2
Razavi, Ali, Aäron van den Oord, and Oriol Vinyals. 2019 · 2019
Earlier work this paper cites.
Avec 2019 workshop and challenge: state-of-mind, detecting depression with ai, and cross-cultural affect recognition
Ringeval, Fabien, Björn Schuller, Michel Valstar, Nicholas Cummins, Roddy Cowie, Leili Tavabi, Maximilian Schmitt, Sina Alisamir, Shahin Amiriparian, Eva-Maria Messner, et al. 2019 · 2019
Earlier work this paper cites.
Artpedia: A new visual-semantic dataset with visual and contextual sentences in the artistic domain
Stefanini, Matteo, Marcella Cornia, Lorenzo Baraldi, Massimiliano Corsini, and Rita Cucchiara. 2019 · 2019
Earlier work this paper cites.
CommonsenseQA: A question answering challenge targeting commonsense knowledge
Talmor, Alon, Jonathan Herzig, Nicholas Lourie, and Jonathan Berant. 2019a · 2019
Earlier work this paper cites.
CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge
Talmor, Alon, Jonathan Herzig, Nicholas Lourie, and Jonathan Berant. 2019b · 2019
Earlier work this paper cites.
RecipeNLG: A cooking recipes dataset for semi-structured text generation
Bień, Michał, Michał Gilski, Martyna Maciejewska, Wojciech Taisner, Dawid Wisniewski, and Agnieszka Lawrynowicz. 2020 · 2020
Earlier work this paper cites.
Visual question answering for cultural heritage
Bongini, Pietro, Federico Becattini, Andrew D Bagdanov, and Alberto Del Bimbo. 2020 · 2020
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Conneau, Alexis, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Earlier work this paper cites.
Artificial Intelligence, Values, and Alignment
Gabriel, Iason. 2020 · 2020
Earlier work this paper cites.
RealToxicityPrompts: Evaluating neural toxic degeneration in language models
Gehman, Samuel, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A. Smith. 2020 · 2020
Earlier work this paper cites.
Mosaic: Finding artistic connections across culture with conditional image retrieval
Hamilton, Mark, Stephanie Fu, Mindren Lu, Johnny Bui, Darius Bopp, Zhenbang Chen, Felix Tran, Margaret Wang, Marina Rogers, Lei Zhang, et al. 2021 · 2020
Earlier work this paper cites.
The state and fate of linguistic diversity and inclusion in the NLP world
Joshi, Pratik, Sebastin Santy, Amar Budhiraja, Kalika Bali, and Monojit Choudhury. 2020 · 2020
Earlier work this paper cites.
Vyaktitv: A multimodal peer-to-peer hindi conversations based dataset for personality assessment
Khan, Shahid Nawaz, Maitree Leekha, Jainendra Shukla, and Rajiv Ratn Shah. 2020 · 2020
Earlier work this paper cites.
ISIA food-500: A dataset for large-scale food recognition via stacked global-local attention network
Min, Weiqing, Linhu Liu, Zhiling Wang, Zhengdong Luo, Xiaoming Wei, Xiaolin Wei, and Shuqiang Jiang. 2020 · 2020
Earlier work this paper cites.
Hate-Speech and Offensive Language Detection in Roman Urdu
Rizwan, Hammad, Muhammad Haroon Shakeel, and Asim Karim. 2020 · 2020
Earlier work this paper cites.
Multimodal fake news detection using a cultural algorithm with situational and normative knowledge
Shah, Priyanshi and Ziad Kobti. 2020 · 2020
Earlier work this paper cites.
ArtEmis: Affective Language for Visual Art
Achlioptas, Panos, Maks Ovsjanikov, Kilichbek Haydarov, Mohamed Elhoseiny, and Leonidas J. Guibas. 2021 · 2021
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?
Bender, Emily M., Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Earlier work this paper cites.
Hegde, Siddhanth U, Adeep Hande, Ruba Priyadharshini, Sajeetha Thavareesan, Ratnasingam Sakuntharaj, Sathiyaraj Thangasamy, B Bharathi, and Bharathi Raja Chakravarthi. 2021 · 2021
Earlier work this paper cites.
Measuring massive multitask language understanding
Hendrycks, Dan, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Earlier work this paper cites.
The importance of modeling social factors of language: Theory and practice
Hovy, Dirk and Diyi Yang. 2021 · 2021
Earlier work this paper cites.
From Culture to Clothing: Discovering the World Events Behind A Century of Fashion Images
Hsiao, Wei-Lin and Kristen Grauman. 2021 · 2021
Earlier work this paper cites.
What changes can large-scale language models bring? intensive study on HyperCLOVA: Billions-scale Korean generative pretrained transformers
Kim, Boseop, HyoungSeok Kim, Sang-Woo Lee, Gichang Lee, Donghyun Kwak, Jeon Dong Hyeon, Sunghyun Park, Sungju Kim, Seonhoon Kim, Dongpil Seo, Heungsub Lee, Minyoung Jeong, Sungjae Lee, Minsub Kim, Suk Hyun Ko, Seokhun Kim, Taeyong Park, Jinuk Kim, Soyoung Kang, Na-Hyeon Ryu, Kang Min Yoo, Minsuk Chang, Soobin Suh, Sookyo In, Jinseong Park, Kyungduk Kim, Hiun Kim, Jisu Jeong, Yong Goo Yeo, Donghoon Ham, Dongju Park, Min Young Lee, Jaewook Kang, Inho Kang, Jung-Woo Ha, Woomyoung Park, and Nako Sung. 2021 · 2021
Earlier work this paper cites.
Few-shot learning with multilingual language models
Lin, Xi Victoria, Todor Mihaylov, Mikel Artetxe, Tianlu Wang, Shuohui Chen, Daniel Simig, Myle Ott, Naman Goyal, Shruti Bhosale, Jingfei Du, et al. 2021 · 2021
Earlier work this paper cites.
Visually grounded reasoning across languages and cultures
Liu, Fangyu, Emanuele Bugliarello, Edoardo Maria Ponti, Siva Reddy, Nigel Collier, and Desmond Elliott. 2021a · 2021
Earlier work this paper cites.
Visually grounded reasoning across languages and cultures
Liu, Fangyu, Emanuele Bugliarello, Edoardo Maria Ponti, Siva Reddy, Nigel Collier, and Desmond Elliott. 2021b · 2021
Earlier work this paper cites.
OpenITI: A Machine-Readable Corpus of Islamicate Texts (2021.2. 5)[Data Set]
Nigst, Lorenz, Maxim Romanov, Sarah Bowen Savant, Masoumeh Seydi, and Peter Verkinderen. 2021 · 2021
Earlier work this paper cites.
Detecting harmful memes and their targets
Pramanick, Shraman, Dimitar Dimitrov, Rituparna Mukherjee, Shivam Sharma, Md. Shad Akhtar, Preslav Nakov, and Tanmoy Chakraborty. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, Alec, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021 · 2021
Earlier work this paper cites.
WIT: wikipedia-based image text dataset for multimodal multilingual machine learning
Srinivasan, Krishna, Karthik Raman, Jiecao Chen, Michael Bendersky, and Marc Najork. 2021 · 2021
Earlier work this paper cites.
Measuring algorithmically infused societies
Wagner, Claudia, Markus Strohmaier, Alexandra Olteanu, Emre Kıcıman, Noshir Contractor, and Tina Eliassi-Rad. 2021 · 2021
Earlier work this paper cites.
mT5: A massively multilingual pre-trained text-to-text transformer
Xue, Linting, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2021
Earlier work this paper cites.
Broaden the Vision: Geo-Diverse Visual Commonsense Reasoning
Yin, Da, Liunian Harold Li, Ziniu Hu, Nanyun Peng, and Kai-Wei Chang. 2021 · 2021
Earlier work this paper cites.
Frechet inception distance (fid) for evaluating gans
Yu, Yu, Weibin Zhang, and Yun Deng. 2021 · 2021
Earlier work this paper cites.
Zhang, Zhu, Jianxin Ma, Chang Zhou, Rui Men, Zhikang Li, Ming Ding, Jie Tang, Jingren Zhou, and Hongxia Yang. 2021 · 2021
Earlier work this paper cites.
Resources for Multilingual Hate Speech Detection
Arango Monnar, Ayme, Jorge Perez, Barbara Poblete, Magdalena Saldaña, and Valentina Proust. 2022 · 2022
Earlier work this paper cites.
How well can text-to-image generative models understand ethical natural language interventions?
Bansal, Hritik, Da Yin, Masoud Monajatipoor, and Kai-Wei Chang. 2022 · 2022
Earlier work this paper cites.
Automatic identification of motivation for code-switching in speech transcripts
Belani, Ritu and Jeffrey Flanigan. 2022 · 2022
Earlier work this paper cites.
Cultural Value Resonance in Folktales: A Transformer-Based Analysis with the World Value Corpus
Benkler, Noam, Scott Friedman, Sonja Schmer-Galunder, Drisana Mosaphir, Vasanth Sarathy, Pavan Kantharaju, Matthew D McLure, and Robert P Goldman. 2022 · 2022
Earlier work this paper cites.
Power to the people? opportunities and challenges for participatory ai
Birhane, Abeba, William Isaac, Vinodkumar Prabhakaran, Mark Diaz, Madeleine Clare Elish, Iason Gabriel, and Shakir Mohamed. 2022 · 2022
Earlier work this paper cites.
Theory-grounded measurement of U.S. social stereotypes in English language models
Cao, Yang Trista, Anna Sotnikova, Hal Daumé III, Rachel Rudinger, and Linda Zou. 2022 · 2022
Earlier work this paper cites.
Pew global attitudes survey
Center, Pew Research. 2022 · 2022
Earlier work this paper cites.
The myth of culturally agnostic ai models
Cetinic, Eva. 2022 · 2022
Earlier work this paper cites.
StereoKG: Data-driven knowledge graph construction for cultural knowledge and stereotypes
Deshpande, Awantee, Dana Ruiter, Marius Mosbach, and Dietrich Klakow. 2022 · 2022
Earlier work this paper cites.
GlobalWoZ: Globalizing MultiWoZ to Develop Multilingual Task-Oriented Dialogue Systems
Ding, Bosheng, Junjie Hu, Lidong Bing, Mahani Aljunied, Shafiq Joty, Luo Si, and Chunyan Miao. 2022 · 2022
Earlier work this paper cites.
Towards artificial general intelligence via a multimodal foundation model
Fei, Nanyi, Zhiwu Lu, Yizhao Gao, Guoxing Yang, Yuqi Huo, Jingyuan Wen, Haoyu Lu, Ruihua Song, Xin Gao, Tao Xiang, et al. 2022 · 2022
Earlier work this paper cites.
Red teaming language models to reduce harms: Methods, scaling behaviors, and lessons learned
Ganguli, Deep, Liane Lovitt, Jackson Kernion, Amanda Askell, Yuntao Bai, Saurav Kadavath, Ben Mann, Ethan Perez, Nicholas Schiefer, Kamal Ndousse, et al. 2022 · 2022
Earlier work this paper cites.
ToxiGen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection
Hartvigsen, Thomas, Saadia Gabriel, Hamid Palangi, Maarten Sap, Dipankar Ray, and Ece Kamar. 2022 · 2022
Earlier work this paper cites.
Imagen video: High definition video generation with diffusion models
Ho, Jonathan, William Chan, Chitwan Saharia, Jay Whang, Ruiqi Gao, Alexey Gritsenko, Diederik P Kingma, Ben Poole, Mohammad Norouzi, David J Fleet, et al. 2022 · 2022
Earlier work this paper cites.
Multi2WOZ: A Robust Multilingual Dataset and Conversational Pretraining for Task-Oriented Dialog
Hung, Chia-Chien, Anne Lauscher, Ivan Vulić, Simone Ponzetto, and Goran Glavaš. 2022 · 2022
Earlier work this paper cites.
KOLD: Korean offensive language dataset
Jeong, Younghoon, Juhyun Oh, Jongwon Lee, Jaimeen Ahn, Jihyung Moon, Sungjoon Park, and Alice Oh. 2022 · 2022
Earlier work this paper cites.
The ghost in the machine has an american accent: value conflict in gpt-3
Johnson, Rebecca L, Giada Pistilli, Natalia Menédez-González, Leslye Denisse Dias Duran, Enrico Panai, Julija Kalpokiene, and Donald Jay Bertulfo. 2022 · 2022
Earlier work this paper cites.
A multi-modal knowledge graph for classical Chinese poetry
Li, Yuqing, Yuxin Zhang, Bin Wu, Ji-Rong Wen, Ruihua Song, and Ting Bai. 2022 · 2022
Earlier work this paper cites.
FigMemes: A dataset for figurative language identification in politically-opinionated memes
Liu, Chen, Gregor Geigle, Robin Krebs, and Iryna Gurevych. 2022a · 2022
Earlier work this paper cites.
Testing the ability of language models to interpret figurative language
Liu, Emmy, Chenxuan Cui, Kenneth Zheng, and Graham Neubig. 2022b · 2022
Earlier work this paper cites.
Counterfactual recipe generation: Exploring compositional generalization in a realistic scenario
Liu, Xiao, Yansong Feng, Jizhi Tang, Chengang Hu, and Dongyan Zhao. 2022c · 2022
Earlier work this paper cites.
Subverting machines, fluctuating identities: Re-learning human categorization
Lu, Christina, Jackie Kay, and Kevin McKee. 2022 · 2022
Earlier work this paper cites.
EnCBP: A new benchmark dataset for finer-grained cultural background prediction in English
Ma, Weicheng, Samiha Datta, Lili Wang, and Soroush Vosoughi. 2022 · 2022
Earlier work this paper cites.
Listening to Affected Communities to Define Extreme Speech: Dataset and Experiments
Maronikolakis, Antonis, Axel Wisiorek, Leah Nann, Haris Jabbar, Sahana Udupa, and Hinrich Schuetze. 2022 · 2022
Earlier work this paper cites.
ArtELingo: A million emotion annotations of WikiArt with emphasis on diversity over language and culture
Mohamed, Youssef, Mohamed Abdelfattah, Shyma Alhuwaider, Feifan Li, Xiangliang Zhang, Kenneth Church, and Mohamed Elhoseiny. 2022 · 2022
Earlier work this paper cites.
CoCoA-MT: A dataset and benchmark for contrastive controlled MT with application to formality
Nadejde, Maria, Anna Currey, Benjamin Hsu, Xing Niu, Marcello Federico, and Georgiana Dinu. 2022 · 2022
Earlier work this paper cites.
French CrowS-pairs: Extending a challenge dataset for measuring social bias in masked language models to a language other than English
Névéol, Aurélie, Yoann Dupont, Julien Bezançon, and Karën Fort. 2022 · 2022
Earlier work this paper cites.
Cards against AI: Predicting humor in a fill-in-the-blank party game
Ofer, Dan and Dafna Shahaf. 2022 · 2022
Earlier work this paper cites.
Social simulacra: Creating populated prototypes for social computing systems
Park, Joon Sung, Lindsay Popowski, Carrie Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein. 2022 · 2022
Earlier work this paper cites.
Red teaming language models with language models
Perez, Ethan, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving. 2022 · 2022
Earlier work this paper cites.
Cultural Incongruencies in Artificial Intelligence
Prabhakaran, Vinodkumar, Rida Qadri, and Ben Hutchinson. 2022 · 2022
Earlier work this paper cites.
Hierarchical text-conditional image generation with clip latents
Ramesh, Aditya, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022 · 2022
Earlier work this paper cites.
The idioms and culture-specific items translation strategy for a classic novel
Rohmawati, Inayah Ahyana, Esti Junining, and Pratnyawati Nuridi Suwarso. 2022 · 2022
Earlier work this paper cites.
The dollar street dataset: Images representing the geographic and socioeconomic diversity of the world
Rojas, William Gaviria, Sudnya Frederick Diamos, Keertan Kini, David Kanter, Vijay Janapa Reddi, and Cody Coleman. 2022 · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Rombach, Robin, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022 · 2022
Earlier work this paper cites.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, Chitwan, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily L. Denton, Seyed Kamyar Seyed Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, Jonathan Ho, David J. Fleet, and Mohammad Norouzi. 2022 · 2022
Earlier work this paper cites.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Sap, Maarten, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Earlier work this paper cites.
Latvian national corpora collection – korpuss.lv
Saulite, Baiba, Roberts Darģis, Normunds Gruzitis, Ilze Auzina, Kristīne Levāne-Petrova, Lauma Pretkalniņa, Laura Rituma, Peteris Paikens, Arturs Znotins, Laine Strankale, Kristīne Pokratniece, Ilmārs Poikāns, Guntis Barzdins, Inguna Skadiņa, Anda Baklāne, Valdis Saulespurēns, and Jānis Ziediņš. 2022 · 2022
Cited alongside, same era.
LAION-5B: an open large-scale dataset for training next generation image-text models
Schuhmann, Christoph, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, Patrick Schramowski, Srivatsa Kundurthy, Katherine Crowson, Ludwig Schmidt, Robert Kaczmarczyk, and Jenia Jitsev. 2022 · 2022
Cited alongside, same era.
Detecting and understanding harmful memes: A survey
Sharma, Shivam, Firoj Alam, Md. Shad Akhtar, Dimitar Dimitrov, Giovanni Da San Martino, Hamed Firooz, Alon Y. Halevy, Fabrizio Silvestri, Preslav Nakov, and Tanmoy Chakraborty. 2022 · 2022
Cited alongside, same era.
Social Norms-Grounded Machine Ethics in Complex Narrative Situation
Shen, Tao, Xiubo Geng, and Daxin Jiang. 2022 · 2022
Cited alongside, same era.
Good night at 4 pm?! time expressions in different cultures
Shwartz, Vered. 2022 · 2022
Benchmarking Cognitive Domains for LLMs: Insights from Taiwanese Hakka Culture
Chang, Chen-Chi, Ching-Yuan Chen, Hung-Shin Lee, and Chih-Cheng Lee. 2024 · 2024
Closest in time.
Chiu, Yu Ying, Liwei Jiang, Maria Antoniak, Chan Young Park, Shuyue Stella Li, Mehar Bhatia, Sahithya Ravi, Yulia Tsvetkov, Vered Shwartz, and Yejin Choi. 2024 · 2024
Closest in time.
The Echoes of Multilinguality: Tracing Cultural Value Shifts during Language Model Fine-tuning
Choenni, Rochelle, Anne Lauscher, and Ekaterina Shutova. 2024 · 2024
Closest in time.
GuyLingo: The Republic of Guyana creole corpora
Clarke, Christopher, Roland Daynauth, Jason Mars, Charlene Wilkinson, and Hubert Devonish. 2024 · 2024
Closest in time.
"they are uncultured": Unveiling covert harms and social threats in llm generated conversations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
World values survey
Survey, World Values. 2022 · 2022
Cited alongside, same era.
Exploring document-level literary machine translation with parallel paragraphs from world literature
Thai, Katherine, Marzena Karpinska, Kalpesh Krishna, Bill Ray, Moira Inghilleri, John Wieting, and Mohit Iyyer. 2022 · 2022
Cited alongside, same era.
Crossmodal-3600: A massively multilingual multimodal evaluation dataset
Thapliyal, Ashish V., Jordi Pont Tuset, Xi Chen, and Radu Soricut. 2022 · 2022
Cited alongside, same era.
Super-NaturalInstructions: Generalization via declarative instructions on 1600+ NLP tasks
Wang, Yizhong, Swaroop Mishra, Pegah Alipoormolabashi, Yeganeh Kordi, Amirreza Mirzaei, Atharva Naik, Arjun Ashok, Arut Selvan Dhanasekaran, Anjana Arunkumar, David Stap, Eshaan Pathak, Giannis Karamanolakis, Haizhi Lai, Ishan Purohit, Ishani Mondal, Jacob Anderson, Kirby Kuznia, Krima Doshi, Kuntal Kumar Pal, Maitreya Patel, Mehrad Moradshahi, Mihir Parmar, Mirali Purohit, Neeraj Varshney, Phani Rohitha Kaza, Pulkit Verma, Ravsehaj Singh Puri, Rushang Karia, Savan Doshi, Shailaja Keyur Sampat, Siddhartha Mishra, Sujan Reddy A, Sumanta Patro, Tanay Dixit, and Xudong Shen. 2022 · 2022
Cited alongside, same era.
M3ED: Multi-modal multi-scene multi-label emotional dialogue database
Zhao, Jinming, Tenggan Zhang, Jingwen Hu, Yuchen Liu, Qin Jin, Xinchao Wang, and Haizhou Li. 2022 · 2022
Cited alongside, same era.
VLUE: A Multi-Task Multi-Dimension Benchmark for Evaluating Vision-Language Pre-training
Zhou, Wangchunshu, Yan Zeng, Shizhe Diao, and Xinsong Zhang. 2022 · 2022
Cited alongside, same era.
The Moral Integrity Corpus: A Benchmark for Ethical Dialogue Systems
Ziems, Caleb, Jane Yu, Yi-Chia Wang, Alon Halevy, and Diyi Yang. 2022 · 2022
Cited alongside, same era.
Dammu, Preetam Prabhu Srikar, Hayoung Jung, Anjali Singh, Monojit Choudhury, and Tanushree Mitra. 2024 · 2024
Closest in time.
Masive: Open-ended affective state identification in english and spanish
Deas, Nicholas, Elsbeth Turcan, Iván Pérez Mejía, and Kathleen McKeown. 2024 · 2024
Closest in time.
Multi-domain Hate Speech Detection Using Dual Contrastive Learning and Paralinguistic Features
Dehghan, Somaiyeh and Berrin Yanıkoğlu. 2024 · 2024
Closest in time.
Research on Generating Cultural Relic Images Based on a Low-Rank Adaptive Diffusion Model
Deng, Juntao, Xu Cao, and Bingqi Cheng. 2024 · 2024
Closest in time.
Chinese tiny llm: Pretraining a chinese-centric large language model
Du, Xinrun, Zhouliang Yu, Songyang Gao, Ding Pan, Yuyang Cheng, Ziyang Ma, Ruibin Yuan, Xingwei Qu, Jiaheng Liu, Tianyu Zheng, et al. 2024 · 2024
Closest in time.
Toucan: Many-to-many translation for 150 african language pairs
Elmadany, AbdelRahim, Ife Adebara, and Muhammad Abdul-Mageed. 2024 · 2024
Closest in time.
Bertaqa: How much do language models know about local culture?
Etxaniz, Julen, Gorka Azkune, Aitor Soroa, Oier Lopez de Lacalle, and Mikel Artetxe. 2024 · 2024
Closest in time.
DIALECTBENCH: A NLP Benchmark for Dialects, Varieties, and Closely-Related Languages
Faisal, Fahim, Orevaoghene Ahia, Aarohi Srivastava, Kabir Ahuja, David Chiang, Yulia Tsvetkov, and Antonios Anastasopoulos. 2024 · 2024
Closest in time.
Modular pluralism: Pluralistic alignment via multi-llm collaboration
Feng, Shangbin, Taylor Sorensen, Yuhan Liu, Jillian Fisher, Chan Young Park, Yejin Choi, and Yulia Tsvetkov. 2024 · 2024
Closest in time.
Biased ai can influence political decision-making
Fisher, Jillian, Shangbin Feng, Robert Aron, Thomas Richardson, Yejin Choi, Daniel W. Fisher, Jennifer Pan, Yulia Tsvetkov, and Katharina Reinecke. 2024 · 2024
Closest in time.
Your stereotypical mileage may vary: Practical challenges of evaluating biases in multiple languages and cultural contexts
Fort, Karen, Laura Alonso Alemany, Luciana Benotti, Julien Bezançon, Claudia Borg, Marthese Borg, Yongjian Chen, Fanny Ducel, Yoann Dupont, Guido Ivetta, Zhijian Li, Margot Mieskes, Marco Naguib, Yuyan Qian, Matteo Radaelli, Wolfgang S. Schmeisser-Nieto, Emma Raimundo Schulz, Thiziri Saci, Sarah Saidi, Javier Torroba Marchante, Shilin Xie, Sergio E. Zanotto, and Aurélie Névéol. 2024 · 2024
Closest in time.
Massively multi-cultural knowledge acquisition & lm benchmarking
Fung, Yi, Ruining Zhao, Jae Doo, Chenkai Sun, and Heng Ji. 2024 · 2024
Closest in time.
Funk, Marius, Shogo Okada, and Elisabeth André. 2024 · 2024
Closest in time.
Khayyam challenge (persianmmlu): Is your llm truly wise to the persian language?
Ghahroodi, Omid, Marzia Nouri, Mohammad Vali Sanian, Alireza Sahebi, Doratossadat Dastgheib, Ehsaneddin Asgari, Mahdieh Soleymani Baghshah, and Mohammad Hossein Rohban. 2024 · 2024
Closest in time.
Statistical challenges with dataset construction: Why you will never have enough images
Goldman, Josh and John K Tsotsos. 2024 · 2024
Closest in time.
RuBia: A Russian language bias detection dataset
Grigoreva, Veronika, Anastasiia Ivanova, Ilseyar Alimova, and Ekaterina Artemova. 2024 · 2024
Closest in time.
Walledeval: A comprehensive safety evaluation toolkit for large language models
Gupta, Prannaya, Le Qi Yau, Hao Han Low, I-Shiang Lee, Hugo Maximus Lim, Yu Xin Teoh, Jia Hng Koh, Dar Win Liew, Rishabh Bhardwaj, Rajat Bhardwaj, and Soujanya Poria. 2024 · 2024
Closest in time.
Nativqa: Multilingual culturally-aligned natural query for llms
Hasan, Md Arid, Maram Hasanain, Fatema Ahmad, Sahinur Rahman Laskar, Sunaya Upadhyay, Vrunda N Sukhadia, Mucahid Kutlu, Shammur Absar Chowdhury, and Firoj Alam. 2024 · 2024
Closest in time.
Building knowledge-guided lexica to model cultural variation
Havaldar, Shreya, Salvatore Giorgi, Sunny Rai, Thomas Talhelm, Sharath Chandra Guntuku, and Lyle Ungar. 2024 · 2024
Closest in time.
Chumor 1.0: A truly funny and challenging chinese humor understanding dataset from ruo zhi ba
He, Ruiqi, Yushu He, Longju Bai, Jiarui Liu, Zhenjie Sun, Zenghao Tang, He Wang, Hanchen Xia, and Naihao Deng. 2024 · 2024
Closest in time.
Recent advances in hate speech moderation: Multimodality and the role of large models
Hee, Ming Shan, Shivam Sharma, Rui Cao, Palash Nandi, Preslav Nakov, Tanmoy Chakraborty, and Roy Ka-Wei Lee. 2024 · 2024
Closest in time.
AceGPT, localizing large language models in Arabic
Huang, Huang, Fei Yu, Jianqing Zhu, Xuening Sun, Hao Cheng, Song Dingjie, Zhihong Chen, Mosen Alharthi, Bang An, Juncai He, Ziche Liu, Junying Chen, Jianquan Li, Benyou Wang, Lian Zhang, Ruoyu Sun, Xiang Wan, Haizhou Li, and Jinchao Xu. 2024 · 2024
Closest in time.
CBBQ: A Chinese bias benchmark dataset curated with human-AI collaboration for large language models
Huang, Yufei and Deyi Xiong. 2024 · 2024
Closest in time.
Cross-cultural inspiration detection and analysis in real and llm-generated social media data
Ignat, Oana, Gayathri Ganesh Lakshmy, and Rada Mihalcea. 2024 · 2024
Closest in time.
Indicvoices: Towards building an inclusive multilingual speech dataset for indian languages
Javed, Tahir, Janki Atul Nawale, Eldho Ittan George, Sakshi Joshi, Kaushal Santosh Bhogale, Deovrat Mehendale, Ishvinder Virender Sethi, Aparna Ananthanarayanan, Hafsah Faquih, Pratiti Palit, et al. 2024 · 2024
Closest in time.
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
Jha, Akshita, Vinodkumar Prabhakaran, Remi Denton, Sarah Laszlo, Shachi Dave, Rida Qadri, Chandan Reddy, and Sunipa Dev. 2024 · 2024
Closest in time.
Bridging discrete and continuous: A multimodal strategy for complex emotion detection
Jia, Jiehui, Huan Zhang, and Jinhua Liang. 2024 · 2024
Closest in time.
CPopQA: Ranking cultural concept popularity by LLMs
Jiang, Ming and Mansi Joshi. 2024 · 2024
Closest in time.
KoBBQ: Korean bias benchmark for question answering
Jin, Jiho, Jiseon Kim, Nayeon Lee, Haneul Yoo, Alice Oh, and Hwaran Lee. 2024 · 2024
Closest in time.
Does cross-cultural alignment change the commonsense morality of language models?
Jinnai, Yuu. 2024 · 2024
Closest in time.
Beyond aesthetics: Cultural competence in text-to-image models
Kannen, Nithish, Arif Ahmad, Marco Andreetto, Vinodkumar Prabhakaran, Utsav Prabhu, Adji Bousso Dieng, Pushpak Bhattacharyya, and Shachi Dave. 2024 · 2024
Closest in time.
Epistemic Injustice in Generative AI
Kay, Jackie, Atoosa Kasirzadeh, and Shakir Mohamed. 2024 · 2024
Closest in time.
Indian-bhed: A dataset for measuring india-centric biases in large language models
Khandelwal, Khyati, Manuel Tonneau, Andrew M Bean, Hannah Rose Kirk, and Scott A Hale. 2024a · 2024
Closest in time.
Indian-bhed: A dataset for measuring india-centric biases in large language models
Khandelwal, Khyati, Manuel Tonneau, Andrew M. Bean, Hannah Rose Kirk, and Scott A. Hale. 2024b · 2024
Closest in time.
Khanuja, Simran, Sathyanarayanan Ramamoorthy, Yueqi Song, and Graham Neubig. 2024 · 2024
Closest in time.
CLIcK: A Benchmark Dataset of Cultural and Linguistic Intelligence in Korean
Kim, Eunsu, Juyoung Suk, Philhoon Oh, Haneul Yoo, James Thorne, and Alice Oh. 2024a · 2024
Closest in time.
K-pop Lyric Translation: Dataset, Analysis, and Neural-Modelling
Kim, Haven, Jongmin Jung, Dasaem Jeong, and Juhan Nam. 2024b · 2024
Closest in time.
How Do Moral Emotions Shape Political Participation? A Cross-Cultural Analysis of Online Petitions Using Language Models
Kim, Jaehong, Chaeyoon Jeong, Seongchan Park, Meeyoung Cha, and Wonjae Lee. 2024c · 2024
Closest in time.
KpopMT: Translation Dataset with Terminology for Kpop Fandom
Kim, JiWoo, Yunsu Kim, and JinYeong Bak. 2024 · 2024
Closest in time.
The Benefits, Risks and Bounds of Personalizing the Alignment of Large Language Models to Individuals
Kirk, Hannah Rose, Bertie Vidgen, Paul Röttger, and Scott A. Hale. 2024 · 2024
Closest in time.
From bytes to borsch: Fine-tuning gemma and mistral for the Ukrainian language representation
Kiulian, Artur, Anton Polishko, Mykola Khandoga, Oryna Chubych, Jack Connor, Raghav Ravishankar, and Adarsh Shirawalmath. 2024 · 2024
Closest in time.
The Challenges of Creating a Parallel Multilingual Hate Speech Corpus: An Exploration
Korre, Katerina, Arianna Muti, and Alberto Barrón-Cedeño. 2024 · 2024
Closest in time.
ArabicMMLU: Assessing massive multitask language understanding in Arabic
Koto, Fajri, Haonan Li, Sara Shatnawi, Jad Doughman, Abdelrahman Sadallah, Aisha Alraeesi, Khalid Almubarak, Zaid Alyafeai, Neha Sengupta, Shady Shehata, Nizar Habash, Preslav Nakov, and Timothy Baldwin. 2024a · 2024
Closest in time.
Exploring cross-cultural differences in English hate speech annotations: From dataset construction to analysis
Lee, Nayeon, Chani Jung, Junho Myung, Jiho Jin, Jose Camacho-Collados, Juho Kim, and Alice Oh. 2024b · 2024
Closest in time.
A diverse multilingual news headlines dataset from around the world
Leeb, Felix and Bernhard Schölkopf. 2024 · 2024
Closest in time.
Genai-bench: A holistic benchmark for compositional text-to-visual generation
Li, Baiqi, Zhiqiu Lin, Deepak Pathak, Jiayao Emily Li, Xide Xia, Graham Neubig, Pengchuan Zhang, and Deva Ramanan. 2024a · 2024
Closest in time.
X-Instruction: Aligning Language Model in Low-resource Languages with Self-curated Cross-lingual Instructions
Li, Chong, Wen Yang, Jiajun Zhang, Jinliang Lu, Shaonan Wang, and Chengqing Zong. 2024d · 2024
Closest in time.
CMMLU: Measuring massive multitask language understanding in Chinese
Li, Haonan, Yixuan Zhang, Fajri Koto, Yifei Yang, Hai Zhao, Yeyun Gong, Nan Duan, and Timothy Baldwin. 2024e · 2024
Closest in time.
An Unsupervised Framework for Adaptive Context-aware Simplified-Traditional Chinese Character Conversion
Li, Wei, Shutan Huang, and Yanqiu Shao. 2024 · 2024
Closest in time.
From text to historical ecological knowledge: The construction and application of the Shan jing knowledge base
Liang, Ke, Chu-Ren Huang, and Xin-Lan Jiang. 2024 · 2024
Closest in time.
Culturally aware and adapted nlp: A taxonomy and a survey of the state of the art
Liu, Chen Cecilia, Iryna Gurevych, and Anna Korhonen. 2024 · 2024
Closest in time.
Funnynet-w: Multimodal learning of funny moments in videos in the wild
Liu, Zhi-Song, Robin Courant, and Vicky Kalogeiton. 2024 · 2024
Closest in time.
Seacrowd: A multilingual multimodal data hub and benchmark suite for southeast asian languages
Lovenia, Holy, Rahmad Mahendra, Salsabil Maulana Akbar, Lester James V Miranda, Jennifer Santoso, Elyanah Aco, Akhdan Fadhilah, Jonibek Mansurov, Joseph Marvin Imperial, Onno P Kampman, et al. 2024 · 2024
Closest in time.
Magomere, Jabez, Shu Ishida, Tejumade Afonja, Aya Salama, Daniel Kochin, Foutse Yuehgoh, Imane Hamzaoui, Raesetje Sefala, Aisha Alaagib, Elizaveta Semenova, et al. 2024 · 2024
Closest in time.
Fairylandai: Personalized fairy tales utilizing chatgpt and dalle-3
Makridis, Georgios, Athanasios Oikonomou, and Vasileios Koukos. 2024 · 2024
Closest in time.
Large language models are geographically biased
Manvi, Rohin, Samar Khanna, Marshall Burke, David B. Lobell, and Stefano Ermon. 2024 · 2024
Closest in time.
Meadows, Gwenyth Isobel, Nicholas Wai Long Lau, Eva Adelina Susanto, Chi Lok Yu, and Aditya Paul. 2024 · 2024
Closest in time.
Disce aut deficere: Evaluating llms proficiency on the invalsi italian benchmark
Mercorio, Fabio, Mario Mezzanzanica, Daniele Potertì, Antonio Serino, and Andrea Seveso. 2024 · 2024
Closest in time.
A Robot Walks into a Bar: Can Language Models Serve as Creativity SupportTools for Comedy? An Evaluation of LLMs’ Humour Alignment with Comedians
Mirowski, Piotr, Juliette Love, Kory Mathewson, and Shakir Mohamed. 2024 · 2024
Closest in time.
Global-liar: Factuality of llms over time and geographic regions
Mirza, Shujaat, Bruno Coelho, Yuyuan Cui, Christina Pöpper, and Damon McCoy. 2024 · 2024
Closest in time.
Navigating text-to-image generative bias across indic languages
Mittal, Surbhi, Arnav Sudan, Mayank Vatsa, Richa Singh, Tamar Glaser, and Tal Hassner. 2024 · 2024
Closest in time.
Advancing cultural inclusivity: Optimizing embedding spaces for balanced music recommendations
Moradi, Armin, Nicola Neophytou, and Golnoosh Farnadi. 2024 · 2024
Closest in time.
Aradice: Benchmarks for dialectal and cultural capabilities in llms
Mousi, Basel, Nadir Durrani, Fatema Ahmad, Md Arid Hasan, Maram Hasanain, Tameem Kabbani, Fahim Dalvi, Shammur Absar Chowdhury, and Firoj Alam. 2024 · 2024
Closest in time.
Global Gallery: The Fine Art of Painting Culture Portraits through Multilingual Instruction Tuning
Mukherjee, Anjishnu, Aylin Caliskan, Ziwei Zhu, and Antonios Anastasopoulos. 2024a · 2024
Closest in time.
Narayan, Malur, John Pasmore, Elton Sampaio, Vijay Raghavan, and Gabriella Waters. 2024 · 2024
Closest in time.
Benchmarking vision language models for cultural understanding
Nayak, Shravan, Kanishk Jain, Rabiul Awal, Siva Reddy, Sjoerd van Steenkiste, Lisa Anne Hendricks, Karolina Stańczak, and Aishwarya Agrawal. 2024 · 2024
Closest in time.
Mbbq: A dataset for cross-lingual comparison of stereotypes in generative llms
Neplenbroek, Vera, Arianna Bisazza, and Raquel Fernández. 2024 · 2024
Closest in time.
CulturaX: A cleaned, enormous, and multilingual dataset for large language models in 167 languages
Nguyen, Thuat, Chien Van Nguyen, Viet Dac Lai, Hieu Man, Nghia Trung Ngo, Franck Dernoncourt, Ryan A. Rossi, and Thien Huu Nguyen. 2024 · 2024
Closest in time.
Multi-cultural commonsense knowledge distillation
Nguyen, Tuan-Phong, Simon Razniewski, and Gerhard Weikum. 2024 · 2024
Closest in time.
Ochieng, Millicent, Varun Gumma, Sunayana Sitaram, Jindong Wang, Vishrav Chaudhary, Keshet Ronen, Kalika Bali, and Jacki O’Neill. 2024 · 2024
Closest in time.
Komodo: A Linguistic Expedition into Indonesia’s Regional Languages
Owen, Louis, Vishesh Tripathi, Abhay Kumar, and Biddwan Ahmed. 2024 · 2024
Closest in time.
Towards cross-lingual explanation of artwork in large-scale vision language models
Ozaki, Shintaro, Kazuki Hayashi, Yusuke Sakai, Hidetaka Kamigaito, Katsuhiko Hayashi, and Taro Watanabe. 2024 · 2024
Closest in time.
Large Language Models Can Infer Psychological Dispositions of Social Media Users
Peters, Heinrich and Sandra C Matz. 2024 · 2024
Closest in time.
Deciphering emotional landscapes in the Iliad: A novel French-annotated dataset for emotion recognition
Picca, Davide and John Pavlopoulos. 2024 · 2024
Closest in time.
Civics: Building a dataset for examining culturally-informed values in large language models
Pistilli, Giada, Alina Leidinger, Yacine Jernite, Atoosa Kasirzadeh, Alexandra Sasha Luccioni, and Margaret Mitchell. 2024 · 2024
Closest in time.
GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives
Prabhakaran, Vinodkumar, Christopher Homan, Lora Aroyo, Aida Mostafazadeh Davani, Alicia Parrish, Alex Taylor, Mark Diaz, Ding Wang, and Gregory Serapio-García. 2024 · 2024
Closest in time.
A cross-cultural analysis of social norms in bollywood and hollywood movies
Rai, Sunny, Khushang Zilesh Zaveri, Shreya Havaldar, Soumna Nema, Lyle Ungar, and Sharath Chandra Guntuku. 2024 · 2024
Closest in time.
Language varieties of Italy: Technology challenges and opportunities
Ramponi, Alan. 2024 · 2024
Closest in time.
Normad: A benchmark for measuring the cultural adaptability of large language models
Rao, Abhinav, Akhila Yerukola, Vishwa Shah, Katharina Reinecke, and Maarten Sap. 2024 · 2024
Closest in time.
CVQA: Culturally-diverse Multilingual Visual Question Answering Benchmark
Romero, David, Chenyang Lyu, Haryo Akbarianto Wibowo, Teresa Lynn, Injy Hamed, Aditya Nanda Kishore, Aishik Mandal, Alina Dragonetti, Artem Abzaliev, Atnafu Lambebo Tonja, et al. 2024 · 2024
Closest in time.
IndiBias: A benchmark dataset to measure social biases in language models for Indian context
Sahoo, Nihar, Pranamya Kulkarni, Arif Ahmad, Tanu Goyal, Narjis Asad, Aparna Garimella, and Pushpak Bhattacharyya. 2024 · 2024
Closest in time.
Revisiting the classics: A study on identifying and rectifying gender stereotypes in rhymes and poems
Sankaran, Aditya Narayan, Vigneshwaran Shankaran, Sampath Lonka, and Rajesh Sharma. 2024 · 2024
Closest in time.
Schneider, Florian and Sunayana Sitaram. 2024 · 2024
Closest in time.
DOSA: A Dataset of Social Artifacts from Different Indian Geographical Subcultures
Seth, Agrima, Sanchit Ahuja, Kalika Bali, and Sunayana Sitaram. 2024 · 2024
Closest in time.
Understanding the capabilities and limitations of large language models for cultural commonsense
Shen, Siqi, Lajanugen Logeswaran, Moontae Lee, Honglak Lee, Soujanya Poria, and Rada Mihalcea. 2024 · 2024
Closest in time.
Shi, Weiyan, Ryan Li, Yutong Zhang, Caleb Ziems, Raya Horesh, Rogério Abreu de Paula, Diyi Yang, et al. 2024 · 2024
Closest in time.
AI models collapse when trained on recursively generated data
Shumailov, Ilia, Zakhar Shumaylov, Yiren Zhao, Nicolas Papernot, Ross Anderson, and Yarin Gal. 2024 · 2024
Closest in time.
Generalizable Multilingual Hate Speech Detection on Low Resource Indian Languages using Fair Selection in Federated Learning
Singh, Akshay and Rahul Thakur. 2024 · 2024
Closest in time.
HAE-RAE Bench: Evaluation of Korean Knowledge in Language Models
Son, Guijin, Hanwool Lee, Suwan Kim, Huiseo Kim, Jae cheol Lee, Je Won Yeom, Jihyu Jung, Jung woo Kim, and Songseong Kim. 2024b · 2024
Closest in time.
The typing cure: Experiences with large language model chatbots for mental health support
Song, Inhwa, Sachin R Pendse, Neha Kumar, and Munmun De Choudhury. 2024 · 2024
Closest in time.
ThatiAR: Subjectivity Detection in Arabic News Sentences
Suwaileh, Reem, Maram Hasanain, Fatema Hubail, Wajdi Zaghouani, and Firoj Alam. 2024 · 2024
Closest in time.
CHisIEC: An information extraction corpus for Ancient Chinese history
Tang, Xuemei, Qi Su, Jun Wang, and Zekun Deng. 2024b · 2024
Closest in time.
Cultural bias and cultural alignment of large language models
Tao, Yan, Olga Viberg, Ryan S Baker, and René F Kizilcec. 2024 · 2024
Closest in time.
A dataset for metaphor detection in early medieval Hebrew poetry
Toker, Michael, Oren Mishali, Ophir Münz-Manor, Benny Kimelfeld, and Yonatan Belinkov. 2024 · 2024
Closest in time.
EthioLLM: Multilingual large language models for Ethiopian languages with task evaluation
Tonja, Atnafu Lambebo, Israel Abebe Azime, Tadesse Destaw Belay, Mesay Gemeda Yigezu, Moges Ahmed Ah Mehamed, Abinew Ali Ayele, Ebrahim Chekol Jibril, Michael Melese Woldeyohannis, Olga Kolesnikova, Philipp Slusallek, Dietrich Klakow, and Seid Muhie Yimam. 2024 · 2024
Closest in time.
From languages to geographies: Towards evaluating cultural bias in hate speech datasets
Tonneau, Manuel, Diyi Liu, Samuel Fraiberger, Ralph Schroeder, Scott Hale, and Paul Röttger. 2024 · 2024
Closest in time.
How to Use Large-Language Models for Text Analysis
Törnberg, Petter. 2024 · 2024
Closest in time.
Uccix: Irish-excellence large language model
Tran, Khanh-Tung, Barry O’Sullivan, and Hoang D Nguyen. 2024 · 2024
Closest in time.
Detecting Cybercrimes in Accordance with Pakistani Law: Dataset and Evaluation Using PLMs
Ullah, Faizad, Ali Faheem, Ubaid Azam, Muhammad Sohaib Ayub, Faisal Kamiran, and Asim Karim. 2024 · 2024
Closest in time.
Aya model: An instruction finetuned open-access multilingual language model
Üstün, Ahmet, Viraat Aryabumi, Zheng-Xin Yong, Wei-Yin Ko, Daniel D’souza, Gbemileke Onilude, Neel Bhandari, Shivalika Singh, Hui-Lee Ooi, Amr Kayid, et al. 2024 · 2024
Closest in time.
A treebank of Asia minor Greek
Vligouridou, Eleni, Inessa Iliadou, and Çağrı Çöltekin. 2024 · 2024
Closest in time.
Sonnet or not, bot? poetry evaluation for large models and datasets
Walsh, Melanie, Anna Preus, and Maria Antoniak. 2024 · 2024
Closest in time.
Wang, Angelina, Jamie Morgenstern, and John P. Dickerson. 2024 · 2024
Closest in time.
SeaEval for multilingual foundation models: From cross-lingual alignment to cultural reasoning
Wang, Bin, Zhengyuan Liu, Xin Huang, Fangkai Jiao, Yang Ding, AiTi Aw, and Nancy Chen. 2024b · 2024
Closest in time.
CMB: A comprehensive medical benchmark in Chinese
Wang, Xidong, Guiming Chen, Song Dingjie, Zhang Zhiyi, Zhihong Chen, Qingying Xiao, Junying Chen, Feng Jiang, Jianquan Li, Xiang Wan, Benyou Wang, and Haizhou Li. 2024e · 2024
Closest in time.
PARIKSHA: A Large-Scale Investigation of Human-LLM Evaluator Agreement on Multilingual and Multi-Cultural Data
Watts, Ishaan, Varun Gumma, Aditya Yadavalli, Vivek Seshadri, Manohar Swaminathan, and Sunayana Sitaram. 2024 · 2024
Closest in time.
Muchomusic: Evaluating music understanding in multimodal audio-language models
Weck, Benno, Ilaria Manco, Emmanouil Benetos, Elio Quinton, George Fazekas, and Dmitry Bogdanov. 2024 · 2024
Closest in time.
Ac-eval: Evaluating ancient chinese language understanding in large language models
Wei, Yuting, Yuanxing Xu, Xinru Wei, Simin Yang, Yangfu Zhu, Yuqing Li, Di Liu, and Bin Wu. 2024 · 2024
Closest in time.
COPAL-ID: Indonesian Language Reasoning with Local Culture and Nuances
Wibowo, Haryo, Erland Fuadi, Made Nityasya, Radityo Eko Prasojo, and Alham Aji. 2024 · 2024
Closest in time.
Revealing fine-grained values and opinions in large language models
Wright, Dustin, Arnav Arora, Nadav Borenstein, Srishti Yadav, Serge Belongie, and Isabelle Augenstein. 2024 · 2024
Closest in time.
Wu, Minghao, Yulin Yuan, Gholamreza Haffari, and Longyue Wang. 2024 · 2024
Closest in time.
Rtp-lx: Can llms evaluate toxicity in multilingual scenarios?
de Wynter, Adrian, Ishaan Watts, Nektar Ege Altıntoprak, Tua Wongsangaroonsri, Minghui Zhang, Noura Farra, Lena Baur, Samantha Claudet, Pavel Gajdusek, Can Gören, Qilong Gu, Anna Kaminska, Tomasz Kaminski, Ruby Kuo, Akiko Kyuba, Jongho Lee, Kartik Mathur, Petter Merok, Ivana Milovanović, Nani Paananen, Vesa-Matti Paananen, Anna Pavlenko, Bruno Pereira Vidal, Luciano Strika, Yueh Tsao, Davide Turcato, Oleksandr Vakhno, Judit Velcsov, Anna Vickers, Stéphanie Visser, Herdyan Widarmanto, Andrey Zaikin, and Si-Qing Chen. 2024 · 2024
Closest in time.
When search engine services meet large language models: Visions and challenges
Xiong, Haoyi, Jiang Bian, Yuchen Li, Xuhong Li, Mengnan Du, Shuaiqiang Wang, Dawei Yin, and Sumi Helal. 2024 · 2024
Closest in time.
Analyzing social biases in japanese large language models
Yanaka, Hitomi, Namgi Han, Ryoma Kumon, Jie Lu, Masashi Takeshita, Ryo Sekizawa, Taisei Kato, and Hiromi Arai. 2024 · 2024
Closest in time.
Social Skill Training with Large Language Models
Yang, Diyi, Caleb Ziems, William Held, Omar Shaikh, Michael S. Bernstein, and John Mitchell. 2024 · 2024
Closest in time.
Value FULCRA: Mapping large language models to the multidimensional spectrum of basic human value
Yao, Jing, Xiaoyuan Yi, Yifan Gong, Xiting Wang, and Xing Xie. 2024 · 2024
Closest in time.
GOLEM: GOld standard for learning and evaluation of motifs
Yarlott, W. Victor, Anurag Acharya, Diego Castro Estrada, Diana Gomez, and Mark Finlayson. 2024 · 2024
Closest in time.
Altdiffusion: A multilingual text-to-image diffusion model
Ye, Fulong, Guang Liu, Xinya Wu, and Ledell Wu. 2024 · 2024
Closest in time.
Chinese morpheme-informed evaluation of large language models
Yin, Yaqi, Yue Wang, and Yang Liu. 2024 · 2024
Closest in time.
Yin, Ziqi, Hao Wang, Kaito Horio, Daisuke Kawahara, and Satoshi Sekine. 2024 · 2024
Closest in time.
Yoo, Kang Min, Jaegeun Han, Sookyo In, Heewon Jeon, Jisu Jeong, Jaewook Kang, Hyunwook Kim, Kyung-Min Kim, Munhyong Kim, Sungju Kim, et al. 2024 · 2024
Closest in time.
CMoralEval: A Moral Evaluation Benchmark for Chinese Large Language Models
Yu, Linhao, Yongqi Leng, Yufei Huang, Shang Wu, Haixin Liu, Xinmeng Ji, Jiahui Zhao, Jinwang Song, Tingting Cui, Xiaoqing Cheng, Liutao Liutao, and Deyi Xiong. 2024 · 2024
Closest in time.
Measuring Social Norms of Large Language Models
Yuan, Ye, Kexin Tang, Jianhao Shen, Ming Zhang, and Chenguang Wang. 2024 · 2024
Closest in time.
Cic: A framework for culturally-aware image captioning
Yun, Youngsik and Jihie Kim. 2024 · 2024
Closest in time.
Turkishmmlu: Measuring massive multitask language understanding in turkish
Yüksel, Arda, Abdullatif Köksal, Lütfi Kerem Şenel, Anna Korhonen, and Hinrich Schütze. 2024 · 2024
Closest in time.
RENOVI: A benchmark towards remediating norm violations in socio-cultural conversations
Zhan, Haolan, Zhuang Li, Xiaoxi Kang, Tao Feng, Yuncheng Hua, Lizhen Qu, Yi Ying, Mei Rianto Chandra, Kelly Rosalin, Jureynolds Jureynolds, Suraj Sharma, Shilin Qu, Linhao Luo, Ingrid Zukerman, Lay-Ki Soon, Zhaleh Semnani Azad, and Reza Haf. 2024 · 2024
Closest in time.
Partiality and Misconception: Investigating Cultural Representativeness in Text-to-Image Models
Zhang, Lili, Xi Liao, Zaijia Yang, Baihang Gao, Chunjie Wang, Qiuling Yang, and Deshun Li. 2024d · 2024
Closest in time.
Interpreting themes from educational stories
Zhang, Yigeng, Fabio Gonzalez, and Thamar Solorio. 2024 · 2024
Closest in time.
WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models
Zhao, Wenlong, Debanjan Mondal, Niket Tandon, Danica Dillion, Kurt Gray, and Yuling Gu. 2024 · 2024
Closest in time.
Beyond Preferences in AI Alignment
Zhi-Xuan, Tan, Micah Carroll, Matija Franklin, and Hal Ashton. 2024 · 2024
Closest in time.
Does Mapo Tofu Contain Coffee? Probing LLMs for Food-related Cultural Knowledge
Zhou, Li, Taelin Karidi, Nicolas Garneau, Yong Cao, Wanlong Liu, Wenyu Chen, and Daniel Hershcovich. 2024 · 2024
Closest in time.
Quite good, but not enough: Nationality bias in large language models - a case study of ChatGPT
Zhu, Shucheng, Weikang Wang, and Ying Liu. 2024 · 2024
Closest in time.
Can Large Language Models Transform Computational Social Science?
Ziems, Caleb, William Held, Omar Shaikh, Jiaao Chen, Zhehao Zhang, and Diyi Yang. 2024 · 2024
Closest in time.
Are multilingual LLMs culturally-diverse reasoners? an investigation into multicultural proverbs and sayings
Liu, Chen, Fajri Koto, Timothy Baldwin, and Iryna Gurevych. 2024a · 2039
Closest in time.
GeoMLAMA: Geo-diverse commonsense probing on multilingual pre-trained language models
Yin, Da, Hritik Bansal, Masoud Monajatipoor, Liunian Harold Li, and Kai-Wei Chang. 2022 · 2055
Closest in time.
The (undesired) attenuation of human biases by multilinguality
España-Bonet, Cristina and Alberto Barrón-Cedeño. 2022 · 2077
Closest in time.
BBQ: A hand-built bias benchmark for question answering
Parrish, Alicia, Angelica Chen, Nikita Nangia, Vishakh Padmakumar, Jason Phang, Jana Thompson, Phu Mon Htut, and Samuel Bowman. 2022 · 2086
Closest in time.