Fetching the paper…
Reading the bibliography…
Vision-language models (VLMs) can respond to queries about images in many languages.
On defining culture
Leon J Goldstein · 1957
Earlier work this paper cites.
A coefficient of agreement for nominal scales
Jacob Cohen · 1960
Earlier work this paper cites.
The influence of visual perception on culture 1
Marc H Bornstein · 1975
Earlier work this paper cites.
Culture and systems of thought: holistic versus analytic cognition
Richard E Nisbett, Kaiping Peng, Incheol Choi, and Ara Norenzayan · 2001
Earlier work this paper cites.
Age and culture modulate object processing and object—scene binding in the ventral visual area
Joshua O Goh, Michael W Chee, Jiat Chow Tan, Vinod Venkatraman, Andrew Hebrank, Eric D Leshikar, Lucas Jenkins, Bradley P Sutton, Angela H Gutchess, and Denise C Park · 2007
Earlier work this paper cites.
Cultural neuroscience of consciousness: From visual perception to self-awareness
Joan Chiao and T Harada · 2008
Earlier work this paper cites.
Of the east asian cultural sphere: Theorizing cultural regionalization
JungBong Choi · 2010
Earlier work this paper cites.
Fairness through awareness
Cynthia Dwork, Moritz Hardt, Toniann Pitassi, Omer Reingold, and Richard Zemel · 2012
Earlier work this paper cites.
Western civilization
Jackson J Spielvogel, James T Baker, and Joseph T Robertson · 2012
Earlier work this paper cites.
Two-factor designs: when multiple factors can affect a system, allowing for interaction can increase sensitivity
Martin Krzywinski · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al · 2017
Earlier work this paper cites.
Shreya Shankar, Yoni Halpern, Eric Breck, James Atwood, Jimbo Wilson, and D Sculley · 2017
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Earlier work this paper cites.
Does object recognition work for everyone?
Terrance De Vries, Ishan Misra, Changhan Wang, and Laurens Van der Maaten · 2019
Earlier work this paper cites.
Predictive biases in natural language processing models: A conceptual framework and overview
Deven Shah, H Andrew Schwartz, and Dirk Hovy · 2019
Earlier work this paper cites.
Interpreting gpt: The logit lens
Nostalgebraist · 2020
Earlier work this paper cites.
Artemis: Affective language for visual art
Panos Achlioptas, Maks Ovsjanikov, Kilichbek Haydarov, Mohamed Elhoseiny, and Leonidas J Guibas · 2021
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell · 2021
Earlier work this paper cites.
Quantifying social biases in nlp: A generalization and empirical comparison of extrinsic fairness metrics
Paula Czarnowska, Yogarshi Vyas, and Kashif Shah · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2021
Earlier work this paper cites.
Multilingual lama: Investigating knowledge in multilingual pretrained language models
Nora Kassner, Philipp Dufter, and Hinrich Schütze · 2021
Earlier work this paper cites.
Visually grounded reasoning across languages and cultures
Fangyu Liu, Emanuele Bugliarello, Edoardo Maria Ponti, Siva Reddy, Nigel Collier, and Desmond Elliott · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Multimodal few-shot learning with frozen language models
Maria Tsimpoukelli, Jacob L Menick, Serkan Cabi, SM Eslami, Oriol Vinyals, and Felix Hill · 2021
Cited alongside, same era.
Uc2: Universal cross-lingual cross-modal vision-and-language pre-training
Mingyang Zhou, Luowei Zhou, Shuohang Wang, Yu Cheng, Linjie Li, Zhou Yu, and Jingjing Liu · 2021
Cited alongside, same era.
Language models as agent models
Jacob Andreas · 2022
Cited alongside, same era.
Improved baselines with visual instruction tuning
Haotian Liu, Chunyuan Li, Yuheng Li, and Yong Jae Lee · 2023
Later among the works it cites.
Having beer after prayer? measuring cultural bias in large language models
Tarek Naous, Michael J Ryan, Alan Ritter, and Wei Xu · 2023
Later among the works it cites.
Joan Nwatu, Oana Ignat, and Rada Mihalcea · 2023
Later among the works it cites.
Does progress on object recognition benchmarks improve real-world generalization?
Megan Richards, Polina Kirichenko, Diane Bouchacourt, and Mark Ibrahim · 2023
Later among the works it cites.
Nlpositionality: Characterizing design biases of datasets and models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Language contamination helps explain the cross-lingual capabilities of english pretrained models
Terra Blevins and Luke Zettlemoyer · 2022
Cited alongside, same era.
Artelingo: A million emotion annotations of wikiart with emphasis on diversity over language and culture
Youssef Mohamed, Mohamed Abdelfattah, Shyma Alhuwaider, Feifan Li, Xiangliang Zhang, Kenneth Church, and Mohamed Elhoseiny · 2022
Cited alongside, same era.
Crosslingual generalization through multitask finetuning
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, Adam Roberts, Stella Biderman, Teven Le Scao, M Saiful Bari, Sheng Shen, Zheng-Xin Yong, Hailey Schoelkopf, et al · 2022
Cited alongside, same era.
Multilingual multimodal learning with machine translated text
Chen Qiu, Dan Oneata, Emanuele Bugliarello, Stella Frank, and Desmond Elliott · 2022
Cited alongside, same era.
The dollar street dataset: Images representing the geographic and socioeconomic diversity of the world
William A Gaviria Rojas, Sudnya Diamos, Keertan Ranjan Kini, David Kanter, Vijay Janapa Reddi, and Cody Coleman · 2022
Cited alongside, same era.
A-okvqa: A benchmark for visual question answering using world knowledge
Dustin Schwenk, Apoorv Khandelwal, Christopher Clark, Kenneth Marino, and Roozbeh Mottaghi · 2022
Cited alongside, same era.
Will we run out of data? an analysis of the limits of scaling datasets in machine learning
Pablo Villalobos, Jaime Sevilla, Lennart Heim, Tamay Besiroglu, Marius Hobbhahn, and Anson Ho · 2022
Cited alongside, same era.
Sebastin Santy, Jenny T Liang, Ronan Le Bras, Katharina Reinecke, and Maarten Sap · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Baichuan 2: Open large-scale language models
Aiyuan Yang, Bin Xiao, Bingning Wang, Borong Zhang, Ce Bian, Chao Yin, Chenxu Lv, Da Pan, Dian Wang, Dong Yan, et al · 2023
Later among the works it cites.
Cultural and linguistic diversity improves visual representations
Andre Ye, Sebastin Santy, Jena D Hwang, Amy X Zhang, and Ranjay Krishna · 2023
Later among the works it cites.
Givl: Improving geographical inclusivity of vision-language models with pre-training methods
Da Yin, Feng Gao, Govind Thattai, Michael Johnston, and Kai-Wei Chang · 2023
Later among the works it cites.
Cross-view language modeling: Towards unified cross-lingual cross-modal pre-training
Yan Zeng, Wangchunshu Zhou, Ao Luo, Ziming Cheng, and Xinsong Zhang · 2023
Later among the works it cites.
Investigating cultural alignment of large language models
Badr AlKhamissi, Muhammad ElNokrashy, Mai AlKhamissi, and Mona Diab · 2024
Closest in time.
Prometheus 2: An open source language model specialized in evaluating other language models
Seungone Kim, Juyoung Suk, Shayne Longpre, Bill Yuchen Lin, Jamin Shin, Sean Welleck, Graham Neubig, Moontae Lee, Kyungjae Lee, and Minjoon Seo · 2024
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2024
Closest in time.
No filter: Cultural and socioeconomic diversityin contrastive vision-language models
Angéline Pouget, Lucas Beyer, Emanuele Bugliarello, Xiao Wang, Andreas Peter Steiner, Xiaohua Zhai, and Ibrahim Alabdulmohsin · 2024
Closest in time.
Paella: Parameter-efficient lightweight language-agnostic captioning model
Rita Ramos, Emanuele Bugliarello, Bruno Martins, and Desmond Elliott · 2024
Closest in time.
Cvqa: Culturally-diverse multilingual visual question answering benchmark
David Romero, Chenyang Lyu, Haryo Akbarianto Wibowo, Teresa Lynn, Injy Hamed, Aditya Nanda Kishore, Aishik Mandal, Alina Dragonetti, Artem Abzaliev, Atnafu Lambebo Tonja, et al · 2024
Closest in time.
What is missing in multilingual visual reasoning and how to fix it
Yueqi Song, Simran Khanuja, and Graham Neubig · 2024
Closest in time.
Anthony Meng Huat Tiong, Junqi Zhao, Boyang Li, Junnan Li, Steven CH Hoi, and Caiming Xiong · 2024
Closest in time.
Do llamas work in english? on the latent language of multilingual transformers
Chris Wendler, Veniamin Veselovsky, Giovanni Monea, and Robert West · 2024
Closest in time.
A survey on multilingual large language models: Corpora, alignment, and bias
Yuemei Xu, Ling Hu, Jiayi Zhao, Zihan Qiu, Yuqi Ye, and Hanwen Gu · 2024
Closest in time.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al · 2024
Closest in time.