Fetching the paper…
Reading the bibliography…
To achieve equitable performance across languages, large language models (LLMs) must be able to abstract knowledge beyond the language in which it was learnt.
Multi-source transfer of delexicalized dependency parsers
Ryan McDonald, Slav Petrov, and Keith Hall · 2011
Earlier work this paper cites.
Evaluating question answering evaluation
Anthony Chen, Gabriel Stanovsky, Sameer Singh, and Matt Gardner · 2019
Earlier work this paper cites.
Xor qa: Cross-lingual open-retrieval question answering
Akari Asai, Jungo Kasai, Jonathan H Clark, Kenton Lee, Eunsol Choi, and Hannaneh Hajishirzi · 2020
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Xtreme: A massively multilingual multi-task benchmark for evaluating cross-lingual generalisation
Junjie Hu, Sebastian Ruder, Aditya Siddhant, Graham Neubig, Orhan Firat, and Melvin Johnson · 2020
Earlier work this paper cites.
X-FACTR: Multilingual factual knowledge retrieval from pretrained language models
Zhengbao Jiang, Antonios Anastasopoulos, Jun Araki, Haibo Ding, and Graham Neubig · 2020
Earlier work this paper cites.
Multilingual LAMA: Investigating knowledge in multilingual pretrained language models
Nora Kassner, Philipp Dufter, and Hinrich Schütze · 2021
Earlier work this paper cites.
Revisiting the primacy of english in zero-shot cross-lingual transfer
Iulia Turc, Kenton Lee, Jacob Eisenstein, Ming-Wei Chang, and Kristina Toutanova · 2021
Earlier work this paper cites.
Cl-relkt: Cross-lingual language knowledge transfer for multilingual retrieval question answering
Peerat Limkonchotiwat, Wuttikorn Ponwitayarat, Can Udomcharoenchaikit, Ekapol Chuangsuwanich, and Sarana Nutanong · 2022
Earlier work this paper cites.
A balanced data approach for evaluating cross-lingual transfer: Mapping the linguistic blood bank
Dan Malkin, Tomasz Limisiewicz, and Gabriel Stanovsky · 2022
Earlier work this paper cites.
Disentqa: Disentangling parametric and contextual knowledge with counterfactual question answering
Ella Neeman, Roee Aharoni, Or Honovich, Leshem Choshen, Idan Szpektor, and Omri Abend · 2022
Earlier work this paper cites.
Can large language models be an alternative to human evaluations?
Cheng-Han Chiang and Hung-yi Lee · 2023
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2023
Cited alongside, same era.
Not all languages are created equal in LLMs: Improving multilingual capability by cross-lingual-thought prompting
Haoyang Huang, Tianyi Tang, Dongdong Zhang, Xin Zhao, Ting Song, Yan Xia, and Furu Wei · 2023
Cited alongside, same era.
mokb6: A multilingual open knowledge base completion benchmark
Shubham Mittal, Keshav Kolluru, Soumen Chakrabarti, et al · 2023
Cited alongside, same era.
Separating form and meaning: Using self-consistency to quantify task understanding across multiple senses
Xenia Ohmer, Elia Bruni, and Dieuwke Hupkes · 2023
Cited alongside, same era.
Cross-dialect information retrieval: Information access in low-resource and high-variance languages
Robert Litschko, Oliver Kraus, Verena Blaschke, and Barbara Plank · 2024
Later among the works it cites.
The llama 3 herd of models, 2024
Llama Team · 2024
Later among the works it cites.
OpenAI · 2024
Later among the works it cites.
Analyzing the evaluation of cross-lingual knowledge transfer in multilingual language models
Sara Rajaee and Christof Monz · 2024
Later among the works it cites.
Multilingual instruction tuning with just a pinch of multilinguality
Uri Shaham, Jonathan Herzig, Roee Aharoni, Idan Szpektor, Reut Tsarfaty, and Matan Eyal · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cross-lingual consistency of factual knowledge in multilingual language models
Jirui Qi, Raquel Fernández, and Arianna Bisazza · 2023
Cited alongside, same era.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al · 2023
Cited alongside, same era.
Journey to the center of the knowledge neurons: Discoveries of language-independent knowledge neurons and degenerate knowledge neurons
Yuheng Chen, Pengfei Cao, Yubo Chen, Kang Liu, and Jun Zhao · 2024
Cited alongside, same era.
Crosslingual capabilities and knowledge barriers in multilingual large language models
Lynn Chua, Badih Ghazi, Yangsibo Huang, Pritish Kamath, Ravi Kumar, Pasin Manurangsi, Amer Sinha, Chulin Xie, and Chiyuan Zhang · 2024
Cited alongside, same era.
Gemini: A family of highly capable multimodal models, 2024
Gemini Team · 2024
Cited alongside, same era.
Maxim Ifergan, Leshem Choshen, Roee Aharoni, Idan Szpektor, and Omri Abend · 2024
Cited alongside, same era.
Later among the works it cites.
How many van goghs does it take to van gogh? finding the imitation threshold, 2024
Sahil Verma, Royi Rassin, Arnav Das, Gantavya Bhatt, Preethi Seshadri, Chirag Shah, Jeff Bilmes, Hannaneh Hajishirzi, and Yanai Elazar · 2024
Later among the works it cites.
Mlake: Multilingual knowledge editing benchmark for large language models
Zihao Wei, Jingcheng Deng, Liang Pang, Hanxing Ding, Huawei Shen, and Xueqi Cheng · 2024
Later among the works it cites.
How do large language models handle multilingualism?, 2024
Yiran Zhao, Wenxuan Zhang, Guizhen Chen, Kenji Kawaguchi, and Lidong Bing · 2024
Later among the works it cites.
Itai Mondshine, Tzuf Paz-Argaman, and Reut Tsarfaty · 2025
Closest in time.
Qwen2.5 technical report, 2025
Qwen Team · 2025
Closest in time.
Graders should cheat: privileged information enables expert-level automated evaluations, 2025
Jin Peng Zhou, Sébastien M. R. Arnold, Nan Ding, Kilian Q. Weinberger, Nan Hua, and Fei Sha · 2025
Closest in time.