Fetching the paper…
Reading the bibliography…
Multilingual Large Language Models are capable of using powerful Large Language Models to handle and respond to queries in multiple languages, which achieves remarkable success in multilingual natural language processing tasks.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Paws-x: A cross-lingual adversarial dataset for paraphrase identification
Yinfei Yang, Yuan Zhang, Chris Tar, and Jason Baldridge. 2019b · 1908
Earlier work this paper cites.
Mlqa: Evaluating cross-lingual extractive question answering
Patrick Lewis, Barlas Oğuz, Ruty Rinott, Sebastian Riedel, and Holger Schwenk. 2019 · 1910
Earlier work this paper cites.
Multilingual entity and relation extraction dataset and model
Alessandro Seganti, Klaudia Firląg, Helena Skowronska, Michał Satława, and Piotr Andruszkiewicz. 2021 · 1955
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Xglue: A new benchmark dataset for cross-lingual pre-training, understanding and generation
Yaobo Liang, Nan Duan, Yeyun Gong, Ning Wu, Fenfei Guo, Weizhen Qi, Ming Gong, Linjun Shou, Daxin Jiang, Guihong Cao, et al. 2020 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Bleurt: Learning robust metrics for text generation
Thibault Sellam, Dipanjan Das, and Ankur P Parikh. 2020 · 2004
Earlier work this paper cites.
Europarl: A parallel corpus for statistical machine translation
Philipp Koehn. 2005 · 2005
Earlier work this paper cites.
Xcopa: A multilingual dataset for causal commonsense reasoning
Edoardo Maria Ponti, Goran Glavaš, Olga Majewska, Qianchu Liu, Ivan Vulić, and Anna Korhonen. 2020 · 2005
Earlier work this paper cites.
Gshard: Scaling giant models with conditional computation and automatic sharding
Dmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen, Orhan Firat, Yanping Huang, Maxim Krikun, Noam Shazeer, and Zhifeng Chen. 2020 · 2006
Earlier work this paper cites.
Cosda-ml: Multi-lingual code-switching data augmentation for zero-shot cross-lingual nlp
Libo Qin, Minheng Ni, Yue Zhang, and Wanxiang Che. 2020 · 2006
Earlier work this paper cites.
Mtop: A comprehensive multilingual task-oriented semantic parsing benchmark
Haoran Li, Abhinav Arora, Shuohui Chen, Anchit Gupta, Sonal Gupta, and Yashar Mehdad. 2020 · 2008
Earlier work this paper cites.
mt5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2020 · 2010
Earlier work this paper cites.
Momchil Hardalov, Todor Mihaylov, Dimitrina Zlatkova, Yoan Dinkov, Ivan Koychev, and Preslav Nakov. 2020 · 2011
Earlier work this paper cites.
Parallel data, tools and interfaces in OPUS
Jörg Tiedemann. 2012 · 2012
Earlier work this paper cites.
Creating a massively parallel Bible corpus
Thomas Mayer and Michael Cysouw. 2014 · 2014
Earlier work this paper cites.
The United Nations parallel corpus v1.0
Michał Ziemski, Marcin Junczys-Dowmunt, and Bruno Pouliquen. 2016 · 2016
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
chrF++: words helping character n-grams
Maja Popović. 2017 · 2017
Earlier work this paper cites.
Xnli: Evaluating cross-lingual sentence representations
Alexis Conneau, Guillaume Lample, Ruty Rinott, Adina Williams, Samuel R Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Earlier work this paper cites.
The IIT Bombay English-Hindi parallel corpus
Anoop Kunchukuttan, Pratik Mehta, and Pushpak Bhattacharyya. 2018 · 2018
Earlier work this paper cites.
Shashi Narayan, Shay B Cohen, and Mirella Lapata. 2018 · 2018
Earlier work this paper cites.
Cross-lingual transfer learning for multilingual task oriented dialog
Sebastian Schuster, Sonal Gupta, Rushin Shah, and Mike Lewis. 2018 · 2018
Earlier work this paper cites.
Jw300: A wide-coverage parallel corpus for low-resource languages
Željko Agic and Ivan Vulic. 2019 · 2019
Earlier work this paper cites.
Cross-lingual joint entity and word embedding to improve entity linking and parallel sentence mining
Xiaoman Pan, Thamme Gowda, Heng Ji, Jonathan May, and Scott Miller. 2019 · 2019
Earlier work this paper cites.
Asynchronous pipeline for processing huge corpora on medium to low resource infrastructures
Pedro Javier Ortiz Suárez, Benoît Sagot, and Laurent Romary. 2019 · 2019
Earlier work this paper cites.
Ccmt 2019 machine translation evaluation report
Muyun Yang, Xixin Hu, Hao Xiong, Jiayi Wang, Yiliyaer Jiaermuhamaiti, Zhongjun He, Weihua Luo, and Shujian Huang. 2019a · 2019
Earlier work this paper cites.
On the cross-lingual transferability of monolingual representations
Mikel Artetxe, Sebastian Ruder, and Dani Yogatama. 2020 · 2020
Earlier work this paper cites.
Tydi qa: A benchmark for information-seeking question answering in ty pologically di verse languages
Jonathan H Clark, Eunsol Choi, Michael Collins, Dan Garrette, Tom Kwiatkowski, Vitaly Nikolaev, and Jennimaria Palomaki. 2020 · 2020
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Édouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Earlier work this paper cites.
A survey of multilingual neural machine translation
Raj Dabre, Chenhui Chu, and Anoop Kunchukuttan. 2020 · 2020
Earlier work this paper cites.
Xtreme: A massively multilingual multi-task benchmark for evaluating cross-lingual generalisation
Junjie Hu, Sebastian Ruder, Aditya Siddhant, Graham Neubig, Orhan Firat, and Melvin Johnson. 2020 · 2020
Earlier work this paper cites.
The multilingual amazon reviews corpus
Phillip Keung, Yichao Lu, György Szarvas, and Noah A Smith. 2020 · 2020
Earlier work this paper cites.
Comet: A neural framework for mt evaluation
Ricardo Rei, Craig Stewart, Ana C Farinha, and Alon Lavie. 2020 · 2020
Earlier work this paper cites.
End-to-end slot alignment and recognition for cross-lingual NLU
Weijia Xu, Batool Haider, and Saab Mansour. 2020 · 2020
Earlier work this paper cites.
Improving massively multilingual neural machine translation and zero-shot translation
Biao Zhang, Philip Williams, Ivan Titov, and Rico Sennrich. 2020 · 2020
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Tianyi Zhang*, Varsha Kishore*, Felix Wu*, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 2020
Earlier work this paper cites.
Masakhaner: Named entity recognition for african languages
David Ifeoluwa Adelani, Jade Abbott, Graham Neubig, Daniel D’souza, Julia Kreutzer, Constantine Lignos, Chester Palen-Michel, Happy Buzaaba, Shruti Rijhwani, Sebastian Ruder, et al. 2021 · 2021
Earlier work this paper cites.
Diabla: a corpus of bilingual spontaneous written dialogues for machine translation
Rachel Bawden, Eric Bilinski, Thomas Lavergne, and Sophie Rosset. 2021 · 2021
Earlier work this paper cites.
Abhik Bhattacharjee, Tahmid Hasan, Wasi Uddin Ahmad, Yuan-Fang Li, Yong-Bin Kang, and Rifat Shahriyar. 2021 · 2021
Earlier work this paper cites.
Beyond english-centric multilingual machine translation
Angela Fan, Shruti Bhosale, Holger Schwenk, Zhiyi Ma, Ahmed El-Kishky, Siddharth Goyal, Mandeep Baines, Onur Celebi, Guillaume Wenzek, Vishrav Chaudhary, et al. 2021 · 2021
Earlier work this paper cites.
nmt5–is parallel data still relevant for pre-training massively multilingual language models?
Mihir Kale, Aditya Siddhant, Noah Constant, Melvin Johnson, Rami Al-Rfou, and Linting Xue. 2021 · 2021
Earlier work this paper cites.
What changes can large-scale language models bring? intensive study on hyperclova: Billions-scale korean generative pretrained transformers
Boseop Kim, HyoungSeok Kim, Sang-Woo Lee, Gichang Lee, Donghyun Kwak, Jeon Dong Hyeon, Sunghyun Park, Sungju Kim, Seonhoon Kim, Dongpil Seo, et al. 2021 · 2021
Earlier work this paper cites.
Small data? no problem! exploring the viability of pretrained multilingual language models for low-resourced languages
Kelechi Ogueji, Yuxin Zhu, and Jimmy Lin. 2021 · 2021
Earlier work this paper cites.
The curious case of hallucinations in neural machine translation
Vikas Raunak, Arul Menezes, and Marcin Junczys-Dowmunt. 2021 · 2021
Earlier work this paper cites.
Xtreme-r: Towards more challenging and nuanced multilingual evaluation
Sebastian Ruder, Noah Constant, Jan Botha, Aditya Siddhant, Orhan Firat, Jinlan Fu, Pengfei Liu, Junjie Hu, Dan Garrette, Graham Neubig, et al. 2021 · 2021
Earlier work this paper cites.
WikiMatrix: Mining 135M parallel sentences in 1620 language pairs from Wikipedia
Holger Schwenk, Vishrav Chaudhary, Shuo Sun, Hongyu Gong, and Francisco Guzmán. 2021 · 2021
Earlier work this paper cites.
Ernie 3.0: Large-scale knowledge enhanced pre-training for language understanding and generation
Yu Sun, Shuohuan Wang, Shikun Feng, Siyu Ding, Chao Pang, Junyuan Shang, Jiaxiang Liu, Xuyi Chen, Yanbin Zhao, Yuxiang Lu, et al. 2021 · 2021
Earlier work this paper cites.
Alexey Tikhonov and Max Ryabinin. 2021 · 2021
Earlier work this paper cites.
mt5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2021
Earlier work this paper cites.
Cpm-2: Large-scale cost-effective pre-trained language models
Zhengyan Zhang, Yuxian Gu, Xu Han, Shengqi Chen, Chaojun Xiao, Zhenbo Sun, Yuan Yao, Fanchao Qi, Jian Guan, Pei Ke, et al. 2021 · 2021
Earlier work this paper cites.
David Ifeoluwa Adelani, Jesujoba Oluwadara Alabi, Angela Fan, Julia Kreutzer, Xiaoyu Shen, Machel Reid, Dana Ruiter, Dietrich Klakow, Peter Nabende, Ernie Chang, et al. 2022 · 2022
Earlier work this paper cites.
On the calibration of massively multilingual language models
Kabir Ahuja, Sunayana Sitaram, Sandipan Dandapat, and Monojit Choudhury. 2022 · 2022
Earlier work this paper cites.
Bootstrapping multilingual semantic parsers using large language models
Abhijeet Awasthi, Nitish Gupta, Bidisha Samanta, Shachi Dave, Sunita Sarawagi, and Partha Talukdar. 2022 · 2022
Earlier work this paper cites.
Building machine translation systems for the next thousand languages
Ankur Bapna, Isaac Caswell, Julia Kreutzer, Orhan Firat, Daan van Esch, Aditya Siddhant, Mengmeng Niu, Pallavi Baljekar, Xavier Garcia, Wolfgang Macherey, et al. 2022 · 2022
Earlier work this paper cites.
Gpt-neox-20b: An open-source autoregressive language model
Sid Black, Stella Biderman, Eric Hallahan, Quentin Anthony, Leo Gao, Laurence Golding, Horace He, Connor Leahy, Kyle McDonell, Jason Phang, et al. 2022 · 2022
Earlier work this paper cites.
Language contamination helps explain the cross-lingual capabilities of english pretrained models
Terra Blevins and Luke Zettlemoyer. 2022 · 2022
Earlier work this paper cites.
Ernie-code: Beyond english-centric cross-lingual pretraining for programming languages
Yekun Chai, Shuohuan Wang, Chao Pang, Yu Sun, Hao Tian, and Hua Wu. 2022 · 2022
Earlier work this paper cites.
Towards multi-lingual visual question answering
Soravit Changpinyo, Linting Xue, Idan Szpektor, Ashish V Thapliyal, Julien Amelot, Xi Chen, and Radu Soricut. 2022 · 2022
Earlier work this paper cites.
Pali: A jointly-scaled multilingual language-image model
Xi Chen, Xiao Wang, Soravit Changpinyo, AJ Piergiovanni, Piotr Padlewski, Daniel Salz, Sebastian Goodman, Adam Grycner, Basil Mustafa, Lucas Beyer, et al. 2022 · 2022
Earlier work this paper cites.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al. 2022 · 2022
Earlier work this paper cites.
Knowledge extraction in low-resource scenarios: survey and perspective
Shumin Deng, Ningyu Zhang, Feiyu Xiong, Jeff Z Pan, and Huajun Chen. 2022 · 2022
Earlier work this paper cites.
A survey for in-context learning
Qingxiu Dong, Lei Li, Damai Dai, Ce Zheng, Zhiyong Wu, Baobao Chang, Xu Sun, Jingjing Xu, and Zhifang Sui. 2022 · 2022
Earlier work this paper cites.
Jack FitzGerald, Christopher Hench, Charith Peris, Scott Mackie, Kay Rottmann, Ana Sanchez, Aaron Nash, Liam Urbach, Vishesh Kakarala, Richa Singh, et al. 2022 · 2022
Earlier work this paper cites.
Polyglot prompt: Multilingual multitask promptraining
Jinlan Fu, See-Kiong Ng, and Pengfei Liu. 2022 · 2022
Earlier work this paper cites.
Normsage: Multi-lingual multi-cultural norm discovery from conversations on-the-fly
Yi R Fung, Tuhin Chakraborty, Hao Guo, Owen Rambow, Smaranda Muresan, and Heng Ji. 2022 · 2022
Earlier work this paper cites.
Roscoe: A suite of metrics for scoring step-by-step reasoning
O. Yu. Golovneva, Moya Chen, Spencer Poff, Martin Corredor, Luke Zettlemoyer, Maryam Fazel-Zarandi, and Asli Celikyilmaz. 2022 · 2022
Earlier work this paper cites.
The flores-101 evaluation benchmark for low-resource and multilingual machine translation
Naman Goyal, Cynthia Gao, Vishrav Chaudhary, Peng-Jen Chen, Guillaume Wenzek, Da Ju, Sanjana Krishnan, Marc’Aurelio Ranzato, Francisco Guzmán, and Angela Fan. 2022 · 2022
Earlier work this paper cites.
Do multilingual language models capture differing moral norms?
Katharina Hämmerl, Björn Deiseroth, Patrick Schramowski, Jindřich Libovickỳ, Alexander Fraser, and Kristian Kersting. 2022 · 2022
Earlier work this paper cites.
Challenges and strategies in cross-cultural nlp
Daniel Hershcovich, Stella Frank, Heather Lent, Miryam de Lhoneux, Mostafa Abdou, Stephanie Brandl, Emanuele Bugliarello, Laura Cabello Piqueras, Ilias Chalkidis, Ruixiang Cui, et al. 2022 · 2022
Earlier work this paper cites.
Yeskendir Koishekenov, Vassilina Nikoulina, and Alexandre Berard. 2022 · 2022
Earlier work this paper cites.
The bigscience roots corpus: A 1.6 tb composite multilingual dataset
Hugo Laurençon, Lucile Saulnier, Thomas Wang, Christopher Akiki, Albert Villanova del Moral, Teven Le Scao, Leandro Von Werra, Chenghao Mou, Eduardo González Ponferrada, Huu Nguyen, et al. 2022 · 2022
Earlier work this paper cites.
Few-shot learning with multilingual generative language models
Xi Victoria Lin, Todor Mihaylov, Mikel Artetxe, Tianlu Wang, Shuohui Chen, Daniel Simig, Myle Ott, Naman Goyal, Shruti Bhosale, Jingfei Du, Ramakanth Pasunuru, Sam Shleifer, Punit Singh Koura, Vishrav Chaudhary, Brian O’Horo, Jeff Wang, Luke Zettlemoyer, Zornitsa Kozareva, Mona Diab, Veselin Stoyanov, and Xian Li. 2022a · 2022
Earlier work this paper cites.
Few-shot learning with multilingual generative language models
Xi Victoria Lin, Todor Mihaylov, Mikel Artetxe, Tianlu Wang, Shuohui Chen, Daniel Simig, Myle Ott, Naman Goyal, Shruti Bhosale, Jingfei Du, et al. 2022b · 2022
Earlier work this paper cites.
A balanced data approach for evaluating cross-lingual transfer: Mapping the linguistic blood bank
Dan Malkin, Tomasz Limisiewicz, and Gabriel Stanovsky. 2022 · 2022
Earlier work this paper cites.
Multiconer: a large-scale multilingual dataset for complex named entity recognition
Shervin Malmasi, Anjie Fang, Besnik Fetahu, Sudipta Kar, and Oleg Rokhlenko. 2022 · 2022
Earlier work this paper cites.
Rethinking the role of demonstrations: What makes in-context learning work?
Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe, Mike Lewis, Hannaneh Hajishirzi, and Luke Zettlemoyer. 2022 · 2022
Earlier work this paper cites.
Crosslingual generalization through multitask finetuning
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, Adam Roberts, Stella Biderman, Teven Le Scao, M Saiful Bari, Sheng Shen, Zheng-Xin Yong, Hailey Schoelkopf, et al. 2022 · 2022
Earlier work this paper cites.
Evaluating byte and wordpiece level models for massively multilingual semantic parsing
Massimo Nicosia and Francesco Piccinno. 2022 · 2022
Earlier work this paper cites.
Chatgpt
OpenAI. 2022 · 2022
Earlier work this paper cites.
Bidirectional language models are also few-shot learners
Ajay Patel, Bryan Li, Mohammad Sadegh Rasooli, Noah Constant, Colin Raffel, and Chris Callison-Burch. 2022 · 2022
Earlier work this paper cites.
Libo Qin, Qiguang Chen, Tianbao Xie, Qixin Li, Jian-Guang Lou, Wanxiang Che, and Min-Yen Kan. 2022 · 2022
Earlier work this paper cites.
Evgeniia Razumovskaia, Joshua Maynez, Annie Louis, Mirella Lapata, and Shashi Narayan. 2022 · 2022
Earlier work this paper cites.
Clasp: Few-shot cross-lingual data augmentation for semantic parsing
Andy Rosenbaum, Saleh Soltan, Wael Hamza, Amir Saffari, Marco Damonte, and Isabel Groves. 2022a · 2022
Earlier work this paper cites.
Language modelling with pixels
Phillip Rust, Jonas F Lotz, Emanuele Bugliarello, Elizabeth Salesky, Miryam de Lhoneux, and Desmond Elliott. 2022 · 2022
Earlier work this paper cites.
mgpt: Few-shot learners go multilingual
Oleh Shliazhko, Alena Fenogenova, Maria Tikhonova, Vladislav Mikhailov, Anastasia Kozlova, and Tatiana Shavrina. 2022 · 2022
Earlier work this paper cites.
Alexatm 20b: Few-shot learning using a large-scale multilingual seq2seq model
Saleh Soltan, Shankar Ananthakrishnan, Jack FitzGerald, Rahul Gupta, Wael Hamza, Haidar Khan, Charith Peris, Stephen Rawls, Andy Rosenbaum, Anna Rumshisky, et al. 2022 · 2022
Earlier work this paper cites.
Sling: Sino linguistic evaluation of large language models
Yixiao Song, Kalpesh Krishna, Rajesh Bhatt, and Mohit Iyyer. 2022 · 2022
Earlier work this paper cites.
Tydip: A dataset for politeness classification in nine typologically diverse languages
Anirudh Srinivasan and Eunsol Choi. 2022 · 2022
Earlier work this paper cites.
Welm: A well-read pre-trained language model for chinese
Hui Su, Xiao Zhou, Houjin Yu, Xiaoyu Shen, Yuwen Chen, Zilin Zhu, Yang Yu, and Jie Zhou. 2022 · 2022
Earlier work this paper cites.
Crossmodal-3600: A massively multilingual multimodal evaluation dataset
Ashish V Thapliyal, Jordi Pont-Tuset, Xi Chen, and Radu Soricut. 2022 · 2022
Earlier work this paper cites.
Enhancing natural language inference of cross-lingual n-shot transfer with multilingual data
Kuang Tseng and Chow-Sing Lin. 2022 · 2022
Earlier work this paper cites.
Prompting palm for translation: Assessing strategies and performance
David Vilar, Markus Freitag, Colin Cherry, Jiaming Luo, Viresh Ratnakar, and George Foster. 2022 · 2022
Earlier work this paper cites.
Overcoming catastrophic forgetting in zero-shot cross-lingual generation
Tu Vu, Aditya Barua, Brian Lester, Daniel Cer, Mohit Iyyer, and Noah Constant. 2022 · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Earlier work this paper cites.
Bloom: A 176b-parameter open-access multilingual language model
BigScience Workshop, Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, et al. 2022 · 2022
Earlier work this paper cites.
Byt5: Towards a token-free future with pre-trained byte-to-byte models
Linting Xue, Aditya Barua, Noah Constant, Rami Al-Rfou, Sharan Narang, Mihir Kale, Adam Roberts, and Colin Raffel. 2022 · 2022
Earlier work this paper cites.
Geomlama: Geo-diverse commonsense probing on multilingual pre-trained language models
Da Yin, Hritik Bansal, Masoud Monajatipoor, Liunian Harold Li, and Kai-Wei Chang. 2022 · 2022
Earlier work this paper cites.
Bloom+ 1: Adding language support to bloom for zero-shot prompting
Zheng-Xin Yong, Hailey Schoelkopf, Niklas Muennighoff, Alham Fikri Aji, David Ifeoluwa Adelani, Khalid Almubarak, M Saiful Bari, Lintang Sutawika, Jungo Kasai, Ahmed Baruwa, et al. 2022 · 2022
Cited alongside, same era.
Beyond counting datasets: a survey of multilingual dataset construction and necessary resources
Xinyan Velocity Yu, Akari Asai, Trina Chatterjee, Junjie Hu, and Eunsol Choi. 2022 · 2022
Cited alongside, same era.
Universal dependencies 2.10
Daniel Zeman, Joakim Nivre, Mitchell Abrams, Elia Ackermann, Noëmi Aepli, Hamid Aghaei, Željko Agić, Amir Ahmadi, et al. 2022 · 2022
Cited alongside, same era.
Glm-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al. 2022 · 2022
Cited alongside, same era.
Multitude: Large-scale multilingual machine-generated text detection benchmark
Dominik Macko, Robert Moro, Adaku Uchendu, Jason Lucas, Michiharu Yamashita, Matúš Pikuliak, Ivan Srba, Thai Le, Dongwon Lee, Jakub Simko, et al. 2023 · 2023
Later among the works it cites.
Multilingual bias detection and mitigation for indian languages
Ankita Maity, Anubhav Sharma, Rudra Dhar, Tushar Abhishek, Manish Gupta, and Vasudeva Varma. 2023 · 2023
Later among the works it cites.
Structural priming demonstrates abstract grammatical representations in multilingual language models
James Michaelov, Catherine Arnett, Tyler Chang, and Ben Bergen. 2023 · 2023
Later among the works it cites.
Lost in translation, found in spans: Identifying claims in multilingual social media
Shubham Mittal, Megha Sundriyal, and Preslav Nakov. 2023 · 2023
Later among the works it cites.
X-RiSAWOZ: High-quality end-to-end multilingual dialogue datasets and few-shot agents
Mehrad Moradshahi, Tianhao Shen, Kalika Bali, Monojit Choudhury, Gael de Chalendar, Anmol Goel, Sungkyun Kim, Prashant Kodali, Ponnurangam Kumaraguru, Nasredine Semmar, Sina Semnani, Jiwon Seo, Vivek Seshadri, Manish Shrivastava, Michael Sun, Aditya Yadavalli, Chaobin You, Deyi Xiong, and Monica Lam. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al. 2022 · 2022
Cited alongside, same era.
HIT-SCIR at MMNLU-22: Consistency regularization for multilingual spoken language understanding
Bo Zheng, Zhouyang Li, Fuxuan Wei, Qiguang Chen, Libo Qin, and Wanxiang Che. 2022 · 2022
Cited alongside, same era.
Benchmarking arabic ai with large language models
Ahmed Abdelali, Hamdy Mubarak, Shammur Absar Chowdhury, Maram Hasanain, Basel Mousi, Sabri Boughorbel, Yassine El Kheir, Daniel Izham, Fahim Dalvi, Majd Hawasly, et al. 2023 · 2023
Cited alongside, same era.
Jasmine: Arabic gpt models for few-shot learning
Muhammad Abdul-Mageed, Abdelrahim Elmadany, Alcides Inciarte, Md Tawkat Islam Khondaker, et al. 2023 · 2023
Cited alongside, same era.
Masakhanews: News topic classification for african languages
David Ifeoluwa Adelani, Marek Masiak, Israel Abebe Azime, Jesujoba Oluwadara Alabi, Atnafu Lambebo Tonja, Christine Mwase, Odunayo Ogundepo, Bonaventure FP Dossou, Akintunde Oladipo, Doreen Nixdorf, et al. 2023 · 2023
Cited alongside, same era.
Zero-shot cross-lingual reranking with large language models for low-resource languages
Mofetoluwa Adeyemi, Akintunde Oladipo, Ronak Pradeep, and Jimmy Lin. 2023 · 2023
Cited alongside, same era.
Findings of the iwslt 2023 evaluation campaign
Milind Agarwal, Sweta Agarwal, Antonios Anastasopoulos, Luisa Bentivogli, Ondřej Bojar, Claudia Borg, Marine Carpuat, Roldano Cattoni, Mauro Cettolo, Mingda Chen, et al. 2023 · 2023
Cited alongside, same era.
Multilingual summarization with factual consistency evaluation
Roee Aharoni, Shashi Narayan, Joshua Maynez, Jonathan Herzig, Elizabeth Clark, and Mirella Lapata. 2023 · 2023
Cited alongside, same era.
Later among the works it cites.
Aligning neural machine translation models: Human feedback in training and inference
Miguel Moura Ramos, Patrick Fernandes, António Farinhas, and André FT Martins. 2023 · 2023
Later among the works it cites.
Afrisenti: A twitter sentiment analysis benchmark for african languages
Shamsuddeen Hassan Muhammad, Idris Abdulmumin, Abinew Ali Ayele, Nedjma Ousidhoum, David Ifeoluwa Adelani, Seid Muhie Yimam, Ibrahim Sa’id Ahmad, Meriem Beloucif, Saif Mohammad, Sebastian Ruder, et al. 2023 · 2023
Later among the works it cites.
Assessing translation capabilities of large language models involving english and indian languages
Vandan Mujadia, Ashok Urlana, Yash Bhaskar, Penumalla Aditya Pavani, Kukkapalli Shravya, Parameswari Krishnamurthy, and Dipti Misra Sharma. 2023 · 2023
Later among the works it cites.
Evaluating and modeling attribution for cross-lingual question answering
Benjamin Muller, John Wieting, Jonathan H Clark, Tom Kwiatkowski, Sebastian Ruder, Livio Baldini Soares, Roee Aharoni, Jonathan Herzig, and Xinyi Wang. 2023 · 2023
Later among the works it cites.
Cross-lingual transfer of large language model by visually-derived supervision toward low-resource languages
Masayasu Muraoka, Bishwaranjan Bhattacharjee, Michele Merler, Graeme Blackwood, Yulong Li, and Yang Zhao. 2023 · 2023
Later among the works it cites.
Breaking language barriers with a leap: Learning strategies for polyglot llms
Akshay Nambi, Vaibhav Balloli, Mercy Ranjit, Tanuja Ganu, Kabir Ahuja, Sunayana Sitaram, and Kalika Bali. 2023 · 2023
Later among the works it cites.
CoF-CoT: Enhancing large language models with coarse-to-fine chain-of-thought prompting for multi-domain NLU tasks
Hoang Nguyen, Ye Liu, Chenwei Zhang, Tao Zhang, and Philip Yu. 2023a · 2023
Later among the works it cites.
Evaluating task understanding through multilingual consistency: A chatgpt case study
Xenia Ohmer, Elia Bruni, and Dieuwke Hupkes. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
A preliminary evaluation of chatgpt for zero-shot dialogue understanding
Wenbo Pan, Qiguang Chen, Xiao Xu, Wanxiang Che, and Libo Qin. 2023 · 2023
Later among the works it cites.
On the analysis of cross-lingual prompt tuning for decoder-based multilingual model
Nohil Park, Joonsuk Park, Kang Min Yoo, and Sungroh Yoon. 2023 · 2023
Later among the works it cites.
Baolin Peng, Chunyuan Li, Pengcheng He, Michel Galley, and Jianfeng Gao. 2023 · 2023
Later among the works it cites.
Document-level language models for machine translation
Frithjof Petrick, Christian Herold, Pavel Petrushkov, Shahram Khadivi, and Hermann Ney. 2023 · 2023
Later among the works it cites.
Language model tokenizers introduce unfairness between languages
Aleksandar Petrov, Emanuele La Malfa, Philip HS Torr, and Adel Bibi. 2023 · 2023
Later among the works it cites.
mmt5: Modular multilingual pre-training solves source language hallucinations
Jonas Pfeiffer, Francesco Piccinno, Massimo Nicosia, Xinyi Wang, Machel Reid, and Sebastian Ruder. 2023 · 2023
Later among the works it cites.
Fred Philippy, Siwen Guo, and Shohreh Haddadan. 2023 · 2023
Later among the works it cites.
Jonathan Pilault, Xavier Garcia, Arthur Bražinskas, and Orhan Firat. 2023 · 2023
Later among the works it cites.
Sabi \ \backslash ’a: Portuguese large language models
Ramon Pires, Hugo Abonizio, Thales Rogério, and Rodrigo Nogueira. 2023 · 2023
Later among the works it cites.
Decomt: Decomposed prompting for machine translation between related languages using large language models
Ratish Puduppully, Anoop Kunchukuttan, Raj Dabre, Aiti Aw, and Nancy Chen. 2023b · 2023
Later among the works it cites.
Comprehensive evaluation of chatgpt reliability through multilingual inquiries
Poorna Chander Reddy Puttaparthi, Soham Sanjay Deo, Hakan Gul, Yiming Tang, Weiyi Shang, and Zhe Yu. 2023 · 2023
Later among the works it cites.
Cross-lingual consistency of factual knowledge in multilingual language models
Jirui Qi, Raquel Fernández, and Arianna Bisazza. 2023 · 2023
Later among the works it cites.
Detecting and mitigating hallucinations in multilingual summarisation
Yifu Qiu, Yftah Ziser, Anna Korhonen, Edoardo M Ponti, and Shay B Cohen. 2023 · 2023
Later among the works it cites.
Fairness in language models beyond english: Gaps and challenges
Krithika Ramesh, Sunayana Sitaram, and Monojit Choudhury. 2023 · 2023
Later among the works it cites.
Lmcap: Few-shot multilingual image captioning by retrieval augmented language model prompting
Rita Ramos, Bruno Martins, and Desmond Elliott. 2023 · 2023
Later among the works it cites.
Leveraging gpt-4 for automatic translation post-editing
Vikas Raunak, Amr Sharaf, Hany Hassan Awadallah, and Arul Menezes. 2023 · 2023
Later among the works it cites.
Neural machine translation models can learn to be few-shot learners
Raphael Reinauer, Patrick Simianer, Kaden Uhlig, Johannes EM Mosig, and Joern Wuebker. 2023 · 2023
Later among the works it cites.
X-parade: Cross-lingual textual entailment and information divergence across paragraphs
Juan Diego Rodriguez, Katrin Erk, and Greg Durrett. 2023 · 2023
Later among the works it cites.
Revisiting non-english text simplification: A unified multilingual benchmark
Michael J Ryan, Tarek Naous, and Wei Xu. 2023 · 2023
Later among the works it cites.
Gender-specific machine translation with large language models
Eduardo Sánchez, Pierre Andrews, Pontus Stenetorp, Mikel Artetxe, and Marta R Costa-jussà. 2023 · 2023
Later among the works it cites.
Camoscio: An italian instruction-tuned llama
Andrea Santilli and Emanuele Rodolà. 2023 · 2023
Later among the works it cites.
Cross-lingual supervision improves large language models pre-training
Andrea Schioppa, Xavier Garcia, and Orhan Firat. 2023 · 2023
Later among the works it cites.
Polyglot or not? measuring multilingual encyclopedic knowledge in foundation models
Tim Schott, Daniel Furman, and Shreshta Bhat. 2023 · 2023
Later among the works it cites.
Neha Sengupta, Sunil Kumar Sahu, Bokang Jia, Satheesh Katipomu, Haonan Li, Fajri Koto, Osama Mohammed Afzal, Samta Kamboj, Onkar Pandit, Rahul Pal, et al. 2023 · 2023
Later among the works it cites.
Sharegpt
ShareGPT. 2023 · 2023
Later among the works it cites.
Anti-lm decoding for zero-shot in-context machine translation
Suzanna Sia, Alexandra DeLucia, and Kevin Duh. 2023 · 2023
Later among the works it cites.
Hae-rae bench: Evaluation of korean knowledge in language models
Guijin Son, Hanwool Lee, Suwan Kim, Jaecheol Lee, Je Won Yeom, Jihyu Jung, Jung Woo Kim, and Songseong Kim. 2023 · 2023
Later among the works it cites.
Holistic inter-annotator agreement and corpus coherence estimation in a large-scale multilingual annotation campaign
Nicolas Stefanovitch and Jakub Piskorski. 2023 · 2023
Later among the works it cites.
A multi-dimensional evaluation of tokenizer-free multilingual pretrained models
Jimin Sun, Patrick Fernandes, Xinyi Wang, and Graham Neubig. 2023a · 2023
Later among the works it cites.
Multilingual llms are better cross-lingual in-context learners with alignment
Eshaan Tanwar, Manish Borthakur, Subhabrata Dutta, and Tanmoy Chakraborty. 2023 · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al. 2023 · 2023
Later among the works it cites.
Interleaving retrieval with chain-of-thought reasoning for knowledge-intensive multi-step questions
Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2023 · 2023
Later among the works it cites.
Efficiently aligned cross-lingual transfer learning for conversational tasks using prompt-tuning
Lifu Tu, Jin Qu, Semih Yavuz, Shafiq Joty, Wenhao Liu, Caiming Xiong, and Yingbo Zhou. 2023 · 2023
Later among the works it cites.
Anytext: Multilingual visual text generation and editing
Yuxiang Tuo, Wangmeng Xiang, Jun-Yan He, Yifeng Geng, and Xuansong Xie. 2023 · 2023
Later among the works it cites.
Bibek Upadhayay and Vahid Behzadan. 2023 · 2023
Later among the works it cites.
Pmindiasum: Multilingual and cross-lingual headline summarization for languages in india
Ashok Urlana, Pinzhen Chen, Zheng Zhao, Shay B Cohen, Manish Shrivastava, and Barry Haddow. 2023 · 2023
Later among the works it cites.
mlongt5: A multilingual and efficient text-to-text transformer for longer sequences
David Uthus, Santiago Ontañón, Joshua Ainslie, and Mandy Guo. 2023 · 2023
Later among the works it cites.
Large scale multi-lingual multi-modal summarization dataset
Yash Verma, Anubhav Jangra, Raghvendra Kumar, and Sriparna Saha. 2023 · 2023
Later among the works it cites.
Machine translation for ge’ez language
Aman Kassahun Wassie. 2023 · 2023
Later among the works it cites.
Counting the bugs in ChatGPT’s wugs: A multilingual investigation into the morphological capabilities of a large language model
Leonie Weissweiler, Valentin Hofmann, Anjali Kantharuban, Anna Cai, Ritam Dutt, Amey Hengle, Anubha Kabra, Atharva Kulkarni, Abhishek Vijayakumar, Haofei Yu, Hinrich Schuetze, Kemal Oflazer, and David Mortensen. 2023 · 2023
Later among the works it cites.
Hyperpolyglot llms: Cross-lingual interpretability in token embeddings
Andrea Wen-Yi and David Mimno. 2023 · 2023
Later among the works it cites.
Eva-kellm: A new benchmark for evaluating knowledge editing of llms
Suhang Wu, Minlong Peng, Yue Chen, Jinsong Su, and Mingming Sun. 2023 · 2023
Later among the works it cites.
Exploring prompt engineering with gpt language models for document-level machine translation: Insights and findings
Yangjian Wu and Gang Hu. 2023 · 2023
Later among the works it cites.
Task-agnostic low-rank adapters for unseen english dialects
Zedian Xiao, William Held, Yanchen Liu, and Diyi Yang. 2023 · 2023
Later among the works it cites.
Language representation projection: Can we transfer factual knowledge across languages in multilingual language models?
Shaoyang Xu, Junzhuo Li, and Deyi Xiong. 2023d · 2023
Later among the works it cites.
Vnhsge: Vietnamese high school graduation examination dataset for large language models
Dao Xuan-Quy, Le Ngoc-Bich, Vo The-Duy, Phan Xuan-Dung, Ngo Bac-Bien, Nguyen Van-Tien, Nguyen Thi-My-Thanh, and Nguyen Hong-Phuoc. 2023 · 2023
Later among the works it cites.
Lahm: Large annotated dataset for multi-domain and multilingual hate speech identification
Ankit Yadav, Shubham Chandel, Sushant Chatufale, and Anil Bandhakavi. 2023 · 2023
Later among the works it cites.
Prompting multilingual large language models to generate code-mixed texts: The case of south east asian languages
Zheng-Xin Yong, Ruochen Zhang, Jessica Zosa Forde, Skyler Wang, Samuel Cahyawijaya, Holy Lovenia, Genta Indra Winata, Lintang Sutawika, Jan Christian Blaise Cruz, Long Phan, et al. 2023 · 2023
Later among the works it cites.
Lego-mt: Learning detachable models for massively multilingual machine translation
Fei Yuan, Yinquan Lu, Wenhao Zhu, Lingpeng Kong, Lei Li, Yu Qiao, and Jingjing Xu. 2023 · 2023
Later among the works it cites.
Yan Gong Yiping Peng Qiang Niu Lei Zhang Baochang Ma Xiangang Li Yunjie Ji, Yong Deng. 2023 · 2023
Later among the works it cites.
Crocosum: A benchmark dataset for cross-lingual code-switched summarization
Ruochen Zhang and Carsten Eickhoff. 2023 · 2023
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al. 2023 · 2023
Later among the works it cites.
Rc3: Regularized contrastive cross-lingual cross-modal pre-training
Chulun Zhou, Yunlong Liang, Fandong Meng, Jinan Xu, Jinsong Su, and Jie Zhou. 2023 · 2023
Later among the works it cites.
Red ai? inconsistent responses from gpt3. 5 models on political issues in the us and china
Di Zhou and Yinxian Zhang. 2023 · 2023
Later among the works it cites.
Accessible instruction-following agent
Kairui Zhou. 2023 · 2023
Later among the works it cites.
Extrapolating large language models to non-english by aligning languages
Wenhao Zhu, Yunzhe Lv, Qingxiu Dong, Fei Yuan, Jingjing Xu, Shujian Huang, Lingpeng Kong, Jiajun Chen, and Lei Li. 2023 · 2023
Later among the works it cites.
Maple: Multilingual evaluation of parameter efficient finetuning of large language models
Divyanshu Aggarwal, Ashutosh Sathe, and Sunayana Sitaram. 2024 · 2024
Closest in time.
Syed Rameel Ahmad. 2024 · 2024
Closest in time.
Cross-lingual editing in multilingual language models
Himanshu Beniwal, Mayank Singh, et al. 2024 · 2024
Closest in time.
Breaking the curse of multilinguality with cross-lingual expert language models
Terra Blevins, Tomasz Limisiewicz, Suchin Gururangan, Margaret Li, Hila Gonen, Noah A Smith, and Luke Zettlemoyer. 2024 · 2024
Closest in time.
xcot: Cross-lingual instruction tuning for cross-lingual chain-of-thought reasoning
Linzheng Chai, Jian Yang, Tao Sun, Hongcheng Guo, Jiaheng Liu, Bing Wang, Xiannian Liang, Jiaqi Bai, Tongliang Li, Qiyao Peng, et al. 2024 · 2024
Closest in time.
Orion-14b: Open-source multilingual large language models
Du Chen, Yi Huang, Xiaopu Li, Yongqiang Li, Yongqiang Liu, Haihui Pan, Leichao Xu, Dacheng Zhang, Zhipeng Zhang, and Kun Han. 2024 · 2024
Closest in time.
Towards boosting many-to-many multilingual machine translation with large language models
Pengzhi Gao, Zhongjun He, Hua Wu, and Haifeng Wang. 2024 · 2024
Closest in time.
Introducing bode: A fine-tuned large language model for portuguese prompt-based task
Gabriel Lino Garcia, Pedro Henrique Paiola, Luis Henrique Morelli, Giovani Candido, Arnaldo Cândido Júnior, Danilo Samuel Jodas, Luis Afonso, Ivan Rizzo Guilherme, Bruno Elias Penteado, and João Paulo Papa. 2024 · 2024
Closest in time.
Chinese-mixtral-8x7b: An open-source mixture-of-experts llm
HIT-SCIR. 2024 · 2024
Closest in time.
Songbo Hu, Xiaobin Wang, Zhangdie Yuan, Anna Korhonen, and Ivan Vulić. 2024 · 2024
Closest in time.
W Ronny Huang, Cyril Allauzen, Tongzhou Chen, Kilol Gupta, Ke Hu, James Qin, Yu Zhang, Yongqiang Wang, Shuo-Yiin Chang, and Tara N Sainath. 2024 · 2024
Closest in time.
Aiqi Jiang and Arkaitz Zubiaga. 2024 · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, et al. 2024 · 2024
Closest in time.
Lampat: Low-rank adaption for multilingual paraphrasing using adversarial training
Khoi M Le, Trinh Pham, Tho Quan, and Anh Tuan Luu. 2024 · 2024
Closest in time.
Authorship obfuscation in multilingual machine-generated text detection
Dominik Macko, Robert Moro, Adaku Uchendu, Ivan Srba, Jason Samuel Lucas, Michiharu Yamashita, Nafis Irtiza Tripto, Dongwon Lee, Jakub Simko, and Maria Bielikova. 2024 · 2024
Closest in time.
Rrubaa Panchendrarajan and Arkaitz Zubiaga. 2024 · 2024
Closest in time.
Nooshin Pourkamali and Shler Ebrahim Sharifi. 2024 · 2024
Closest in time.
The task of post-editing machine translation for the low-resource language
Diana Rakhimova, Aidana Karibayeva, and Assem Turarbek. 2024 · 2024
Closest in time.
Multilingual instruction tuning with just a pinch of multilinguality
Uri Shaham, Jonathan Herzig, Roee Aharoni, Idan Szpektor, Reut Tsarfaty, and Matan Eyal. 2024 · 2024
Closest in time.
Mapo: Advancing multilingual reasoning through multilingual alignment-as-preference optimization
Shuaijie She, Shujian Huang, Wei Zou, Wenhao Zhu, Xiang Liu, Xiang Geng, and Jiajun Chen. 2024 · 2024
Closest in time.
The language barrier: Dissecting safety challenges of llms in multilingual contexts
Lingfeng Shen, Weiting Tan, Sihao Chen, Yunmo Chen, Jingyu Zhang, Haoran Xu, Boyuan Zheng, Philipp Koehn, and Daniel Khashabi. 2024 · 2024
Closest in time.
Climategpt: Towards ai synthesizing interdisciplinary research on climate change
David Thulke, Yingbo Gao, Petrus Pelser, Rein Brune, Rricha Jalota, Floris Fok, Michael Ramos, Ian van Wyk, Abdallah Nasir, Hayden Goldstein, et al. 2024 · 2024
Closest in time.
Turna: A turkish encoder-decoder language model for enhanced understanding and generation
Gökçe Uludoğan, Zeynep Yirmibeşoğlu Balal, Furkan Akkurt, Melikşah Türker, Onur Güngör, and Susan Üsküdarlı. 2024 · 2024
Closest in time.
Don’t rank, combine! combining machine translation hypotheses using quality estimation
Giorgos Vernikos and Andrei Popescu-Belis. 2024 · 2024
Closest in time.
Langbridge: Multilingual reasoning without multilingual supervision
Dongkeun Yoon, Joel Jang, Sungdong Kim, Seungone Kim, Sheikh Shafayat, and Minjoon Seo. 2024 · 2024
Closest in time.
Question translation training for better multilingual reasoning
Wenhao Zhu, Shujian Huang, Fei Yuan, Shuaijie She, Jiajun Chen, and Alexandra Birch. 2024 · 2024
Closest in time.
Quality and quantity of machine translation references for automated metrics
Vilém Zouhar and Ondřej Bojar. 2024 · 2024
Closest in time.
Meep: Is this engaging? prompting large language models for dialogue evaluation in multilingual settings
Amila Ferron, Amber Shore, Ekata Mitra, and Ameeta Agrawal. 2023 · 2078
Closest in time.