Fetching the paper…
Reading the bibliography…
The development of Large Language Models (LLMs) has predominantly focused on high-resource languages, leaving extremely low-resource languages like Irish with limited representation.
Atlas of the world’s languages in danger
2010
Earlier work this paper cites.
Dirk Goldhahn, Thomas Eckart, and Uwe Quasthoff, ‘Building large monolingual dictionaries at the Leipzig corpora collection: From 100 to 200 languages’, in Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC’12)
2012
Earlier work this paper cites.
Colin Raffel, Noam M. Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu, ‘Exploring the limits of transfer learning with a unified text-to-text transformer’, J. Mach. Learn. Res
2019
Earlier work this paper cites.
Marta Bañón, Pinzhen Chen, Barry Haddow, Kenneth Heafield, Hieu Hoang, Miquel Esplà-Gomis, Mikel L. Forcada, Amir Kamran, Faheem Kirefu, Philipp Koehn, Sergio Ortiz Rojas, Leopoldo Pla Sempere, Gema Ramírez-Sánchez, Elsa Sarrías, Marek Strelec, Brian Thompson, William Waites, Dion Wiggins, and Jaume Zaragoza, ‘ParaCrawl: Web-scale acquisition of parallel corpora’, in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics
2020
Earlier work this paper cites.
Seamus Lankford, Haithem Afli, and Andy Way, ‘Machine translation in the covid domain: an English-Irish case study for LoResMT 2021’, in Proceedings of the 4th Workshop on Technologies for MT of Low Resource Languages (LoResMT2021)
2021
Earlier work this paper cites.
James Barry, Joachim Wagner, Lauren Cassidy, Alan Cowap, Teresa Lynn, Abigail Walsh, Mícheál J. Ó Meachair, and Jennifer Foster, ‘gaBERT — an Irish language model’, in Proceedings of the Thirteenth Language Resources and Evaluation Conference
2022
Earlier work this paper cites.
European Language Resource Coordination, AI for Multilingual Europe – Why Language Data Matters
2022
Earlier work this paper cites.
Jordan Hoffmann et al., ‘An empirical analysis of compute-optimal large language model training’, in Advances in Neural Information Processing Systems
2022
Earlier work this paper cites.
Swaroop Mishra, Daniel Khashabi, Chitta Baral, and Hannaneh Hajishirzi, ‘Cross-task generalization via natural language crowdsourcing instructions’, in Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
2022
Earlier work this paper cites.
Long Ouyang et al., ‘Training language models to follow instructions with human feedback’, in Advances in Neural Information Processing Systems
2022
Earlier work this paper cites.
Bloom: A 176b-parameter open-access multilingual language model, 2022
BigScience Workshop et al · 2022
Cited alongside, same era.
A framework for few-shot language model evaluation, 12 2023
Leo Gao et al · 2023
Cited alongside, same era.
Textbooks are all you need, 2023
Suriya Gunasekar, Yi Zhang, Jyoti Aneja, Caio César Teodoro Mendes, Allie Del Giorno, Sivakanth Gopi, Mojan Javaheripi, Piero Kauffmann, Gustavo de Rosa, Olli Saarikivi, Adil Salim, Shital Shah, Harkirat Singh Behl, Xin Wang, Sébastien Bubeck, Ronen Eldan, Adam Tauman Kalai, Yin Tat Lee, and Yuanzhi Li · 2023
Cited alongside, same era.
Ayyoob ImaniGooghari, Peiqin Lin, Amir Hossein Kargaran, Silvia Severini, Masoud Jalili Sabet, Nora Kassner, Chunlan Ma, Helmut Schmid, André Martins, François Yvon, and Hinrich Schütze, ‘Glot500: Scaling multilingual corpora and language models to 500 languages’, in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
2023
Cited alongside, same era.
Hugo Touvron et al., ‘Llama 2: Open foundation and fine-tuned chat models’, (2023)
2023
Later among the works it cites.
Zheng Xin Yong et al., ‘BLOOM+1: Adding language support to BLOOM for zero-shot prompting’, in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
2023
Later among the works it cites.
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, Hao Zhang, Joseph E. Gonzalez, and Ion Stoica, ‘Judging LLM-as-a-judge with MT-bench and chatbot arena’, in Thirty-seventh Conference on Neural Information Processing Systems Datasets and Benchmarks Track
2023
Later among the works it cites.
https://cloud.google.com/translate?hl=en
Cloud Translation — cloud.google.com · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
Risto Luukkonen et al., ‘FinGPT: Large generative models for a small language’, in Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing
2023
Cited alongside, same era.
Simon Mille, Elaine Uí Dhonnchadha, Lauren Cassidy, Brian Davis, Stamatia Dasiopoulou, and Anya Belz, ‘Generating Irish text with a flexible plug-and-play architecture’, in Proceedings of the 2nd Workshop on Pattern-based Approaches to NLP in the Age of Deep Learning
2023
Cited alongside, same era.
Niklas Muennighoff, Alexander M Rush, Boaz Barak, Teven Le Scao, Nouamane Tazi, Aleksandra Piktus, Sampo Pyysalo, Thomas Wolf, and Colin Raffel, ‘Scaling data-constrained language models’, in Thirty-seventh Conference on Neural Information Processing Systems
2023
Cited alongside, same era.
Culturax: A cleaned, enormous, and multilingual dataset for large language models in 167 languages, 2023
Thuat Nguyen, Chien Van Nguyen, Viet Dac Lai, Hieu Man, Nghia Trung Ngo, Franck Dernoncourt, Ryan A. Rossi, and Thien Huu Nguyen · 2023
Cited alongside, same era.
David Adelani, Hannah Liu, Xiaoyu Shen, Nikita Vassilyev, Jesujoba Alabi, Yanke Mao, Haonan Gao, and En-Shiun Lee, ‘SIB-200: A simple, inclusive, and big evaluation dataset for topic classification in 200+ languages and dialects’, in Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers)
2024
Closest in time.
Zhangir Azerbayev, Hailey Schoelkopf, Keiran Paster, Marco Dos Santos, Stephen Marcus McAleer, Albert Q. Jiang, Jia Deng, Stella Biderman, and Sean Welleck, ‘Llemma: An open language model for mathematics’, in The Twelfth International Conference on Learning Representations
2024
Closest in time.
Yupeng Chang, Xu Wang, Jindong Wang, Yuan Wu, Linyi Yang, Kaijie Zhu, Hao Chen, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, Wei Ye, Yue Zhang, Yi Chang, Philip S. Yu, Qiang Yang, and Xing Xie, ‘A survey on evaluation of large language models’, ACM Trans. Intell. Syst. Technol
2024
Closest in time.
Rephrasing the web: A recipe for compute and data-efficient language modeling, 2024
Pratyush Maini, Skyler Seto, He Bai, David Grangier, Yizhe Zhang, and Navdeep Jaitly · 2024
Closest in time.
Haoran Xu, Young Jin Kim, Amr Sharaf, and Hany Hassan Awadalla, ‘A paradigm shift in machine translation: Boosting translation performance of large language models’, in The Twelfth International Conference on Learning Representations
2024
Closest in time.