Fetching the paper…
Reading the bibliography…
While reasoning and multilingual capabilities in language models (LMs) have achieved remarkable progress in recent years, their integration into a unified paradigm - multilingual reasoning - is at a nascent stage.
Opus: An open source parallel corpus
Jörg Tiedemann. 2012 · 2012
Earlier work this paper cites.
Continuous multilinguality with language vectors
Robert Östling and Jörg Tiedemann. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
A Vaswani. 2017 · 2017
Earlier work this paper cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel R Bowman. 2017 · 2017
Earlier work this paper cites.
Xnli: Evaluating cross-lingual sentence representations
Alexis Conneau, Guillaume Lample, Ruty Rinott, Adina Williams, Samuel R Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Earlier work this paper cites.
Mathqa: Towards interpretable math word problem solving with operation-based formalisms
Aida Amini, Saadia Gabriel, Peter Lin, Rik Koncel-Kedziorski, Yejin Choi, and Hannaneh Hajishirzi. 2019 · 2019
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2020 · 2020
Earlier work this paper cites.
Xcopa: A multilingual dataset for causal commonsense reasoning
Edoardo Maria Ponti, Goran Glavaš, Olga Majewska, Qianchu Liu, Ivan Vulić, and Anna Korhonen. 2020 · 2020
Earlier work this paper cites.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman. 2021 · 2021
Earlier work this paper cites.
Xcsqa: A benchmark for cross-lingual conversational question answering
Yankai Lin, Jiapeng Zhou, Yiming Shen, Wenxuan Zhou, Zhiyuan Liu, Peng Li, Maosong Sun, and Jie Zhou. 2021 · 2021
Earlier work this paper cites.
The flores-200 evaluation benchmark for low-resource and multilingual machine translation
Naman Goyal, Cynthia Gao, Vishrav Chaudhary, Peng-Jen Chen, Guillaume Wenzek, Da Ju, Sanjana Krishnan, Marc’Aurelio Ranzato, Francisco Guzmán, and Angela Fan. 2022 · 2022
Earlier work this paper cites.
A survey on medical document summarization
Raghav Jain, Anubhav Jangra, Sriparna Saha, and Adam Jatowt. 2022 · 2022
Earlier work this paper cites.
Language models are multilingual chain-of-thought reasoners
Freda Shi, Mirac Suzgun, Markus Freitag, Xuezhi Wang, Suraj Srivats, Soroush Vosoughi, Hyung Won Chung, Yi Tay, Sebastian Ruder, Denny Zhou, and 1 others. 2022 · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, and 1 others. 2022 · 2022
Earlier work this paper cites.
Lego-mt: Learning detachable models for massively multilingual machine translation
Fei Yuan, Yinquan Lu, WenHao Zhu, Lingpeng Kong, Lei Li, Yu Qiao, and Jingjing Xu. 2022 · 2022
Earlier work this paper cites.
Breaking language barriers in multilingual mathematical reasoning: Insights and observations
Nuo Chen, Zinan Zheng, Ning Wu, Ming Gong, Dongmei Zhang, and Jia Li. 2023 · 2023
Earlier work this paper cites.
Contrastive chain-of-thought prompting
Yew Ken Chia, Guizhen Chen, Luu Anh Tuan, Soujanya Poria, and Lidong Bing. 2023 · 2023
Earlier work this paper cites.
Okapi: Instruction-tuned large language models in multiple languages with reinforcement learning from human feedback
Viet Dac Lai, Chien Van Nguyen, Nghia Trung Ngo, Thuat Nguyen, Franck Dernoncourt, Ryan A Rossi, and Thien Huu Nguyen. 2023 · 2023
Earlier work this paper cites.
Evaluating multilingual long-context models for retrieval and reasoning
Ameeta Agrawal, Andy Dang, Sina Bagheri Nezhad, Rhitabrat Pokharel, and Russell Scheinberg. 2024 · 2024
Earlier work this paper cites.
sphinx: Sample efficient multilingual instruction fine-tuning through n-shot guided prompting
Sanchit Ahuja, Kumar Tanmay, Hardik Hansrajbhai Chauhan, Barun Patra, Kriti Aggarwal, Luciano Del Corro, Arindam Mitra, Tejas Indulal Dhamecha, Ahmed Awadallah, Monojit Choudhary, Vishrav Chaudhary, and Sunayana Sitaram. 2024 · 2024
Earlier work this paper cites.
Multilingual mathematical reasoning: Advancing open-source llms in hindi and english
Avinash Anand, Kritarth Prasad, Chhavi Kirtani, Ashwin R Nair, Manvendra Kumar Nema, Raj Jaiswal, and Rajiv Ratn Shah. 2024 · 2024
Earlier work this paper cites.
Towards robust knowledge representations in multilingual llms for equivalence and inheritance based consistent reasoning
Gaurav Arora, Srujana Merugu, Shreya Jain, and Vaibhav Saxena. 2024 · 2024
Earlier work this paper cites.
Multilingual llms inherently reward in-language time-sensitive semantic alignment for low-resource languages
Ashutosh Bajpai and Tanmoy Chakraborty. 2024 · 2024
Earlier work this paper cites.
xcot: Cross-lingual instruction tuning for cross-lingual chain-of-thought reasoning, 2024
Linzheng Chai, Jian Yang, Tao Sun, Hongcheng Guo, Jiaheng Liu, Bing Wang, Xiannian Liang, Jiaqi Bai, Tongliang Li, Qiyao Peng, and 1 others. 2024 · 2024
Earlier work this paper cites.
Transfer q star: Principled decoding for llm alignment
Souradip Chakraborty, Soumya Suvra Ghosal, Ming Yin, Dinesh Manocha, Mengdi Wang, Amrit Singh Bedi, and Furong Huang. 2024 · 2024
Earlier work this paper cites.
Exams-v: A multi-discipline multilingual multimodal exam benchmark for evaluating vision language models
Rocktim Jyoti Das, Simeon Emilov Hristov, Haonan Li, Dimitar Iliyanov Dimitrov, Ivan Koychev, and Preslav Nakov. 2024 · 2024
Earlier work this paper cites.
Why not transform chat large language models to non-english?
Xiang Geng, Ming Zhu, Jiahuan Li, Zhejian Lai, Wei Zou, Shuaijie She, Jiaxin Guo, Xiaofeng Zhao, Yinglu Li, Yuang Li, and 1 others. 2024 · 2024
Earlier work this paper cites.
Healthalignsumm: Utilizing alignment for multimodal summarization of code-mixed healthcare dialogues
Akash Ghosh, Arkadeep Acharya, Sriparna Saha, Gaurav Pandey, Dinesh Raghu, and Setu Sinha. 2024d · 2024
Earlier work this paper cites.
M-rewardbench: Evaluating reward models in multilingual settings
Srishti Gureja, Lester James V Miranda, Shayekh Bin Islam, Rishabh Maheshwary, Drishti Sharma, Gusti Winata, Nathan Lambert, Sebastian Ruder, Sara Hooker, and Marzieh Fadaee. 2024 · 2024
Cited alongside, same era.
Mexa: Multilingual evaluation of english-centric llms via cross-lingual alignment
Amir Hossein Kargaran, Ali Modarressi, Nafiseh Nikeghbal, Jana Diesner, François Yvon, and Hinrich Schütze. 2024 · 2024
Cited alongside, same era.
Do moral judgment and reasoning capability of llms change with language? a study using the multilingual defining issues test
Aditi Khandelwal, Utkarsh Agarwal, Kumar Tanmay, and Monojit Choudhury. 2024 · 2024
Cited alongside, same era.
Args: Alignment as reward-guided search
Maxim Khanov, Jirayu Burapacheep, and Yixuan Li. 2024 · 2024
Cited alongside, same era.
mcot: Multilingual instruction tuning for reasoning consistency in language models
Huiyuan Lai and Malvina Nissim. 2024 · 2024
Memla: Enhancing multilingual knowledge editing with neuron-masked low-rank adaptation
Jiakuan Xie, Pengfei Cao, Yuheng Chen, Yubo Chen, Kang Liu, and Jun Zhao. 2024 · 2024
Later among the works it cites.
Cruxeval-x: A benchmark for multilingual code reasoning, understanding and execution
Ruiyang Xu, Jialun Cao, Yaojie Lu, Hongyu Lin, Xianpei Han, Ben He, Shing-Chi Cheung, and Le Sun. 2024 · 2024
Later among the works it cites.
Famma: A benchmark for financial domain multilingual multimodal question answering
Siqiao Xue, Tingting Chen, Fan Zhou, Qingyang Dai, Zhixuan Chu, and Hongyuan Mei. 2024 · 2024
Later among the works it cites.
Language imbalance driven rewarding for multilingual self-improving
Wen Yang, Junhong Wu, Chen Wang, Chengqing Zong, and Jiajun Zhang. 2024 · 2024
Later among the works it cites.
Hdflow: Enhancing llm complex problem-solving with hybrid thinking and dynamic workflows
Wenlin Yao, Haitao Mi, and Dong Yu. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Mufu: Multilingual fused learning for low-resource translation with llm
Zheng Wei Lim, Nitish Gupta, Honglin Yu, and Trevor Cohn. 2024 · 2024
Cited alongside, same era.
Is translation all you need? a study on solving multilingual tasks with large language models
Chaoqun Liu, Wenxuan Zhang, Yiran Zhao, Anh Tuan Luu, and Lidong Bing. 2024 · 2024
Cited alongside, same era.
On the impact of fine-tuning on chain-of-thought reasoning
Elita Lobo, Chirag Agarwal, and Himabindu Lakkaraju. 2024 · 2024
Cited alongside, same era.
Dictionary insertion prompting for multilingual reasoning on multilingual large language models
Hongyuan Lu, Zixuan Li, and Wai Lam. 2024 · 2024
Cited alongside, same era.
Python is not always the best choice: Embracing multilingual program of thoughts
Xianzhen Luo, Qingfu Zhu, Zhiming Zhang, Libo Qin, Xuanyu Zhang, Qing Yang, Dongliang Xu, and Wanxiang Che. 2024 · 2024
Cited alongside, same era.
Regional bias in monolingual english language models
Jiachen Lyu, Katharina Dost, Yun Sing Koh, and Jörg Wicker. 2024 · 2024
Cited alongside, same era.
Can llms learn by teaching for better reasoning? a preliminary study
Xuefei Ning, Zifu Wang, Shiyao Li, Zinan Lin, Peiran Yao, Tianyu Fu, Matthew B Blaschko, Guohao Dai, Huazhong Yang, and Yu Wang. 2024 · 2024
Cited alongside, same era.
Later among the works it cites.
Langbridge: Multilingual reasoning without multilingual supervision
Dongkeun Yoon, Joel Jang, Sungdong Kim, Seungone Kim, Sheikh Shafayat, and Minjoon Seo. 2024 · 2024
Later among the works it cites.
Question translation training for better multilingual reasoning
Wenhao Zhu, Shujian Huang, Fei Yuan, Shuaijie She, Jiajun Chen, and Alexandra Birch. 2024c · 2024
Later among the works it cites.
Multilingual mathematical reasoning: Advancing open-source llms in hindi and english
Avinash Anand, Kritarth Prasad, Chhavi Kirtani, Ashwin R Nair, Manvendra Kumar Nema, Raj Jaiswal, and Rajiv Ratn Shah. 2025 · 2025
Closest in time.
Multilingual llms inherently reward in-language time-sensitive semantic alignment for low-resource languages
Ashutosh Bajpai and Tanmoy Chakraborty. 2025 · 2025
Closest in time.
Thinking machines: A survey of llm based reasoning strategies
Dibyanayan Bandyopadhyay, Soham Bhattacharjee, and Asif Ekbal. 2025 · 2025
Closest in time.
Towards reasoning era: A survey of long chain-of-thought for reasoning large language models
Qiguang Chen, Libo Qin, Jinhao Liu, Dengyun Peng, Jiannan Guan, Peng Wang, Mengkang Hu, Yuhang Zhou, Te Gao, and Wanxiang Che. 2025 · 2025
Closest in time.
Slam: Towards efficient multilingual reasoning via selective language alignment
Yuchun Fan, Yongyu Mu, Yilin Wang, Lei Huang, Junhao Ruan, Bei Li, Tong Xiao, Shujian Huang, Xiaocheng Feng, and Jingbo Zhu. 2025 · 2025
Closest in time.
Pm4bench: A parallel multilingual multi-modal multi-task benchmark for large vision language model
Junyuan Gao, Jiahe Song, Jiang Wu, Runchuan Zhu, Guanlin Shen, Shasha Wang, Xingjian Wei, Haote Yang, Songyang Zhang, Weijia Li, and 1 others. 2025 · 2025
Closest in time.
Relic: Enhancing reward model generalization for low-resource indic languages with few-shot examples
Soumya Suvra Ghosal, Vaibhav Singh, Akash Ghosh, Soumyabrata Pal, Subhadip Baidya, Sriparna Saha, and Dinesh Manocha. 2025 · 2025
Closest in time.
Infogen: Generating complex statistical infographics from documents
Akash Ghosh, Aparna Garimella, Pritika Ramu, Sambaran Bandyopadhyay, and Sriparna Saha. 2025 · 2025
Closest in time.
Pensez: Less data, better reasoning–rethinking french llm
Huy Hoang Ha. 2025 · 2025
Closest in time.
Amey Hengle, Prasoon Bajpai, Soham Dan, and Tanmoy Chakraborty. 2025 · 2025
Closest in time.
Understand, solve and translate: Bridging the multilingual mathematical reasoning gap
Hyunwoo Ko, Guijin Son, and Dasol Choi. 2025 · 2025
Closest in time.
When natural language is not enough: The limits of in-context learning demonstrations in multilingual reasoning
Leonardo Ranaldi, Barry Haddow, and Alexandra Birch. 2025a · 2025
Closest in time.
Multilingual reasoning via self-training
Leonardo Ranaldi and Giulia Pucci. 2025 · 2025
Closest in time.
Zhiwen Ruan, Yixia Li, He Zhu, Longyue Wang, Weihua Luo, Kaifu Zhang, Yun Chen, and Guanhua Chen. 2025 · 2025
Closest in time.
Scaling test-time compute for low-resource languages: Multilingual reasoning in llms
Khanh-Tung Tran, Barry O’Sullivan, and Hoang D Nguyen. 2025 · 2025
Closest in time.
Demystifying multilingual chain-of-thought in process reward modeling
Weixuan Wang, Minghao Wu, Barry Haddow, and Alexandra Birch. 2025 · 2025
Closest in time.
A survey on multilingual large language models: Corpora, alignment, and bias
Yuemei Xu, Ling Hu, Jiayi Zhao, Zihan Qiu, Kexin Xu, Yuqi Ye, and Hanwen Gu. 2025 · 2025
Closest in time.
Mmlu-prox: A multilingual benchmark for advanced large language model evaluation
Weihao Xuan, Rui Yang, Heli Qi, Qingcheng Zeng, Yunze Xiao, Yun Xing, Junjue Wang, Huitao Li, Xin Li, Kunyu Yu, and 1 others. 2025 · 2025
Closest in time.
Mr. guard: Multilingual reasoning guardrail using curriculum learning
Yahan Yang, Soham Dan, Shuo Li, Dan Roth, and Insup Lee. 2025 · 2025
Closest in time.
Crosslingual reasoning through test-time scaling
Zheng-Xin Yong, M Farid Adilazuarda, Jonibek Mansurov, Ruochen Zhang, Niklas Muennighoff, Carsten Eickhoff, Genta Indra Winata, Julia Kreutzer, Stephen H Bach, and Alham Fikri Aji. 2025 · 2025
Closest in time.