Fetching the paper…
Reading the bibliography…
In this paper, we analyze the capabilities of the multi-lingual Dense Passage Retriever (mDPR) for extremely low-resource languages.
Extending multilingual bert to low-resource languages
Zihan Wang, Stephen Mayhew, Dan Roth, et al. 2020 · 2004
Earlier work this paper cites.
Bfqa: A bengali factoid question answering system
Somnath Banerjee, Sudip Kumar Naskar, and Sivaji Bandyopadhyay. 2014 · 2014
Earlier work this paper cites.
Simulating clir translation resource scarcity using high-resource languages
Hamed Bonab, James Allan, and Ramesh Sitaraman. 2019 · 2019
Earlier work this paper cites.
Cross-lingual language model pretraining
Alexis Conneau and Guillaume Lample. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Natural questions: A benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M. Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov. 2019 · 2019
Earlier work this paper cites.
Xlda: Cross-lingual data augmentation for natural language inference and question answering
Jasdeep Singh, Bryan McCann, Nitish Shirish Keskar, Caiming Xiong, and Richard Socher. 2019 · 2019
Earlier work this paper cites.
Robust document representations for cross-lingual information retrieval in low-resource settings
Mahsa Yarmohammadi, Xutai Ma, Sorami Hisamoto, Muhammad Rahman, Yiming Wang, Hainan Xu, Daniel Povey, Philipp Koehn, and Kevin Duh. 2019 · 2019
Earlier work this paper cites.
On the cross-lingual transferability of monolingual representations
Mikel Artetxe, Sebastian Ruder, and Dani Yogatama. 2020 · 2020
Earlier work this paper cites.
TyDi QA: A benchmark for information-seeking question answering in typologically diverse languages
Jonathan H. Clark, Eunsol Choi, Michael Collins, Dan Garrette, Tom Kwiatkowski, Vitaly Nikolaev, and Jennimaria Palomaki. 2020 · 2020
Earlier work this paper cites.
CCAligned: A massive collection of cross-lingual web-document pairs
Ahmed El-Kishky, Vishrav Chaudhary, Francisco Guzmán, and Philipp Koehn. 2020 · 2020
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020 · 2020
Earlier work this paper cites.
Avadhan: System for open-domain telugu question answering
Priyanka Ravva, Ashok Urlana, and Manish Shrivastava. 2020 · 2020
Earlier work this paper cites.
One question answering model for many languages with cross-lingual dense passage retrieval
Akari Asai, Xinyan Yu, Jungo Kasai, and Hanna Hajishirzi. 2021 · 2021
Earlier work this paper cites.
Developing a Vietnamese Tourism Question Answering System Using Knowledge Graph and Deep Learning
Phuc Do, Truong H. V. Phan, and Brij B. Gupta. 2021 · 2021
Cited alongside, same era.
Cross-lingual cross-modal pretraining for multimodal retrieval
Hongliang Fei, Tan Yu, and Ping Li. 2021 · 2021
Cited alongside, same era.
Mkqa: A linguistically diverse benchmark for multilingual open domain question answering
Shayne Longpre, Yi Lu, and Joachim Daiber. 2021 · 2021
Cited alongside, same era.
Multilingual BERT post-pretraining alignment
Lin Pan, Chung-Wei Hang, Haode Qi, Abhishek Shah, Saloni Potdar, and Mo Yu. 2021 · 2021
Cited alongside, same era.
Building a Vietnamese Question Answering System Based on Knowledge Graph and Distributed CNN
Trung Phan and Phuc Do. 2021 · 2021
Cited alongside, same era.
Synthetic data augmentation for zero-shot cross-lingual question answering
Africlirmatrix: Enabling cross-lingual information retrieval for african languages
Odunayo Ogundepo, Xinyu Zhang, Shuo Sun, Kevin Duh, and Jimmy Lin. 2022 · 2022
Later among the works it cites.
Lifting the curse of multilinguality by pre-training modular transformers
Jonas Pfeiffer, Naman Goyal, Xi Lin, Xian Li, James Cross, Sebastian Riedel, and Mikel Artetxe. 2022 · 2022
Later among the works it cites.
Improving low-resource question answering with cross-lingual data augmentation strategies
Ryan Pramana and Radityo Eko Prasojo. 2022 · 2022
Later among the works it cites.
TeQuAD:Telugu question answering dataset
Rakesh Vemula, Mani Nuthi, and Manish Srivastava. 2022 · 2022
Later among the works it cites.
Enhancing low-resource languages question answering with syntactic graph
Linjuan Wu, Jiazheng Zhu, Xiaowang Zhang, Zhiqiang Zhuang, and ZhiYong Feng. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Arij Riabi, Thomas Scialom, Rachel Keraron, Benoît Sagot, Djamé Seddah, and Jacopo Staiano. 2021 · 2021
Cited alongside, same era.
iapp_wiki_qa_squad
Kobkrit Viriyayudhakorn and Charin Polpanumas. 2021 · 2021
Cited alongside, same era.
Cross-lingual language model pretraining for retrieval
Puxuan Yu, Hongliang Fei, and Ping Li. 2021 · 2021
Cited alongside, same era.
MIA 2022 shared task: Evaluating cross-lingual open-retrieval question answering for 16 diverse languages
Akari Asai, Shayne Longpre, Jungo Kasai, Chia-Hsuan Lee, Rui Zhang, Junjie Hu, Ikuya Yamada, Jonathan H. Clark, and Eunsol Choi. 2022 · 2022
Cited alongside, same era.
An improvement of bengali factoid question answering system using unsupervised statistical methods
Arijit Das, Jaydeep Mandal, Zargham Danial, Alok Ranjan Pal, and Diganta Saha. 2022 · 2022
Cited alongside, same era.
Question answering system using deep learning in the low resource language bengali
Arijit Das and Diganta Saha. 2022 · 2022
Cited alongside, same era.
MuCoT: Multilingual contrastive training for question-answering in low-resource languages
Gokul Karthik Kumar, Abhishek Gehlot, Sahal Shaji Mullappilly, and Karthik Nandakumar. 2022 · 2022
Cited alongside, same era.
Tilahun Abedissa, Ricardo Usbeck, and Yaregal Assabie. 2023 · 2023
Later among the works it cites.
Question answering system for tamil using deep learning
Betina Antony and NR Rejin Paul. 2023 · 2023
Later among the works it cites.
Pquad: A persian question answering dataset
Kasra Darvishi, Newsha Shahbodaghkhan, Zahra Abbasiantaeb, and Saeedeh Momtazi. 2023 · 2023
Later among the works it cites.
Improving cross-lingual information retrieval on low-resource languages via optimal transport distillation
Zhiqi Huang, Puxuan Yu, and James Allan. 2023 · 2023
Later among the works it cites.
Thai vs. cambodian: How similar are these 2 incredible languages? - ling app
Michael Thompson. 2023 · 2023
Later among the works it cites.
Impact of tokenization on language models: An analysis for turkish
Cagri Toraman, Eyup Halit Yilmaz, Furkan Şahinuç, and Oguzhan Ozcelik. 2023 · 2023
Later among the works it cites.
Improving Vietnamese Legal Question–Answering System based on Automatic Data Enrichment
Thi-Hai-Yen Vuong, Ha-Thanh Nguyen, Quang-Huy Nguyen, Le-Minh Nguyen, and Xuan-Hieu Phan. 2023 · 2023
Later among the works it cites.
Kenswquad—a question answering dataset for swahili low-resource language
Barack W. Wanjawa, Lilian D. A. Wanzare, Florence Indede, Owen Mconyango, Lawrence Muchemi, and Edward Ombui. 2023 · 2023
Later among the works it cites.