Fetching the paper…
Reading the bibliography…
In order for NLP technology to be widely applicable, fair, and useful, it needs to serve a diverse set of speakers across the world's languages, be equitable, i.e., not unduly biased towards any particular language, and be inclusive of all users, particularly in low-resource settings where compute constraints are common.
Albert: A lite bert for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2019 · 1909
Earlier work this paper cites.
A constructive prediction of the generalization error across scales
Jonathan S Rosenfeld, Amir Rosenfeld, Yonatan Belinkov, and Nir Shavit. 2019 · 1909
Earlier work this paper cites.
On the cross-lingual transferability of monolingual representations
Mikel Artetxe, Sebastian Ruder, and Dani Yogatama. 2019 · 1910
Earlier work this paper cites.
The measurement of the inequality of incomes
Hugh Dalton. 1920 · 1920
Earlier work this paper cites.
Cross-lingual name tagging and linking for 282 languages
Xiaoman Pan, Boliang Zhang, Jonathan May, Joel Nothman, Kevin Knight, and Heng Ji. 2017 · 1958
Earlier work this paper cites.
A formula for the gini coefficient
Robert Dorfman. 1979 · 1979
Earlier work this paper cites.
A new dataset for natural language inference from code-mixed conversations
Simran Khanuja, Sandipan Dandapat, Sunayana Sitaram, and Monojit Choudhury. 2020a · 2004
Earlier work this paper cites.
Gluecos: An evaluation benchmark for code-switched nlp
Simran Khanuja, Sandipan Dandapat, Anirudh Srinivasan, Sunayana Sitaram, and Monojit Choudhury. 2020b · 2004
Earlier work this paper cites.
The gini index of speech
Scott Rickard and Maurice Fallon. 2004 · 2004
Earlier work this paper cites.
Anne Lauscher, Vinit Ravishankar, Ivan Vulić, and Goran Glavaš. 2020 · 2005
Earlier work this paper cites.
Income inequality measures
Fernando G De Maio. 2007 · 2007
Earlier work this paper cites.
Comparing measures of sparsity
Niall Hurley and Scott Rickard. 2009 · 2009
Earlier work this paper cites.
The real wealth of nations: pathways to human development
Jeni Klugman and Development Programme United Nations. 2010 · 2010
Earlier work this paper cites.
Census of india 2011 provisional population totals
Chandramouli. 2011 · 2011
Earlier work this paper cites.
Poverty, growth and inequality over the next 50 years
Evan Hillebrand et al. 2009 · 2012
Earlier work this paper cites.
Multi-task learning for multiple language translation
Daxiang Dong, Hua Wu, Wei He, Dianhai Yu, and Haifeng Wang. 2015 · 2015
Earlier work this paper cites.
Cross-lingual, character-level neural morphological tagging
Ryan Cotterell and Georg Heigold. 2017 · 2017
Earlier work this paper cites.
The iit bombay english-hindi parallel corpus
Anoop Kunchukuttan, Pratik Mehta, and Pushpak Bhattacharyya. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Xnli: Evaluating cross-lingual sentence representations
Alexis Conneau, Guillaume Lample, Ruty Rinott, Adina Williams, Samuel R Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Universal Dependencies 2.2
Joakim Nivre, Mitchell Abrams, Željko Agić, Lars Ahrenberg, and Antonsen et al. 2018 · 2018
Cited alongside, same era.
Taskonomy: Disentangling Task Transfer Learning
Amir R Zamir, Alexander Sax, William Shen, Leonidas Guibas, Jitendra Malik, and Silvio Savarese. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Processing South Asian languages written in the Latin script: the Dakshina dataset
Brian Roark, Lawrence Wolf-Sonkin, Christo Kirov, Sabrina J. Mielke, Cibu Johny, Işin Demirşahin, and Keith Hall. 2020 · 2020
Later among the works it cites.
Include: A large scale dataset for indian sign language recognition
Advaith Sridhar, Rohith Gandhi Ganesan, Pratyush Kumar, and Mitesh Khapra. 2020 · 2020
Later among the works it cites.
The Low-Resource Double Bind: An Empirical Study of Pruning for Low-Resource Machine Translation
Orevaoghene Ahia, Julia Kreutzer, and Sara Hooker. 2021 · 2021
Later among the works it cites.
How linguistically fair are multilingual pre-trained language models
Monojit Choudhury and Amit Deshpande. 2021 · 2021
Later among the works it cites.
US constitution, 2021
US Constitution. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
BLT4All: Language Technologies for All
European Language Resources Association. 2019 · 2019
Cited alongside, same era.
Gender-preserving debiasing for pre-trained word embeddings
Masahiro Kaneko and Danushka Bollegala. 2019 · 2019
Cited alongside, same era.
Choosing Transfer Languages for Cross-Lingual Learning
Yu-Hsiang Lin, Chian-Yu Chen, Jean Lee, Zirui Li, Yuyan Zhang, Mengzhou Xia, Shruti Rijhwani, Junxian He, Zhisong Zhang, Xuezhe Ma, Antonios Anastasopoulos, Patrick Littell, and Graham Neubig. 2019 · 2019
Cited alongside, same era.
Massively multilingual transfer for ner
Afshin Rahimi, Yuan Li, and Trevor Cohn. 2019 · 2019
Cited alongside, same era.
Transfer learning in natural language processing
Sebastian Ruder, Matthew E. Peters, Swabha Swayamdipta, and Thomas Wolf. 2019 · 2019
Cited alongside, same era.
Deep model transferability from attribution maps
Jie Song, Yixin Chen, Xinchao Wang, Chengchao Shen, and Mingli Song. 2019 · 2019
Cited alongside, same era.
Crowdsourcing speech data for low-resource languages from low-income workers
Basil Abraham, Danish Goel, Divya Siddarth, Kalika Bali, Manu Chopra, Monojit Choudhury, Pratik Joshi, Preethi Jyoti, Sunayana Sitaram, and Vivek Seshadri. 2020 · 2020
Cited alongside, same era.
Arnab Debnath, Navid Rajabi, Fardina Fathmiul Alam, and Antonios Anastasopoulos. 2021 · 2021
Later among the works it cites.
Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity
William Fedus, Barret Zoph, and Noam Shazeer. 2021 · 2021
Later among the works it cites.
Muril: Multilingual representations for indian languages
Simran Khanuja, Diksha Bansal, Sarvesh Mehtani, Savya Khosla, Atreyee Dey, Balaji Gopalan, Dilip Kumar Margam, Pooja Aggarwal, Rajiv Teja Nagipogu, Shachi Dave, et al. 2021 · 2021
Later among the works it cites.
Dynaboard: An evaluation-as-a-service platform for holistic next-generation benchmarking
Zhiyi Ma, Kawin Ethayarajh, Tristan Thrush, Somya Jain, Ledell Wu, Robin Jia, Christopher Potts, Adina Williams, and Douwe Kiela. 2021 · 2021
Later among the works it cites.
Samanantar: The largest publicly available parallel corpora collection for 11 indic languages
Gowtham Ramesh, Sumanth Doddapaneni, Aravinth Bheemaraj, Mayank Jobanputra, Raghavan AK, Ajitesh Sharma, Sujit Sahoo, Harshita Diddee, Divyanshu Kakwani, Navneet Kumar, et al. 2021 · 2021
Later among the works it cites.
Handbook of statistics on indian economy
Reserve Bank of India RBI. 2021 · 2021
Later among the works it cites.
XTREME-R: Towards More Challenging and Nuanced Multilingual Evaluation
Sebastian Ruder, Noah Constant, Jan Botha, Aditya Siddhant, Orhan Firat, Jinlan Fu, Pengfei Liu, Junjie Hu, Graham Neubig, and Melvin Johnson. 2021 · 2021
Later among the works it cites.
Re-imagining algorithmic fairness in india and beyond
Nithya Sambasivan, Erin Arnesen, Ben Hutchinson, Tulsee Doshi, and Vinodkumar Prabhakaran. 2021 · 2021
Later among the works it cites.
Revisiting the primacy of english in zero-shot cross-lingual transfer
Iulia Turc, Kenton Lee, Jacob Eisenstein, Ming-Wei Chang, and Kristina Toutanova. 2021 · 2021
Later among the works it cites.
Multi-view Subword Regularization
Xinyi Wang, Sebastian Ruder, and Graham Neubig. 2021 · 2021
Later among the works it cites.
Multi task learning for zero shot performance prediction of multilingual models
Kabir Ahuja, Shanu Kumar, Sandipan Dandapat, and Monojit Choudhury. 2022 · 2022
Closest in time.
Systematic inequalities in language technology performance across the world’s languages
Damian Blasi, Antonios Anastasopoulos, and Graham Neubig. 2022 · 2022
Closest in time.
KinyaBERT: a morphology-aware Kinyarwanda language model
Antoine Nzeyimana and Andre Niyongabo Rubungo. 2022 · 2022
Closest in time.