Fetching the paper…
Reading the bibliography…
The recent advances in Natural Language Processing have only been a boon for well represented languages, negating research in lesser known global languages.
Reuters-21578 text categorization collection data set, 1997
David D Lewis · 1997
Earlier work this paper cites.
Explaining global patterns of language diversity
Daniel Nettle · 1998
Earlier work this paper cites.
The new york times annotated corpus
Evan Sandhaus · 2008
Cited alongside, same era.
Lorelei language packs: Data, tools, and resources for technology development in low resource languages
Stephanie Strassel and Jennifer Tracey · 2016
Cited alongside, same era.
Google’s multilingual neural machine translation system: Enabling zero-shot translation
Melvin Johnson, Mike Schuster, Quoc V Le, Maxim Krikun, Yonghui Wu, Zhifeng Chen, Nikhil Thorat, Fernanda Viégas, Martin Wattenberg, Greg Corrado, et al · 2017
Later among the works it cites.
Improving short text classification through global augmentation methods
Vukosi Marivate and Tshephisho Sefara · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…