Fetching the paper…
Reading the bibliography…
We present the Multilingual Amazon Reviews Corpus (MARC), a large-scale collection of Amazon reviews for multilingual text classification.
Phillip Keung, Yichao Lu, and Vikas Bhardwaj. 2019 · 1909
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1911
Earlier work this paper cites.
Cross-lingual text categorization
Nuria Bel, Cornelis HA Koster, and Marta Villegas. 2003 · 2003
Earlier work this paper cites.
On the evaluation of contextual embeddings for zero-shot cross-lingual transfer learning
Phillip Keung, Yichao Lu, Julian Salazar, and Vikas Bhardwaj. 2020 · 2004
Earlier work this paper cites.
Rcv1: A new benchmark collection for text categorization research
David D Lewis, Yiming Yang, Tony G Rose, and Fan Li. 2004 · 2004
Earlier work this paper cites.
Reuters corpus, volume 2, multilingual corpus
Reuters Ltd. 2005 · 2005
Earlier work this paper cites.
Multilingual text classification using ontologies
Gerard De Melo and Stefan Siersdorfer. 2007 · 2007
Earlier work this paper cites.
Twitter as a corpus for sentiment analysis and opinion mining
Alexander Pak and Patrick Paroubek. 2010 · 2010
Earlier work this paper cites.
Cross-language text classification using structural correspondence learning
Peter Prettenhofer and Benno Stein. 2010 · 2010
Cited alongside, same era.
Learning word vectors for sentiment analysis
Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. 2011 · 2011
Cited alongside, same era.
Amazon customer reviews dataset
Amazon Inc. 2015 · 2015
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Cited alongside, same era.
Hierarchical attention networks for document classification
Zichao Yang, Diyi Yang, Chris Dyer, Xiaodong He, Alex Smola, and Eduard Hovy. 2016 · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
A neural interlingua for multilingual machine translation
Yichao Lu, Phillip Keung, Faisal Ladhak, Vikas Bhardwaj, Shaonan Zhang, and Jason Sun. 2018 · 2018
Later among the works it cites.
A corpus for multilingual document classification in eight languages
Holger Schwenk and Xian Li. 2018 · 2018
Later among the works it cites.
Massively multilingual sentence embeddings for zero-shot cross-lingual transfer and beyond
Mikel Artetxe and Holger Schwenk. 2019 · 2019
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Justifying recommendations using distantly-labeled reviews and fine-grained aspects
Jianmo Ni, Jiacheng Li, and Julian McAuley. 2019 · 2019
Later among the works it cites.
Beto, bentz, becas: The surprising cross-lingual effectiveness of bert
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Xnli: Evaluating cross-lingual sentence representations
Alexis Conneau, Ruty Rinott, Guillaume Lample, Adina Williams, Samuel R. Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Cited alongside, same era.
Shijie Wu and Mark Dredze. 2019 · 2019
Later among the works it cites.
Yelp open dataset
Yelp Inc. 2019 · 2019
Later among the works it cites.