Fetching the paper…
Reading the bibliography…
Traditional text classification approaches often require a good amount of labeled data, which is difficult to obtain, especially in restricted domains or less widespread languages.
Joining statistics with NLP for text categorization
Paul S Jacobs · 1992
Earlier work this paper cites.
Text classification from labeled and unlabeled documents using EM
Kamal Nigam, Andrew Kachites McCallum, Sebastian Thrun, and Tom Mitchell · 2000
Earlier work this paper cites.
Evaluating Cross-Language Annotation Transfer in the MultiSemCor Corpus
Luisa Bentivogli, Pamela Forner, and Emanuele Pianta · 2004
Earlier work this paper cites.
Importance of Semantic Representation: Dataless Classification
Ming-Wei Chang, Lev-Arie Ratinov, Dan Roth, and Vivek Srikumar · 2008
Earlier work this paper cites.
Cross-validation
Payam Refaeilzadeh, Lei Tang, and Huan Liu · 2009
Earlier work this paper cites.
Zero-Shot Learning Through Cross-Modal Transfer
Richard Socher, Milind Ganjoo, Christopher D Manning, and Andrew Ng · 2013
Earlier work this paper cites.
Representation Learning for Text-level Discourse Parsing
Yangfeng Ji and Jacob Eisenstein · 2014
Earlier work this paper cites.
Semi-Supervised Zero-Shot Classification with Label Representation Learning
Xin Li, Yuhong Guo, and Dale Schuurmans · 2015
Earlier work this paper cites.
Short text classification based on LDA topic model
Qiuxing Chen, Lixiu Yao, and Jie Yang · 2016
Earlier work this paper cites.
ASSIN: Avaliacao de similaridade semantica e inferencia textual
E Fonseca, L Santos, Marcelo Criscuolo, and S Aluisio · 2016
Earlier work this paper cites.
Accelerated hierarchical density based clustering
Leland McInnes and John Healy · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
XNLI: Evaluating Cross-lingual Sentence Representations
Alexis Conneau, Ruty Rinott, Guillaume Lample, Adina Williams, Samuel Bowman, Holger Schwenk, and Veselin Stoyanov · 2018
Cited alongside, same era.
Defining textual entailment
Daniel Z Korman, Eric Mack, Jacob Jett, and Allen H Renear · 2018
Cited alongside, same era.
A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference
Adina Williams, Nikita Nangia, and Samuel Bowman · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Unsupervised Cross-lingual Representation Learning at Scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Édouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov · 2020
Later among the works it cites.
BERTopic: Leveraging BERT and c-TF-IDF to create easily interpretable topics., 2020
Maarten Grootendorst · 2020
Later among the works it cites.
Few-shot sequence learning with transformers
Lajanugen Logeswaran, Ann Lee, Myle Ott, Honglak Lee, Marc’Aurelio Ranzato, and Arthur Szlam · 2020
Later among the works it cites.
Contextualized weak supervision for text classification
Dheeraj Mekala and Jingbo Shang · 2020
Later among the works it cites.
Text Classification Using Label Names Only: A Language Model Self-Training Approach
Yu Meng, Yunyi Zhang, Jiaxin Huang, Chenyan Xiong, Heng Ji, Chao Zhang, and Jiawei Han · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf · 2019
Cited alongside, same era.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R Bowman · 2019
Cited alongside, same era.
Benchmarking Zero-shot Text Classification: Datasets, Evaluation and Entailment Approach
Wenpeng Yin, Jamaal Hay, and Dan Roth · 2019
Cited alongside, same era.
Integrating Semantic Knowledge to Tackle Zero-shot Text Classification
Jingqing Zhang, Piyawat Lertvittayakumjorn, and Yike Guo · 2019
Cited alongside, same era.
Longformer: The long-document transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
The ASSIN 2 Shared Task: a quick overview
Livy Real, Erick Fonseca, and Hugo Gonçalo Oliveira · 2020
Later among the works it cites.
BERTimbau: pretrained BERT models for Brazilian Portuguese
Fábio Souza, Rodrigo Nogueira, and Roberto Lotufo · 2020
Later among the works it cites.
Big bird: Transformers for longer sequences
Manzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie, Chris Alberti, Santiago Ontanon, Philip Pham, Anirudh Ravula, Qifan Wang, Li Yang, et al · 2020
Later among the works it cites.
DEBACER: a method for slicing moderated debates
Thomas Palmeira Ferraz, Alexandre Alcoforado, Enzo Bustos, André Seidel Oliveira, Rodrigo Gerber, Naíde Müller, André Corrêa d’Almeida, Bruno Miguel Veloso, and Anna Helena Reali Costa · 2021
Later among the works it cites.
A Survey on Recent Approaches for Natural Language Processing in Low-Resource Scenarios
Michael A Hedderich, Lukas Lange, Heike Adel, Jannik Strötgen, and Dietrich Klakow · 2021
Later among the works it cites.