Fetching the paper…
Reading the bibliography…
In real-world scenarios, a text classification task often begins with a cold start, when labeled data is scarce.
Multi-task deep neural networks for natural language understanding
Xiaodong Liu, Pengcheng He, W. Chen, and Jianfeng Gao. 2019a · 1901
Earlier work this paper cites.
Rodrigo Nogueira and Kyunghyun Cho. 2019 · 1901
Earlier work this paper cites.
How to fine-tune BERT for text classification?
Chi Sun, Xipeng Qiu, Yige Xu, and Xuanjing Huang. 2019 · 1905
Earlier work this paper cites.
RoBERTa: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
To tune or not to tune? how about the best of both worlds?
Ran Wang, Haibo Su, Chunye Wang, Kailin Ji, and Jupeng Ding. 2019b · 1907
Earlier work this paper cites.
Alexander Rietzler, Sebastian Stabinger, Paul Opitz, and Stefan Engl. 2020 · 1908
Earlier work this paper cites.
Domain adaptive training BERT for response selection
Taesun Whang, Dongyub Lee, Chanhee Lee, Kisu Yang, Dongsuk Oh, and Heuiseok Lim. 2019 · 1908
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, W. Li, and Peter J. Liu. 2019 · 1910
Earlier work this paper cites.
Unsupervised transfer learning via BERT neuron selection
Mehrdad Valipour, En-Shiun Annie Lee, Jaime R. Jamacaro, and Carolina Bessega. 2019 · 1912
Earlier work this paper cites.
Revisiting self-supervised visual representation learning
Alexander Kolesnikov, Xiaohua Zhai, and Lucas Beyer. 2019 · 1929
Earlier work this paper cites.
Statistical methods for research workers
Ronald. A. Fisher. 1971 · 1971
Earlier work this paper cites.
Least squares quantization in PCM
Stuart Lloyd. 1982 · 1982
Earlier work this paper cites.
Newsweeder: Learning to filter netnews
Ken Lang. 1995 · 1995
Earlier work this paper cites.
Combining labeled and unlabeled data with co-training
Avrim Blum and Tom Mitchell. 1998 · 1998
Earlier work this paper cites.
The class imbalance problem: A systematic study
Nathalie Japkowicz and Shaju Stephen. 2002 · 2002
Earlier work this paper cites.
Unsupervised document classification using sequential information maximization
Noam Slonim, Nir Friedman, and Naftali Tishby. 2002 · 2002
Earlier work this paper cites.
Improving BERT fine-tuning via self-ensemble and self-distillation
Yige Xu, Xipeng Qiu, L. Zhou, and X. Huang. 2020 · 2002
Earlier work this paper cites.
Don’t stop pretraining: Adapt language models to domains and tasks
Suchin Gururangan, Ana Marasović, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A. Smith. 2020 · 2004
Earlier work this paper cites.
Pre-training text representations as meta learning
Shangwen Lv, Yuechen Wang, Daya Guo, Duyu Tang, N. Duan, F. Zhu, Ming Gong, Linjun Shou, Ryan Ma, Daxin Jiang, G. Cao, M. Zhou, and Songlin Hu. 2020 · 2004
Earlier work this paper cites.
A sentimental education: Sentiment analysis using subjectivity summarization based on minimum cuts
Bo Pang and Lillian Lee. 2004 · 2004
Earlier work this paper cites.
Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales
Bo Pang and Lillian Lee. 2005 · 2005
Earlier work this paper cites.
Yada Pruksachatkun, Jason Phang, Haokun Liu, Phu Mon Htut, Xiaoyi Zhang, Richard Yuanzhe Pang, Clara Vania, Katharina Kann, and Samuel R Bowman. 2020 · 2005
Earlier work this paper cites.
Visualizing data using t-SNE
Laurens van der Maaten and Geoffrey Hinton. 2008 · 2008
Cited alongside, same era.
Parsing with multilingual BERT, a small corpus, and a small treebank
Ethan C. Chau, Lucy H. Lin, and Noah A. Smith. 2020 · 2009
Cited alongside, same era.
Text classification using label names only: A language model self-training approach
Yu Meng, Yunyi Zhang, Jiaxin Huang, Chenyan Xiong, Heng Ji, Chao Zhang, and Jiawei Han. 2020 · 2010
Cited alongside, same era.
Adaptive self-training for few-shot neural sequence labeling
Yaqing Wang, Subhabrata Mukherjee, H. Chu, Yuancheng Tu, Miaonan Wu, Jing Gao, and Ahmed Hassan Awadallah. 2020b · 2010
Cited alongside, same era.
Contributions to the study of sms spam filtering: new collection and results
Tiago A. Almeida, José María Gómez Hidalgo, and Akebo Yamakami. 2011 · 2011
Weakly supervised cyberbullying detection with participant-vocabulary consistency
Elaheh Raisi and Bert Huang. 2018 · 2018
Later among the works it cites.
Will it blend? blending weak and strong labeled data in a neural network for argumentation mining
Eyal Shnarch, Carlos Alzate, Lena Dankin, Martin Gleize, Yufang Hou, Leshem Choshen, Ranit Aharonov, and Noam Slonim. 2018 · 2018
Later among the works it cites.
Diverse few-shot text classification with multiple metrics
Mo Yu, Xiaoxiao Guo, Jinfeng Yi, Shiyu Chang, Saloni Potdar, Yu Cheng, Gerald Tesauro, Haoyu Wang, and Bowen Zhou. 2018 · 2018
Later among the works it cites.
Pivot based language modeling for improved neural domain adaptation
Yftah Ziser and Roi Reichart. 2018 · 2018
Later among the works it cites.
RIPPED: Recursive intent propagation using pretrained embedding distances
Michael Ball. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Recognizing textual entailment: Models and applications
Ido Dagan, Dan Roth, Mark Sammons, and Fabio Massimo Zanzotto. 2013 · 2013
Cited alongside, same era.
Hartigan’s K-means vs. Lloyd’s K-means – is it time for a change?
Noam Slonim, Ehud Aharoni, and Koby Crammer. 2013 · 2013
Cited alongside, same era.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
Siamese neural networks for one-shot image recognition
Gregory R. Koch. 2015 · 2015
Cited alongside, same era.
LSHTC: A benchmark for large-scale text classification
Ioannis Partalas, Aris Kosmopoulos, Nicolas Baskiotis, Thierry Artieres, George Paliouras, Eric Gaussier, Ion Androutsopoulos, Massih-Reza Amini, and Patrick Galinari. 2015 · 2015
Cited alongside, same era.
Universality versus cultural specificity of three emotion domains: Some evidence based on the cascading model of emotional intelligence
Bo Shao, Lorna Doucet, and David R. Caruso. 2015 · 2015
Cited alongside, same era.
Ruiying Geng, Binhua Li, Yongbin Li, Xiaodan Zhu, Ping Jian, and Jian Sun. 2019 · 2019
Later among the works it cites.
Noisy channel for low resource grammatical error correction
Ophélie Lacroix, Simon Flachs, and Anders Søgaard. 2019 · 2019
Later among the works it cites.
Combining unsupervised pre-training and annotator rationales to improve low-shot text classification
Oren Melamud, Mihaela Bornea, and Ken Barker. 2019 · 2019
Later among the works it cites.
Analyzing BERT with pre-train on SQuAD 2.0
Chenchen Pan. 2019 · 2019
Later among the works it cites.
Cross-lingual alignment of contextual word embeddings, with applications to zero-shot dependency parsing
Tal Schuster, Ori Ram, Regina Barzilay, and Amir Globerson. 2019 · 2019
Later among the works it cites.
Pre-training BERT on domain resources for short answer grading
Chul Sung, Tejas Dhamecha, Swarnadeep Saha, Tengfei Ma, Vinay Reddy, and Rishi Arora. 2019 · 2019
Later among the works it cites.
Cl-aff deep semisupervised clustering
Johnny Torres and Carmen Vaca. 2019 · 2019
Later among the works it cites.
BERT Post-Training for Review Reading Comprehension and Aspect-based Sentiment Analysis
Hu Xu, Bing Liu, Lei Shu, and Philip Yu. 2019 · 2019
Later among the works it cites.
XLNet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
Do not have enough data? Deep learning to the rescue!
Ateret Anaby-Tavor, Boaz Carmeli, Esther Goldbraich, Amir Kantor, George Kour, Segev Shlomov, N. Tepper, and Naama Zwerdling. 2020 · 2020
Later among the works it cites.
From arguments to key points: Towards automatic argument summarization
Roy Bar-Haim, Lilach Eden, Roni Friedman, Yoav Kantor, Dan Lahav, and Noam Slonim. 2020 · 2020
Later among the works it cites.
Corpus wide argument mining-a working solution
Liat Ein-Dor, Eyal Shnarch, Lena Dankin, Alon Halfon, Benjamin Sznajder, Ariel Gera, Carlos Alzate, Martin Gleize, Leshem Choshen, Yufang Hou, et al. 2020 · 2020
Later among the works it cites.
Injecting numerical reasoning skills into language models
Mor Geva, Ankit Gupta, and Jonathan Berant. 2020 · 2020
Later among the works it cites.
Few-shot relation classification by context attention-based prototypical networks with BERT
Bei Hui, Liang Liu, Jia Chen, Xue Zhou, and Yuhui Nian. 2020 · 2020
Later among the works it cites.
ALBERT: A lite BERT for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2020 · 2020
Later among the works it cites.
BioBERT: a pre-trained biomedical language representation model for biomedical text mining
Jinhyuk Lee, Wonjin Yoon, Sungdong Kim, D. Kim, Sunkyu Kim, Chan Ho So, and Jaewoo Kang. 2020 · 2020
Later among the works it cites.
Clusterfit: Improving generalization of visual representations
Xueting Yan, Ishan Misra, Abhinav Gupta, Deepti Ghadiyaram, and Dhruv Mahajan. 2020 · 2020
Later among the works it cites.