Fetching the paper…
Reading the bibliography…
Deep learning based models have surpassed classical machine learning based approaches in various text classification tasks, including sentiment analysis, news categorization, question answering, and natural language inference.
K. Fukushima, “Neocognitron: A self-organizing neural network model for a mechanism of pattern recognition unaffected by shift in position,” Biological cybernetics , vol. 36, no. 4, pp. 193–202, 1980
1980
Earlier work this paper cites.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning internal representations by error propagation,” California Univ San Diego La Jolla Inst for Cognitive Science, Tech. Rep., 1985
1985
Earlier work this paper cites.
D. E. Rumelhart, G. E. Hinton, R. J. Williams et al. , “Learning representations by back-propagating errors,” Cognitive modeling , vol. 5, no. 3, p. 1, 1988
1988
Earlier work this paper cites.
S. Deerwester, S. T. Dumais, G. W. Furnas, T. K. Landauer, and R. Harshman, “Indexing by latent semantic analysis,” Journal of the American society for information science , vol. 41, no. 6, pp. 391–407, 1990
1990
Earlier work this paper cites.
J. BROMLEY, J. W. BENTZ, L. BOTTOU, I. GUYON, Y. LECUN, C. MOORE, E. SÄCKINGER, and R. SHAH, “Signature verification using a Siamese time delay neural network,” International Journal of Pattern Recognition and Artificial Intelligence , 1993
1993
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
D. Jurafsky and J. H. Martin, Speech and Language Processing: An Introduction to Natural Language Processing, Computational Linguistics, and Speech Recognition , 1st ed. USA: Prentice Hall PTR, 2000
2000
Earlier work this paper cites.
B. Pang, L. Lee, and S. Vaithyanathan, “Thumbs up?: sentiment classification using machine learning techniques,” in Proceedings of the ACL conference on Empirical methods in natural language processing , 2002, pp. 79–86
2002
Earlier work this paper cites.
Y. Bengio, R. Ducharme, P. Vincent, and C. Jauvin, “A neural probabilistic language model,” Journal of machine learning research , vol. 3, no. Feb, pp. 1137–1155, 2003
2003
Earlier work this paper cites.
R. Mihalcea and P. Tarau, “Textrank: Bringing order into text,” in Proceedings of the 2004 conference on empirical methods in natural language processing , 2004, pp. 404–411
2004
Earlier work this paper cites.
B. Dolan, C. Quirk, and C. Brockett, “Unsupervised construction of large paraphrase corpora: Exploiting massively parallel news sources,” in Proceedings of the 20th international conference on Computational Linguistics . ACL, 2004, p. 350
2004
Earlier work this paper cites.
D. Greene and P. Cunningham, “Practical solutions to the problem of diagonal dominance in kernel document clustering,” in Proc. 23rd International Conference on Machine learning (ICML’06) . ACM Press, 2006, pp. 377–384
2006
Earlier work this paper cites.
I. Dagan, O. Glickman, and B. Magnini, “The PASCAL Recognising Textual Entailment Challenge,” in Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) , 2006
2006
Earlier work this paper cites.
A. S. Das, M. Datar, A. Garg, and S. Rajaram, “Google news personalization: scalable online collaborative filtering,” in Proceedings of the 16th international conference on World Wide Web . ACM, 2007, pp. 271–280
2007
Earlier work this paper cites.
C. D. Manning, H. Schütze, and P. Raghavan, Introduction to information retrieval . Cambridge university press, 2008
2008
Earlier work this paper cites.
D. Jurasky and J. H. Martin, “Speech and language processing: An introduction to natural language processing,” Computational Linguistics and Speech Recognition. Prentice Hall, New Jersey , 2008
2008
Earlier work this paper cites.
E. L. Mencia and J. Fürnkranz, “Efficient pairwise multilabel classification for large-scale problems in the legal domain,” in Joint European Conference on Machine Learning and Knowledge Discovery in Databases . Springer, 2008, pp. 50–65
2008
Earlier work this paper cites.
J. C. Martineau and T. Finin, “Delta tfidf: An improved feature space for sentiment analysis,” in Third international AAAI conference on weblogs and social media , 2009
2009
Earlier work this paper cites.
G. E. Hinton, A. Krizhevsky, and S. D. Wang, “Transforming auto-encoders,” in International conference on artificial neural networks . Springer, 2011, pp. 44–51
2011
Earlier work this paper cites.
W. tau Yih, K. Toutanova, J. C. Platt, and C. Meek, “Learning discriminative projections for text similarity measures,” in CoNLL 2011 - Fifteenth Conference on Computational Natural Language Learning, Proceedings of the Conference , 2011
2011
Earlier work this paper cites.
R. Collobert, J. Weston, L. Bottou, M. Karlen, K. Kavukcuoglu, and P. Kuksa, “Natural language processing (almost) from scratch,” Journal of machine learning research , vol. 12, no. Aug, pp. 2493–2537, 2011
2011
Earlier work this paper cites.
Z. Lu, “Pubmed and beyond: a survey of web tools for searching biomedical literature,” Database , vol. 2011, 2011
2011
Earlier work this paper cites.
A. L. Maas, R. E. Daly, P. T. Pham, D. Huang, A. Y. Ng, and C. Potts, “Learning word vectors for sentiment analysis,” in Proceedings of the 49th annual meeting of the association for computational linguistics: Human language technologies-volume 1 , 2011, pp. 142–150
2011
Earlier work this paper cites.
S. Wang and C. D. Manning, “Baselines and bigrams: Simple, good sentiment and topic classification,” in Proceedings of the 50th annual meeting of the association for computational linguistics: Short papers-volume 2 . Association for Computational Linguistics, 2012, pp. 90–94
2012
Earlier work this paper cites.
M. Schuster and K. Nakajima, “Japanese and korean voice search,” in 2012 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2012, pp. 5149–5152
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in Advances in neural information processing systems , 2013, pp. 3111–3119
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
R. Socher, A. Perelygin, J. Wu, J. Chuang, C. D. Manning, A. Y. Ng, and C. Potts, “Recursive deep models for semantic compositionality over a sentiment treebank,” in Proceedings of the 2013 conference on empirical methods in natural language processing , 2013, pp. 1631–1642
2013
Earlier work this paper cites.
P.-S. Huang, X. He, J. Gao, L. Deng, A. Acero, and L. Heck, “Learning deep structured semantic models for web search using clickthrough data,” in Proceedings of the 22nd ACM international conference on Information & Knowledge Management , 2013, pp. 2333–2338
2013
Earlier work this paper cites.
M. Richardson, C. J. Burges, and E. Renshaw, “Mctest: A challenge dataset for the open-domain machine comprehension of text,” in Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing , 2013, pp. 193–203
2013
Earlier work this paper cites.
M. Marelli, L. Bentivogli, M. Baroni, R. Bernardi, S. Menini, and R. Zamparelli, “Semeval-2014 task 1: Evaluation of compositional distributional semantic models on full sentences through semantic relatedness and textual entailment,” in Proceedings of the 8th international workshop on semantic evaluation (SemEval 2014) , 2014, pp. 1–8
2014
Earlier work this paper cites.
J. Pennington, R. Socher, and C. Manning, “Glove: Global vectors for word representation,” in Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP) , 2014, pp. 1532–1543
2014
Earlier work this paper cites.
Q. Le and T. Mikolov, “Distributed representations of sentences and documents,” in International conference on machine learning , 2014, pp. 1188–1196
2014
Earlier work this paper cites.
N. Kalchbrenner, E. Grefenstette, and P. Blunsom, “A convolutional neural network for modelling sentences,” in 52nd Annual Meeting of the Association for Computational Linguistics, ACL 2014 - Proceedings of the Conference , 2014
2014
Earlier work this paper cites.
Y. Kim, “Convolutional neural networks for sentence classification,” in EMNLP 2014 - 2014 Conference on Empirical Methods in Natural Language Processing, Proceedings of the Conference , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
Y. Shen, X. He, J. Gao, L. Deng, and G. Mesnil, “A latent semantic model with convolutional-pooling structure for information retrieval,” in ACM International Conference on Conference on Information and Knowledge Management . ACM, 2014, pp. 101–110
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in 2nd International Conference on Learning Representations, ICLR 2014 - Conference Track Proceedings , 2014
2014
Earlier work this paper cites.
D. J. Rezende, S. Mohamed, and D. Wierstra, “Stochastic backpropagation and approximate inference in deep generative models,” ICML , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
F. Wang, Z. Wang, Z. Li, and J.-R. Wen, “Concept-based short text classification and ranking,” in Proceedings of the 23rd ACM International Conference on Conference on Information and Knowledge Management . ACM, 2014, pp. 1069–1078
2014
Earlier work this paper cites.
B. C. Wallace, L. Kertz, E. Charniak et al. , “Humans require context to infer ironic intent (so computers probably do, too),” in Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , 2014, pp. 512–516
2014
Earlier work this paper cites.
O. Abdel-Hamid, A.-r. Mohamed, H. Jiang, L. Deng, G. Penn, and D. Yu, “Convolutional neural networks for speech recognition,” IEEE/ACM Transactions on audio, speech, and language processing , vol. 22, no. 10, pp. 1533–1545, 2014
2014
Earlier work this paper cites.
M. Iyyer, V. Manjunatha, J. Boyd-Graber, and H. Daumé III, “Deep unordered composition rivals syntactic methods for text classification,” in Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) , 2015, pp. 1681–1691
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
X. Zhu, P. Sobihani, and H. Guo, “Long short-term memory over recursive structures,” in International Conference on Machine Learning , 2015, pp. 1604–1612
2015
Earlier work this paper cites.
P. Liu, X. Qiu, X. Chen, S. Wu, and X.-J. Huang, “Multi-timescale long short-term memory neural network for modelling sentences and documents,” in Proceedings of the 2015 conference on empirical methods in natural language processing , 2015, pp. 2326–2335
2015
Earlier work this paper cites.
R. Johnson and T. Zhang, “Effective use of word order for text categorization with convolutional neural networks,” in NAACL HLT 2015 - 2015 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Proceedings of the Conference , 2015
2015
Earlier work this paper cites.
X. Zhang, J. Zhao, and Y. LeCun, “Character-level convolutional networks for text classification,” in Advances in neural information processing systems , 2015, pp. 649–657
2015
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in 3rd International Conference on Learning Representations, ICLR 2015 - Conference Track Proceedings , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. Rios and R. Kavuluru, “Convolutional neural networks for biomedical text classification: Application in indexing biomedical articles,” in BCB 2015 - 6th ACM Conference on Bioinformatics, Computational Biology, and Health Informatics , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
J. Weston, S. Chopra, and A. Bordes, “Memory networks,” in 3rd International Conference on Learning Representations, ICLR 2015 - Conference Track Proceedings , 2015
2015
Earlier work this paper cites.
S. Sukhbaatar, J. Weston, R. Fergus et al. , “End-to-end memory networks,” in Advances in neural information processing systems , 2015, pp. 2440–2448
2015
Earlier work this paper cites.
A. Severyn and A. Moschittiy, “Learning to rank short text pairs with convolutional deep neural networks,” in SIGIR 2015 - Proceedings of the 38th International ACM SIGIR Conference on Research and Development in Information Retrieval , 2015
2015
Earlier work this paper cites.
H. He, K. Gimpel, and J. Lin, “Multi-perspective sentence similarity modeling with convolutional neural networks,” in Conference Proceedings - EMNLP 2015: Conference on Empirical Methods in Natural Language Processing , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
D. Tang, B. Qin, and T. Liu, “Document modeling with gated recurrent neural network for sentiment classification,” in Proceedings of the 2015 conference on empirical methods in natural language processing , 2015, pp. 1422–1432
2015
Earlier work this paper cites.
S. Lai, L. Xu, K. Liu, and J. Zhao, “Recurrent convolutional neural networks for text classification,” in Twenty-ninth AAAI conference on artificial intelligence , 2015
2015
Earlier work this paper cites.
R. Srivastava, K. Greff, and J. Schmidhuber, “Training very deep networks,” in Advances in Neural Information Processing Systems , 2015
2015
Earlier work this paper cites.
R. Kiros, Y. Zhu, R. R. Salakhutdinov, R. Zemel, R. Urtasun, A. Torralba, and S. Fidler, “Skip-thought vectors,” in Advances in neural information processing systems , 2015, pp. 3294–3302
2015
Earlier work this paper cites.
A. M. Dai and Q. V. Le, “Semi-supervised sequence learning,” in Advances in Neural Information Processing Systems , 2015
2015
Earlier work this paper cites.
L. Deng and J. Wiebe, “Mpqa 3.0: An entity/event-level sentiment corpus,” in Proceedings of the 2015 conference of the North American chapter of the association for computational linguistics: human language technologies , 2015, pp. 1323–1328
2015
Earlier work this paper cites.
J. Lehmann, R. Isele, M. Jakob, A. Jentzsch, D. Kontokostas, P. N. Mendes, S. Hellmann, M. Morsey, P. Van Kleef, S. Auer et al. , “Dbpedia–a large-scale, multilingual knowledge base extracted from wikipedia,” Semantic Web , vol. 6, no. 2, pp. 167–195, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
M. Kusner, Y. Sun, N. Kolkin, and K. Weinberger, “From word embeddings to document distances,” in International conference on machine learning , 2015, pp. 957–966
2015
Earlier work this paper cites.
R. Sennrich, B. Haddow, and A. Birch, “Neural machine translation of rare words with subword units,” arXiv preprint:1508.07909 , 2015
2015
Earlier work this paper cites.
http://colah.github.io/posts/2015-08-Understanding-LSTMs/
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
I. Goodfellow, Y. Bengio, and A. Courville, Deep learning . MIT press, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2018
Later among the works it cites.
C. Tan, F. Wei, W. Wang, W. Lv, and M. Zhou, “Multiway attention networks for modeling sentence pairs,” in IJCAI , 2018, pp. 4411–4417
2018
Later among the works it cites.
S. Wang, M. Huang, and Z. Deng, “Densely connected cnn with multi-scale feature attention for text classification.” in IJCAI , 2018, pp. 4468–4474
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Wan, Y. Lan, J. Guo, J. Xu, L. Pang, and X. Cheng, “A deep architecture for semantic matching with multiple positional sentence representations,” in Thirtieth AAAI Conference on Artificial Intelligence , 2016
2016
Cited alongside, same era.
Y. Kim, Y. Jernite, D. Sontag, and A. M. Rush, “Character-aware neural language models,” in Thirtieth AAAI Conference on Artificial Intelligence , 2016
2016
Cited alongside, same era.
J. D. Prusa and T. M. Khoshgoftaar, “Designing a better data representation for deep neural networks and text classification,” in Proceedings - 2016 IEEE 17th International Conference on Information Reuse and Integration, IRI 2016 , 2016
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
L. Pang, Y. Lan, J. Guo, J. Xu, S. Wan, and X. Cheng, “Text matching as image recognition,” in 30th AAAI Conference on Artificial Intelligence, AAAI 2016 , 2016
2016
Cited alongside, same era.
S. Peng, R. You, H. Wang, C. Zhai, H. Mamitsuka, and S. Zhu, “DeepMeSH: Deep semantic representation for improving large-scale MeSH indexing,” Bioinformatics , 2016
2016
Cited alongside, same era.
Z. Yang, D. Yang, C. Dyer, X. He, A. Smola, and E. Hovy, “Hierarchical attention networks for document classification,” in Proceedings of the 2016 conference of the North American chapter of the association for computational linguistics: human language technologies , 2016, pp. 1480–1489
2016
Cited alongside, same era.
H. Peng, J. Li, Y. He, Y. Liu, M. Bao, L. Wang, Y. Song, and Q. Yang, “Large-scale hierarchical text classification with recursively regularized deep graph-cnn,” in Proceedings of the 2018 World Wide Web Conference . International World Wide Web Conferences Steering Committee, 2018, pp. 1063–1072
2018
Later among the works it cites.
Y. Tay, L. A. Tuan, and S. C. Hui, “Hyperbolic representation learning for fast and efficient neural question answering,” in Proceedings of the Eleventh ACM International Conference on Web Search and Data Mining , 2018, pp. 583–591
2018
Later among the works it cites.
Y. Meng, J. Shen, C. Zhang, and J. Han, “Weakly-supervised neural text classification,” in CIKM , 2018
2018
Later among the works it cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
Y. Li, Q. Pan, S. Wang, T. Yang, and E. Cambria, “A generative model for category text generation,” Information Sciences , vol. 450, pp. 301–315, 2018
2018
Later among the works it cites.
T. Zhang, M. Huang, and L. Zhao, “Learning structured representation for text classification via reinforcement learning,” in Thirty-Second AAAI Conference on Artificial Intelligence , 2018
2018
Later among the works it cites.
P. Rajpurkar, R. Jia, and P. Liang, “Know what you don’t know: Unanswerable questions for squad,” arXiv preprint:1806.03822 , 2018
2018
Later among the works it cites.
Y. Yang, W.-t. Yih, and C. Meek, “Wikiqa: A challenge dataset for open-domain question answering,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing , 2015, pp. 2013–2018
2018
Later among the works it cites.
2018
Later among the works it cites.
T. Khot, A. Sabharwal, and P. Clark, “Scitail: A textual entailment dataset from science question answering,” in 32nd AAAI Conference on Artificial Intelligence, AAAI 2018 , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
T. Kudo, “Subword regularization: Improving neural network translation models with multiple subword candidates,” in ACL 2018 - 56th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference (Long Papers) , 2018
2018
Later among the works it cites.
G. Marcus and E. Davis, Rebooting AI: Building artificial intelligence we can trust . Pantheon, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
K. Kowsari, K. Jafari Meimandi, M. Heidarysafa, S. Mendu, L. Barnes, and D. Brown, “Text classification algorithms: A survey,” Information , vol. 10, no. 4, p. 150, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
A. B. Duque, L. L. J. Santos, D. Macêdo, and C. Zanchettin, “Squeezed Very Deep Convolutional Neural Networks for Text Classification,” in Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) , 2019
2019
Later among the works it cites.
B. Guo, C. Zhang, J. Liu, and X. Ma, “Improving text classification with weighted word embeddings via a multi-channel TextCNN model,” Neurocomputing , 2019
2019
Later among the works it cites.
M. Yang, W. Zhao, L. Chen, Q. Qu, Z. Zhao, and Y. Shen, “Investigating the transferring capability of capsule networks for text classification,” Neural Networks , vol. 118, pp. 247–261, 2019
2019
Later among the works it cites.
W. Zhao, H. Peng, S. Eger, E. Cambria, and M. Yang, “Towards scalable and reliable capsule networks for challenging NLP applications,” in ACL , 2019, pp. 1549–1559
2019
Later among the works it cites.
R. Aly, S. Remus, and C. Biemann, “Hierarchical multi-label classification of text with capsule networks,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics: Student Research Workshop , 2019, pp. 323–330
2019
Later among the works it cites.
S. Kim, I. Kang, and N. Kwak, “Semantic sentence matching with densely-connected recurrent and co-attentive information,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, 2019, pp. 6586–6593
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
L. Yao, C. Mao, and Y. Luo, “Graph convolutional networks for text classification,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, 2019, pp. 7370–7377
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
P. Liu, S. Chang, X. Huang, J. Tang, and J. C. K. Cheung, “Contextualized non-local neural networks for sequence learning,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, 2019, pp. 6762–6769
2019
Later among the works it cites.
J. Gao, M. Galley, and L. Li, “Neural approaches to conversational ai,” Foundations and Trends® in Information Retrieval , vol. 13, no. 2-3, pp. 127–298, 2019
2019
Later among the works it cites.
N. Reimers and I. Gurevych, “Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks,” 2019
2019
Later among the works it cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,” OpenAI Blog , vol. 1, no. 8, p. 9, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
C. Sun, X. Qiu, Y. Xu, and X. Huang, “How to fine-tune bert for text classification?” in China National Conference on Chinese Computational Linguistics . Springer, 2019, pp. 194–206
2019
Later among the works it cites.
2019
Later among the works it cites.
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. R. Salakhutdinov, and Q. V. Le, “Xlnet: Generalized autoregressive pretraining for language understanding,” in Advances in neural information processing systems , 2019, pp. 5754–5764
2019
Later among the works it cites.
L. Dong, N. Yang, W. Wang, F. Wei, X. Liu, Y. Wang, J. Gao, M. Zhou, and H.-W. Hon, “Unified language model pre-training for natural language understanding and generation,” in Advances in Neural Information Processing Systems , 2019, pp. 13 042–13 054
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Zhang, Y. Wu, W. Li, and W. Li, “Learning Universal Sentence Representations with Mean-Max Attention Autoencoder,” 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
D. S. Sachan, M. Zaheer, and R. Salakhutdinov, “Revisiting lstm networks for semi-supervised text classification via mixed objective function,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, 2019, pp. 6940–6948
2019
Later among the works it cites.
2019
Later among the works it cites.
A. Ratner, B. Hancock, J. Dunnmon, F. Sala, S. Pandey, and C. Ré, “Training complex models with multi-task weak supervision,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, 2019, pp. 4763–4771
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
J. Kim, S. Jang, E. Park, and S. Choi, “Text classification using capsules,” Neurocomputing , vol. 376, pp. 214–221, 2020
2020
Closest in time.
M. E. Basiri, S. Nemati, M. Abdar, E. Cambria, and U. R. Acharya, “Abcdm: An attention-based bidirectional cnn-rnn deep model for sentiment analysis,” Future Generation Computer Systems , vol. 115, pp. 279–294, 2020
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
Y. Sun, S. Wang, Y.-K. Li, S. Feng, H. Tian, H. Wu, and H. Wang, “Ernie 2.0: A continual pre-training framework for language understanding.” in AAAI , 2020, pp. 8968–8975
2020
Closest in time.
2020
Closest in time.
J. Chen, Z. Yang, and D. Yang, “Mixtext: Linguistically-informed interpolation of hidden space for semi-supervised text classification,” in ACL , 2020
2020
Closest in time.
X. Liu, L. Mou, H. Cui, Z. Lu, and S. Song, “Finding decision jumps in text classification,” Neurocomputing , vol. 371, pp. 177–187, 2020
2020
Closest in time.
2020
Closest in time.
S. Mukherjee and A. H. Awadallah, “Xtremedistil: Multi-stage distillation for massive multilingual models,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , 2020, pp. 2221–2234
2020
Closest in time.
2020
Closest in time.