Fetching the paper…
Reading the bibliography…
CNNs, RNNs, GCNs, and CapsNets have shown significant insights in representation learning and are widely used in various text mining tasks such as large-scale multi-label text classification.
D. D. Lewis, “An evaluation of phrasal and clustered representations on a text categorization task,” in SIGIR , 1992, pp. 37–50
1992
Earlier work this paper cites.
Y. Yang, “An evaluation of statistical approaches to text categorization,” Information retrieval , vol. 1, no. 1-2, pp. 69–90, 1999
1999
Earlier work this paper cites.
A. Sun and E.-P. Lim, “Hierarchical text classification and evaluation,” in ICDM . IEEE, 2001, pp. 521–528
2001
Earlier work this paper cites.
Y. Bengio, R. Ducharme, P. Vincent, and C. Jauvin, “A neural probabilistic language model,” JMLR , vol. 3, no. Feb, pp. 1137–1155, 2003
2003
Earlier work this paper cites.
M. W. Berry, Survey of Text Mining . Berlin, Heidelberg: Springer-Verlag, 2003
2003
Earlier work this paper cites.
D. M. Blei, A. Y. Ng, and M. I. Jordan, “Latent dirichlet allocation,” JMLR , vol. 3, pp. 993–1022, 2003
2003
Earlier work this paper cites.
D. D. Lewis, Y. Yang, T. G. Rose, and F. Li, “RCV1: A new benchmark collection for text categorization research,” JMLR , vol. 5, pp. 361–397, 2004
2004
Earlier work this paper cites.
T. Liu, Y. Yang, H. Wan, H. Zeng, Z. Chen, and W. Ma, “Support vector machines classification with a very large-scale taxonomy,” SIGKDD Explorations , vol. 7, no. 1, pp. 36–43, 2005
2005
Earlier work this paper cites.
I. Tsochantaridis, T. Joachims, T. Hofmann, and Y. Altun, “Large margin methods for structured and interdependent output variables,” JMLR , vol. 6, pp. 1453–1484, 2005
2005
Earlier work this paper cites.
G.-R. Xue, D. Xing, Q. Yang, and Y. Yu, “Deep classification in large-scale text hierarchies,” in Proceedings of the 31st annual international ACM SIGIR conference on Research and development in information retrieval . ACM, 2008, pp. 619–626
2008
Earlier work this paper cites.
E. Loza Mencía and J. Fürnkranz, “Efficient pairwise multilabel classification for large-scale problems in the legal domain,” in ECML/PKDD , 2008, pp. 50–65
2008
Earlier work this paper cites.
G. E. Hinton, A. Krizhevsky, and S. D. Wang, “Transforming auto-encoders,” in ICANN , 2011, pp. 44–51
2011
Earlier work this paper cites.
C. C. Aggarwal and C. Zhai, “A survey of text classification algorithms,” in Mining Text Data , 2012, pp. 163–222
2012
Earlier work this paper cites.
S. Gopal, Y. Yang, B. Bai, and A. Niculescu-Mizil, “Bayesian models for large-scale hierarchical classification,” in NIPS , 2012, pp. 2420–2428
2012
Earlier work this paper cites.
G. Siddharth and Y. Yiming, “Recursive regularization for large-scale classification with hierarchical and graphical dependencies,” in KDD , 2013, pp. 257–265
2013
Earlier work this paper cites.
S. Xie, X. Kong, J. Gao, W. Fan, and S. Y. Philip, “Multilabel consensus classification,” in 2013 IEEE 13th International Conference on Data Mining . IEEE, 2013, pp. 1241–1246
2013
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in NIPS , 2013, pp. 3111–3119
2013
Earlier work this paper cites.
T. Mikolov, K. Chen, G. Corrado, and J. Dean, “Efficient estimation of word representations in vector space,” Computer Science , 2013
2013
Earlier work this paper cites.
D. I. Shuman, S. K. Narang, P. Frossard, A. Ortega, and P. Vandergheynst, “The emerging field of signal processing on graphs: Extending high-dimensional data analysis to networks and other irregular domains,” IEEE Signal Process. Mag. , vol. 30, no. 3, pp. 83–98, 2013
2013
Earlier work this paper cites.
J. Bruna, W. Zaremba, A. Szlam, and Y. Lecun, “Spectral networks and locally connected networks on graphs,” Computer Science , 2013
2013
Earlier work this paper cites.
Y. Kim, “Convolutional neural networks for sentence classification,” in EMNLP , 2014, pp. 1746–1751
2014
Cited alongside, same era.
M.-L. Zhang and Z.-H. Zhou, “A review on multi-label learning algorithms,” IEEE transactions on knowledge and data engineering , vol. 26, no. 8, pp. 1819–1837, 2014
2014
Cited alongside, same era.
B. Perozzi, R. Al-Rfou, and S. Skiena, “Deepwalk: Online learning of social representations,” in Proceedings of the 20th ACM SIGKDD international conference on Knowledge discovery and data mining . ACM, 2014, pp. 701–710
2014
Cited alongside, same era.
N. Kalchbrenner, E. Grefenstette, and P. Blunsom, “A convolutional neural network for modelling sentences,” in ACL , 2014, pp. 655–665
2014
Cited alongside, same era.
Y. Lecun, Y. Bengio, and G. Hinton, “Deep learning,” Nature , vol. 521, no. 7553, p. 436, 2015
2015
2016
Later among the works it cites.
Z. Yang, D. Yang, C. Dyer, X. He, A. Smola, and E. Hovy, “Hierarchical attention networks for document classification,” in NAACL , 2017, pp. 1480–1489
2017
Later among the works it cites.
J. Liu, W. Chang, Y. Wu, and Y. Yang, “Deep learning for extreme multi-label text classification,” in SIGIR , 2017, pp. 115–124
2017
Later among the works it cites.
P. Liu, X. Qiu, and X. Huang, “Adversarial multi-task learning for text classification,” in ACL , 2017, pp. 1–10
2017
Later among the works it cites.
S. Sabour, N. Frosst, and G. E. Hinton, “Dynamic routing between capsules,” in NIPS , 2017, pp. 3859–3869
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
K. S. Tai, R. Socher, and C. D. Manning, “Improved semantic representations from tree-structured long short-term memory networks,” in ACL , 2015, pp. 1556–1566
2015
Cited alongside, same era.
D. Tang, B. Qin, and T. Liu, “Document modeling with gated recurrent neural network for sentiment classification,” in EMNLP , 2015, pp. 1422–1432
2015
Cited alongside, same era.
S. Gopal and Y. Yang, “Hierarchical bayesian inference and recursive regularization for large-scale classification,” TKDD , pp. 18:1–18:23, 2015
2015
Cited alongside, same era.
S. Lai, L. Xu, K. Liu, and J. Zhao, “Recurrent convolutional neural networks for text classification,” in AAAI , 2015, pp. 2267–2273
2015
Cited alongside, same era.
F. Rousseau, E. Kiagias, and M. Vazirgiannis, “Text categorization as a graph classification problem,” in ACL , 2015, pp. 1702–1712
2015
Cited alongside, same era.
X. Zhang, J. J. Zhao, and Y. LeCun, “Character-level convolutional networks for text classification,” in NIPS , 2015, pp. 649–657
2015
Cited alongside, same era.
2015
Cited alongside, same era.
Y. Dong, N. V. Chawla, and A. Swami, “metapath2vec: Scalable representation learning for heterogeneous networks,” in Proceedings of the 23rd ACM SIGKDD international conference on knowledge discovery and data mining . ACM, 2017, pp. 135–144
2017
Later among the works it cites.
Z. Lin, M. Feng, C. N. d. Santos, M. Yu, B. Xiang, B. Zhou, and Y. Bengio, “A structured self-attentive sentence embedding,” in ICLR , 2017
2017
Later among the works it cites.
T. Shen, T. Zhou, G. Long, J. Jiang, and C. Zhang, “Bi-directional block self-attention for fast and memory-efficient sequence modeling,” in ICLR , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
H. Peng, J. Li, Y. He, Y. Liu, M. Bao, L. Wang, Y. Song, and Q. Yang, “Large-scale hierarchical text classification with recursively regularized deep graph-cnn,” in WWW , 2018, pp. 1063–1072
2018
Later among the works it cites.
M. Yang, W. Zhao, J. Ye, Z. Lei, Z. Zhao, and S. Zhang, “Investigating capsule networks with dynamic routing for text classification,” in EMNLP , 2018, pp. 3110–3119
2018
Later among the works it cites.
T. Shen, T. Zhou, G. Long, J. Jiang, S. Pan, and C. Zhang, “Disan: Directional self-attention network for rnn/cnn-free language understanding,” in Thirty-Second AAAI Conference on Artificial Intelligence , 2018
2018
Later among the works it cites.
L. Xiao, H. Zhang, W. Chen, Y. Wang, and Y. Jin, “Mcapsnet: Capsule network for text with multi-task learning,” in Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing , 2018, pp. 4565–4574
2018
Later among the works it cites.
G. Hinton, N. Frosst, and S. Sabour, “Matrix capsules with em routing,” 2018
2018
Later among the works it cites.
A. Jiménez-Sánchez, S. Albarqouni, and D. Mateus, “Capsule networks against medical imaging data challenges,” in Intravascular Imaging and Computer Assisted Stenting and Large-Scale Annotation of Biomedical Data and Expert Label Synthesis . Springer, 2018, pp. 150–160
2018
Later among the works it cites.
P. Yang, X. Sun, W. Li, S. Ma, W. Wu, and H. Wang, “Sgm: sequence generation model for multi-label classification,” in ACL , 2018, pp. 3915–3926
2018
Later among the works it cites.
N. Zhang, S. Deng, Z. Sun, X. Chen, W. Zhang, and H. Chen, “Attention-based capsule networks with dynamic routing for relation extraction,” in Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing , 2018, pp. 986–992
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
M. Niepert, M. Ahmed, and K. Kutzkov, “Learning convolutional neural networks for graphs,” in ICML , 2016, pp. 2014–2023
2023
Closest in time.