Fetching the paper…
Reading the bibliography…
Short text clustering is a challenging task due to the lack of signal contained in such short texts.
A. K. Jain, R. C. Dubes, Algorithms for Clustering Data, Prentice-Hall, Inc., Upper Saddle River, NJ, USA, 1988
1988
Earlier work this paper cites.
V. Kumar, An introduction to cluster analysis for data mining, Tech. rep., Dept. of Computer Science, Univ. of Minnesota, Minneapolis, MN, https://www-users.cs.umn.edu/~hanxx023/dmclass/cluster_survey_10_02_00.pdf (2000)
2000
Earlier work this paper cites.
L. M. Manevitz, M. Yousef, One-class svms for document classification, Journal of Machine Learning Research 2 (2002) 139–154
2002
Earlier work this paper cites.
S. Shekhar, C.-T. Lu, P. Zhang, A unified approach to detecting spatial outliers, GeoInformatica 7 (2) (2003) 139–166
2003
Earlier work this paper cites.
Li Maokuan, Cheng Yusheng, Zhao Honghai, Unlabeled data classification via support vector machines and k-means clustering, in: Proceedings. International Conference on Computer Graphics, Imaging and Visualization, 2004. CGIV 2004., 2004, pp. 183–186
2004
Earlier work this paper cites.
S. Banerjee, K. Ramanathan, A. Gupta, Clustering short texts using wikipedia, in: Proceedings of the 30th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, SIGIR ’07, ACM, Amsterdam, The Netherlands, 2007, pp. 787–788
2007
Earlier work this paper cites.
D. Arthur, S. Vassilvitskii, K-means++: The advantages of careful seeding, in: Proceedings of the Eighteenth Annual ACM-SIAM Symposium on Discrete Algorithms, SODA ’07, New Orleans, Louisiana, 2007, pp. 1027–1035
2007
Earlier work this paper cites.
F. T. Liu, K. M. Ting, Z.-H. Zhou, Isolation forest, in: Proceedings of the 2008 Eighth IEEE International Conference on Data Mining, ICDM ’08, IEEE Computer Society, Washington, DC, USA, 2008, pp. 413–422
2008
Earlier work this paper cites.
X. Phan, L. Nguyen, S. Horiguchi, Learning to classify short and sparse text & web with hidden topics from large-scale data collections, in: Proceedings of the 17th International Conference on World Wide Web, Beijing, China, 2008, pp. 91–100
2008
Earlier work this paper cites.
M. Mathioudakis, N. Koudas, Twittermonitor: Trend detection over the twitter stream, in: Proceedings of the 2010 ACM SIGMOD International Conference on Management of Data, SIGMOD ’10, Association for Computing Machinery, New York, USA, 2010, pp. 1155–1158
2010
Earlier work this paper cites.
T. Gollub, B. Stein, Unsupervised sparsification of similarity graphs, in: H. Locarek-Junge, C. Weihs (Eds.), Classification as a Tool for Research, Springer, Berlin, Heidelberg, 2010, pp. 71–79
2010
Earlier work this paper cites.
S. Na, L. Xumin, G. Yong, Research on k-means clustering algorithm: An improved k-means clustering algorithm, in: 2010 Third International Symposium on Intelligent Information Technology and Security Informatics, Jinggangshan, China, 2010, pp. 63–67
2010
Cited alongside, same era.
R. Navigli, S. P. Ponzetto, Babelnet: The automatic construction, evaluation and application of a wide-coverage multilingual semantic network, Artificial Intelligence 193 (2012) 217–250
2012
Cited alongside, same era.
C. C. Aggarwal, C. Zhai, A Survey of Text Classification Algorithms, Springer US, Boston, MA, 2012, pp. 163–222
2012
Cited alongside, same era.
T. Mikolov, K. Chen, G. Corrado, J. Dean, Efficient estimation of word representations in vector space, Proceedings of Workshop at ICLR 2013 (01 2013)
2013
Cited alongside, same era.
J. Xu, B. Xu, P. Wang, S. Zheng, G. Tian, J. Zhao, B. Xu, Self-taught convolutional neural networks for short text clustering, Neural Networks 88 (2017) 22–31
2017
Later among the works it cites.
M. R. H. Rakib, M. Jankowska, N. Zeh, E. Milios, Improving short text clustering by similarity matrix sparsification, in: Proceedings of the ACM Symposium on Document Engineering 2018, DocEng ’18, ACM, Halifax, NS, Canada, 2018, pp. 50:1–50:4
2018
Later among the works it cites.
C. Zheng, C. Liu, H. Wong, Corpus-based topic diffusion for short text clustering, Neurocomputing 275 (2018) 2444–2458
2018
Later among the works it cites.
S. Kanj, T. Brüls, S. Gazut, Shared nearest neighbor clustering in a locality sensitive hashing framework, Journal of Computational Biology 25 (2) (2018) 236–250
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
X. Cheng, X. Yan, Y. Lan, J. Guo, Btm: Topic modeling over short texts, IEEE Transactions on Knowledge and Data Engineering 26 (12) (2014) 2928–2941
2014
Cited alongside, same era.
J. Pennington, R. Socher, C. D. Manning, Glove: Global vectors for word representation, in: Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing, Doha, Qatar, 2014, pp. 1532–1543
2014
Cited alongside, same era.
X. Zhang, Y. LeCun, Text understanding from scratch (2015). URL http://arxiv.org/abs/1502.01710
2015
Cited alongside, same era.
J. Yin, J. Wang, A model-based approach for text clustering with outlier detection, in: 2016 IEEE 32nd International Conference on Data Engineering (ICDE), 2016, pp. 625–636
2016
Cited alongside, same era.
M. Sanchiz, J. Chin, A. Chevalier, W. Fu, F. Amadieu, J. He, Searching for information on the web, Information Processing & Management 53 (1) (2017) 281–294
2017
Cited alongside, same era.
M. Allahyari, S. Pouriyeh, M. Assefi, S. Safaei, E. D. Trippe, J. B. Gutierrez, K. Kochut, A brief survey of text mining: Classification, clustering and extraction techniques (2017) · 2017
Cited alongside, same era.
2018
Later among the works it cites.
C. Orăsan, Automatic summarization: 25 years on, Natural Language Engineering 25 (6) (2019) 735–751
2019
Later among the works it cites.
C.-H. Chee, J. Jaafar, I. A. Aziz, M. H. Hasan, W. Yeoh, Algorithms for frequent itemset mining: a literature review, Artificial Intelligence Review 52 (4) (2019) 2603–2621
2019
Later among the works it cites.
A. Hadifar, L. Sterckx, T. Demeester, C. Develder, A self-training approach for short text clustering, in: Proceedings of the 4th Workshop on Representation Learning for NLP, Association for Computational Linguistics, Florence, Italy, 2019, pp. 194–199
2019
Later among the works it cites.
M. Kozlowski, H. Rybinski, Clustering of semantically enriched short texts, Journal of Intelligent Information Systems 53 (1) (2019) 69–92
2019
Later among the works it cites.
S. Yang, G. Huang, B. Cai, Discovering topic representative terms for short text clustering, IEEE Access 7 (2019) 92037–92047
2019
Later among the works it cites.
S. A. Curiskis, B. Drake, T. R. Osborn, P. J. Kennedy, An evaluation of document clustering and topic modelling in two online social networks: Twitter and reddit, Information Processing & Management 57 (2) (2020) 102034
2020
Closest in time.