Fetching the paper…
Reading the bibliography…
Despite the remarkable success of Large Language Models (LLMs) in text understanding and generation, their potential for text clustering tasks remains underexplored.
Comparing partitions
Lawrence Hubert and Phipps Arabie · 1985
Earlier work this paper cites.
Correlation clustering
Nikhil Bansal, Avrim Blum, and Shuchi Chawla · 2002
Earlier work this paper cites.
Discriminative training methods for hidden markov models: Theory and experiments with perceptron algorithms
Michael Collins · 2002
Earlier work this paper cites.
Support vector machine learning for interdependent and structured output spaces
Ioannis Tsochantaridis, Thomas Hofmann, Thorsten Joachims, and Yasemin Altun · 2004
Earlier work this paper cites.
Supervised clustering with support vector machines
Thomas Finley and Thorsten Joachims · 2005
Earlier work this paper cites.
Supervised clustering of streaming data for email batch detection
Peter Haider, Ulf Brefeld, and Tobias Scheffer · 2007
Earlier work this paper cites.
Supervised k-means clustering
Thomas Finley and Thorsten Joachims · 2008
Earlier work this paper cites.
Information theoretic measures for clusterings comparison: is a correction for chance necessary?
Nguyen Xuan Vinh, Julien Epps, and James Bailey · 2009
Earlier work this paper cites.
Learning structural svms with latent variables
Chun-Nam John Yu and Thorsten Joachims · 2009
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
Latent structure perceptron with feature induction for unrestricted coreference resolution
Eraldo R. Fernandes, Cícero Nogueira dos Santos, and Ruy Luiz Milidiú · 2012
Earlier work this paper cites.
Facenet: A unified embedding for face recognition and clustering
Florian Schroff, Dmitry Kalenichenko, and James Philbin · 2015
Earlier work this paper cites.
Supervised clustering of questions into intents for dialog system applications
Iryna Haponchyk, Antonio Uva, Seunghak Yu, Olga Uryupina, and Alessandro Moschitti · 2018
Cited alongside, same era.
Justifying recommendations using distantly-labeled reviews and fine-grained aspects
Jianmo Ni, Jiacheng Li, and Julian McAuley · 2019
Cited alongside, same era.
ETC: encoding long and structured inputs in transformers
Joshua Ainslie, Santiago Ontañón, Chris Alberti, Vaclav Cvicek, Zachary Fisher, Philip Pham, Anirudh Ravula, Sumit Sanghai, Qifan Wang, and Li Yang · 2020
Cited alongside, same era.
Longformer: The long-document transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan · 2020
Cited alongside, same era.
Reformer: The efficient transformer
Nikita Kitaev, Lukasz Kaiser, and Anselm Levskaya · 2020
Cited alongside, same era.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, et al · 2022
Later among the works it cites.
Bertopic: Neural topic modeling with a class-based TF-IDF procedure
Maarten Grootendorst · 2022
Later among the works it cites.
Topic discovery via latent space clustering of pretrained language model representations
Yu Meng, Yunyi Zhang, Jiaxin Huang, Yu Zhang, and Jiawei Han · 2022
Later among the works it cites.
New intent discovery with pre-training and contrastive learning
Yuwei Zhang, Haode Zhang, Li-Ming Zhan, Xiao-Ming Wu, and Albert Y. S. Lam · 2022
Later among the works it cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Discovering new intents via constrained deep adaptive clustering with cluster refinement
Ting-En Lin, Hua Xu, and Hanlei Zhang · 2020
Cited alongside, same era.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush · 2020
Cited alongside, same era.
Supervised neural clustering via latent structured output learning: Application to question intents
Iryna Haponchyk and Alessandro Moschitti · 2021
Cited alongside, same era.
Text data augmentation for deep learning
Connor Shorten, Taghi M Khoshgoftaar, and Borko Furht · 2021
Cited alongside, same era.
Supporting clustering with contrastive learning
Dejiao Zhang, Feng Nan, Xiaokai Wei, Shang-Wen Li, Henghui Zhu, Kathleen R. McKeown, Ramesh Nallapati, Andrew O. Arnold, and Bing Xiang · 2021
Cited alongside, same era.
Short text clustering with a deep multi-embedded self-supervised model
Kai Zhang, Zheng Lian, Jiangmeng Li, Haichang Li, and Xiaohui Hu · 2021
Cited alongside, same era.
Short text clustering algorithms, application and challenges: A survey
Majid Hameed Ahmed, Sabrina Tiun, Nazlia Omar, and Nor Samsiah Sani · 2022
Cited alongside, same era.
Later among the works it cites.
Generalized category discovery with decoupled prototypical network
Wenbin An, Feng Tian, Qinghua Zheng, Wei Ding, QianYing Wang, and Ping Chen · 2023
Later among the works it cites.
Supervised clustering loss for clustering-friendly sentence embeddings: An application to intent clustering
Giorgio Barnabo, Antonio Uva, Sandro Pollastrini, Chiara Rubagotti, and Davide Bernardi · 2023
Later among the works it cites.
Text is all you need: Learning language representations for sequential recommendation
Jiacheng Li, Ming Wang, Jin Li, Jinmiao Fu, Xin Shen, Jingbo Shang, and Julian J. McAuley · 2023
Later among the works it cites.
Using LLM for improving key event discovery: Temporal-guided news stream clustering with event summaries
Nishanth Sridhar Nakshatri, Siyi Liu, Sihao Chen, Dan Roth, Dan Goldwasser, and Daniel Hopkins · 2023
Later among the works it cites.
Large language models enable few-shot clustering
Vijay Viswanathan, Kiril Gashteovski, Carolin Lawrence, Tongshuang Wu, and Graham Neubig · 2023
Later among the works it cites.
ClusterLLM: Large language models as a guide for text clustering
Yuwei Zhang, Zihan Wang, and Jingbo Shang · 2023
Later among the works it cites.