Fetching the paper…
Reading the bibliography…
Unlike traditional unsupervised clustering, semi-supervised clustering allows users to provide meaningful structure to the data, which helps the clustering algorithm to match the user's intent.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
The hungarian method for the assignment problem
Harold W. Kuhn. 1955 · 1955
Earlier work this paper cites.
Least squares quantization in pcm
S. Lloyd. 1982 · 1982
Earlier work this paper cites.
Clustering with instance-level constraints
Kiri L. Wagstaff and Claire Cardie. 2000 · 2000
Earlier work this paper cites.
Semi-supervised clustering by seeding
Sugato Basu, Arindam Banerjee, and Raymond J. Mooney. 2002 · 2002
Earlier work this paper cites.
Active semi-supervision for pairwise constrained clustering
Sugato Basu, Arindam Banerjee, and Raymond J. Mooney. 2004 · 2004
Earlier work this paper cites.
Using encyclopedic knowledge for named entity disambiguation
Razvan C. Bunescu and Marius Pasca. 2006 · 2006
Earlier work this paper cites.
k-means++: the advantages of careful seeding
David Arthur and Sergei Vassilvitskii. 2007 · 2007
Earlier work this paper cites.
Open information extraction from the web
Michele Banko, Michael J. Cafarella, Stephen Soderland, Matthew Broadhead, and Oren Etzioni. 2007 · 2007
Earlier work this paper cites.
Learning to link with wikipedia
David N. Milne and Ian H. Witten. 2008 · 2008
Earlier work this paper cites.
Which clustering do you want? inducing your ideal clustering with minimal feedback
Sajib Dasgupta and Vincent Ng. 2010 · 2010
Earlier work this paper cites.
Identifying relations for open information extraction
Anthony Fader, Stephen Soderland, and Oren Etzioni. 2011 · 2011
Earlier work this paper cites.
A survey of text clustering algorithms
Charu C. Aggarwal and ChengXiang Zhai. 2012 · 2012
Cited alongside, same era.
Local algorithms for interactive clustering
Pranjal Awasthi, Maria-Florina Balcan, and Konstantin Voevodski. 2013 · 2013
Cited alongside, same era.
Translating embeddings for modeling multi-relational data
Antoine Bordes, Nicolas Usunier, Alberto Garcia-Durán, Jason Weston, and Oksana Yakhnenko. 2013 · 2013
Cited alongside, same era.
Clustering: Probably approximately useless?
Rich Caruana. 2013 · 2013
Cited alongside, same era.
Unsupervised deep embedding for clustering analysis
Junyuan Xie, Ross B. Girshick, and Ali Farhadi. 2015 · 2015
Cited alongside, same era.
A model-based approach for text clustering with outlier detection
Jianhua Yin and Jianyong Wang. 2016 · 2016
Cited alongside, same era.
An evaluation dataset for intent classification and out-of-scope prediction
Stefan Larson, Anish Mahendran, Joseph J. Peper, Christopher Clarke, Andrew Lee, Parker Hill, Jonathan K. Kummerfeld, Kevin Leach, Michael A. Laurenzano, Lingjia Tang, and Jason Mars. 2019 · 2019
Later among the works it cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Later among the works it cites.
A framework for deep constrained clustering - algorithms and advances
Hongjing Zhang, Sugato Basu, and Ian Davidson. 2019 · 2019
Later among the works it cites.
Interactive clustering: A comprehensive review
Juhee Bae, Tove Helldin, Maria Riveiro, Sławomir Nowaczyk, Mohamed-Rafik Bouguelia, and Göran Falkman. 2020 · 2020
Later among the works it cites.
Efficient intent detection with dual sentence encoders
Iñigo Casanueva, Tadas Temčinas, Daniela Gerz, Matthew Henderson, and Ivan Vulić. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A method to accelerate human in the loop clustering
Anni Coden, Marina Danilevsky, Daniel F. Gruhl, Linda Kato, and Meena Nagarajan. 2017 · 2017
Cited alongside, same era.
Minie: Minimizing facts in open information extraction
Kiril Gashteovski, Rainer Gemulla, and Luciano Del Corro. 2017 · 2017
Cited alongside, same era.
A data-driven analysis of workers’ earnings on amazon mechanical turk
Kotaro Hara, Abigail Adams, Kristy Milland, Saiph Savage, Chris Callison-Burch, and Jeffrey P. Bigham. 2017 · 2018
Cited alongside, same era.
Cesi: Canonicalizing open knowledge bases using embeddings and side information
Shikhar Vashishth, Prince Jain, and Partha Talukdar. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Opiec: An open information extraction corpus
Kiril Gashteovski, Sebastian Wanner, Sven Hertling, Samuel Broscheit, and Rainer Gemulla. 2019 · 2019
Cited alongside, same era.
Supporting clustering with contrastive learning
Dejiao Zhang, Feng Nan, Xiaokai Wei, Shang-Wen Li, Henghui Zhu, Kathleen McKeown, Ramesh Nallapati, Andrew O. Arnold, and Bing Xiang. 2021 · 2021
Later among the works it cites.
Multi-view clustering for open knowledge base canonicalization
Wei Shen, Yang Yang, and Yinan Liu. 2022 · 2022
Later among the works it cites.
One embedder, any task: Instruction-finetuned text embeddings
Hongjin Su, Weijia Shi, Jungo Kasai, Yizhong Wang, Yushi Hu, Mari Ostendorf, Wen-tau Yih, Noah A. Smith, Luke Zettlemoyer, and Tao Yu. 2022 · 2022
Later among the works it cites.
GPTscore: Evaluate as you desire
Jinlan Fu, See-Kiong Ng, Zhengbao Jiang, and Pengfei Liu. 2023 · 2023
Closest in time.
Large language models as simulated economic agents: What can we learn from homo silicus?
John J Horton. 2023 · 2023
Closest in time.
Generative agents: Interactive simulacra of human behavior
Joon Sung Park, Joseph C O’Brien, Carrie J Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2023 · 2023
Closest in time.
Clusterllm: Large language models as a guide for text clustering
Yuwei Zhang, Zihan Wang, and Jingbo Shang. 2023 · 2023
Closest in time.