Fetching the paper…
Reading the bibliography…
Unsupervised clustering is widely used to explore large corpora, but existing formulations neither consider the users' goals nor explain clusters' meanings.
Nettaxo: Automated topic taxonomy construction from text-rich network
Jingbo Shang, Xinyang Zhang, Liyuan Liu, Sha Li, and Jiawei Han. 2020 · 1919
Earlier work this paper cites.
Latent dirichlet allocation
David M Blei, Andrew Y Ng, and Michael I Jordan. 2003 · 2003
Earlier work this paper cites.
Unsupervised domain clusters in pretrained language models
Roee Aharoni and Yoav Goldberg. 2020b · 2004
Earlier work this paper cites.
Stability-based validation of clustering solutions
Tilman Lange, Volker Roth, Mikio L Braun, and Joachim M Buhmann. 2004 · 2004
Earlier work this paper cites.
Automatically labeling hierarchical clusters
Pucktada Treeratpituk and Jamie Callan. 2006 · 2006
Earlier work this paper cites.
Enhancing cluster labeling using wikipedia
David Carmel, Haggai Roitman, and Naama Zwerdling. 2009 · 2009
Earlier work this paper cites.
Reading tea leaves: How humans interpret topic models
Jonathan Chang, Sean Gerrish, Chong Wang, Jordan Boyd-Graber, and David Blei. 2009 · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al. 2009 · 2009
Earlier work this paper cites.
Pulp : A linear programming toolkit for python
Stuart Mitchell, Michael J. O’Sullivan, and Iain Dunning. 2011 · 2011
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay. 2011 · 2011
Earlier work this paper cites.
A survey of text clustering algorithms
Charu C Aggarwal and ChengXiang Zhai. 2012 · 2012
Earlier work this paper cites.
Interactive topic modeling
Yuening Hu, Jordan Boyd-Graber, Brianna Satinoff, and Alison Smith. 2014 · 2014
Earlier work this paper cites.
Taxonomy construction using syntactic contextual evidence
Anh Tuan Luu, Jung-jae Kim, and See Kiong Ng. 2014 · 2014
Earlier work this paper cites.
Efficient methods for inferring large sparse topic hierarchies
Doug Downey, Chandra Bhagavatula, and Yi Yang. 2015 · 2015
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015 · 2015
Earlier work this paper cites.
What makes a convincing argument? empirical analysis and detecting attributes of convincingness in web argumentation
Ivan Habernal and Iryna Gurevych. 2016 · 2016
Earlier work this paper cites.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering
Ruining He and Julian McAuley. 2016 · 2016
Earlier work this paper cites.
Learning multi-modal grounded linguistic semantics by playing" i spy"
Jesse Thomason, Jivko Sinapov, Maxwell Svetlik, Peter Stone, and Raymond J Mooney. 2016 · 2016
Cited alongside, same era.
Happydb: A corpus of 100,000 crowdsourced happy moments
Akari Asai, Sara Evensen, Behzad Golshan, Alon Halevy, Vivian Li, Andrei Lopatenko, Daniela Stepanov, Yoshihiko Suhara, Wang-Chiew Tan, and Yinzhan Xu. 2018 · 2018
Cited alongside, same era.
A Million News Headlines
Rohit Kulkarni. 2018 · 2018
Cited alongside, same era.
Automated phrase mining from massive text corpora
Jingbo Shang, Jialu Liu, Meng Jiang, Xiang Ren, Clare R Voss, and Jiawei Han. 2018 · 2018
Cited alongside, same era.
Unseen class discovery in open-world classification
Lei Shu, Hu Xu, and Bing Liu. 2018 · 2018
Cited alongside, same era.
Domino: Discovering systematic errors with cross-modal embeddings
Sabri Eyuboglu, Maya Varma, Khaled Saab, Jean-Benoit Delbrouck, Christopher Lee-Messer, Jared Dunnmon, James Zou, and Christopher Ré. 2022 · 2022
Later among the works it cites.
Explaining patterns in data with language models via interpretable autoprompting
Chandan Singh, John X Morris, Jyoti Aneja, Alexander M Rush, and Jianfeng Gao. 2022 · 2022
Later among the works it cites.
One embedder, any task: Instruction-finetuned text embeddings
Hongjin Su, Jungo Kasai, Yizhong Wang, Yushi Hu, Mari Ostendorf, Wen-tau Yih, Noah A Smith, Luke Zettlemoyer, Tao Yu, et al. 2022 · 2022
Later among the works it cites.
Generalized category discovery
Sagar Vaze, Kai Han, Andrea Vedaldi, and Andrew Zisserman. 2022 · 2022
Later among the works it cites.
Text embeddings by weakly-supervised contrastive pre-training
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Taxogen: Unsupervised topic taxonomy construction by adaptive term embedding and clustering
Chao Zhang, Fangbo Tao, Xiusi Chen, Jiaming Shen, Meng Jiang, Brian Sadler, Michelle Vanni, and Jiawei Han. 2018 · 2018
Cited alongside, same era.
Justifying recommendations using distantly-labeled reviews and fine-grained aspects
Jianmo Ni, Jiacheng Li, and Julian McAuley. 2019 · 2019
Cited alongside, same era.
Big Data Set from RateMyProfessor.com for Professors’ Teaching Evaluation
Jibo He. 2020 · 2020
Cited alongside, same era.
The Examiner - Spam Clickbait Catalog
Rohit Kulkarni. 2020 · 2020
Cited alongside, same era.
Contextualized weak supervision for text classification
Dheeraj Mekala and Jingbo Shang. 2020 · 2020
Cited alongside, same era.
Discriminative topic mining via category-name guided text embedding
Yu Meng, Jiaxin Huang, Guangyuan Wang, Zihan Wang, Chao Zhang, Yu Zhang, and Jiawei Han. 2020 · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Cited alongside, same era.
Liang Wang, Nan Yang, Xiaolong Huang, Binxing Jiao, Linjun Yang, Daxin Jiang, Rangan Majumder, and Furu Wei. 2022 · 2022
Later among the works it cites.
Towards open-set object detection and discovery
Jiyang Zheng, Weihao Li, Jie Hong, Lars Petersson, and Nick Barnes. 2022 · 2022
Later among the works it cites.
Describing differences between text distributions with natural language
Ruiqi Zhong, Charlie Snell, Dan Klein, and Jacob Steinhardt. 2022 · 2022
Later among the works it cites.
Gsclip: A framework for explaining distribution shifts in natural language
Zhiying Zhu, Weixin Liang, and James Zou. 2022 · 2022
Later among the works it cites.
Scaling laws for generative mixed-modal language models
Armen Aghajanyan, Lili Yu, Alexis Conneau, Wei-Ning Hsu, Karen Hambardzumyan, Susan Zhang, Stephen Roller, Naman Goyal, Omer Levy, and Luke Zettlemoyer. 2023 · 2023
Closest in time.
Chatgpt outperforms crowd-workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Training language models with language feedback at scale
Jérémy Scheurer, Jon Ander Campos, Tomasz Korbak, Jun Shern Chan, Angelica Chen, Kyunghyun Cho, and Ethan Perez. 2023 · 2023
Closest in time.
Explaining black box text modules in natural language with language models
Chandan Singh, Aliyah R Hsu, Richard Antonello, Shailee Jain, Alexander G Huth, Bin Yu, and Jianfeng Gao. 2023 · 2023
Closest in time.
Wot-class: Weakly supervised open-world text classification
Tianle Wang, Zihan Wang, Weitang Liu, and Jingbo Shang. 2023 · 2023
Closest in time.
Open-world weakly-supervised object localization
Jinheng Xie, Zhaochuan Luo, Yuexiang Li, Haozhe Liu, Linlin Shen, and Mike Zheng Shou. 2023 · 2023
Closest in time.
Incremental generalized category discovery
Bingchen Zhao and Oisin Mac Aodha. 2023 · 2023
Closest in time.
Goal driven discovery of distributional differences via language descriptions
Ruiqi Zhong, Peter Zhang, Steve Li, Jinwoo Ahn, Dan Klein, and Jacob Steinhardt. 2023 · 2023
Closest in time.