Fetching the paper…
Reading the bibliography…
Short Text Classification (STC) is crucial for processing and understanding the brief but substantial content prevalent on contemporary digital platforms.
Ohsumed: An interactive retrieval evaluation and new large test collection for research
William Hersh, Chris Buckley, TJ Leone, and David Hickam · 1994
Earlier work this paper cites.
Support-vector networks
Corinna Cortes and Vladimir Vapnik · 1995
Earlier work this paper cites.
Latent dirichlet allocation
David M Blei, Andrew Y Ng, and Michael I Jordan · 2003
Earlier work this paper cites.
Exploitingclassrelationshipsforsentimentcate gorizationwithrespectratingsales
L PaNgB · 2005
Earlier work this paper cites.
Learning to classify short and sparse text & web with hidden topics from large-scale data collections
Xuan-Hieu Phan, Le-Minh Nguyen, and Susumu Horiguchi · 2008
Earlier work this paper cites.
Short text classification improved by learning multi-granularity topics
Mengen Chen, Xiaoming Jin, and Dou Shen · 2011
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim · 2014
Earlier work this paper cites.
Pte: Predictive text embedding through large-scale heterogeneous text networks
Jian Tang, Meng Qu, and Qiaozhu Mei · 2015
Earlier work this paper cites.
Semantic clustering and convolutional neural network for short text categorization
Peng Wang, Jiaming Xu, Bo Xu, Chenglin Liu, Heng Zhang, Fangyuan Wang, and Hongwei Hao · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun · 2015
Earlier work this paper cites.
Recurrent neural network for text classification with multi-task learning
Pengfei Liu, Xipeng Qiu, and Xuanjing Huang · 2016
Earlier work this paper cites.
Semantic expansion using word embedding clustering and convolutional neural network for improving short text classification
Peng Wang, Bo Xu, Jiaming Xu, Guanhua Tian, Cheng-Lin Liu, and Hongwei Hao · 2016
Earlier work this paper cites.
Improving short text classification by learning vector representations of both words and hidden topics
Heng Zhang and Guoqiang Zhong · 2016
Earlier work this paper cites.
Combining knowledge with deep convolutional neural networks for short text classification
Jin Wang, Zhongyuan Wang, Dawei Zhang, and Jun Yan · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Topic memory networks for short text classification
Jichuan Zeng, Jing Li, Yan Song, Cuiyun Gao, Michael R Lyu, and Irwin King · 2018
Earlier work this paper cites.
Graph convolutional networks for text classification
Liang Yao, Chengsheng Mao, and Yuan Luo · 2019
Earlier work this paper cites.
Multi-task deep neural networks for natural language understanding
Xiaodong Liu, Pengcheng He, Weizhu Chen, and Jianfeng Gao · 2019
Earlier work this paper cites.
Deep short text classification with knowledge powered attention
Jindong Chen, Yizhou Hu, Jingping Liu, Yanghua Xiao, and Haiyun Jiang · 2019
Earlier work this paper cites.
Text level graph neural network for text classification
Lianzhe Huang, Dehong Ma, Sujian Li, Xiaodong Zhang, and Houfeng Wang · 2019
Earlier work this paper cites.
Heterogeneous graph attention networks for semi-supervised short text classification
Hu Linmei, Tianchi Yang, Chuan Shi, Houye Ji, and Xiaoli Li · 2019
Cited alongside, same era.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu · 2020
Cited alongside, same era.
A survey on text classification: From shallow to deep learning
Qian Li, Hao Peng, Jianxin Li, Congying Xia, Renyu Yang, Lichao Sun, Philip S Yu, and Lifang He · 2020
Cited alongside, same era.
Every document owns its structure: Inductive text classification via graph neural networks
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Self-consistency improves chain of thought reasoning in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou · 2022
Later among the works it cites.
Star: Bootstrapping reasoning with reasoning
Eric Zelikman, Yuhuai Wu, Jesse Mu, and Noah Goodman · 2022
Later among the works it cites.
Explanations from large language models make small reasoners better
Shiyang Li, Jianshu Chen, Yelong Shen, Zhiyu Chen, Xinlu Zhang, Zekun Li, Hong Wang, Jing Qian, Baolin Peng, Yi Mao, et al · 2022
Later among the works it cites.
Distilling reasoning capabilities into smaller language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yufeng Zhang, Xueli Yu, Zeyu Cui, Shu Wu, Zhongzhen Wen, and Liang Wang · 2020
Cited alongside, same era.
Be more with less: Hypergraph attention networks for inductive text classification
Kaize Ding, Jianling Wang, Jundong Li, Dingcheng Li, and Huan Liu · 2020
Cited alongside, same era.
Tensor graph convolutional networks for text classification
Xien Liu, Xinxin You, Xiao Zhang, Ji Wu, and Ping Lv · 2020
Cited alongside, same era.
Incorporating context-relevant concepts into convolutional neural networks for short text classification
Jingyun Xu, Yi Cai, Xin Wu, Xue Lei, Qingbao Huang, Ho-fung Leung, and Qing Li · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
Generated knowledge prompting for commonsense reasoning
Jiacheng Liu, Alisa Liu, Ximing Lu, Sean Welleck, Peter West, Ronan Le Bras, Yejin Choi, and Hannaneh Hajishirzi · 2021
Cited alongside, same era.
Hierarchical heterogeneous graph representation learning for short text classification
Yaqing Wang, Song Wang, Quanming Yao, and Dejing Dou · 2021
Cited alongside, same era.
Hgat: Heterogeneous graph attention networks for semi-supervised short text classification
Tianchi Yang, Linmei Hu, Chuan Shi, Houye Ji, Xiaoli Li, and Liqiang Nie · 2021
Cited alongside, same era.
Kumar Shridhar, Alessandro Stolfo, and Mrinmaya Sachan · 2022
Later among the works it cites.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Eric Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, et al · 2022
Later among the works it cites.
Gpt-ner: Named entity recognition via large language models
Shuhe Wang, Xiaofei Sun, Xiaoya Li, Rongbin Ouyang, Fei Wu, Tianwei Zhang, Jiwei Li, and Guoyin Wang · 2023
Later among the works it cites.
Revisiting relation extraction in the era of large language models
Somin Wadhwa, Silvio Amir, and Byron C Wallace · 2023
Later among the works it cites.
Exploring the feasibility of chatgpt for event extraction
Jun Gao, Huan Zhao, Changlong Yu, and Ruifeng Xu · 2023
Later among the works it cites.
Taja Kuzman, Igor Mozetic, and Nikola Ljubešic · 2023
Later among the works it cites.
Will affective computing emerge from foundation models and general ai? a first evaluation on chatgpt
Mostafa M Amin, Erik Cambria, and Björn W Schuller · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Pal: Program-aided language models
Luyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon, Pengfei Liu, Yiming Yang, Jamie Callan, and Graham Neubig · 2023
Later among the works it cites.
Reasoning implicit sentiment with chain-of-thought prompting
Hao Fei, Bobo Li, Qian Liu, Lidong Bing, Fei Li, and Tat-Seng Chua · 2023
Later among the works it cites.
Anni Zou, Zhuosheng Zhang, Hai Zhao, and Xiangru Tang · 2023
Later among the works it cites.
Tree of thoughts: Deliberate problem solving with large language models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Thomas L Griffiths, Yuan Cao, and Karthik Narasimhan · 2023
Later among the works it cites.
Graph of thoughts: Solving elaborate problems with large language models
Maciej Besta, Nils Blach, Ales Kubicek, Robert Gerstenberger, Lukas Gianinazzi, Joanna Gajda, Tomasz Lehmann, Michal Podstawski, Hubert Niewiadomski, Piotr Nyczyk, et al · 2023
Later among the works it cites.
Algorithm of thoughts: Enhancing exploration of ideas in large language models
Bilgehan Sel, Ahmad Al-Tawaha, Vanshaj Khattar, Lu Wang, Ruoxi Jia, and Ming Jin · 2023
Later among the works it cites.
Gkd: Generalized knowledge distillation for auto-regressive sequence models
Rishabh Agarwal, Nino Vieillard, Piotr Stanczyk, Sabela Ramos, Matthieu Geist, and Olivier Bachem · 2023
Later among the works it cites.
Knowledge distillation of large language models
Yuxian Gu, Li Dong, Furu Wei, and Minlie Huang · 2023
Later among the works it cites.