Fetching the paper…
Reading the bibliography…
Text classification is a fundamental problem in information retrieval with many real-world applications, such as predicting the topics of online articles and the categories of e-commerce product descriptions.
A. K. McCallum, K. Nigam, J. Rennie, and K. Seymore, “Automating the construction of internet portals with machine learning,” Information Retrieval , vol. 3, no. 2, pp. 127–163, 2000
2000
Earlier work this paper cites.
T. Mikolov, K. Chen, G. Corrado, and J. Dean, “Efficient estimation of word representations in vector space,” arXiv , 2013
2013
Earlier work this paper cites.
B. Perozzi, R. Al-Rfou, and S. Skiena, “DeepWalk: Online learning of social representations,” in KDD , 2014, pp. 701–710
2014
Earlier work this paper cites.
Z. Yang, W. Cohen, and R. Salakhudinov, “Revisiting semi-supervised learning with graph embeddings,” in ICML , 2016, pp. 40–48
2016
Earlier work this paper cites.
K. Sohn, “Improved deep metric learning with multi-class n-pair loss objective,” NeurIPS , vol. 29, 2016
2016
Earlier work this paper cites.
R. Sennrich, B. Haddow, and A. Birch, “Neural machine translation of rare words with subword units,” in ACL , 2016, pp. 1715–1725
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” NeurIPS , vol. 30, 2017
2017
Earlier work this paper cites.
T. N. Kipf and M. Welling, “Semi-supervised classification with graph convolutional networks,” in ICLR . OpenReview.net, 2017
2017
Earlier work this paper cites.
T. Miyato, A. M. Dai, and I. J. Goodfellow, “Adversarial training methods for semi-supervised text classification,” in ICLR . OpenReview.net, 2017
2017
Earlier work this paper cites.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” in ICML , 2017, pp. 1126–1135
2017
Earlier work this paper cites.
D. Ha, A. M. Dai, and Q. V. Le, “Hypernetworks,” in ICLR , 2017
2017
Earlier work this paper cites.
W. Hamilton, Z. Ying, and J. Leskovec, “Inductive representation learning on large graphs,” in NeurIPS , 2017
2017
Earlier work this paper cites.
Y. Xian, B. Schiele, and Z. Akata, “Zero-shot learning-the good, the bad and the ugly,” in CVPR , 2017, pp. 4582–4591
2017
Earlier work this paper cites.
A. Radford, K. Narasimhan, T. Salimans, I. Sutskever et al. , “Improving language understanding by generative pre-training,” OpenAI blog , 2018
2018
Earlier work this paper cites.
P. Veličković, G. Cucurull, A. Casanova, A. Romero, P. Liò, and Y. Bengio, “Graph attention networks,” in ICLR , 2018
2018
Earlier work this paper cites.
M. Yu, X. Guo, J. Yi, S. Chang, S. P. Y. C. G. Tesauro, H. W. B. Zhou, and A. Foundations-Learning, “Diverse few-shot text classification with multiple metrics,” in NAACL , 2018, pp. 1206–1215
2018
Earlier work this paper cites.
X. Han, H. Zhu, P. Yu, Z. Wang, Y. Yao, Z. Liu, and M. Sun, “FewRel: A large-scale supervised few-shot relation classification dataset with state-of-the-art evaluation,” in EMNLP , 2018, pp. 4803–4809
2018
Earlier work this paper cites.
J. Liu, Z. He, L. Wei, and Y. Huang, “Content to node: Self-translation network embedding,” in KDD , 2018, pp. 1794–1802
2018
Earlier work this paper cites.
G. Wang, C. Li, W. Wang, Y. Zhang, D. Shen, X. Zhang, R. Henao, and L. Carin, “Joint embedding of words and labels for text classification,” in ACL , 2018, pp. 2321–2331
2018
Earlier work this paper cites.
Y. Wu and T. Lee, “Reducing model complexity for dnn based large-scale audio classification,” in ICASSP , 2018, pp. 331–335
2018
Earlier work this paper cites.
H. Xu, B. Liu, L. Shu, and P. Yu, “Open-world learning and application to product classification,” in WWW , 2019, pp. 3413–3419
2019
Earlier work this paper cites.
J. D. M.-W. C. Kenton and L. K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in NAACL , 2019, pp. 4171–4186
2019
Earlier work this paper cites.
P. Velickovic, W. Fedus, W. L. Hamilton, P. Liò, Y. Bengio, and R. D. Hjelm, “Deep graph infomax.” ICLR , vol. 2, no. 3, p. 4, 2019
2019
Earlier work this paper cites.
W. Hu, B. Liu, J. Gomes, M. Zitnik, P. Liang, V. Pande, and J. Leskovec, “Strategies for pre-training graph neural networks,” in ICLR , 2019
2019
Earlier work this paper cites.
K. Xu, W. Hu, J. Leskovec, and S. Jegelka, “How powerful are graph neural networks?” in ICLR . OpenReview.net, 2019
2019
Earlier work this paper cites.
H. Linmei, T. Yang, C. Shi, H. Ji, and X. Li, “Heterogeneous graph attention networks for semi-supervised short text classification,” in EMNLP , 2019, pp. 4821–4830
2019
Earlier work this paper cites.
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. R. Salakhutdinov, and Q. V. Le, “Xlnet: Generalized autoregressive pretraining for language understanding,” NeurIPS , vol. 32, 2019
2019
Earlier work this paper cites.
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov, “RoBERTa: A robustly optimized bert pretraining approach,” arXiv , 2019
2019
Earlier work this paper cites.
F. Zhou, C. Cao, K. Zhang, G. Trajcevski, T. Zhong, and J. Geng, “Meta-GNN: On few-shot node classification in graph meta-learning,” in CIKM , 2019, pp. 2357–2360
2019
Earlier work this paper cites.
J. Ni, J. Li, and J. McAuley, “Justifying recommendations using distantly-labeled reviews and fine-grained aspects,” in EMNLP , 2019, pp. 188–197
2019
Cited alongside, same era.
L. Yao, C. Mao, and Y. Luo, “Graph convolutional networks for text classification,” in AAAI , vol. 33, no. 01, 2019, pp. 7370–7377
2019
Cited alongside, same era.
2019
Cited alongside, same era.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” NeurIPS , vol. 33, pp. 1877–1901, 2020
2020
Cited alongside, same era.
Z. Wu, S. Pan, F. Chen, G. Long, C. Zhang, and S. Y. Philip, “A comprehensive survey on graph neural networks,” TNNLS , vol. 32, no. 1, pp. 4–24, 2020
W. Zhang, Y. Deng, X. Li, Y. Yuan, L. Bing, and W. Lam, “Aspect sentiment quad prediction as paraphrase generation,” in EMNLP , 2021, pp. 9209–9219
2021
Later among the works it cites.
X. Han, W. Zhao, N. Ding, Z. Liu, and M. Sun, “PTR: Prompt tuning with rules for text classification,” arXiv , 2021
2021
Later among the works it cites.
O. Sainz, O. L. de Lacalle, G. Labaka, A. Barrena, and E. Agirre, “Label verbalization and entailment for effective zero and few-shot relation extraction,” in EMNLP , 2021, pp. 1199–1212
2021
Later among the works it cites.
Z. Wen, Y. Fang, and Z. Liu, “Meta-inductive node classification across graphs,” in SIGIR , 2021, pp. 1219–1228
2021
Later among the works it cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in ICML , 2021, pp. 8748–8763
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
Z. Hu, Y. Dong, K. Wang, K.-W. Chang, and Y. Sun, “Gpt-gnn: Generative pre-training of graph neural networks,” in KDD , 2020, pp. 1857–1867
2020
Cited alongside, same era.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, P. J. Liu et al. , “Exploring the limits of transfer learning with a unified text-to-text transformer.” J. Mach. Learn. Res. , vol. 21, no. 140, pp. 1–67, 2020
2020
Cited alongside, same era.
T. Shin, Y. Razeghi, R. L. Logan IV, E. Wallace, and S. Singh, “AutoPrompt: Eliciting knowledge from language models with automatically generated prompts,” in EMNLP , 2020, pp. 4222–4235
2020
Cited alongside, same era.
Q. Xie, Z. Dai, E. Hovy, T. Luong, and Q. Le, “Unsupervised data augmentation for consistency training,” NeurIPS , vol. 33, pp. 6256–6268, 2020
2020
Cited alongside, same era.
J. Chen, Z. Yang, and D. Yang, “MixText: Linguistically-informed interpolation of hidden space for semi-supervised text classification,” in ACL , 2020, pp. 2147–2157
2020
Cited alongside, same era.
T. Bansal, R. Jha, T. Munkhdalai, and A. McCallum, “Self-supervised meta-learning for few-shot natural language classification tasks,” in EMNLP , 2020, pp. 522–534
2020
Cited alongside, same era.
Y. Bao, M. Wu, S. Chang, and R. Barzilay, “Few-shot text classification with distributional signatures,” in ICLR , 2020
2020
Cited alongside, same era.
2021
Later among the works it cites.
Z. Wang, J. Wang, Y. Guo, and Z. Gong, “Zero-shot node classification with decomposed graph prototype network,” in KDD , 2021, pp. 1769–1779
2021
Later among the works it cites.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen, “LoRA: Low-rank adaptation of large language models,” arXiv , 2021
2021
Later among the works it cites.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Conditional prompt learning for vision-language models,” in CVPR , 2022, pp. 16 816–16 825
2022
Later among the works it cites.
S. Hu, N. Ding, H. Wang, Z. Liu, J. Wang, J. Li, W. Wu, and M. Sun, “Knowledgeable prompt-tuning: Incorporating knowledge into prompt verbalizer for text classification,” in ACL , 2022, pp. 2225–2240
2022
Later among the works it cites.
S. Min, M. Lewis, H. Hajishirzi, and L. Zettlemoyer, “Noisy channel language model prompting for few-shot text classification,” in ACL , 2022, pp. 5316–5330
2022
Later among the works it cites.
Z. Tan, X. Zhang, S. Wang, and Y. Liu, “MSP: Multi-stage prompting for making pre-trained language models better translators,” in ACL , 2022, pp. 6131–6142
2022
Later among the works it cites.
X. Chen, N. Zhang, X. Xie, S. Deng, Y. Yao, C. Tan, F. Huang, L. Si, and H. Chen, “KnowPrompt: Knowledge-aware prompt-tuning with synergistic optimization for relation extraction,” in WWW , 2022, pp. 2778–2788
2022
Later among the works it cites.
M. Sun, K. Zhou, X. He, Y. Wang, and X. Wang, “GPPT: Graph pre-training and prompt tuning to generalize graph neural networks,” in KDD , 2022, pp. 1717–1727
2022
Later among the works it cites.
T. Fang, Y. Zhang, Y. Yang, and C. Wang, “Prompt tuning for graph neural networks,” arXiv , 2022
2022
Later among the works it cites.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,” IJCV , vol. 130, no. 9, pp. 2337–2348, 2022
2022
Later among the works it cites.
X. Liu, K. Ji, Y. Fu, W. Tam, Z. Du, Z. Yang, and J. Tang, “P-tuning: Prompt tuning can be comparable to fine-tuning across scales and tasks,” in ACL , 2022, pp. 61–68
2022
Later among the works it cites.
N. Muennighoff, T. Wang, L. Sutawika, A. Roberts, S. Biderman, T. L. Scao, M. S. Bari, S. Shen, Z.-X. Yong, H. Schoelkopf et al. , “Crosslingual generalization through multitask finetuning,” arXiv , 2022
2022
Later among the works it cites.
Z. Wen and Y. Fang, “Augmenting low-resource text classification with graph-grounded pre-training and prompting,” arXiv , 2023
2023
Closest in time.
J. Zhao, M. Qu, C. Li, H. Yan, Q. Liu, R. Li, X. Xie, and J. Tang, “Learning on large-scale text-attributed graphs via variational inference,” in ICLR , 2023
2023
Closest in time.
Z. Tan, R. Guo, K. Ding, and H. Liu, “Virtual node tuning for few-shot node classification,” arXiv , 2023
2023
Closest in time.
Z. Liu, X. Yu, Y. Fang, and X. Zhang, “GraphPrompt: Unifying pre-training and downstream tasks for graph neural networks,” in WWW , 2023, pp. 417–428
2023
Closest in time.
Z. Wen, “Generalizing graph neural network across graphs and time,” in WSDM , 2023, pp. 1214–1215
2023
Closest in time.
J. Bai, S. Bai, Y. Chu, Z. Cui, K. Dang, X. Deng, Y. Fan, W. Ge, Y. Han, F. Huang et al. , “Qwen technical report,” arXiv , 2023
2023
Closest in time.
J. Achiam, S. Adler, S. Agarwal, L. Ahmad, I. Akkaya, F. L. Aleman, D. Almeida, J. Altenschmidt, S. Altman, S. Anadkat et al. , “GPT-4 technical report,” arXiv , 2023
2023
Closest in time.
“Llama 3 model card,” https://llama.meta.com/docs/model-cards-and-prompt-formats/meta-llama-3/Llama 3 Model Card , 2024
2024
Closest in time.
“A 13b large language model developed by baichuan intelligent technology.” https://github.com/baichuan-inc/Baichuan-13B , 2024
2024
Closest in time.
L. Zheng, W.-L. Chiang, Y. Sheng, S. Zhuang, Z. Wu, Y. Zhuang, Z. Lin, Z. Li, D. Li, E. Xing et al. , “Judging LLM-as-a-judge with MT-bench and chatbot arena,” NeuraIPS , vol. 36, 2024
2024
Closest in time.
“An overview of bard: an early experiment with generative ai,” https://ai.google/static/documents/google-about-bard.pdf , 2024
2024
Closest in time.
“Alpaca: A strong, replicable instruction-following model,” https://crfm.stanford.edu/2023/03/13/alpaca.html , 2024
2024
Closest in time.