Fetching the paper…
Reading the bibliography…
Scientific document classification is a critical task for a wide range of applications, but the cost of obtaining massive amounts of human-labeled data can be prohibitive.
Selecting good expansion terms for pseudo-relevance feedback. In SIGIR . 243–250
Guihong Cao, Jian-Yun Nie, Jianfeng Gao, and Stephen Robertson. 2008 · 2008
Earlier work this paper cites.
Importance of Semantic Representation: Dataless Classification.. In AAAI . 830–835
Ming-Wei Chang, Lev-Arie Ratinov, Dan Roth, and Vivek Srikumar. 2008 · 2008
Earlier work this paper cites.
Reciprocal rank fusion outperforms condorcet and individual rank learning methods. In SIGIR . 758–759
Gordon V Cormack, Charles LA Clarke, and Stefan Buettcher. 2009 · 2009
Earlier work this paper cites.
The probabilistic relevance framework: BM25 and beyond
Stephen Robertson and Hugo Zaragoza. 2009 · 2009
Earlier work this paper cites.
KNN with TF-IDF based framework for text categorization
Bruno Trstenjak, Sasa Mikac, and Dzenana Donko. 2014 · 2014
Earlier work this paper cites.
Character-level Convolutional Networks for Text Classification. In NIPS
Xiang Zhang, Junbo Jake Zhao, and Yann LeCun. 2015 · 2015
Earlier work this paper cites.
Query Expansion with Locally-Trained Word Embeddings. In ACL
Fernando Diaz, Bhaskar Mitra, and Nick Craswell. 2016 · 2016
Earlier work this paper cites.
Paper2vec: Combining graph and text information for scientific paper representation. In ECIR . 383–395
Soumyajit Ganguly and Vikram Pudi. 2017 · 2017
Earlier work this paper cites.
Weakly-supervised neural text classification. In CIKM . 983–992
Yu Meng, Jiaming Shen, Chao Zhang, and Jiawei Han. 2018 · 2018
Earlier work this paper cites.
SciBERT: A Pretrained Language Model for Scientific Text. In EMNLP-IJCNLP . 3615–3620
Iz Beltagy, Kyle Lo, and Arman Cohan. 2019 · 2019
Earlier work this paper cites.
On the Use of ArXiv as a Dataset
Colin B Clement, Matthew Bierbaum, Kevin P O’Keeffe, and Alexander A Alemi. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In NAACL-HLT
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks. In EMNLP-IJCNLP . 3982–3992
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
SPECTER: Document-level Representation Learning using Citation-informed Transformers. In ACL . 2270–2282
Arman Cohan, Sergey Feldman, Iz Beltagy, Doug Downey, and Daniel S Weld. 2020 · 2020
Cited alongside, same era.
Dense Passage Retrieval for Open-Domain Question Answering. In EMNLP . 6769–6781
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020 · 2020
Cited alongside, same era.
Text classification using label names only: A language model self-training approach
Yu Meng, Yunyi Zhang, Jiaxin Huang, Chenyan Xiong, Heng Ji, Chao Zhang, and Jiawei Han. 2020 · 2020
Cited alongside, same era.
Understanding contrastive representation learning through alignment and uniformity on the hypersphere. In ICML
Tongzhou Wang and Phillip Isola. 2020 · 2020
Cited alongside, same era.
BERTopic: Neural topic modeling with a class-based TF-IDF procedure
Maarten Grootendorst. 2022 · 2022
Later among the works it cites.
Unsupervised Dense Information Retrieval with Contrastive Learning
Gautier Izacard, Mathilde Caron, Lucas Hosseini, Sebastian Riedel, Piotr Bojanowski, Armand Joulin, and Edouard Grave. 2022 · 2022
Later among the works it cites.
Fbnetgen: Task-aware gnn-based fmri analysis via functional brain network generation. In MIDL
Xuan Kan, Hejie Cui, Joshua Lukemire, Ying Guo, and Carl Yang. 2022 · 2022
Later among the works it cites.
FastClass: A Time-Efficient Approach to Weakly-Supervised Text Classification
Tingyu Xia, Yue Wang, Yuan Tian, and Yi Chang. 2022 · 2022
Later among the works it cites.
Counterfactual and factual reasoning over hypergraphs for interpretable clinical predictions on ehr. In Machine Learning for Health . PMLR, 259–278
R. Xu, Y. Yu, C. Zhang, M. K Ali, JC. Ho, and C. Yang. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
X-Class: Text Classification with Extremely Weak Supervision. In NAACL . 3043–3053
Zihan Wang, Dheeraj Mekala, and Jingbo Shang. 2021 · 2021
Cited alongside, same era.
Learning domain semantics and cross-domain correlations for paper recommendation. In SIGIR . 706–715
Yi Xie, Yuqing Sun, and Elisa Bertino. 2021 · 2021
Cited alongside, same era.
Fine-Tuning Pre-trained Language Model with Weak Supervision: A Contrastive-Regularized Self-Training Approach. In NAACL . 1063–1077
Yue Yu, Simiao Zuo, Haoming Jiang, Wendi Ren, Tuo Zhao, and Chao Zhang. 2021 · 2021
Cited alongside, same era.
WRENCH: A Comprehensive Benchmark for Weak Supervision. In NeurIPS
Jieyu Zhang, Yue Yu, Yinghao Li, Yujing Wang, Yaming Yang, Mao Yang, and Alexander Ratner. 2021 · 2021
Cited alongside, same era.
How Can Graph Neural Networks Help Document Retrieval: A Case Study on CORD19 with Concept Map Generation. In ECIR
Hejie Cui, Jiaying Lu, Yao Ge, and Carl Yang. 2022 · 2022
Cited alongside, same era.
Unsupervised Corpus Aware Language Model Pre-training for Dense Passage Retrieval. In ACL . 2843–2853
Luyu Gao and Jamie Callan. 2022 · 2022
Cited alongside, same era.
AcTune: Uncertainty-Based Active Self-Training for Active Fine-Tuning of Pretrained Language Models. In NAACL . 1422–1436
Yue Yu, Lingkai Kong, Jieyu Zhang, Rongzhi Zhang, and Chao Zhang. 2022a
Cited in the paper.
Later among the works it cites.
Pre-train Graph Neural Networks for Brain Network Analysis. In IEEE-Big Data
Yi Yang, Hejie Cui, and Carl Yang. 2022 · 2022
Later among the works it cites.
Structure-enhanced heterogeneous graph contrastive learning. In SDM
Yanqiao Zhu, Yichen Xu, Hejie Cui, Carl Yang, Qiang Liu, and Shu Wu. 2022 · 2022
Later among the works it cites.
ReSel: N-ary Relation Extraction from Scientific Text and Tables by Learning to Retrieve and Select. In EMNLP . 730–744
Yuchen Zhuang, Yinghao Li, Junyang Zhang, Yue Yu, Yingjun Mou, Xiang Chen, Le Song, and Chao Zhang. 2022 · 2022
Later among the works it cites.
Zihan Wang, Tianle Wang, Dheeraj Mekala, and Jingbo Shang. 2023 · 2023
Closest in time.
Neighborhood-Regularized Self-Training for Learning with Few Labels. In AAAI , Vol. 37
Ran Xu, Yue Yu, Hejie Cui, Xuan Kan, Yanqiao Zhu, Joyce Ho, Chao Zhang, and Carl Yang. 2023 · 2023
Closest in time.
Pre-training Multi-task Contrastive Learning Models for Scientific Literature Understanding
Yu Zhang, Hao Cheng, Zhihong Shen, Xiaodong Liu, Ye-Yi Wang, and Jianfeng Gao. 2023a · 2023
Closest in time.