Fetching the paper…
Reading the bibliography…
Recently, large language models (LLMs) have demonstrated superior capabilities in understanding and zero-shot learning on textual data, promising significant advances for many text-related domains.
Language models are few-shot learners. In NeurIPS’20 , Vol. 33. 1877–1901
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach. In arXiv preprint arXiv:1907.11692
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019a · 1907
Earlier work this paper cites.
Sentence-bert: sentence embeddings using siamese bert-networks. In arXiv preprint arXiv:1908.10084
Nils Reimers and Iryna Gurevych. 2019 · 1908
Earlier work this paper cites.
Zhenghao Liu, Chenyan Xiong, Maosong Sun, and Zhiyuan Liu. 2019b · 1910
Earlier work this paper cites.
The theory of graphs. In Courier Corporation
Claude Berge. 2001 · 2001
Earlier work this paper cites.
Random graph models of social networks. In Proceedings of the national academy of Sciences , Vol. 99. 2566–2572
Mark EJ Newman, Duncan J Watts, and Steven H Strogatz. 2002 · 2002
Earlier work this paper cites.
Deberta: decoding-enhanced bert with disentangled attention. In arXiv preprint arXiv:2006.03654
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen. 2020 · 2006
Earlier work this paper cites.
Text and structural data mining of influenza mentions in web and social media. In International journal of environmental research and public health , Vol. 7. 596–615
Courtney D Corley, Diane J Cook, Armin R Mikler, and Karan P Singh. 2010 · 2010
Earlier work this paper cites.
Inductive representation learning on large graphs. In NeurIPS’17 , Vol. 30
Will Hamilton, Zhitao Ying, and Jure Leskovec. 2017 · 2017
Earlier work this paper cites.
Attention is all you need. In NeurIPS’17 , Vol. 30
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Graph attention networks. In arXiv preprint arXiv:1710.10903
Petar Veličković, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Lio, and Yoshua Bengio. 2017 · 2017
Earlier work this paper cites.
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
Johannes Gasteiger, Aleksandar Bojchevski, and Stephan Günnemann. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training. In OpenAI
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al · 2018
Earlier work this paper cites.
How powerful are graph neural networks?. In arXiv preprint arXiv:1810.00826
Keyulu Xu, Weihua Hu, Jure Leskovec, and Stefanie Jegelka. 2018 · 2018
Cited alongside, same era.
Graph convolutional neural networks for web-scale recommender systems. In KDD’18 . 974–983
Rex Ying, Ruining He, Kaifeng Chen, Pong Eksombatchai, William L Hamilton, and Jure Leskovec. 2018 · 2018
Cited alongside, same era.
Language models are unsupervised multitask learners. In OpenAI
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Cited alongside, same era.
Semantic understanding of scenes through the ade20k dataset. In IJCV’19 , Vol. 127. 302–321
Bolei Zhou, Hang Zhao, Xavier Puig, Tete Xiao, Sanja Fidler, Adela Barriuso, and Antonio Torralba. 2019 · 2019
Cited alongside, same era.
Open graph benchmark: datasets for machine learning on graphs. In NeurIPS’20 , Vol. 33. 22118–22133
Weihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong, Hongyu Ren, Bowen Liu, Michele Catasta, and Jure Leskovec. 2020 · 2020
Bert-enhanced text graph neural network for classification. In Entropy , Vol. 23. 1–1
Yiping Yang and Xiaohui Cui. 2021 · 2021
Later among the works it cites.
Promptbert: improving bert sentence embeddings with prompts. In arXiv preprint arXiv:2201.04337
Ting Jiang, Jian Jiao, Shaohan Huang, Zihan Zhang, Deqing Wang, Fuzhen Zhuang, Furu Wei, Haizhen Huang, Denvy Deng, and Qi Zhang. 2022 · 2022
Later among the works it cites.
Glm-130b: an open bilingual pre-trained model. In arXiv preprint arXiv:2210.02414
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al · 2022
Later among the works it cites.
Jianan Zhao, Meng Qu, Chaozhuo Li, Hao Yan, Qian Liu, Rui Li, Xing Xie, and Jian Tang. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multimodal post attentive profiling for influencer marketing. In WWW’20 . 2878–2884
Seungbae Kim, Jyun-Yu Jiang, Masaki Nakada, Jinyoung Han, and Wei Wang. 2020 · 2020
Cited alongside, same era.
Ernie 2.0: a continual pre-training framework for language understanding. In AAAI’20 , Vol. 34. 8968–8975
Yu Sun, Shuohuan Wang, Yukun Li, Shikun Feng, Hao Tian, Hua Wu, and Haifeng Wang. 2020 · 2020
Cited alongside, same era.
Eli Chien, Wei-Cheng Chang, Cho-Jui Hsieh, Hsiang-Fu Yu, Jiong Zhang, Olgica Milenkovic, and Inderjit S Dhillon. 2021 · 2021
Cited alongside, same era.
Simcse: simple contrastive learning of sentence embeddings. In arXiv preprint arXiv:2104.08821
Tianyu Gao, Xingcheng Yao, and Danqi Chen. 2021 · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models. In arXiv preprint arXiv:2106.09685
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
The power of scale for parameter-efficient prompt tuning. In arXiv preprint arXiv:2104.08691
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021 · 2021
Cited alongside, same era.
Prefix-tuning: optimizing continuous prompts for generation. In arXiv preprint arXiv:2101.00190
Xiang Lisa Li and Percy Liang. 2021 · 2021
Cited alongside, same era.
Zhikai Chen, Haitao Mao, Hang Li, Wei Jin, Hongzhi Wen, Xiaochi Wei, Shuaiqiang Wang, Dawei Yin, Wenqi Fan, Hui Liu, et al · 2023
Later among the works it cites.
Keyu Duan, Qian Liu, Tat-Seng Chua, Shuicheng Yan, Wei Tsang Ooi, Qizhe Xie, and Junxian He. 2023 · 2023
Later among the works it cites.
Jiayan Guo, Lun Du, and Hengyu Liu. 2023 · 2023
Later among the works it cites.
Xiaoxin He, Xavier Bresson, Thomas Laurent, and Bryan Hooi. 2023 · 2023
Later among the works it cites.
Patton: language model pretraining on text-Rich networks. In arXiv preprint arXiv:2305.12268
Bowen Jin, Wentao Zhang, Yu Zhang, Yu Meng, Xinyang Zhang, Qi Zhu, and Jiawei Han. 2023 · 2023
Later among the works it cites.
Costas Mavromatis, Vassilis N Ioannidis, Shen Wang, Da Zheng, Soji Adeshina, Jun Ma, Han Zhao, Christos Faloutsos, and George Karypis. 2023 · 2023
Later among the works it cites.
Graph neural prompting with large language models. In arXiv preprint arXiv:2309.15427
Yijun Tian, Huan Song, Zichen Wang, Haozhu Wang, Ziqing Hu, Fang Wang, Nitesh V Chawla, and Panpan Xu. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models. In arXiv preprint arXiv:2307.09288
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Evaluating generative models for graph-to-text generation. In arXiv preprint arXiv:2307.14712
Shuzhou Yuan and Michael Färber. 2023 · 2023
Later among the works it cites.