Roberta: A robustly optimized bert pretraining approach
Original
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
Probability inequalities for sums of bounded random variables
Wassily Hoeffding · 1994
Earlier work this paper cites.
Birds of a feather: Homophily in social networks
Miller McPherson, Lynn Smith-Lovin, and James M Cook · 2001
Earlier work this paper cites.
Semi-supervised learning literature survey
Xiaojin Jerry Zhu · 2005
Earlier work this paper cites.
Link prediction in complex networks: A survey
Linyuan Lü and Tao Zhou · 2011
Earlier work this paper cites.
Weisfeiler-lehman graph kernels
Nino Shervashidze, Pascal Schweitzer, Erik Jan Van Leeuwen, Kurt Mehlhorn, and Karsten M Borgwardt · 2011
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean · 2013
Earlier work this paper cites.
FastXML: A fast, accurate and stable tree-classifier for extreme multi-label learning
Yashoteja Prabhu and Manik Varma · 2014
Earlier work this paper cites.
A coarse-to-fine approach for fast deformable object detection
Marco Pedersoli, Andrea Vedaldi, Jordi Gonzalez, and Xavier Roca · 2015
Earlier work this paper cites.
Learning local feature descriptors with triplets and shallow convolutional neural networks
Vassileios Balntas, Edgar Riba, Daniel Ponsa, and Krystian Mikolajczyk · 2016
Earlier work this paper cites.
BiRank: Towards ranking on bipartite graphs
Xiangnan He, Ming Gao, Min-Yen Kan, and Dingxian Wang · 2016
Earlier work this paper cites.
Variational graph auto-encoders
Original
Thomas N Kipf and Max Welling · 2016
Earlier work this paper cites.
Inductive representation learning on large graphs
Will Hamilton, Zhitao Ying, and Jure Leskovec · 2017
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Thomas N. Kipf and Max Welling · 2017
Earlier work this paper cites.
Deep Laplacian pyramid networks for fast and accurate super-resolution
Wei-Sheng Lai, Jia-Bin Huang, Narendra Ahuja, and Ming-Hsuan Yang · 2017
Earlier work this paper cites.
Contextual stochastic block models
Yash Deshpande, Andrea Montanari, Elchanan Mossel, and Subhabrata Sen · 2018
Earlier work this paper cites.
Progressive growing of gans for improved quality, stability, and variation
Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen · 2018
Earlier work this paper cites.
Predict then propagate: Graph neural networks meet personalized pagerank
Johannes Klicpera, Aleksandar Bojchevski, and Stephan Günnemann · 2018
Earlier work this paper cites.
Parabel: Partitioned label trees for extreme classification with application to dynamic search advertising
Yashoteja Prabhu, Anil Kag, Shrutendra Harsola, Rahul Agrawal, and Manik Varma · 2018
Earlier work this paper cites.
Graph attention networks
Petar Velickovic, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Lia, and Yoshua Bengio · 2018
Earlier work this paper cites.
GraphRNN: Generating realistic graphs with deep auto-regressive models
Jiaxuan You, Rex Ying, Xiang Ren, William Hamilton, and Jure Leskovec · 2018
Earlier work this paper cites.
Cluster-gcn: An efficient algorithm for training deep and large graph convolutional networks
Wei-Lin Chiang, Xuanqing Liu, Si Si, Yang Li, Samy Bengio, and Cho-Jui Hsieh · 2019
Earlier work this paper cites.