Improved sample complexities for deep networks and robust classification via an all-layer margin
Original
Colin Wei and Tengyu Ma · 1910
Earlier work this paper cites.
A lower bound for the smallest eigenvalue of the laplacian
Jeff Cheeger · 1969
Earlier work this paper cites.
Eigenvalues in combinatorial optimization
Bojan Mohar and Svatopluk Poljak · 1993
Earlier work this paper cites.
Isoperimetric problems for convex bodies and a localization lemma
Ravi Kannan, László Lovász, and Miklós Simonovits · 1995
Earlier work this paper cites.
The nature of statistical learning theory
Vladimir Vapnik · 1995
Earlier work this paper cites.
Unsupervised word sense disambiguation rivaling supervised methods
David Yarowsky · 1995
Earlier work this paper cites.
An isoperimetric inequality on the discrete cube, and an elementary proof of the isoperimetric inequality in gauss space
Sergey G Bobkov et al · 1997
Earlier work this paper cites.
Spectral graph theory
Fan RK Chung and Fan Chung Graham · 1997
Earlier work this paper cites.
Combining labeled and unlabeled data with co-training
Avrim Blum and Tom Mitchell · 1998
Earlier work this paper cites.
Isoperimetric and analytic inequalities for log-concave probability measures
Sergey G Bobkov et al · 1999
Earlier work this paper cites.
Unsupervised learning: foundations of neural computation
Geoffrey E Hinton, Terrence Joseph Sejnowski, Tomaso A Poggio, et al · 1999
Earlier work this paper cites.
Learning with labeled and unlabeled data
Matthias Seeger · 2000
Earlier work this paper cites.
A simple framework for contrastive learning of visual representations
Original
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton · 2002
Earlier work this paper cites.
Pac generalization bounds for co-training
Sanjoy Dasgupta, Michael L Littman, and David A McAllester · 2002
Earlier work this paper cites.
Improved baselines with momentum contrastive learning
Original
Xinlei Chen, Haoqi Fan, Ross Girshick, and Kaiming He · 2003
Earlier work this paper cites.
Error bounds for transductive learning via compression and clustering
Philip Derbeko, Ran El-Yaniv, and Ron Meir · 2004
Earlier work this paper cites.
Co-training and expansion: Towards bridging theory and practice
Maria-Florina Balcan, Avrim Blum, and Ke Yang · 2005
Earlier work this paper cites.
Semi-supervised learning by entropy minimization
Yves Grandvalet and Yoshua Bengio · 2005
Earlier work this paper cites.
Self-training avoids using spurious features under domain shift
Original
Yining Chen, Colin Wei, Ananya Kumar, and Tengyu Ma · 2006
Earlier work this paper cites.
Expander graphs and their applications
Shlomo Hoory, Nathan Linial, and Avi Wigderson · 2006
Earlier work this paper cites.
The geometry of logconcave functions and sampling algorithms
László Lovász and Santosh Vempala · 2007
Earlier work this paper cites.
Generalization error bounds in semi-supervised classification under the cluster assumption
Philippe Rigollet · 2007
Earlier work this paper cites.
Does unlabeled data provably help? worst-case analysis of the sample complexity of semi-supervised learning
Shai Ben-David, Tyler Lu, and Dávid Pál · 2008
Earlier work this paper cites.
Unlabeled data: Now it helps, now it doesn’t
Aarti Singh, Robert Nowak, and Jerry Zhu · 2009
Earlier work this paper cites.
A discriminative model for semi-supervised learning
Maria-Florina Balcan and Avrim Blum · 2010
Earlier work this paper cites.
A theory of learning from different domains
Shai Ben-David, John Blitzer, Koby Crammer, Alex Kulesza, Fernando Pereira, and Jennifer Wortman Vaughan · 2010
Earlier work this paper cites.
Semi-Supervised Learning
Olivier Chapelle, Bernhard Schlkopf, and Alexander Zien · 2010
Earlier work this paper cites.