Fetching the paper…
Reading the bibliography…
Self-Supervised Learning (SSL) surmises that inputs and pairwise positive relationships are enough to learn meaningful representations.
The approximation of one matrix by another of lower rank
Carl Eckart and Gale Young · 1936
Earlier work this paper cites.
The use of multiple measurements in taxonomic problems
Ronald A Fisher · 1936
Earlier work this paper cites.
Further aspects of the theory of multiple regression
Maurice S Bartlett · 1938
Earlier work this paper cites.
Theory of reproducing kernels
Nachman Aronszajn · 1950
Earlier work this paper cites.
A rotation method for computing canonical correlations
MJR Healy · 1957
Earlier work this paper cites.
Inversion of matrices by biorthogonalization and related results
Magnus R Hestenes · 1958
Earlier work this paper cites.
Multidimensional scaling by optimizing goodness of fit to a nonmetric hypothesis
Joseph B Kruskal · 1964
Earlier work this paper cites.
Singular value decomposition and least squares solutions
Gene H Golub and Christian Reinsch · 1971
Earlier work this paper cites.
Canonical ridge and econometrics of joint production
Hrishikesh D Vinod · 1976
Earlier work this paper cites.
Discrete hubbard-stratonovich transformation for fermion lattice models
Jorge E Hirsch · 1983
Earlier work this paper cites.
Manova method for analyzing repeated measures designs: an extensive primer
Ralph G O’Brien and Mary K Kaiser · 1985
Earlier work this paper cites.
Convergence rate estimates for iterative methods for a mesh symmetrie eigenvalue problem
Andrew V Knyazev · 1987
Earlier work this paper cites.
Radial basis functions, multi-variable functional interpolation and adaptive networks
David S Broomhead and David Lowe · 1988
Earlier work this paper cites.
Computing a nearest symmetric positive semidefinite matrix
Nicholas J Higham · 1988
Earlier work this paper cites.
Canonical correlations and generalized svd: applications and new algorithms
L Magnus Ewerbring and Franklin T Luk · 1989
Earlier work this paper cites.
Dynamical systems that sort lists, diagonalize matrices, and solve linear programming problems
Roger W Brockett · 1991
Earlier work this paper cites.
A general regression neural network
Donald F Specht et al · 1991
Earlier work this paper cites.
Signature verification using a" siamese" time delay neural network
Jane Bromley, Isabelle Guyon, Yann LeCun, Eduard Säckinger, and Roopak Shah · 1993
Earlier work this paper cites.
Penalized discriminant analysis
Trevor Hastie, Andreas Buja, and Robert Tibshirani · 1995
Earlier work this paper cites.
Multidimensional scaling by iterative majorization using radial basis functions
Andrew R Webb · 1995
Earlier work this paper cites.
Nonlinear canonical analysis and independence tests
Jacques Dauxois and Guy Martial Nkiet · 1998
Earlier work this paper cites.
Nonlinear component analysis as a kernel eigenvalue problem
Bernhard Schölkopf, Alexander Smola, and Klaus-Robert Müller · 1998
Earlier work this paper cites.
Fisher discriminant analysis with kernels
Sebastian Mika, Gunnar Ratsch, Jason Weston, Bernhard Scholkopf, and Klaus-Robert Mullers · 1999
Earlier work this paper cites.
Kernel and nonlinear canonical correlation analysis
Pei Ling Lai and Colin Fyfe · 2000
Earlier work this paper cites.
Nonlinear dimensionality reduction by locally linear embedding
Sam T Roweis and Lawrence K Saul · 2000
Earlier work this paper cites.
A global geometric framework for nonlinear dimensionality reduction
Joshua B Tenenbaum, Vin de Silva, and John C Langford · 2000
Earlier work this paper cites.
On a connection between kernel pca and metric multidimensional scaling
Christopher Williams · 2000
Earlier work this paper cites.
Global versus local methods in nonlinear dimensionality reduction
Vin Silva and Joshua Tenenbaum · 2002
Earlier work this paper cites.
Distance metric learning with application to clustering with side-information
Eric Xing, Michael Jordan, Stuart J Russell, and Andrew Ng · 2002
Earlier work this paper cites.
Laplacian eigenmaps for dimensionality reduction and data representation
Mikhail Belkin and Partha Niyogi · 2003
Earlier work this paper cites.
Out-of-sample extensions for lle, isomap, mds, eigenmaps, and spectral clustering
Yoshua Bengio, Jean-françcois Paiement, Pascal Vincent, Olivier Delalleau, Nicolas Roux, and Marie Ouimet · 2003
Earlier work this paper cites.
A generalized foley–sammon transform based on generalized fisher discriminant criterion and its application to face recognition
Yue-Fei Guo, Shi-Jin Li, Jing-Yu Yang, Ting-Ting Shu, and Li-De Wu · 2003
Earlier work this paper cites.
Locality preserving projections
Xiaofei He and Partha Niyogi · 2003
Earlier work this paper cites.
The geometry of kernel canonical correlation analysis
Malte Kuss and Thore Graepel · 2003
Earlier work this paper cites.
Statistical pattern recognition
Andrew R Webb · 2003
Cited alongside, same era.
Canonical correlation analysis: An overview with application to learning methods
David R Hardoon, Sandor Szedmak, and John Shawe-Taylor · 2004
Cited alongside, same era.
Outlier detection using k-nearest neighbour graph
Ville Hautamaki, Ismo Karkkainen, and Pasi Franti · 2004
Cited alongside, same era.
Supervised kernel locality preserving projections for face recognition
Jian Cheng, Qingshan Liu, Hanqing Lu, and Yen-Wei Chen · 2005
Cited alongside, same era.
Kernel methods for measuring independence
Arthur Gretton, Ralf Herbrich, Alexander Smola, Olivier Bousquet, Bernhard Schölkopf, et al · 2005
Cited alongside, same era.
Dimensionality reduction by learning an invariant mapping
Raia Hadsell, Sumit Chopra, and Yann LeCun · 2006
Cited alongside, same era.
Time-contrastive networks: Self-supervised learning from video
Pierre Sermanet, Corey Lynch, Yevgen Chebotar, Jasmine Hsu, Eric Jang, Stefan Schaal, Sergey Levine, and Google Brain · 2018
Later among the works it cites.
A theoretical analysis of contrastive unsupervised representation learning
Sanjeev Arora, Hrishikesh Khandeparkar, Mikhail Khodak, Orestis Plevrakis, and Nikunj Saunshi · 2019
Later among the works it cites.
Breaking the softmax bottleneck via learnable monotonic pointwise non-linearities
Octavian Ganea, Sylvain Gelly, Gary Bécigneul, and Aliaksei Severyn · 2019
Later among the works it cites.
Scaling and benchmarking self-supervised visual representation learning
Priya Goyal, Dhruv Mahajan, Abhinav Gupta, and Ishan Misra · 2019
Later among the works it cites.
Self-supervised video representation learning with space-time cubic puzzles
Dahun Kim, Donghyeon Cho, and In So Kweon · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Applied MANOVA and discriminant analysis
Carl J Huberty and Stephen Olejnik · 2006
Cited alongside, same era.
Like Hui and Mikhail Belkin · 2006
Cited alongside, same era.
Probabilistic linear discriminant analysis
Sergey Ioffe · 2006
Cited alongside, same era.
Statistical consistency of kernel canonical correlation analysis
Kenji Fukumizu, Francis R Bach, and Arthur Gretton · 2007
Cited alongside, same era.
Kernel fisher discriminant functions–a concise and rigorous introduction
H Knaf · 2007
Cited alongside, same era.
A tutorial on spectral clustering
Ulrike Von Luxburg · 2007
Cited alongside, same era.
Azade Nazi, Will Hang, Anna Goldie, Sujith Ravi, and Azalia Mirhoseini · 2019
Later among the works it cites.
Spectral inference networks: Unifying deep and spectral learning
David Pfau, Stig Petersen, Ashish Agarwal, David G. T. Barrett, and Kimberly L. Stachenfeld · 2019
Later among the works it cites.
Self-supervised spatiotemporal learning via video clip order prediction
Dejing Xu, Jun Xiao, Zhou Zhao, Jian Shao, Di Xie, and Yueting Zhuang · 2019
Later among the works it cites.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Alexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, and Michael Auli · 2020
Later among the works it cites.
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton · 2020
Later among the works it cites.
On the surprising similarities between supervised and self-supervised models
Robert Geirhos, Kantharaju Narayanappa, Benjamin Mitzkus, Matthias Bethge, Felix A Wichmann, and Wieland Brendel · 2020
Later among the works it cites.
Self-supervised learning of pretext-invariant representations
Ishan Misra and Laurens van der Maaten · 2020
Later among the works it cites.
Run away from your teacher: Understanding byol by a novel self-supervised approach
Haizhou Shi, Dongliang Luo, Siliang Tang, Jian Wang, and Yueting Zhuang · 2020
Later among the works it cites.
Understanding contrastive representation learning through alignment and uniformity on the hypersphere
Tongzhou Wang and Phillip Isola · 2020
Later among the works it cites.
Vicreg: Variance-invariance-covariance regularization for self-supervised learning
Adrien Bardes, Jean Ponce, and Yann LeCun · 2021
Later among the works it cites.
High fidelity visualization of what your self-supervised representation knows about
Florian Bordes, Randall Balestriero, and Pascal Vincent · 2021
Later among the works it cites.
Emerging properties in self-supervised vision transformers
Mathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou, Julien Mairal, Piotr Bojanowski, and Armand Joulin · 2021
Later among the works it cites.
With a little help from my friends: Nearest-neighbor contrastive learning of visual representations
Debidatta Dwibedi, Yusuf Aytar, Jonathan Tompson, Pierre Sermanet, and Andrew Zisserman · 2021
Later among the works it cites.
How well do self-supervised models transfer?
Linus Ericsson, Henry Gouk, and Timothy M Hospedales · 2021
Later among the works it cites.
Provable guarantees for self-supervised deep learning with spectral contrastive loss
Jeff Z HaoChen, Colin Wei, Adrien Gaidon, and Tengyu Ma · 2021
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2021
Later among the works it cites.
On feature decorrelation in self-supervised learning
Tianyu Hua, Wenxiao Wang, Zihui Xue, Sucheng Ren, Yue Wang, and Hang Zhao · 2021
Later among the works it cites.
Understanding dimensional collapse in contrastive self-supervised learning
Li Jing, Pascal Vincent, Yann LeCun, and Yuandong Tian · 2021
Later among the works it cites.
Mean shift for self-supervised learning
Soroush Abbasi Koohpayegani, Ajinkya Tejankar, and Hamed Pirsiavash · 2021
Later among the works it cites.
On generalizing trace minimization
Xin Liang, Li Wang, Lei-Hong Zhang, and Ren-Cang Li · 2021
Later among the works it cites.
Understanding self-supervised learning dynamics without contrastive pairs
Yuandong Tian, Xinlei Chen, and Surya Ganguli · 2021
Later among the works it cites.
Contrastive learning, multi-view redundancy, and linear models
Christopher Tosh, Akshay Krishnamurthy, and Daniel Hsu · 2021
Later among the works it cites.
Towards demystifying representation learning with non-contrastive self-supervision
Xiang Wang, Xinlei Chen, Simon S Du, and Yuandong Tian · 2021
Later among the works it cites.
Toward understanding the feature learning process of self-supervised contrastive learning
Zixin Wen and Yuanzhi Li · 2021
Later among the works it cites.
Barlow twins: Self-supervised learning via redundancy reduction
Jure Zbontar, Li Jing, Ishan Misra, Yann LeCun, and Stéphane Deny · 2021
Later among the works it cites.
Jeff Z HaoChen, Colin Wei, Ananya Kumar, and Tengyu Ma · 2022
Closest in time.
Contrasting the landscape of contrastive and non-contrastive learning
Ashwini Pokle, Jinjin Tian, Yuchen Li, and Andrej Risteski · 2022
Closest in time.
Kernelized supervised laplacian eigenmap for visualization and classification of multi-label data
Mariko Tai, Mineichi Kudo, Akira Tanaka, Hideyuki Imai, and Keigo Kimura · 2022
Closest in time.
Deep contrastive learning is provably (almost) principal component analysis
Yuandong Tian · 2022
Closest in time.