Fetching the paper…
Reading the bibliography…
Self-Supervised learning (SSL) with Joint-Embedding Architectures (JEA) has led to outstanding performances.
“CheXpert: A Large Chest Radiograph Dataset with Uncertainty Labels and Expert Comparison”, 2019
Jeremy Irvin et al · 1901
Earlier work this paper cites.
“A critical analysis of self-supervision, or what we can learn from a single image”, 2020
Yuki. Asano, Christian Rupprecht and Andrea Vedaldi · 1904
Earlier work this paper cites.
“Scaling Laws for Neural Language Models”, 2020
Jared Kaplan et al · 2001
Earlier work this paper cites.
Lucas Beyer et al · 2006
Earlier work this paper cites.
“RP2K: A Large-Scale Retail Product Dataset for Fine-Grained Image Classification”, 2021
Jingtian Peng, Chang Xiao and Yifan Li · 2006
Earlier work this paper cites.
Senthil Purushwalkam and Abhinav Gupta · 2007
Earlier work this paper cites.
“Products-10K: A Large-scale Product Recognition Dataset”, 2020
Yalong Bai et al · 2008
Earlier work this paper cites.
“What Should Not Be Contrastive in Contrastive Learning”, 2021
Tete Xiao, Xiaolong Wang, Alexei. Efros and Trevor Darrell · 2008
Earlier work this paper cites.
“ImageNet: A Large-Scale Hierarchical Image Database”
J. Deng et al · 2009
Earlier work this paper cites.
“Scalable logo recognition in real-world images”, 2011, pp. 25
Stefan Romberg, Lluis Pueyo, Rainer Lienhart and Roelof Zwol · 2011
Earlier work this paper cites.
“VinDr-CXR: An open dataset of chest X-rays with radiologist’s annotations”, 2020
Ha. Nguyen et al · 2012
Earlier work this paper cites.
“Indoor segmentation and support inference from rgbd images”
Nathan Silberman, Derek Hoiem, Pushmeet Kohli and Rob Fergus · 2012
Earlier work this paper cites.
“Man vs. computer: Benchmarking machine learning algorithms for traffic sign recognition”
J. Stallkamp, M. Schlipsing, J. Salmen and C. Igel · 2012
Earlier work this paper cites.
“Training data-efficient image transformers & distillation through attention”, 2021
Hugo Touvron et al · 2012
Earlier work this paper cites.
“Unsupervised visual representation learning by context prediction”
Carl Doersch, Abhinav Gupta and Alexei Efros · 2015
Earlier work this paper cites.
“Imagenet large scale visual recognition challenge”
Olga Russakovsky et al · 2015
Earlier work this paper cites.
“Discriminative unsupervised feature learning with exemplar convolutional neural networks”
Alexey Dosovitskiy et al · 2016
Earlier work this paper cites.
“Unsupervised learning by predicting Noise”
Piotr Bojanowski and Armand Joulin · 2017
Earlier work this paper cites.
“Remote Sensing Image Scene Classification: Benchmark and State of the Art”
Gong Cheng, Junwei Han and Xiaoqiang Lu · 2017
Earlier work this paper cites.
“ChestX-Ray8: Hospital-Scale Chest X-Ray Database and Benchmarks on Weakly-Supervised Classification and Localization of Common Thorax Diseases”
Xiaosong Wang et al · 2017
Cited alongside, same era.
“Scene parsing through ade20k dataset”
Bolei Zhou et al · 2017
Cited alongside, same era.
“Deep clustering for unsupervised learning of visual features”
Mathilde Caron, Piotr Bojanowski, Armand Joulin and Matthijs Douze · 2018
Cited alongside, same era.
“Unsupervised Representation Learning by Predicting Image Rotations”, 2018
Spyros Gidaris, Praveer Singh and Nikos Komodakis · 2018
Cited alongside, same era.
“The inaturalist species classification and detection dataset”
Grant Van et al · 2018
Cited alongside, same era.
“Masked autoencoders are scalable vision learners”
Kaiming He et al · 2021
Later among the works it cites.
“Are Large-scale Datasets Necessary for Self-Supervised Pre-training?”, 2021
Alaaeldin El-Nouby et al · 2021
Later among the works it cites.
“How to Train Vision Transformer on Small-scale Datasets?”, 2022
Hanan Gani, Muzammal Naseer and Mohammad Yaqub · 2022
Later among the works it cites.
“Self-Supervised Learning with Data Augmentations Provably Isolates Content from Style”, 2022
Julius von Kügelgen et al · 2022
Later among the works it cites.
“Understanding contrastive learning requires incorporating inductive biases”
Nikunj Saunshi et al · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alistair.. Johnson et al · 2019
Cited alongside, same era.
“Do ImageNet Classifiers Generalize to ImageNet?”
Benjamin Recht, Rebecca Roelofs, Ludwig Schmidt and Vaishaal Shankar · 2019
Cited alongside, same era.
“Unsupervised learning of visual features by contrasting cluster assignments”
Mathilde Caron et al · 2020
Cited alongside, same era.
“A simple framework for contrastive learning of visual representations”
Ting Chen, Simon Kornblith, Mohammad Norouzi and Geoffrey Hinton · 2020
Cited alongside, same era.
“Big self-supervised models are strong semi-supervised learners”
Ting Chen et al · 2020
Cited alongside, same era.
“Improved baselines with momentum contrastive learning”
Xinlei Chen, Haoqi Fan, Ross Girshick and Kaiming He · 2020
Cited alongside, same era.
“An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale”
Alexey Dosovitskiy et al · 2020
Cited alongside, same era.
Later among the works it cites.
“Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks”
Wenhui Wang et al · 2022
Later among the works it cites.
“Rethinking the augmentation module in contrastive learning: Learning hierarchical augmentation invariance with expanded views”
Junbo Zhang and Kaisheng Ma · 2022
Later among the works it cites.
“Opt: Open pre-trained transformer language models”
Susan Zhang et al · 2022
Later among the works it cites.
“Self-Supervised Learning from Images with a Joint-Embedding Predictive Architecture”
Mahmoud Assran et al · 2023
Later among the works it cites.
“No Free Lunch in Self Supervised Representation Learning”, 2023
Ihab Bendidi et al · 2023
Later among the works it cites.
“Self-Supervised Disentanglement by Leveraging Structure in Data Augmentations”
Cian Eastwood et al · 2023
Later among the works it cites.
Jonas Geiping et al · 2023
Later among the works it cites.
“Dinov2: Learning robust visual features without supervision”
Maxime Oquab et al · 2023
Later among the works it cites.
Shuchang Shen, Sachith Seneviratne, Xinye Wanyan and Michael Kirley · 2023
Later among the works it cites.
“LLaMA: Open and Efficient Foundation Language Models”
Hugo Touvron et al · 2023
Later among the works it cites.
“Better & Faster Large Language Models via Multi-token Prediction”, 2024
Fabian Gloeckle et al · 2024
Closest in time.
“Scalable Pre-training of Large Autoregressive Image Models”
Alaaeldin El-Nouby et al · 2024
Closest in time.
“Scalable Pre-training of Large Autoregressive Image Models”, 2024
Alaaeldin El-Nouby et al · 2024
Closest in time.