Fetching the paper…
Reading the bibliography…
While many unsupervised learning models focus on one family of tasks, either generative or discriminative, we explore the possibility of a unified representation learner: a model which uses a single pre-training stage to address both families of tasks simultaneously.
M.-E. Nilsback and A. Zisserman, “Automated flower classification over a large number of classes,” in
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in
2009
Earlier work this paper cites.
C. Wah, S. Branson, P. Welinder, P. Perona, and S. Belongie, “The Caltech-UCSD Birds-200-2011 Dataset,” Tech. Rep. CNS-TR-2011-001, California Institute of Technology, 2011
2011
Earlier work this paper cites.
A. Khosla, N. Jayadevaprakash, B. Yao, and L. Fei-Fei, “Novel dataset for fine-grained image categorization,” in
2011
Earlier work this paper cites.
S. Maji, J. Kannala, E. Rahtu, M. Blaschko, and A. Vedaldi, “Fine-grained visual classification of aircraft,” tech. rep., 2013
2013
Earlier work this paper cites.
J. Krause, M. Stark, J. Deng, and L. Fei-Fei, “3d object representations for fine-grained categorization,” in
2013
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” 2015
2015
Earlier work this paper cites.
G. Van Horn, S. Branson, R. Farrell, S. Haber, J. Barry, P. Ipeirotis, P. Perona, and S. Belongie, “Building a bird recognition app and large scale dataset with citizen scientists: The fine print in fine-grained dataset collection,” in
2015
Earlier work this paper cites.
J. Donahue, P. Krähenbühl, and T. Darrell, “Adversarial feature learning,”
2016
Earlier work this paper cites.
J.-Y. Zhu, P. Krähenbühl, E. Shechtman, and A. A. Efros, “Generative visual manipulation on the natural image manifold,” in
2016
Earlier work this paper cites.
R. Zhang, P. Isola, and A. A. Efros, “Colorful image colorization,” in
2016
Earlier work this paper cites.
M. Noroozi and P. Favaro, “Unsupervised learning of visual representations by solving jigsaw puzzles,” in
2016
Earlier work this paper cites.
D. Pathak, P. Krahenbuhl, J. Donahue, T. Darrell, and A. A. Efros, “Context encoders: Feature learning by inpainting,” in
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
X. Chen, Y. Duan, R. Houthooft, J. Schulman, I. Sutskever, and P. Abbeel, “Infogan: Interpretable representation learning by information maximizing generative adversarial nets,”
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
M. Caron, P. Bojanowski, A. Joulin, and M. Douze, “Deep clustering for unsupervised learning of visual features,” in
2018
Earlier work this paper cites.
J. Donahue and K. Simonyan, “Large scale adversarial representation learning,”
2019
Earlier work this paper cites.
S. Kornblith, M. Norouzi, H. Lee, and G. Hinton, “Similarity of neural network representations revisited,” in
2019
Earlier work this paper cites.
T. Karras, S. Laine, and T. Aila, “A style-based generator architecture for generative adversarial networks,” in
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
A. Razavi, A. Van den Oord, and O. Vinyals, “Generating diverse high-fidelity images with vq-vae-2,”
2019
Earlier work this paper cites.
W. Van Gansbeke, S. Vandenhende, S. Georgoulis, M. Proesmans, and L. Van Gool, “Scan: Learning to classify images without labels,” in
2020
Earlier work this paper cites.
K. Gupta, S. Singh, and A. Shrivastava, “Patchvae: Learning local latent codes for recognition,” in
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Cited alongside, same era.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial networks,”
2020
Cited alongside, same era.
T. Karras, S. Laine, M. Aittala, J. Hellsten, J. Lehtinen, and T. Aila, “Analyzing and improving the image quality of stylegan,” in
2020
Cited alongside, same era.
T. Karras, M. Aittala, J. Hellsten, S. Laine, J. Lehtinen, and T. Aila, “Training generative adversarial networks with limited data,”
2020
Cited alongside, same era.
K. Gupta, G. Somepalli, A. Anubhav, V. Y. J. Magalle Hewa, M. Zwicker, and A. Shrivastava, “Patchgame: Learning to signal mid-level patches in referential games,” in
2021
Later among the works it cites.
A. Q. Nichol and P. Dhariwal, “Improved denoising diffusion probabilistic models,” in
2021
Later among the works it cites.
P. Goyal, Q. Duval, J. Reizenstein, M. Leavitt, M. Xu, B. Lefaudeux, M. Singh, V. Reis, M. Caron, P. Bojanowski, A. Joulin, and I. Misra, “Vissl.”
2021
Later among the works it cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” 2022
2022
Later among the works it cites.
H. Chang, H. Zhang, L. Jiang, C. Liu, and W. T. Freeman, “Maskgit: Masked generative image transformer,” in
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
J. Zhu, Y. Shen, D. Zhao, and B. Zhou, “In-domain gan inversion for real image editing,” in
2020
Cited alongside, same era.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,”
2020
Cited alongside, same era.
I. Misra and L. v. d. Maaten, “Self-supervised learning of pretext-invariant representations,” in
2020
Cited alongside, same era.
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton, “A simple framework for contrastive learning of visual representations,” in
2020
Cited alongside, same era.
T. Chen, S. Kornblith, K. Swersky, M. Norouzi, and G. E. Hinton, “Big self-supervised models are strong semi-supervised learners,”
2020
Cited alongside, same era.
X. Chen, H. Fan, R. B. Girshick, and K. He, “Improved baselines with momentum contrastive learning,”
2020
Cited alongside, same era.
X. Chen, H. Fan, R. Girshick, and K. He, “Improved baselines with momentum contrastive learning,”
2020
Cited alongside, same era.
T. Li, H. Chang, S. K. Mishra, H. Zhang, D. Katabi, and D. Krishnan, “Mage: Masked generative encoder to unify representation learning and image synthesis,” 2022
2022
Later among the works it cites.
M. Gwilliam and A. Shrivastava, “Beyond supervised vs. unsupervised: Representative benchmarking and analysis of image representation learning,” in
2022
Later among the works it cites.
2022
Later among the works it cites.
D. Roich, R. Mokady, A. H. Bermano, and D. Cohen-Or, “Pivotal tuning for latent-based editing of real images,”
2022
Later among the works it cites.
Y. Alaluf, O. Tov, R. Mokady, R. Gal, and A. Bermano, “Hyperstyle: Stylegan inversion with hypernetworks for real image editing,” in
2022
Later among the works it cites.
2022
Later among the works it cites.
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans,
2022
Later among the works it cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in
2022
Later among the works it cites.
S. Chen, P. Sun, Y. Song, and P. Luo, “Diffusiondet: Diffusion model for object detection,”
2022
Later among the works it cites.
2022
Later among the works it cites.
N. Tomasev, I. Bica, B. McWilliams, L. Buesing, R. Pascanu, C. Blundell, and J. Mitrovic, “Pushing the limits of self-supervised resnets: Can we outperform supervised learning without labels on imagenet?,” 2022
2022
Later among the works it cites.
P. Goyal, Q. Duval, I. Seessel, M. Caron, I. Misra, L. Sagun, A. Joulin, and P. Bojanowski, “Vision models are more robust and fair when pretrained on uncurated images without supervision,” 2022
2022
Later among the works it cites.
B. Pang, Y. Zhang, Y. Li, J. Cai, and C. Lu, “Unsupervised visual representation learning by synchronous momentum grouping,” 2022
2022
Later among the works it cites.
P. Zhou, Y. Zhou, C. Si, W. Yu, T. K. Ng, and S. Yan, “Mugs: A multi-granular self-supervised learning framework,” 2022
2022
Later among the works it cites.
C. Li, J. Yang, P. Zhang, M. Gao, B. Xiao, X. Dai, L. Yuan, and J. Gao, “Efficient self-supervised vision transformers for representation learning,” 2022
2022
Later among the works it cites.
J. Zhou, C. Wei, H. Wang, W. Shen, C. Xie, A. Yuille, and T. Kong, “ibot: Image bert pre-training with online tokenizer,” 2022
2022
Later among the works it cites.
M. Assran, M. Caron, I. Misra, P. Bojanowski, F. Bordes, P. Vincent, A. Joulin, M. Rabbat, and N. Ballas, “Masked siamese networks for label-efficient learning,” 2022
2022
Later among the works it cites.
H. Bao, L. Dong, S. Piao, and F. Wei, “Beit: Bert pre-training of image transformers,” 2022
2022
Later among the works it cites.
Z. Huang, X. Jin, C. Lu, Q. Hou, M.-M. Cheng, D. Fu, X. Shen, and J. Feng, “Contrastive masked autoencoders are stronger vision learners,” 2022
2022
Later among the works it cites.
M. Oquab, T. Darcet, T. Moutakanni, H. Vo, M. Szafraniec, V. Khalidov, P. Fernandez, D. Haziza, F. Massa, A. El-Nouby, M. Assran, N. Ballas, W. Galuba, R. Howes, P.-Y. Huang, S.-W. Li, I. Misra, M. Rabbat, V. Sharma, G. Synnaeve, H. Xu, H. Jegou, J. Mairal, P. Labatut, A. Joulin, and P. Bojanowski, “Dinov2: Learning robust visual features without supervision,” 2023
2023
Closest in time.
M. Walmer, S. Suri, K. Gupta, and A. Shrivastava, “Teaching matters: Investigating the role of supervision in vision transformers,” 2023
2023
Closest in time.