Fetching the paper…
Reading the bibliography…
We observe that the mapping between an image's representation in one model to its representation in another can be learned surprisingly well with just a linear layer, even across diverse models.
Principal component analysis
Wold, S., Esbensen, K., and Geladi, P · 1987
Earlier work this paper cites.
BREEDS: benchmarks for subpopulation shift
Santurkar, S., Tsipras, D., and Madry, A · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A., Hinton, G., et al · 2009
Earlier work this paper cites.
Mnist handwritten digit database
LeCun, Y., Cortes, C., and Burges, C · 2010
Earlier work this paper cites.
An analysis of single-layer networks in unsupervised feature learning
Coates, A., Ng, A., and Lee, H · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Netzer, Y., Wang, T., Coates, A., Bissacco, A., Wu, B., and Ng, A. Y · 2011
Earlier work this paper cites.
Describing textures in the wild
Cimpoi, M., Maji, S., Kokkinos, I., Mohamed, S., , and Vedaldi, A · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Hinton, G. E., Vinyals, O., and Dean, J · 2015
Earlier work this paper cites.
Understanding image representations by measuring their equivariance and equivalence
Lenc, K. and Vedaldi, A · 2015
Earlier work this paper cites.
Deep learning face attributes in the wild
Liu, Z., Luo, P., Wang, X., and Tang, X · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Earlier work this paper cites.
Yfcc100m: The new data in multimedia research
Thomee, B., Shamma, D. A., Friedland, G., Elizalde, B., Ni, K., Poland, D., Borth, D., and Li, L.-J · 2016
Earlier work this paper cites.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
Xiao, H., Rasul, K., and Vollgraf, R · 2017
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Zhu, J.-Y., Park, T., Isola, P., and Efros, A. A · 2017
Earlier work this paper cites.
Multi-modal cycle-consistent generalized zero-shot learning
Felix, R., Kumar, B. V., Reid, I. D., and Carneiro, G · 2018
Earlier work this paper cites.
Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav)
Kim, B., Wattenberg, M., Gilmer, J., Cai, C., Wexler, J., Viegas, F., et al · 2018
Cited alongside, same era.
Objectnet: A large-scale bias-controlled dataset for pushing the limits of object recognition models
Barbu, A., Mayo, D., Alverio, J., Luo, W., Wang, C., Gutfreund, D., Tenenbaum, J. B., and Katz, B · 2019
Cited alongside, same era.
Towards automatic concept-based explanations
Ghorbani, A., Wexler, J., Zou, J. Y., and Kim, B · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., Desmaison, A., Kopf, A., Yang, E., DeVito, Z., Raison, M., Tejani, A., Chilamkurthy, S., Steiner, B., Fang, L., Bai, J., and Chintala, S · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., and Sutskever, I · 2019
Cited alongside, same era.
Similarity and matching of neural network representations
Csiszárik, A., Kőrösi-Szabó, P., Matszangosz, Á., Papp, G., and Varga, D · 2021
Later among the works it cites.
Magma - multimodal augmentation of generative models through adapter-based finetuning
Eichenberg, C., Black, S., Weinbach, S., Parcalabescu, L., and Frank, A · 2021
Later among the works it cites.
Open clip, 7 2021
Ilharco, G., Wortsman, M., Carlini, N., Taori, R., Dave, A., Shankar, V., Namkoong, H., Miller, J., Hajishirzi, H., Farhadi, A., and Schmidt, L · 2021
Later among the works it cites.
Clipcap: Clip prefix for image captioning
Mokady, R., Hertz, A., and Bermano, A. H · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pytorch image models
Wightman, R · 2019
Cited alongside, same era.
Concept bottleneck models
Koh, P. W., Nguyen, T., Tang, Y. S., Mussmann, S., Pierson, E., Kim, B., and Liang, P · 2020
Cited alongside, same era.
2d geometric shapes dataset – for machine learning and pattern recognition
Korchi, A. and Ghanou, Y · 2020
Cited alongside, same era.
Do adversarially robust imagenet models transfer better?
Salman, H., Ilyas, A., Engstrom, L., Kapoor, A., and Madry, A · 2020
Cited alongside, same era.
Noise or signal: The role of image backgrounds in object recognition
Xiao, K., Engstrom, L., Ilyas, A., and Madry, A · 2020
Cited alongside, same era.
Invertible concept-based explanations for cnn models with non-negative concept activation vectors
Zhang, R., Madumal, P., Miller, T., Ehinger, K. A., and Rubinstein, B. I. P · 2020
Cited alongside, same era.
Revisiting model stitching to compare neural representations
Bansal, Y., Nakkiran, P., and Barak, B · 2021
Cited alongside, same era.
Zerocap: Zero-shot image-to-text generation for visual-semantic arithmetic
Tewel, Y., Shalev, Y., Schwartz, I., and Wolf, L · 2021
Later among the works it cites.
Multimodal few-shot learning with frozen language models
Tsimpoukelli, M., Menick, J., Cabi, S., Eslami, S. M. A., Vinyals, O., and Hill, F · 2021
Later among the works it cites.
Lit: Zero-shot transfer with locked-image text tuning
Zhai, X., Wang, X., Mustafa, B., Steiner, A., Keysers, D., Kolesnikov, A., and Beyer, L · 2021
Later among the works it cites.
Craft: Concept recursive activation factorization for explainability
Fel, T., Picard, A., Béthune, L., Boissin, T., Vigouroux, D., Colin, J., Cadene, R., and Serre, T · 2022
Later among the works it cites.
Distilling model failures as directions in latent space
Jain, S., Lawrence, H., Moitra, A., and Madry, A · 2022
Later among the works it cites.
Linearly mapping from image to text space
Merullo, J., Castricato, L., Eickhoff, C., and Pavlick, E · 2022
Later among the works it cites.
A comprehensive study of image classification model sensitivity to foregrounds, backgrounds, and visual attributes
Moayeri, M., Pope, P. E., Balaji, Y., and Feizi, S · 2022
Later among the works it cites.
Relative representations enable zero-shot latent space communication
Moschella, L., Maiorca, V., Fumero, M., Norelli, A., Locatello, F., and Rodolà, E · 2022
Later among the works it cites.
Clip-dissect: Automatic description of neuron representations in deep vision networks
Oikarinen, T. and Weng, T.-W · 2022
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Schuhmann, C., Beaumont, R., Vencu, R., Gordon, C., Wightman, R., Cherti, M., Coombes, T., Katta, A., Mullis, C., Wortsman, M., et al · 2022
Later among the works it cites.