Fetching the paper…
Reading the bibliography…
Learned representations are a central component in modern ML systems, serving a multitude of downstream tasks.
Analysis of a complex of statistical variables into principal components
H. Hotelling · 1933
Earlier work this paper cites.
Modeling by shortest data description
J. Rissanen · 1978
Earlier work this paper cites.
An algorithm for vector quantizer design
Y. Linde, A. Buzo, and R. Gray · 1980
Earlier work this paper cites.
Extensions of lipschitz mappings into a hilbert space
W. B. Johnson · 1984
Earlier work this paper cites.
K-d trees for semidynamic point sets
J. L. Bentley · 1990
Earlier work this paper cites.
Solving multiclass learning problems via error-correcting output codes
T. G. Dietterich and G. Bakiri · 1994
Earlier work this paper cites.
On the use of neighbourhood-based non-parametric classifiers
J. S. Sánchez, F. Pla, and F. J. Ferri · 1997
Earlier work this paper cites.
The anatomy of a large-scale hypertextual web search engine
S. Brin and L. Page · 1998
Earlier work this paper cites.
Approximate nearest neighbors: towards removing the curse of dimensionality
P. Indyk and R. Motwani · 1998
Earlier work this paper cites.
Coarse-grained information dominates fine-grained information in judgments of time-to-contact from retinal flow
M. G. Harris and C. D. Giachritsis · 2000
Earlier work this paper cites.
Rapid object detection using a boosted cascade of simple features
P. Viola and M. Jones · 2001
Earlier work this paper cites.
Unsupervised feature selection using feature similarity
P. Mitra, C. Murthy, and S. K. Pal · 2002
Earlier work this paper cites.
Locality-sensitive hashing scheme based on p-stable distributions
M. Datar, N. Immorlica, P. Indyk, and V. S. Mirrokni · 2004
Earlier work this paper cites.
Cover trees for nearest neighbor
A. Beygelzimer, S. Kakade, and J. Langford · 2006
Earlier work this paper cites.
Learning a nonlinear embedding by preserving class neighbourhood structure
R. Salakhutdinov and G. Hinton · 2007
Earlier work this paper cites.
Time course of visual perception: coarse-to-fine processing and beyond
J. Hegdé · 2008
Earlier work this paper cites.
Challenges in building large-scale information retrieval systems
J. Dean · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Fast similarity search for learned metrics
B. Kulis, P. Jain, and K. Grauman · 2009
Earlier work this paper cites.
Semantic hashing
R. Salakhutdinov and G. Hinton · 2009
Earlier work this paper cites.
Dimensionality reduction: a comparative
L. Van Der Maaten, E. Postma, J. Van den Herik, et al · 2009
Earlier work this paper cites.
Label embedding trees for large multi-class tasks
S. Bengio, J. Weston, and D. Grangier · 2010
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
M. Gutmann and A. Hyvärinen · 2010
Earlier work this paper cites.
Product quantization for nearest neighbor search
H. Jegou, M. Douze, and C. Schmid · 2010
Earlier work this paper cites.
Hierarchical semantic indexing for large scale image retrieval
J. Deng, A. C. Berg, and L. Fei-Fei · 2011
Earlier work this paper cites.
Stacked convolutional auto-encoders for hierarchical feature extraction
J. Masci, U. Meier, D. Cireşan, and J. Schmidhuber · 2011
Earlier work this paper cites.
Deep learning of representations for unsupervised and transfer learning
Y. Bengio · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
On the importance of initialization and momentum in deep learning
I. Sutskever, J. Martens, G. Dahl, and G. Hinton · 2013
Earlier work this paper cites.
Learning everything about anything: Webly-supervised visual concept learning
S. K. Divvala, A. Farhadi, and C. Guestrin · 2014
Earlier work this paper cites.
Learning ordered representations with nested dropout
O. Rippel, M. Gelbart, and R. Adams · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
How transferable are features in deep neural networks?
J. Yosinski, J. Clune, Y. Bengio, and H. Lipson · 2014
Earlier work this paper cites.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. S. Corrado, A. Davis, J. Dean, M. Devin, S. Ghemawat, I. Goodfellow, A. Harp, G. Irving, M. Isard, Y. Jia, R. Jozefowicz, L. Kaiser, M. Kudlur, J. Levenberg, D. Mané, R. Monga, S. Moore, D. Murray, C. Olah, M. Schuster, J. Shlens, B. Steiner, I. Sutskever, K. Talwar, P. Tucker, V. Vanhoucke, V. Vasudevan, F. Viégas, O. Vinyals, P. Warden, M. Wattenberg, M. Wicke, Y. Yu, and X. Zheng · 2015
Cited alongside, same era.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Cited alongside, same era.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al · 2015
Cited alongside, same era.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books
Y. Zhu, R. Kiros, R. Zemel, R. Salakhutdinov, R. Urtasun, A. Torralba, and S. Fidler · 2015
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Understanding searches better than ever before
P. Nayak · 2019
Later among the works it cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, et al · 2019
Later among the works it cites.
Do imagenet classifiers generalize to imagenet?
B. Recht, R. Roelofs, L. Schmidt, and V. Shankar · 2019
Later among the works it cites.
Transfer learning in natural language processing
S. Ruder, M. E. Peters, S. Swayamdipta, and T. Wolf · 2019
Later among the works it cites.
Efficientnet: Rethinking model scaling for convolutional neural networks
M. Tan and Q. Le · 2019
Later among the works it cites.
Extreme classification
M. Varma · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A baseline for detecting misclassified and out-of-distribution examples in neural networks
D. Hendrycks and K. Gimpel · 2016
Cited alongside, same era.
Stochastic multiple choice learning for training diverse deep ensembles
S. Lee, S. Purushwalkam Shiva Prakash, M. Cogswell, V. Ranjan, D. Crandall, and D. Batra · 2016
Cited alongside, same era.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
A. G. Howard, M. Zhu, B. Chen, D. Kalenichenko, W. Wang, T. Weyand, M. Andreetto, and H. Adam · 2017
Cited alongside, same era.
In-datacenter performance analysis of a tensor processing unit
N. P. Jouppi, C. Young, N. Patil, D. Patterson, G. Agrawal, R. Bajwa, S. Bates, S. Bhatia, N. Boden, A. Borchers, et al · 2017
Cited alongside, same era.
Decoupled weight decay regularization
I. Loshchilov and F. Hutter · 2017
Cited alongside, same era.
Grad-cam: Visual explanations from deep networks via gradient-based localization
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra · 2017
Cited alongside, same era.
Cyclical learning rates for training neural networks
L. N. Smith · 2017
Cited alongside, same era.
As search needs evolve, microsoft makes ai tools for better search available to researchers and developers
C. Waldburger · 2019
Later among the works it cites.
Learning robust global representations by penalizing local predictive power
H. Wang, S. Ge, Z. Lipton, and E. P. Xing · 2019
Later among the works it cites.
Language models are few-shot learners
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Later among the works it cites.
Pre-training tasks for embedding-based large-scale retrieval
W.-C. Chang, F. X. Yu, Y.-W. Chang, Y. Yang, and S. Kumar · 2020
Later among the works it cites.
A simple framework for contrastive learning of visual representations
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton · 2020
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, et al · 2020
Later among the works it cites.
Momentum contrast for unsupervised visual representation learning
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick · 2020
Later among the works it cites.
Characterising bias in compressed models
S. Hooker, N. Moorosi, G. Clark, S. Bengio, and E. Denton · 2020
Later among the works it cites.
Soft threshold weight reparameterization for learnable sparsity
A. Kusupati, V. Ramanujan, R. Somani, M. Wortsman, P. Jain, S. Kakade, and A. Farhadi · 2020
Later among the works it cites.
Extreme regression for dynamic search advertising
Y. Prabhu, A. Kusupati, N. Gupta, and M. Varma · 2020
Later among the works it cites.
Are we overfitting to experimental setups in recognition?
M. Wallingford, A. Kusupati, K. Alizadeh-Vahid, A. Walsman, A. Kembhavi, and A. Farhadi · 2020
Later among the works it cites.
Multiple networks are more efficient than one: Fast and accurate models via ensembles and cascades
X. Wang, D. Kondratyuk, K. M. Kitani, Y. Movshovitz-Attias, and E. Eban · 2020
Later among the works it cites.
Extreme multi-label learning for semantic matching in product search
W.-C. Chang, D. Jiang, H.-F. Yu, C. H. Teo, J. Zhang, K. Zhong, K. Kolluri, Q. Hu, N. Shandilya, V. Ievgrafov, et al · 2021
Later among the works it cites.
Meta-baseline: exploring simple meta-learning for few-shot learning
Y. Chen, Z. Liu, H. Xu, T. Darrell, and X. Wang · 2021
Later among the works it cites.
Virtex: Learning visual representations from textual annotations
K. Desai and J. Johnson · 2021
Later among the works it cites.
A survey of quantization methods for efficient neural network inference
A. Gholami, S. Kim, Z. Dong, Z. Yao, M. W. Mahoney, and K. Keutzer · 2021
Later among the works it cites.
Masked autoencoders are scalable vision learners
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. Girshick · 2021
Later among the works it cites.
Scaling up visual and vision-language representation learning with noisy text supervision
C. Jia, Y. Yang, Y. Xia, Y.-T. Chen, Z. Parekh, H. Pham, Q. Le, Y.-H. Sung, Z. Li, and T. Duerig · 2021
Later among the works it cites.
Vertex ai matching engine
T. C. Kaz Sato · 2021
Later among the works it cites.
Llc: Accurate, multi-purpose learnt low-dimensional binary codes
A. Kusupati, M. Wallingford, V. Ramanujan, R. Somani, J. S. Park, K. Pillutla, P. Jain, S. Kakade, and A. Farhadi · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Later among the works it cites.
Robust fine-tuning of zero-shot models
M. Wortsman, G. Ilharco, M. Li, J. W. Kim, H. Hajishirzi, A. Farhadi, H. Namkoong, and L. Schmidt · 2021
Later among the works it cites.
Hers: Homomorphically encrypted representation search
J. J. Engelsma, A. K. Jain, and V. N. Boddeti · 2022
Closest in time.
Task adaptive parameter sharing for multi-task learning
M. Wallingford, H. Li, A. Achille, A. Ravichandran, C. Fowlkes, R. Bhotika, and S. Soatto · 2022
Closest in time.
Pecos: Prediction for enormous and correlated output spaces
H.-F. Yu, K. Zhong, J. Zhang, W.-C. Chang, and I. S. Dhillon · 2022
Closest in time.
Merlot reserve: Neural script knowledge through vision and language and sound
R. Zellers, J. Lu, X. Lu, Y. Yu, Y. Zhao, M. Salehi, A. Kusupati, J. Hessel, A. Farhadi, and Y. Choi · 2022
Closest in time.
Diffused redundancy in pre-trained representations
V. Nanda, T. Speicher, J. P. Dickerson, S. Feizi, K. P. Gummadi, and A. Weller · 2023
Closest in time.