Fetching the paper…
Reading the bibliography…
Human learning benefits from multi-modal inputs that often appear as rich semantics (e.g., description of an object's attributes while learning about it).
Learning to recognize objects
Linda B Smith · 2003
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Caltech-UCSD Birds 200
P. Welinder, S. Branson, T. Mita, C. Wah, F. Schroff, S. Belongie, and P. Perona · 2010
Earlier work this paper cites.
Annotator rationales for visual recognition
Jeff Donahue and Kristen Grauman · 2011
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, G. S. Corrado, and J. Dean · 2013
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D Manning · 2014
Earlier work this paper cites.
Learning deep representations of fine-grained visual descriptions
Scott Reed, Zeynep Akata, Bernt Schiele, and Honglak Lee · 2016
Earlier work this paper cites.
Matching networks for one shot learning
Oriol Vinyals, Charles Blundell, Timothy Lillicrap, koray kavukcuoglu, and Daan Wierstra · 2016
Earlier work this paper cites.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov · 2017
Earlier work this paper cites.
Contrastive constraints guide explanation-based category learning
Seth Chin-Parker and Julie Cantelon · 2017
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Earlier work this paper cites.
Shapeworld - a new test methodology for multimodal language understanding
Alexander Kuhnle and Ann A. Copestake · 2017
Earlier work this paper cites.
Prototypical networks for few-shot learning
Jake Snell, Kevin Swersky, and Richard Zemel · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Learning with latent language
Jacob Andreas, Dan Klein, and Sergey Levine · 2018
Cited alongside, same era.
A guide to convolutional neural networks for computer vision
Salman Khan, Hossein Rahmani, Syed Afaq Ali Shah, and Mohammed Bennamoun · 2018
Cited alongside, same era.
Tadam: Task dependent adaptive metric for improved few-shot learning
Boris Oreshkin, Pau Rodríguez López, and Alexandre Lacoste · 2018
Cited alongside, same era.
Learning to compare: Relation network for few-shot learning
Flood Sung, Yongxin Yang, L. Zhang, T. Xiang, P. Torr, and Timothy M. Hospedales · 2018
Cited alongside, same era.
Boosting few-shot visual learning with self-supervision
Spyros Gidaris, Andrei Bursuc, Nikos Komodakis, P. Pérez, and M. Cord · 2019
Gaussian error linear units (gelus), 2020
Dan Hendrycks and Kevin Gimpel · 2020
Later among the works it cites.
Hyperbolic image embeddings
Valentin Khrulkov, L. Mirvakhabova, E. Ustinova, I. Oseledets, and V. Lempitsky · 2020
Later among the works it cites.
Alice: Active learning with contrastive natural language explanations
Weixin Liang, James Zou, and Zhou Yu · 2020
Later among the works it cites.
Shaping visual representations with language for few-shot classification
Jesse Mu, Percy Liang, and Noah Goodman · 2020
Later among the works it cites.
Self-supervised knowledge distillation for few-shot learning
Jathushan Rajasegaran, Salman Khan, Munawar Hayat, Fahad Shahbaz Khan, and Mubarak Shah · 2020
Later among the works it cites.
When does self-supervision improve few-shot learning?
Jong-Chyi Su, Subhransu Maji, and Bharath Hariharan · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Meta-learning with differentiable convex optimization
Kwonjoon Lee, Subhransu Maji, A. Ravichandran, and Stefano Soatto · 2019
Cited alongside, same era.
Revisiting local descriptor based image-to-class measure for few-shot learning
Wenbin Li, Lei Wang, J. Xu, Jing Huo, Y. Gao, and Jiebo Luo · 2019
Cited alongside, same era.
Explanation in artificial intelligence: Insights from the social sciences
Tim Miller · 2019
Cited alongside, same era.
Baby steps towards few-shot learning with multiple semantics
Eli Schwartz, Leonid Karlinsky, Rogerio Feris, Raja Giryes, and Alex M Bronstein · 2019
Cited alongside, same era.
Simpleshot: Revisiting nearest-neighbor classification for few-shot learning
Yan Wang, Wei-Lun Chao, Kilian Q. Weinberger, and Laurens van der Maaten · 2019
Cited alongside, same era.
Adaptive cross-modal few-shot learning
Chen Xing, Negar Rostamzadeh, Boris Oreshkin, and Pedro O O. Pinheiro · 2019
Cited alongside, same era.
Later among the works it cites.
Rethinking few-shot image classification: A good embedding is all you need?
Yonglong Tian, Yue Wang, Dilip Krishnan, Joshua B. Tenenbaum, and Phillip Isola · 2020
Later among the works it cites.
Meta-baseline: Exploring simple meta-learning for few-shot learning
Yinbo Chen, Zhuang Liu, Huijuan Xu, Trevor Darrell, and Xiaolong Wang · 2021
Closest in time.
Virtex: Learning visual representations from textual annotations
Karan Desai and Justin Johnson · 2021
Closest in time.
Transformers in vision: A survey
Salman Khan, Muzammal Naseer, Munawar Hayat, Syed Waqas Zamir, Fahad Shahbaz Khan, and Mubarak Shah · 2021
Closest in time.
Exploring complementary strengths of invariant and equivariant representations for few-shot learning
Mamshad Nayeem Rizve, Salman Khan, Fahad Shahbaz Khan, and Mubarak Shah · 2021
Closest in time.
Open-vocabulary object detection using captions
Alireza Zareian, Kevin Dela Rosa, Derek Hao Hu, and Shih-Fu Chang · 2021
Closest in time.