Fetching the paper…
Reading the bibliography…
Recent approaches have shown that training deep neural networks directly on large-scale image-text pair collections enables zero-shot transfer on various recognition tasks.
Solving multiclass learning problems via error-correcting output codes
Thomas G. Dietterich and Ghulum Bakiri · 1995
Earlier work this paper cites.
Wordnet: a lexical database for english
George A. Miller · 1995
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Describing objects by their attributes
A. Farhadi, I. Endres, D. Hoiem, and D. Forsyth · 2009
Earlier work this paper cites.
Learning to detect unseen object classes by between-class attribute transfer
C. H. Lampert, H. Nickisch, and S. Harmeling · 2009
Earlier work this paper cites.
Zero-shot learning with semantic output codes
M. Palatucci, D. Pomerleau, and G. E. Hinton · 2009
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
Mark Everingham, Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman · 2010
Earlier work this paper cites.
Object detection with discriminatively trained part based models
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, and D. Ramanan · 2010
Earlier work this paper cites.
3d object representations for fine-grained categorization
Jonathan Krause, Michael Stark, Jia Deng, and Li Fei-Fei · 2013
Earlier work this paper cites.
Fine-grained visual classification of aircraft
Subhransu Maji, Esa Rahtu, Juho Kannala, Matthew B. Blaschko, and Andrea Vedaldi · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean · 2013
Earlier work this paper cites.
Birdsnap: Large-scale fine-grained visual categorization of birds
Thomas Berg, Jiongxin Liu, Seung Woo Lee, Michelle L. Alexander, David W. Jacobs, and Peter N. Belhumeur · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
Ross Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, Lubomir Bourdev, Ross Girshick, James Hays, Pietro Perona, Deva Ramanan, C. Lawrence Zitnick, and Piotr Dollár · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollar, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning · 2014
Earlier work this paper cites.
Fast r-rcnn
Ross Girshick · 2015
Earlier work this paper cites.
Real-time analysis and visualization of the YFCC100m dataset
Sebastian Kalkowski, Christian Schulze, Andreas Dengel, and Damian Borth · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Cited alongside, same era.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei · 2015
Cited alongside, same era.
Weakly supervised deep detection networks
Hakan Bilen and Andrea Vedaldi · 2016
Cited alongside, same era.
Weakly supervised object localization with multi-fold multiple instance learning
Ramazan Gokberk Cinbis, Jakob Verbeek, and Cordelia Schmid · 2016
Cited alongside, same era.
Ssd: Single shot multibox detector
Wei Liu, Dragomir Anguelov, Dumitru Erhan, Christian Szegedy, Scott Reed, Cheng-Yang Fu, and Alexander C. Berg · 2016
Cited alongside, same era.
You only look once: Unified, real-time object detection
Zero-shot learning—a comprehensive evaluation of the good, the bad and the ugly
Xian Y, Lampert C H, Schiele B, and Akata Z · 2019
Later among the works it cites.
Object detection in 20 years: A survey
Zhengxia Zou, Zhenwei Shi, Yuhong Guo, and Jieping Ye · 2019
Later among the works it cites.
Synthesizing the unseen for zero-shot object detection
Nasir Hayat, Munawar Hayat, Shafin Rahman, Salman Khan, Syed Waqas Zamir, and Fahad Shahbaz Khan · 2020
Later among the works it cites.
Improved visual-semantic alignment for zero-shot object detection
S. Rahman, S. Khan, and N. Barne · 2020
Later among the works it cites.
1st place solution of lvis challenge 2020: A good box is not a guarantee of a good mask
Jingru Tan, Gang Zhang, Hanming Deng, Changbao Wang, Lewei Lu, Quanquan Li, and Jifeng Dai · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Joseph Redmon, Santosh Divvala, Ross Girshick, and Ali Farhadi · 2016
Cited alongside, same era.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollar, and Ross Girshick · 2017
Cited alongside, same era.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
Ranjay Krishna, Yuke Zhu, Oliver Groth, Justin Johnson, Kenji Hata, Joshua Kravitz, Stephanie Chen, Yannis Kalanditis, Li-Jia Li, David A Shamma, Michael Bernstein, and Li Fei-Fei · 2017
Cited alongside, same era.
Yolo9000: Better, faster, stronger
Joseph Redmon and Ali Farhadi · 2017
Cited alongside, same era.
Zero-shot object detection
Ankan Bansal, Karan Sikka, Gaurav Sharma, Rama Chellappa, and Ajay Divakaran · 2018
Cited alongside, same era.
Zero-shot object detection: Learning to simultaneously recognize and localize novel concepts
Shafin Rahman, Salman H. Khan, and Fatih Porikli · 2018
Cited alongside, same era.
Yolov3: An incremental improvement
Joseph Redmon and Ali Farhadi · 2018
Cited alongside, same era.
Learning latent semantic attributes for zero-shot object detection
K. Wang, L. Zhang, Y. Tan, J. Zhao, and S. Zhou · 2020
Later among the works it cites.
Zero-shot object detection via learning an embedding from semantic space to visual space
Licheng Zhang, Xianzhi Wang, Lina Yao, Lin Wu, and Feng Zheng · 2020
Later among the works it cites.
Gtnet: Generative transfer network for zero-shot object detection
S. Zhao, C. Gao, Y. Shao, L. Li, C. Yu, Z. Ji, and et al · 2020
Later among the works it cites.
Background learnable cascade for zero-shot object detection
Y. Zheng, R. Huang, C. Han, X. Huang, and L. Cui · 2020
Later among the works it cites.
Don’t even look once: Synthesizing features for zero-shot detection
Pengkai Zhu, Hanxiao Wang, and Venkatesh Saligrama · 2020
Later among the works it cites.
Dont even look once: Synthesizing features for zero-shot detection
P. Zhu, H. Wang, and V. Saligrama · 2020
Later among the works it cites.
Open-vocabulary object detection via vision and language knowledge distillation, 2021
Xiuye Gu, Tsung-Yi Lin, Weicheng Kuo, and Yin Cui · 2021
Closest in time.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Closest in time.
Open-vocabulary object detection using captions
Alireza Zareian, Kevin Dela Rosa, Derek Hao Hu, and Shih-Fu Chang · 2021
Closest in time.
Zero-shot instances segmentation
Ye Zheng, Jiahong Wu, Yongqiang Qin, Faen Zhang, and Li Cui · 2021
Closest in time.
Learning to prompt for vision-language models
Kaiyang Zhou, Jingkang Yang, Chen Change Loy, and Ziwei Liu · 2021
Closest in time.
ultralytics/yolov5: v6.1 - TensorRT, TensorFlow Edge TPU and OpenVINO Export and Inference, Feb. 2022
Glenn Jocher, Ayush Chaurasia, Alex Stoken, Jirka Borovec, NanoCode012, Yonghye Kwon, TaoXie, Jiacong Fang, imyhxy, Kalen Michael, Lorna, Abhiram V, Diego Montes, Jebastin Nadar, Laughing, tkianai, yxNONG, Piotr Skalski, Zhiqiang Wang, Adam Hogan, Cristi Fati, Lorenzo Mammana, AlexWang1900, Deep Patel, Ding Yiwei, Felix You, Jan Hajek, Laurentiu Diaconu, and Mai Thanh Minh · 2022
Closest in time.