Fetching the paper…
Reading the bibliography…
Pre-training has become a standard paradigm in many computer vision tasks.
Language models are few-shot learners
Brown, T. B.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 1901
Earlier work this paper cites.
Multi-view 3D object detection network for autonomous driving
Chen, X.; Ma, H.; Wan, J.; Li, B.; and Xia, T. 2017 · 1915
Earlier work this paper cites.
Bootstrap your own latent: A new approach to self-supervised learning
Grill, J.-B.; Strub, F.; Altché, F.; Tallec, C.; Richemond, P. H.; Buchatskaya, E.; Doersch, C.; Pires, B. A.; Guo, Z. D.; Azar, M. G.; et al. 2020 · 2006
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
Deng, J.; Dong, W.; Socher, R.; Li, L.-J.; Li, K.; and Fei-Fei, L. 2009 · 2009
Earlier work this paper cites.
Multi-view adaptive fusion network for 3D object detection
Wang, G.; Tian, B.; Zhang, Y.; Chen, L.; Cao, D.; and Wu, J. 2020 · 2011
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
Geiger, A.; Lenz, P.; and Urtasun, R. 2012 · 2012
Earlier work this paper cites.
Multi-modality cut and paste for 3D object detection
Zhang, W.; Wang, Z.; and Change Loy, C. 2020 · 2012
Earlier work this paper cites.
Learning to see by moving
Agrawal, P.; Carreira, J.; and Malik, J. 2015 · 2015
Earlier work this paper cites.
Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture
Eigen, D.; and Fergus, R. 2015 · 2015
Earlier work this paper cites.
Learning image representations tied to ego-motion
Jayaraman, D.; and Grauman, K. 2015 · 2015
Earlier work this paper cites.
U-Net: Convolutional networks for biomedical image segmentation
Ronneberger, O.; Fischer, P.; and Brox, T. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Unsupervised learning of visual representations by solving jigsaw puzzles
Noroozi, M.; and Favaro, P. 2016 · 2016
Earlier work this paper cites.
Context encoders: Feature learning by inpainting
Pathak, D.; Krahenbuhl, P.; Donahue, J.; Darrell, T.; and Efros, A. A. 2016 · 2016
Cited alongside, same era.
Colorful image colorization
Zhang, R.; Isola, P.; and Efros, A. A. 2016 · 2016
Cited alongside, same era.
PointNet++: Deep hierarchical feature learning on point sets in a metric space
Qi, C. R.; Yi, L.; Su, H.; and Guibas, L. J. 2017 · 2017
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2018 · 2018
Cited alongside, same era.
Unsupervised representation learning by predicting image rotations
Gidaris, S.; Singh, P.; and Komodakis, N. 2018 · 2018
Cited alongside, same era.
Self-supervised relative depth learning for urban scene understanding
Jiang, H.; Larsson, G.; Shakhnarovich, M. M. G.; and Learned-Miller, E. 2018 · 2018
Momentum contrast for unsupervised visual representation learning
He, K.; Fan, H.; Wu, Y.; Xie, S.; and Girshick, R. 2020 · 2020
Later among the works it cites.
HR-Depth: high resolution self-supervised monocular depth estimation
Lyu, X.; Liu, L.; Wang, M.; Kong, X.; Liu, L.; Liu, Y.; Chen, X.; and Yuan, Y. 2020 · 2020
Later among the works it cites.
Scalability in perception for autonomous driving: Waymo open dataset
Sun, P.; Kretzschmar, H.; Dotiwalla, X.; Chouard, A.; Patnaik, V.; Tsui, P.; Guo, J.; Zhou, Y.; Chai, Y.; Caine, B.; et al. 2020 · 2020
Later among the works it cites.
PointContrast: Unsupervised pre-training for 3D point cloud understanding
Xie, S.; Gu, J.; Guo, D.; Qi, C. R.; Guibas, L.; and Litany, O. 2020 · 2020
Later among the works it cites.
3DSSD: Point-based 3D single stage object detector
Yang, Z.; Sun, Y.; Liu, S.; and Jia, J. 2020 · 2020
Later among the works it cites.
Adabins: Depth estimation using adaptive bins
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep contextualized word representations
Peters, M. E.; Neumann, M.; Iyyer, M.; Gardner, M.; Clark, C.; Lee, K.; and Zettlemoyer, L. 2018 · 2018
Cited alongside, same era.
Visuomotor Understanding for Representation Learning of Driving Scenes
Lee, S.; Kim, J.; Oh, T.-H.; Jeong, Y.; Yoo, D.; Lin, S.; and Kweon, I. S. 2019 · 2019
Cited alongside, same era.
Multi-task multi-sensor fusion for 3D object detection
Liang, M.; Yang, B.; Chen, Y.; Hu, R.; and Urtasun, R. 2019 · 2019
Cited alongside, same era.
Mvx-net: Multimodal voxelnet for 3D object detection
Sindagi, V. A.; Zhou, Y.; and Tuzel, O. 2019 · 2019
Cited alongside, same era.
A simple framework for contrastive learning of visual representations
Chen, T.; Kornblith, S.; Norouzi, M.; and Hinton, G. 2020 · 2020
Cited alongside, same era.
MMDetection3D: OpenMMLab next-generation platform for general 3D object detection
Contributors, M. 2020 · 2020
Cited alongside, same era.
Bhat, S. F.; Alhashim, I.; and Wonka, P. 2021 · 2021
Closest in time.
A survey on contrastive self-supervised learning
Jaiswal, A.; Babu, A. R.; Zadeh, M. Z.; Banerjee, D.; and Makedon, F. 2021 · 2021
Closest in time.
CoCoNets: Continuous contrastive 3D scene representations
Lal, S.; Prabhudesai, M.; Mediratta, I.; Harley, A. W.; and Fragkiadaki, K. 2021 · 2021
Closest in time.
Learning from 2D: Pixel-to-point knowledge transfer for 3D pretraining
Liu, Y.-C.; Huang, Y.-K.; Chiang, H.-Y.; Su, H.-T.; Liu, Z.-Y.; Chen, C.-T.; Tseng, C.-Y.; and Hsu, W. H. 2021 · 2021
Closest in time.
Self-supervised pillar motion learning for autonomous driving
Luo, C.; Yang, X.; and Yuille, A. 2021 · 2021
Closest in time.
FCOS3D: Fully convolutional one-stage monocular 3D object detection
Wang, T.; Zhu, X.; Pang, J.; and Lin, D. 2021 · 2021
Closest in time.
Multimodal contrastive training for visual representation learning
Yuan, X.; Lin, Z.; Kuen, J.; Zhang, J.; Wang, Y.; Maire, M.; Kale, A.; and Faieta, B. 2021 · 2021
Closest in time.