Fetching the paper…
Reading the bibliography…
Large-scale pre-trained models have demonstrated impressive performance in vision and language tasks within open-world scenarios.
G. Bradski and S. Grossberg, “Recognition of 3-d objects from multiple 2-d views by a self-organizing neural architecture,” in From Statistics to Neural Networks: Theory and Pattern Recognition Applications
1994
Earlier work this paper cites.
H. Su, S. Maji, E. Kalogerakis, and E. Learned-Miller, “Multi-view convolutional neural networks for 3d shape recognition,” in Proceedings of the IEEE international conference on computer vision
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
Z. Wu, S. Song, A. Khosla, F. Yu, L. Zhang, X. Tang, and J. Xiao, “3d shapenets: A deep representation for volumetric shapes,” in Proceedings of the IEEE conference on computer vision and pattern recognition
2015
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Kanezaki, Y. Matsushita, and Y. Nishida, “Rotationnet: Joint object categorization and pose estimation using multiviews from unsupervised viewpoints,” in Proceedings of the IEEE conference on computer vision and pattern recognition
2018
Earlier work this paper cites.
J.-C. Su, M. Gadelha, R. Wang, and S. Maji, “A deeper look at 3d shape classifiers,” in Second Workshop on 3D Reconstruction Meets Semantics, ECCV
2018
Earlier work this paper cites.
Y. Su, Y. Li, W. Nie, D. Song, and A.-A. Liu, “Joint heterogeneous feature learning and distribution alignment for 2d image-based 3d object retrieval,” IEEE Transactions on Circuits and Systems for Video Technology
2019
Earlier work this paper cites.
A. Cheraghian, S. Rahman, and L. Petersson, “Zero-shot learning of 3d point cloud objects,” in 2019 16th International Conference on Machine Vision Applications (MVA)
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al
2020
Earlier work this paper cites.
X. Wei, R. Yu, and J. Sun, “View-gcn: View-based graph convolutional network for 3d shape analysis,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2020
Earlier work this paper cites.
Z. Jiang, F. F. Xu, J. Araki, and G. Neubig, “How can we know what language models know?,” Transactions of the Association for Computational Linguistics
2020
Earlier work this paper cites.
J. Sun, Z. Wang, W. Wang, H. Li, F. Sun, and Z. Ding, “Joint adaptive dual graph and feature selection for domain adaptation,” IEEE Transactions on Circuits and Systems for Video Technology
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al
2021
Earlier work this paper cites.
J. Huang, W. Yan, G. Li, T. Li, and S. Liu, “Learning disentangled representation for multi-view 3d object recognition,” IEEE Transactions on Circuits and Systems for Video Technology
2021
Earlier work this paper cites.
A. Hamdi, S. Giancola, and B. Ghanem, “Mvtn: Multi-view transformation network for 3d shape recognition,” in Proceedings of the IEEE/CVF International Conference on Computer Vision
2021
Earlier work this paper cites.
A. Goyal, H. Law, B. Liu, A. Newell, and J. Deng, “Revisiting point cloud shape classification with a simple and effective baseline,” in International Conference on Machine Learning
2021
Earlier work this paper cites.
H. Zhao, L. Jiang, J. Jia, P. H. Torr, and V. Koltun, “Point transformer,” in Proceedings of the IEEE/CVF international conference on computer vision
2021
Cited alongside, same era.
G. Ilharco, M. Wortsman, R. Wightman, C. Gordon, N. Carlini, R. Taori, A. Dave, V. Shankar, H. Namkoong, J. Miller, et al
2021
Cited alongside, same era.
A. Cheraghian, S. Rahman, T. F. Chowdhury, D. Campbell, and L. Petersson, “Zero-shot learning on 3d point cloud objects and beyond,” International Journal of Computer Vision
2022
Cited alongside, same era.
Y. Su, J. Li, W. Li, Z. Gao, H. Chen, X. Li, and A.-A. Liu, “Semantically guided projection for zero-shot 3d model classification and retrieval,” Multimedia Systems
2022
Cited alongside, same era.
R. Zhang, Z. Guo, W. Zhang, K. Li, X. Miao, B. Cui, Y. Qiao, P. Gao, and H. Li, “Pointclip: Point cloud understanding by CLIP,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2022, New Orleans, LA, USA, June 18-24, 2022
2023
Closest in time.
T. Huang, B. Dong, Y. Yang, X. Huang, R. W. Lau, W. Ouyang, and W. Zuo, “Clip2point: Transfer clip to point cloud classification with image-depth pre-training,” in Proceedings of the IEEE/CVF International Conference on Computer Vision
2023
Closest in time.
H. Wang, J. Tang, J. Ji, X. Sun, R. Zhang, Y. Ma, M. Zhao, L. Li, Z. Zhao, T. Lv, et al
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
R. Zhang, Z. Guo, P. Gao, R. Fang, B. Zhao, D. Wang, Y. Qiao, and H. Li, “Point-m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training,” Advances in neural information processing systems
2022
Cited alongside, same era.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Conditional prompt learning for vision-language models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2022
Cited alongside, same era.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,” International Journal of Computer Vision
2022
Cited alongside, same era.
M. Jia, L. Tang, B.-C. Chen, C. Cardie, S. Belongie, B. Hariharan, and S.-N. Lim, “Visual prompt tuning,” in European Conference on Computer Vision
2022
Cited alongside, same era.
2022
Cited alongside, same era.
A. Hamdi, F. AlZahrani, S. Giancola, and B. Ghanem, “Mvtn: Learning multi-view transformations for 3d understanding,” 2022
2022
Cited alongside, same era.
N. Mu, A. Kirillov, D. Wagner, and S. Xie, “Slip: Self-supervision meets language-image pre-training,” in European Conference on Computer Vision
2022
Cited alongside, same era.
2023
Closest in time.
Y. Zeng, C. Jiang, J. Mao, J. Han, C. Ye, Q. Huang, D.-Y. Yeung, Z. Yang, X. Liang, and H. Xu, “Clip2: Contrastive language-image-point pretraining from real-world point cloud data,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
R. Zhang, L. Wang, Y. Qiao, P. Gao, and H. Li, “Learning 3d representations from 2d pre-trained models via image-to-point masked autoencoders,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2023
Closest in time.
P. Liu, W. Yuan, J. Fu, Z. Jiang, H. Hayashi, and G. Neubig, “Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing,” ACM Computing Surveys
2023
Closest in time.
M. U. Khattak, H. Rasheed, M. Maaz, S. Khan, and F. S. Khan, “Maple: Multi-modal prompt learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2023
Closest in time.
Z. Guo, Y. Tang, R. Zhang, D. Wang, Z. Wang, B. Zhao, and X. Li, “Viewrefer: Grasp the multi-view knowledge for 3d visual grounding,” pp. 15372–15383, 2023
2023
Closest in time.
S. Pratt, I. Covert, R. Liu, and A. Farhadi, “What does a platypus look like? generating customized prompts for zero-shot image classification,” in Proceedings of the IEEE/CVF International Conference on Computer Vision
2023
Closest in time.
R. Zhang, X. Hu, B. Li, S. Huang, H. Deng, Y. Qiao, P. Gao, and H. Li, “Prompt, generate, then cache: Cascade of foundation models makes strong few-shot learners,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2023
Closest in time.
Z. Novack, J. McAuley, Z. C. Lipton, and S. Garg, “Chils: Zero-shot image classification with hierarchical label sets,” in International Conference on Machine Learning
2023
Closest in time.
M. Deitke, D. Schwenk, J. Salvador, L. Weihs, O. Michel, E. VanderBilt, L. Schmidt, K. Ehsani, A. Kembhavi, and A. Farhadi, “Objaverse: A universe of annotated 3d objects,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2023
Closest in time.