Fetching the paper…
Reading the bibliography…
Contemporary autonomous vehicle (AV) benchmarks have advanced techniques for training 3D detectors.
Reed, W.J.: The pareto, zipf and other power laws. Economics letters (2001)
2001
Earlier work this paper cites.
Chawla, N.V., Bowyer, K.W., Hall, L.O., Kegelmeyer, W.P.: Smote: synthetic minority over-sampling technique. Journal of artificial intelligence research (2002)
2002
Earlier work this paper cites.
Drummond, C., Holte, R.C., et al
2003
Earlier work this paper cites.
Han, H., Wang, W.-Y., Mao, B.-H.: Borderline-smote: a new over-sampling method in imbalanced data sets learning. In: International Conference on Intelligent Computing (2005). Springer
2005
Earlier work this paper cites.
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., Fei-Fei, L.: Imagenet: A large-scale hierarchical image database. In: CVPR (2009)
2009
Earlier work this paper cites.
Felzenszwalb, P.F., Girshick, R.B., McAllester, D., Ramanan, D.: Object detection with discriminatively trained part-based models. IEEE transactions on pattern analysis and machine intelligence (2009)
2009
Earlier work this paper cites.
Geiger, A., Lenz, P., Urtasun, R.: Are we ready for autonomous driving? the kitti vision benchmark suite. In: 2012 IEEE Conference on Computer Vision and Pattern Recognition (2012). IEEE
2012
Earlier work this paper cites.
Girshick, R., Donahue, J., Darrell, T., Malik, J.: Rich feature hierarchies for accurate object detection and semantic segmentation. In: CVPR (2014)
2014
Earlier work this paper cites.
Lin, T.-Y., Maire, M., Belongie, S.J., Hays, J., Perona, P., Ramanan, D., Dollár, P., Zitnick, C.L.: Microsoft coco: Common objects in context. In: European Conference on Computer Vision (2014)
2014
Earlier work this paper cites.
Everingham, M., Eslami, S.A., Van Gool, L., Williams, C.K., Winn, J., Zisserman, A.: The pascal visual object classes challenge: A retrospective. IJCV (2015)
2015
Earlier work this paper cites.
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M., et al.: Imagenet large scale visual recognition challenge. International journal of computer vision (2015)
2015
Earlier work this paper cites.
Ren, S., He, K., Girshick, R., Sun, J.: Faster r-cnn: Towards real-time object detection with region proposal networks. In: Advances in Neural Information Processing Systems (2015)
2015
Earlier work this paper cites.
Ren, S., He, K., Girshick, R., Sun, J.: Faster r-cnn: Towards real-time object detection with region proposal networks. In: Advances in Neural Information Processing Systems (2015)
2015
Earlier work this paper cites.
Liu, W., Anguelov, D., Erhan, D., Szegedy, C., Reed, S., Fu, C.-Y., Berg, A.C.: Ssd: Single shot multibox detector. In: ECCV (2016)
2016
Earlier work this paper cites.
Redmon, J., Divvala, S., Girshick, R., Farhadi, A.: You only look once: Unified, real-time object detection. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2016)
2016
Earlier work this paper cites.
Bodla, N., Singh, B., Chellappa, R., Davis, L.S.: Soft-nms–improving object detection with one line of code. In: ICCV (2017)
2017
Earlier work this paper cites.
Guo, C., Pleiss, G., Sun, Y., Weinberger, K.Q.: On calibration of modern neural networks. In: ICML (2017)
2017
Earlier work this paper cites.
Khan, S.H., Hayat, M., Bennamoun, M., Sohel, F.A., Togneri, R.: Cost-sensitive learning of deep feature representations from imbalanced data. IEEE transactions on neural networks and learning systems (2017)
2017
Earlier work this paper cites.
Lin, T.-Y., Goyal, P., Girshick, R., He, K., Dollár, P.: Focal loss for dense object detection. In: ICCV (2017)
2017
Earlier work this paper cites.
Qi, C.R., Su, H., Mo, K., Guibas, L.J.: Pointnet: Deep learning on point sets for 3d classification and segmentation. In: CVPR (2017)
2017
Earlier work this paper cites.
Redmon, J., Farhadi, A.: Yolo9000: better, faster, stronger. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2017)
2017
Earlier work this paper cites.
Qi, C.R., Liu, W., Wu, C., Su, H., Guibas, L.J.: Frustum pointnets for 3d object detection from rgb-d data. In: IEEE Conference on Computer Vision and Pattern Recognition (2018)
2018
Earlier work this paper cites.
Van Horn, G., Mac Aodha, O., Song, Y., Cui, Y., Sun, C., Shepard, A., Adam, H., Perona, P., Belongie, S.: The inaturalist species classification and detection dataset. In: CVPR (2018)
2018
Earlier work this paper cites.
Xu, D., Anguelov, D., Jain, A.: Pointfusion: Deep sensor fusion for 3d bounding box estimation. In: IEEE Conference on Computer Vision and Pattern Recognition (2018)
2018
Earlier work this paper cites.
Cui, Y., Jia, M., Lin, T.-Y., Song, Y., Belongie, S.: Class-balanced loss based on effective number of samples. In: CVPR (2019)
2019
Earlier work this paper cites.
Chang, M.-F., Lambert, J., Sangkloy, P., Singh, J., Bak, S., Hartnett, A., Wang, D., Carr, P., Lucey, S., Ramanan, D., et al
2019
Earlier work this paper cites.
Cao, K., Wei, C., Gaidon, A., Arechiga, N., Ma, T.: Learning imbalanced datasets with label-distribution-aware margin loss. In: NeurIPS (2019)
2019
Earlier work this paper cites.
Gupta, A., Dollar, P., Girshick, R.: Lvis: A dataset for large vocabulary instance segmentation. In: CVPR (2019)
2019
Earlier work this paper cites.
Huang, C., Li, Y., Loy, C.C., Tang, X.: Deep imbalanced learning for face recognition and attribute prediction. PAMI (2019)
2019
Earlier work this paper cites.
Liu, Z., Miao, Z., Zhan, X., Wang, J., Gong, B., Yu, S.X.: Large-scale long-tailed recognition in an open world. In: CVPR (2019)
2019
Earlier work this paper cites.
Liu, Z., Miao, Z., Zhan, X., Wang, J., Gong, B., Yu, S.X.: Large-scale long-tailed recognition in an open world. In: CVPR (2019)
2019
Cited alongside, same era.
Lang, A.H., Vora, S., Caesar, H., Zhou, L., Yang, J., Beijbom, O.: Pointpillars: Fast encoders for object detection from point clouds. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2019)
2019
Cited alongside, same era.
Taeihagh, A., Lim, H.S.M.: Governing autonomous vehicles: emerging responses for safety, liability, privacy, cybersecurity, and industry risks. Transport Reviews (2019)
2019
Cited alongside, same era.
Tian, Z., Shen, C., Chen, H., He, T.: Fcos: Fully convolutional one-stage object detection. In: ICCV (2019)
2019
Cited alongside, same era.
Wang, X., Cai, Z., Gao, D., Vasconcelos, N.: Towards universal object detection by domain attention. In: CVPR (2019)
Bai, X., Hu, Z., Zhu, X., Huang, Q., Chen, Y., Fu, H., Tai, C.-L.: Transfusion: Robust lidar-camera fusion for 3d object detection with transformers. In: CVPR (2022)
2022
Later among the works it cites.
Chen, Y.-T., Shi, J., Ye, Z., Mertz, C., Ramanan, D., Kong, S.: Multimodal object detection via probabilistic ensembling. In: ECCV (2022)
2022
Later among the works it cites.
Gupta, S., Kanjani, J., Li, M., Ferroni, F., Hays, J., Ramanan, D., Kong, S.: Far3det: Towards far-field 3d detection. In: NeurIPS (2022)
2022
Later among the works it cites.
Grauman, K., Westbury, A., Byrne, E., Chavis, Z., Furnari, A., Girdhar, R., Hamburger, J., al.: Ego4d: Around the world in 3, 000 hours of egocentric video. In: Computer Vision and Pattern Recognition (2022)
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Wu, C.J., Tygert, M., LeCun, Y.: A hierarchical loss and its problems when classifying non-hierarchically. PLoS ONE 14
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Caesar, H., Bankiti, V., Lang, A.H., Vora, S., Liong, V.E., Xu, Q., Krishnan, A., Pan, Y., Baldan, G., Beijbom, O.: nuscenes: A multimodal dataset for autonomous driving. In: CVPR (2020)
2020
Cited alongside, same era.
Carion, N., Massa, F., Synnaeve, G., Usunier, N., Kirillov, A., Zagoruyko, S.: End-to-end object detection with transformers. In: ECCV (2020)
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Guizilini, V., Ambrus, R., Pillai, S., Raventos, A., Gaidon, A.: 3d packing for self-supervised monocular depth estimation. In: CVPR (2020)
2020
Cited alongside, same era.
Li, Y., Wang, T., Kang, B., Tang, S., Wang, C., Li, J., Feng, J.: Overcoming classifier imbalance for long-tail object detection with balanced group softmax. In: 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2020)
2020
Cited alongside, same era.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
Lin, Z., Pathak, D., Wang, Y.-X., Ramanan, D., Kong, S.: Continual learning with evolving class ontologies. Advances in Neural Information Processing Systems (2022)
2022
Later among the works it cites.
2022
Later among the works it cites.
Li, Z., Wang, W., Li, H., Xie, E., Sima, C., Lu, T., Qiao, Y., Dai, J.: Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers. In: ECCV (2022). Springer
2022
Later among the works it cites.
Li, L.H., Zhang, P., Zhang, H., Yang, J., Li, C., Zhong, Y., Wang, L., Yuan, L., Zhang, L., Hwang, J.-N., et al
2022
Later among the works it cites.
Peri, N., Dave, A., Ramanan, D., Kong, S.: Towards long-tailed 3d detection. In: Conference on Robot Learning (CoRL) (2022)
2022
Later among the works it cites.
Savage, N.: Robots rise to meet the challenge of caring for old people. Nature (2022)
2022
Later among the works it cites.
Shi, S., Jiang, L., Deng, J., Wang, Z., Guo, C., Shi, J., Wang, X., Li, H.: Pv-rcnn++: Point-voxel feature set abstraction with local vector representation for 3d object detection. IJCV (2022)
2022
Later among the works it cites.
2022
Later among the works it cites.
Yang, Z., Chen, J., Miao, Z., Li, W., Zhu, X., Zhang, L.: Deepinteraction: 3d object detection via modality interaction. In: NeurIPS (2022)
2022
Later among the works it cites.
Zhou, X., Koltun, V., Krähenbühl, P.: Simple multi-dataset detection. In: CVPR (2022)
2022
Later among the works it cites.
Zhang, H., Li, F., Liu, S., Zhang, L., Su, H., Zhu, J., Ni, L., Shum, H.: Dino: Detr with improved denoising anchor boxes for end-to-end object detection. In: ICLR (2022)
2022
Later among the works it cites.
2023
Closest in time.
2023
Closest in time.
Peri, N., Li, M., Wilson, B., Wang, Y.-X., Hays, J., Ramanan, D.: An empirical analysis of range for 3d object detection. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (2023)
2023
Closest in time.
Yan, J., Liu, Y., Sun, J., Jia, F., Li, S., Wang, T., Zhang, X.: Cross modal transformer: Towards fast and robust 3d object detection. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (2023)
2023
Closest in time.
Shi, J., Gare, G., Tian, J., Chai, S., Lin, Z., Vasudevan, A., Feng, D., Ferroni, F., Kong, S.: Lca-on-the-line: Benchmarking out-of-distribution generalization with class taxonomies. In: International Conference on Machine Learning (ICML) (2024)
2024
Closest in time.
Yin, J., Shen, J., Chen, R., Li, W., Yang, R., Frossard, P., Wang, W.: Is-fusion: Instance-scene collaborative fusion for multimodal 3d object detection. In: CVPR (2024)
2024
Closest in time.
Cai, H., Yin, D., Yu, F.R., Xiong, S.: Dstr: Dual scenes transformer for cross-modal fusion in 3d object detection. In: IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), pp. 3064–3073 (2025). IEEE
2025
Closest in time.
Jin, X., Su, H., Liu, K., Ma, C., Wu, W., Hui, F., Yan, J.: Unimamba: Unified spatial-channel representation learning with group-efficient mamba for lidar-based 3d object detection. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition Conference, pp. 1407–1417 (2025)
2025
Closest in time.
Liu, S., Cui, M., Li, B., Liang, Q., Hong, T., Huang, K., Shan, Y.: Fshnet: Fully sparse hybrid network for 3d object detection. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition Conference, pp. 8900–8909 (2025)
2025
Closest in time.
Robicheaux, P., Gallagher, J., Nelson, J., Robinson, I.: Rf-detr: A sota real-time object detection model. Roboflow Blog.(Introducing the RF-DETR model (2025)
2025
Closest in time.