Fetching the paper…
Reading the bibliography…
We propose the SAL (Segment Anything in Lidar) method consisting of a text-promptable zero-shot model for segmenting and classifying any object in Lidar, and a pseudo-labeling engine that facilitates model training without manual supervision.
Thorpe, C., Herbert, M., Kanade, T., Shafer, S.: Toward autonomous driving: the cmu navlab. i. perception. IEEE expert
1991
Earlier work this paper cites.
Ester, M., Kriegel, H.P., Sander, J., Xu, X., et al.: A density-based algorithm for discovering clusters in large spatial databases with noise. In: Rob. Sci. Sys. (1996)
1996
Earlier work this paper cites.
Thrun, S., Montemerlo, M., Dahlkamp, H., Stavens, D., Aron, A., Diebel, J., Fong, P., Gale, J., Halpenny, M., Hoffmann, G.: Stanley: The robot that won the darpa grand challenge. Journal of field Robotics (2006)
2006
Earlier work this paper cites.
Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., Fei-Fei, L.: ImageNet: A large-scale hierarchical image database. In: IEEE Conf. Comput. Vis. Pattern Recog. (2009)
2009
Earlier work this paper cites.
Petrovskaya, A., Thrun, S.: Model based vehicle detection and tracking for autonomous urban driving. Aut. Rob
2009
Earlier work this paper cites.
Teichman, A., Levinson, J., Thrun, S.: Towards 3D object recognition via classification of arbitrary object tracks. In: Int. Conf. Rob. Automat. (2011)
2011
Earlier work this paper cites.
Xiong, X., Munoz, D., Bagnell, J.A., Hebert, M.: 3-D Scene Analysis via Sequenced Predictions over Points and Regions. In: Int. Conf. Rob. Automat. pp. 2609–2616 (2011)
2011
Earlier work this paper cites.
Achanta, R., Shaji, A., Smith, K., Lucchi, A., Fua, P., Süsstrunk, S.: Slic superpixels compared to state-of-the-art superpixel methods. IEEE Trans. Pattern Anal. Mach. Intell
2012
Earlier work this paper cites.
Prest, A., Leistner, C., Civera, J., Schmid, C., Ferrari, V.: Learning object class detectors from weakly annotated video. In: IEEE Conf. Comput. Vis. Pattern Recog. (2012)
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
Moosmann, F., Stiller, C.: Joint self-localization and tracking of generic objects in 3d range data. In: Int. Conf. Rob. Automat. (2013)
2013
Earlier work this paper cites.
Held, D., Levinson, J., Thrun, S., Savarese, S.: Combining 3d shape, color, and motion for robust anytime tracking. In: Rob. Sci. Sys. (2014)
2014
Earlier work this paper cites.
Lin, T., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., Zitnick, C.L.: Microsoft COCO: Common objects in context. In: Eur. Conf. Comput. Vis. (2014)
2014
Earlier work this paper cites.
Cordts, M., Omran, M., Ramos, S., Rehfeld, T., Enzweiler, M., Benenson, R., Franke, U., Roth, S., Schiele, B.: The cityscapes dataset for semantic urban scene understanding. In: IEEE Conf. Comput. Vis. Pattern Recog. (2016)
2016
Earlier work this paper cites.
Held, D., Guillory, D., Rebsamen, B., Thrun, S., Savarese, S.: A probabilistic framework for real-time 3d segmentation using spatial, temporal, and semantic cues. In: Rob. Sci. Sys. (2016)
2016
Earlier work this paper cites.
Dai, A., Chang, A.X., Savva, M., Halber, M., Funkhouser, T., Nießner, M.: Scannet: Richly-annotated 3d reconstructions of indoor scenes. In: IEEE Conf. Comput. Vis. Pattern Recog. (2017)
2017
Earlier work this paper cites.
Qi, C.R., Su, H., Mo, K., Guibas, L.J.: Pointnet: Deep learning on point sets for 3d classification and segmentation. In: IEEE Conf. Comput. Vis. Pattern Recog. (2017)
2017
Earlier work this paper cites.
Qi, C.R., Yi, L., Su, H., Guibas, L.J.: Pointnet++: Deep hierarchical feature learning on point sets in a metric space. In: Adv. Neural Inform. Process. Syst. (2017)
2017
Earlier work this paper cites.
Bansal, A., Sikka, K., Sharma, G., Chellappa, R., Divakaran, A.: Zero-shot object detection. In: Eur. Conf. Comput. Vis. (2018)
2018
Earlier work this paper cites.
Kirillov, A., He, K., Girshick, R.B., Rother, C., Dollár, P.: Panoptic segmentation. IEEE Conf. Comput. Vis. Pattern Recog. (2018)
2018
Earlier work this paper cites.
Miller, D., Nicholson, L., Dayoub, F., Sünderhauf, N.: Dropout sampling for robust object detection in open-set conditions. In: Int. Conf. Rob. Automat. (2018)
2018
Earlier work this paper cites.
Osep, A., Voigtlaender, P., Luiten, J., Breuers, S., Leibe, B.: Towards large-scale video video object mining. In: ECCV Workshop on Interactive and Adaptive Learning in an Open World (2018)
2018
Earlier work this paper cites.
Ošep, A., Mehner, W., Voigtlaender, P., Leibe, B.: Track, then decide: Category-agnostic vision-based multi-object tracking. In: Int. Conf. Rob. Automat. (2018)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Rahman, S., Khan, S.H., Porikli, F.: Zero-shot object detection: Learning to simultaneously recognize and localize novel concepts. Asian Conf. Comput. Vis. (2018)
2018
Earlier work this paper cites.
Wu, B., Wan, A., Yue, X., Keutzer, K.: Squeezeseg: Convolutional neural nets with recurrent crf for real-time road-object segmentation from 3d lidar point cloud. In: Int. Conf. Rob. Automat. (2018)
2018
Earlier work this paper cites.
Xian, Y., Lampert, C.H., Schiele, B., Akata, Z.: Zero-shot learning - a comprehensive evaluation of the good, the bad and the ugly. IEEE Trans. Pattern Anal. Mach. Intell. (2018)
2018
Earlier work this paper cites.
Yan, Y., Mao, Y., Li, B.: Second: Sparsely embedded convolutional detection. Sensors
2018
Earlier work this paper cites.
Zhou, Y., Tuzel, O.: Voxelnet: End-to-end learning for point cloud based 3d object detection. In: IEEE Conf. Comput. Vis. Pattern Recog. (2018)
2018
Earlier work this paper cites.
Behley, J., Garbade, M., Milioto, A., Quenzel, J., Behnke, S., Stachniss, C., Gall, J.: SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR Sequences. In: Int. Conf. Comput. Vis. (2019)
2019
Earlier work this paper cites.
Bucher, M., Vu, T.H., Cord, M., Pérez, P.: Zero-shot semantic segmentation. Adv. Neural Inform. Process. Syst. (2019)
2019
Earlier work this paper cites.
Choy, C., Gwak, J., Savarese, S.: 4D spatio-temporal convnets: Minkowski convolutional neural networks. In: IEEE Conf. Comput. Vis. Pattern Recog. (2019)
2019
Earlier work this paper cites.
Kirillov, A., He, K., Girshick, R., Rother, C., Dollár, P.: Panoptic segmentation. In: IEEE Conf. Comput. Vis. Pattern Recog. (2019)
2019
Earlier work this paper cites.
Lang, A.H., Vora, S., Caesar, H., Zhou, L., Yang, J., Beijbom, O.: Pointpillars: Fast encoders for object detection from point clouds. In: IEEE Conf. Comput. Vis. Pattern Recog. (2019)
2019
Earlier work this paper cites.
Milioto, A., Vizzo, I., Behley, J., Stachniss, C.: RangeNet++: Fast and Accurate LiDAR Semantic Segmentation. In: Int. Conf. Intel. Rob. Sys. (2019)
2019
Cited alongside, same era.
Ošep, A., Voigtlaender, P., Luiten, J., Breuers, S., Leibe, B.: Large-scale object mining for object discovery from unlabeled video. In: Int. Conf. Rob. Automat. (2019)
2019
Cited alongside, same era.
Thomas, H., Qi, C.R., Deschaud, J.E., Marcotegui, B., Goulette, F., Guibas, L.J.: Kpconv: Flexible and deformable convolution for point clouds. In: Int. Conf. Comput. Vis. (2019)
2019
Cited alongside, same era.
Wu, B., Zhou, X., Zhao, S., Yue, X., Keutzer, K.: Squeezesegv2: Improved model structure and unsupervised domain adaptation for road-object segmentation from a lidar point cloud. In: Int. Conf. Rob. Automat. (2019)
2019
Cited alongside, same era.
Kreuzberg, L., Zulfikar, I.E., Mahadevan, S., Engelmann, F., Leibe, B.: 4d-stop: Panoptic segmentation of 4d lidar using spatio-temporal object proposal generation and aggregation. In: ECCV AVVision Workshop (2022)
2022
Later among the works it cites.
Lee, S., Lim, H., Myung, H.: Patchwork++: Fast and robust ground segmentation solving partial under-segmentation using 3d point cloud. In: Int. Conf. Intel. Rob. Sys. (2022)
2022
Later among the works it cites.
Li, B., Weinberger, K.Q., Belongie, S., Koltun, V., Ranftl, R.: Language-driven semantic segmentation. In: Int. Conf. Learn. Represent. (2022)
2022
Later among the works it cites.
Li, J., He, X., Wen, Y., Gao, Y., Cheng, Y., Zhang, D.: Panoptic-phnet: Towards real-time and high-precision lidar panoptic segmentation via clustering pseudo heatmap. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Aksoy, E.E., Baci, S., Cavdar, S.: Salsanet: Fast road and vehicle segmentation in lidar point clouds for autonomous driving. In: Intel. Veh. Symp. (2020)
2020
Cited alongside, same era.
Carion, N., Massa, F., Synnaeve, G., Usunier, N., Kirillov, A., Zagoruyko, S.: End-to-end object detection with transformers. In: Eur. Conf. Comput. Vis. (2020)
2020
Cited alongside, same era.
Caron, M., Misra, I., Mairal, J., Goyal, P., Bojanowski, P., Joulin, A.: Unsupervised learning of visual features by contrasting cluster assignments. Adv. Neural Inform. Process. Syst. (2020)
2020
Cited alongside, same era.
He, K., Fan, H., Wu, Y., Xie, S., Girshick, R.: Momentum contrast for unsupervised visual representation learning. In: IEEE Conf. Comput. Vis. Pattern Recog. (2020)
2020
Cited alongside, same era.
Hu, P., Held, D., Ramanan, D.: Learning to optimally segment point clouds. IEEE Robotics and Automation Letters
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Sun, P., Kretzschmar, H., Dotiwalla, X., Chouard, A., Patnaik, V., Tsui, P., Guo, J., Zhou, Y., Chai, Y., Caine, B., et al.: Scalability in perception for autonomous driving: Waymo open dataset. In: IEEE Conf. Comput. Vis. Pattern Recog. (2020)
2020
Cited alongside, same era.
Tang, H., Liu, Z., Zhao, S., Lin, Y., Lin, J., Wang, H., Han, S.: Searching efficient 3d architectures with sparse point-voxel convolution. In: Eur. Conf. Comput. Vis. (2020)
2020
Cited alongside, same era.
Lin, Z., Pathak, D., Wang, Y.X., Ramanan, D., Kong, S.: Continual learning with evolving class ontologies. Adv. Neural Inform. Process. Syst. (2022)
2022
Later among the works it cites.
Marcuzzi, R., Nunes, L., Wiesmann, L., Vizzo, I., Behley, J., Stachniss, C.: Contrastive instance association for 4d panoptic segmentation using sequences of 3d lidar scans. IEEE Rob. Automat. Letters (2022)
2022
Later among the works it cites.
Najibi, M., Ji, J., Zhou, Y., Qi, C.R., Yan, X., Ettinger, S., Anguelov, D.: Motion inspired unsupervised perception and prediction in autonomous driving. In: Eur. Conf. Comput. Vis. (2022)
2022
Later among the works it cites.
Nunes, L., Marcuzzi, R., Chen, X., Behley, J., Stachniss, C.: Segcontrast: 3d point cloud feature representation learning through self-supervised segment discrimination. IEEE Rob. Automat. Letters
2022
Later among the works it cites.
Peri, N., Luiten, J., Li, M., Ošep, A., Leal-Taixé, L., Ramanan, D.: Forecasting from lidar via future object detection. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Rao, Y., Zhao, W., Chen, G., Tang, Y., Zhu, Z., Huang, G., Zhou, J., Lu, J.: Denseclip: Language-guided dense prediction with context-aware prompting. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Sautier, C., Puy, G., Gidaris, S., Boulch, A., Bursuc, A., Marlet, R.: Image-to-lidar self-supervised distillation for autonomous driving data. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Zhong, Y., Yang, J., Zhang, P., Li, C., Codella, N., Li, L.H., Zhou, L., Dai, X., Yuan, L., Li, Y., et al.: Regionclip: Region-based language-image pretraining. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Zhou, C., Loy, C.C., Dai, B.: Extract free dense labels from clip. In: Eur. Conf. Comput. Vis. (2022)
2022
Later among the works it cites.
Agarwalla, A., Huang, X., Ziglar, J., Ferroni, F., Laura, L.T., Hays, J., Osep, A., Ramanan, D.: Lidar panoptic segmentation and tracking without bells and whistles. In: Int. Conf. Intel. Rob. Sys. (2023)
2023
Later among the works it cites.
Ding, Z., Wang, J., Tu, Z.: Open-vocabulary universal image segmentation with maskclip. In: Int. Conf. Mach. Learn. (2023)
2023
Later among the works it cites.
Kirillov, A., Mintun, E., Ravi, N., Mao, H., Rolland, C., Gustafson, L., Xiao, T., Whitehead, S., Berg, A.C., Lo, W.Y., et al.: Segment anything. In: Int. Conf. Comput. Vis. (2023)
2023
Later among the works it cites.
Liang, F., Wu, B., Dai, X., Li, K., Zhao, Y., Zhang, H., Zhang, P., Vajda, P., Marculescu, D.: Open-vocabulary semantic segmentation with mask-adapted clip. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Lu, Y., Jiang, Q., Chen, R., Hou, Y., Zhu, X., Ma, Y.: See more and know more: Zero-shot point cloud segmentation via multi-modal visual data. In: Int. Conf. Comput. Vis. (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Marcuzzi, R., Nunes, L., Wiesmann, L., Behley, J., Stachniss, C.: Mask-based panoptic lidar segmentation for autonomous driving. IEEE Rob. Automat. Letters
2023
Later among the works it cites.
Marcuzzi, R., Nunes, L., Wiesmann, L., Marks, E., Behley, J., Stachniss, C.: Mask4d: End-to-end mask-based 4d panoptic segmentation for lidar sequences. IEEE Rob. Automat. Letters (2023)
2023
Later among the works it cites.
Najibi, M., Ji, J., Zhou, Y., Qi, C.R., Yan, X., Ettinger, S., Anguelov, D.: Unsupervised 3d perception with 2d vision-language distillation for autonomous driving. In: Int. Conf. Comput. Vis. (2023)
2023
Later among the works it cites.
Peng, S., Genova, K., Jiang, C., Tagliasacchi, A., Pollefeys, M., Funkhouser, T.: Openscene: 3d scene understanding with open vocabularies. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Later among the works it cites.
Peri, N., Dave, A., Ramanan, D., Kong, S.: Towards long-tailed 3d detection. In: Conf. Rob. Learn. (2023)
2023
Later among the works it cites.
Peri, N., Li, M., Wilson, B., Wang, Y.X., Hays, J., Ramanan, D.: An empirical analysis of range for 3d object detection. In: ICCV Workshops (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Xu, J., Liu, S., Vahdat, A., Byeon, W., Wang, X., De Mello, S.: Open-vocabulary panoptic segmentation with text-to-image diffusion models. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Later among the works it cites.
Xu, M., Zhang, Z., Wei, F., Hu, H., Bai, X.: Side adapter network for open-vocabulary semantic segmentation. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhang, L., Yang, A.J., Xiong, Y., Casas, S., Yang, B., Ren, M., Urtasun, R.: Towards unsupervised object detection from lidar point clouds. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Later among the works it cites.
Zhu, M., Han, S., Cai, H., Borse, S., Ghaffari, M., Porikli, F.: 4d panoptic segmentation as invariant and equivariant field prediction. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Later among the works it cites.
Seidenschwarz, J., Ošep, A., Ferroni, F., Lucey, S., Leal-Taixé, L.: Semoli: What moves together belongs together. IEEE Conf. Comput. Vis. Pattern Recog. (2024)
2024
Closest in time.