Fetching the paper…
Reading the bibliography…
Panoptic segmentation is an important computer vision task, where the current state-of-the-art solutions require specialized components to perform well.
The hungarian method for the assignment problem
Kuhn, H. W · 1955
Earlier work this paper cites.
Dynamic Programming
Bellman, R. and Corporation, R · 1959
Earlier work this paper cites.
Adaptive Control Processes: A Guided Tour
Bellman, R · 1961
Earlier work this paper cites.
Nonlinear total variation based noise removal algorithms
Rudin, L. I., Osher, S., and Fatemi, E · 1992
Earlier work this paper cites.
Generalized overlap measures for evaluation and validation in medical image analysis
Crum, W. R., Camara, O., and Hill, D. L · 2006
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
Silberman, N., Hoiem, D., Kohli, P., and Fergus, R · 2012
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
Eigen, D., Puhrsch, C., and Fergus, R · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., and Zitnick, C. L · 2014
Earlier work this paper cites.
The multimodal brain tumor image segmentation benchmark (brats)
Menze, B. H., Jakab, A., Bauer, S., Kalpathy-Cramer, J., Farahani, K., Kirby, J., Burren, Y., Porz, N., Slotboom, J., Wiest, R., et al · 2014
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation
Long, J., Shelhamer, E., and Darrell, T · 2015
Earlier work this paper cites.
The cityscapes dataset for semantic urban scene understanding
Cordts, M., Omran, M., Ramos, S., Rehfeld, T., Enzweiler, M., Benenson, R., Franke, U., Roth, S., and Schiele, B · 2016
Earlier work this paper cites.
Sgdr: Stochastic gradient descent with warm restarts
Loshchilov, I. and Hutter, F · 2016
Earlier work this paper cites.
Mask r-cnn
He, K., Gkioxari, G., Dollár, P., and Girshick, R · 2017
Earlier work this paper cites.
Generalised dice overlap as a deep learning loss function for highly unbalanced segmentations
Sudre, C. H., Li, W., Vercauteren, T., Ourselin, S., and Jorge Cardoso, M · 2017
Earlier work this paper cites.
Cascade r-cnn delving into high quality object detection
Vasconcelos, N. and Cai, Z · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Earlier work this paper cites.
Lateral inhibition and image processing
Jernigan, M. and McLean, G · 2018
Cited alongside, same era.
Multi-task learning using uncertainty to weigh losses for scene geometry and semantics
Kendall, A., Gal, Y., and Cipolla, R · 2018
Cited alongside, same era.
Panoptic segmentation with an end-to-end cell r-cnn for pathology image analysis
Zhang, D., Song, Y., Liu, D., Jia, H., Liu, S., Xia, Y., Huang, H., and Cai, W · 2018
Cited alongside, same era.
Panoptic segmentation
Kirillov, A., He, K., Girshick, R., Rother, C., and Dollár, P · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Loshchilov, I. and Hutter, F · 2019
Cited alongside, same era.
Deep gated attention networks for large-scale street-level scene segmentation
Zhang, P., Liu, W., Wang, H., Lei, Y., and Lu, H · 2019
Cited alongside, same era.
Vision transformers for dense prediction
Ranftl, R., Bochkovskiy, A., and Koltun, V · 2021
Later among the works it cites.
Masked-attention mask transformer for universal image segmentation
Cheng, B., Misra, I., Schwing, A. G., Kirillov, A., and Girdhar, R · 2022
Later among the works it cites.
Vision and art (updated and expanded edition)
Livingstone, M. S · 2022
Later among the works it cites.
Unified-io: A unified model for vision, language, and multi-modal tasks
Lu, J., Clark, C., Zellers, R., Mottaghi, R., and Kembhavi, A · 2022
Later among the works it cites.
kmax-deeplab: k-means mask transformer
Yu, Q., Wang, H., Qiao, S., Collins, M., Zhu, Y., Adam, H., Yuille, A., and Chen, L.-C · 2022
Later among the works it cites.
Midas v3. 1–a model zoo for robust monocular relative depth estimation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning fully dense neural networks for image semantic segmentation
Zhen, M., Wang, J., Zhou, L., Fang, T., and Quan, L · 2019
Cited alongside, same era.
Improving semantic segmentation via video propagation and label relaxation
Zhu, Y., Sapra, K., Reda, F. A., Shih, K. J., Newsam, S., Tao, A., and Catanzaro, B · 2019
Cited alongside, same era.
End-to-end object detection with transformers
Carion, N., Massa, F., Synnaeve, G., Usunier, N., Kirillov, A., and Zagoruyko, S · 2020
Cited alongside, same era.
Panoptic-deeplab: A simple, strong, and fast baseline for bottom-up panoptic segmentation
Cheng, B., Collins, M. D., Zhu, Y., Liu, T., Huang, T. S., Adam, H., and Chen, L.-C · 2020
Cited alongside, same era.
Agriculture-vision: A large aerial image database for agricultural pattern analysis
Chiu, M. T., Xu, X., Wei, Y., Huang, Z., Schwing, A. G., Brunner, R., Khachatrian, H., Karapetyan, H., Dozier, I., Rose, G., et al · 2020
Cited alongside, same era.
Per-pixel classification is not all you need for semantic segmentation
Cheng, B., Schwing, A., and Kirillov, A · 2021
Cited alongside, same era.
Birkl, R., Wofk, D., and Müller, M · 2023
Later among the works it cites.
A generalist framework for panoptic segmentation of images and videos
Chen, T., Li, L., Saxena, S., Hinton, G., and Fleet, D. J · 2023
Later among the works it cites.
Hierarchical mask2former: Panoptic segmentation of crops, weeds and leaves
Darbyshire, M., Sklar, E., and Parsons, S · 2023
Later among the works it cites.
Oneformer: One transformer to rule universal image segmentation
Jain, J., Li, J., Chiu, M. T., Hassani, A., Orlov, N., and Shi, H · 2023
Later among the works it cites.
Mask dino: Towards a unified transformer-based framework for object detection and segmentation
Li, F., Zhang, H., Xu, H., Liu, S., Zhang, L., Ni, L. M., and Shum, H.-Y · 2023
Later among the works it cites.
Dinov2: Learning robust visual features without supervision
Oquab, M., Darcet, T., Moutakanni, T., Vo, H., Szafraniec, M., Khalidov, V., Fernandez, P., Haziza, D., Massa, F., El-Nouby, A., et al · 2023
Later among the works it cites.
Panoptic swiftnet: Pyramidal fusion for real-time panoptic segmentation
Šarić, J., Oršić, M., and Šegvić, S · 2023
Later among the works it cites.
Images speak in images: A generalist painter for in-context visual learning
Wang, X., Wang, W., Cao, Y., Shen, C., and Huang, T · 2023
Later among the works it cites.
Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action
Lu, J., Clark, C., Lee, S., Zhang, Z., Khosla, S., Marten, R., Hoiem, D., and Kembhavi, A · 2024
Closest in time.
A simple latent diffusion approach for panoptic segmentation and mask inpainting
Van Gansbeke, W. and De Brabandere, B · 2024
Closest in time.