Fetching the paper…
Reading the bibliography…
This paper presents EdgeSAM, an accelerated variant of the Segment Anything Model (SAM), optimized for efficient execution on edge devices with minimal compromise in performance.
Huber PJ (1992) Robust estimation of a location parameter. Breakthroughs in statistics: Methodology and distribution
1992
Earlier work this paper cites.
Hinton G, Vinyals O, Dean J (2014) Distilling the knowledge in a neural network. NeurIPSW
2014
Earlier work this paper cites.
Lin TY, Maire M, Belongie S, Hays J, Perona P, Ramanan D, Dollár P, Zitnick CL (2014) Microsoft coco: Common objects in context. In: ECCV
2014
Earlier work this paper cites.
Ren S, He K, Girshick R, Sun J (2015) Faster r-cnn: Towards real-time object detection with region proposal networks. NeurIPS
2015
Earlier work this paper cites.
Iandola FN, Han S, Moskewicz MW, Ashraf K, Dally WJ, Keutzer K (2016) Squeezenet: Alexnet-level accuracy with 50x fewer parameters and¡ 0.5 mb model size. arXiv preprint
2016
Earlier work this paper cites.
Howard AG, Zhu M, Chen B, Kalenichenko D, Wang W, Weyand T, Andreetto M, Adam H (2017) Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv preprint
2017
Earlier work this paper cites.
Sudre CH, Li W, Vercauteren T, Ourselin S, Jorge Cardoso M (2017) Generalised dice overlap as a deep learning loss function for highly unbalanced segmentations. In: MICCAIW
2017
Earlier work this paper cites.
Cai Z, Vasconcelos N (2018) Cascade r-cnn: Delving into high quality object detection. In: CVPR
2018
Earlier work this paper cites.
Ma N, Zhang X, Zheng HT, Sun J (2018) Shufflenet v2: Practical guidelines for efficient cnn architecture design. In: ECCV
2018
Earlier work this paper cites.
Mehta S, Rastegari M, Caspi A, Shapiro L, Hajishirzi H (2018) Espnet: Efficient spatial pyramid of dilated convolutions for semantic segmentation. In: ECCV
2018
Earlier work this paper cites.
Sandler M, Howard A, Zhu M, Zhmoginov A, Chen LC (2018) Mobilenetv2: Inverted residuals and linear bottlenecks. In: CVPR
2018
Earlier work this paper cites.
Xie J, Shuai B, Hu JF, Lin J, Zheng WS (2018) Improving fast segmentation with teacher-student learning. BMVC
2018
Earlier work this paper cites.
Yu C, Wang J, Peng C, Gao C, Yu G, Sang N (2018) Bisenet: Bilateral segmentation network for real-time semantic segmentation. In: ECCV
2018
Earlier work this paper cites.
Zhang X, Zhou X, Lin M, Sun J (2018) Shufflenet: An extremely efficient convolutional neural network for mobile devices. In: CVPR
2018
Earlier work this paper cites.
Zhao H, Qi X, Shen X, Shi J, Jia J (2018) Icnet for real-time semantic segmentation on high-resolution images. In: ECCV
2018
Earlier work this paper cites.
Bolya D, Zhou C, Xiao F, Lee YJ (2019) Yolact: Real-time instance segmentation. In: ICCV
2019
Earlier work this paper cites.
Chen K, Wang J, Pang J, Cao Y, Xiong Y, Li X, Sun S, Feng W, Liu Z, Xu J, et al (2019) Mmdetection: Open mmlab detection toolbox and benchmark. arXiv preprint
2019
Earlier work this paper cites.
Gupta A, Dollar P, Girshick R (2019) Lvis: A dataset for large vocabulary instance segmentation. In: CVPR
2019
Earlier work this paper cites.
Howard A, Sandler M, Chu G, Chen LC, Chen B, Tan M, Wang W, Zhu Y, Pang R, Vasudevan V, et al (2019) Searching for mobilenetv3. In: ICCV
2019
Earlier work this paper cites.
Liu Y, Chen K, Liu C, Qin Z, Luo Z, Wang J (2019) Structured knowledge distillation for semantic segmentation. In: CVPR
2019
Earlier work this paper cites.
Loshchilov I, Hutter F (2019) Decoupled weight decay regularization. In: ICLR
2019
Earlier work this paper cites.
Mehta S, Rastegari M, Shapiro L, Hajishirzi H (2019) Espnetv2: A light-weight, power efficient, and general purpose convolutional neural network. In: CVPR
2019
Earlier work this paper cites.
Wang T, Yuan L, Zhang X, Feng J (2019) Distilling object detectors with fine-grained feature imitation. In: CVPR
2019
Cited alongside, same era.
Wu Y, Kirillov A, Massa F, Lo WY, Girshick R (2019) Detectron2
2019
Cited alongside, same era.
Carion N, Massa F, Synnaeve G, Usunier N, Kirillov A, Zagoruyko S (2020) End-to-end object detection with transformers. In: ECCV
2020
Cited alongside, same era.
Forte M, Price B, Cohen S, Xu N, Pitié F (2020) Getting to 99% accuracy in interactive segmentation. arXiv preprint
2020
Cited alongside, same era.
Guan Y, Zhao P, Wang B, Zhang Y, Yao C, Bian K, Tang J (2020) Differentiable feature aggregation search for knowledge distillation. In: ECCV
2020
Cited alongside, same era.
Ma H, Xia X, Wang X, Xiao X, Li J, Zheng M (2022) Mocovit: Mobile convolutional vision transformer. arXiv preprint
2022
Later among the works it cites.
Maaz M, Shaker A, Cholakkal H, Khan S, Zamir SW, Anwer RM, Khan FS (2022) Edgenext: efficiently amalgamated cnn-transformer architecture for mobile vision applications. In: ECCVW
2022
Later among the works it cites.
Mehta S, Rastegari M (2022) Mobilevit: Light-weight, general-purpose, and mobile-friendly vision transformer. In: ICLR
2022
Later among the works it cites.
Pan J, Bulat A, Tan F, Zhu X, Dudziak L, Li H, Tzimiropoulos G, Martinez B (2022) Edgevits: Competing light-weight cnns on mobile devices with vision transformers. In: ECCV
2022
Later among the works it cites.
Sofiiuk K, Petrov IA, Konushin A (2022) Reviving iterative training with mask guidance for interactive segmentation. In: ICIP
2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
Li X, You A, Zhu Z, Zhao H, Yang M, Yang K, Tong Y (2020) Semantic flow for fast and accurate scene parsing. In: ECCV
2020
Cited alongside, same era.
Tan M, Pang R, Le QV (2020) Efficientdet: Scalable and efficient object detection. In: CVPR
2020
Cited alongside, same era.
Tancik M, Srinivasan P, Mildenhall B, Fridovich-Keil S, Raghavan N, Singhal U, Ramamoorthi R, Barron J, Ng R (2020) Fourier features let networks learn high frequency functions in low dimensional domains. NeurIPS
2020
Cited alongside, same era.
Wang Y, Zhou W, Jiang T, Bai X, Xu Y (2020) Intra-class feature variation distillation for semantic segmentation. In: ECCV
2020
Cited alongside, same era.
Xie Q, Luong MT, Hovy E, Le QV (2020) Self-training with noisy student improves imagenet classification. In: CVPR
2020
Cited alongside, same era.
Yuan L, Tay FE, Li G, Wang T, Feng J (2020) Revisiting knowledge distillation via label smoothing regularization. In: CVPR
2020
Cited alongside, same era.
Later among the works it cites.
Wu K, Zhang J, Peng H, Liu M, Xiao B, Fu J, Yuan L (2022) Tinyvit: Fast pretraining distillation for small vision transformers. In: ECCV
2022
Later among the works it cites.
Zhang W, Huang Z, Luo G, Chen T, Wang X, Liu W, Yu G, Shen C (2022) Topformer: Token pyramid transformer for mobile semantic segmentation. In: CVPR
2022
Later among the works it cites.
Zhou X, Girdhar R, Joulin A, Krähenbühl P, Misra I (2022) Detecting twenty-thousand classes using image-level supervision. In: ECCV
2022
Later among the works it cites.
(2023) Nanosam
2023
Closest in time.
Cai H, Li J, Hu M, Gan C, Han S (2023) Efficientvit: Lightweight multi-scale attention for on-device semantic segmentation. In: ICCV
2023
Closest in time.
Chang J, Wang S, Xu HM, Chen Z, Yang C, Zhao F (2023) Detrdistill: A universal knowledge distillation framework for detr-families. In: ICCV
2023
Closest in time.
Hu J, Huang L, Ren T, Zhang S, Ji R, Cao L (2023) You only segment once: Towards real-time panoptic segmentation. In: CVPR
2023
Closest in time.
Kirillov A, Mintun E, Ravi N, Mao H, Rolland C, Gustafson L, Xiao T, Whitehead S, Berg AC, Lo WY, et al (2023) Segment anything. ICCV
2023
Closest in time.
Li X, Zhang J, Yang Y, Cheng G, Yang K, Tong Y, Tao D (2023) Sfnet: Faster and accurate semantic segmentation via semantic flow. IJCV
2023
Closest in time.
Wan Q, Huang Z, Lu J, Yu G, Zhang L (2023) Seaformer: Squeeze-enhanced axial transformer for mobile semantic segmentation. In: ICLR
2023
Closest in time.
Xiong Y, Varadarajan B, Wu L, Xiang X, Xiao F, Zhu C, Dai X, Wang D, Sun F, Iandola F, et al (2023) Efficientsam: Leveraged masked image pretraining for efficient segment anything. arXiv preprint
2023
Closest in time.
Zhao X, Ding W, An Y, Du Y, Yu T, Li M, Tang M, Wang J (2023) Fast segment anything. arXiv preprint
2023
Closest in time.
Ravi N, Gabeur V, Hu YT, Hu R, Ryali C, Ma T, Khedr H, Rädle R, Rolland C, Gustafson L, et al (2024) Sam 2: Segment anything in images and videos. arXiv preprint
2024
Closest in time.
Sun X, Liu J, Shen HT, Zhu X, Hu P (2024) On efficient variants of segment anything model: A survey. arXiv preprint
2024
Closest in time.
Xiong Y, Zhou C, Xiang X, Wu L, Zhu C, Liu Z, Suri S, Varadarajan B, Akula R, Iandola F, et al (2024) Efficient track anything. arXiv preprint
2024
Closest in time.
Zhang Z, Cai H, Han S (2024) Efficientvit-sam: Accelerated segment anything model without performance loss. In: CVPRW
2024
Closest in time.
Zhou C, Zhu C, Xiong Y, Suri S, Xiao F, Wu L, Krishnamoorthi R, Dai B, Loy CC, Chandra V, et al (2025) Edgetam: On-device track anything model. arXiv preprint
2025
Closest in time.