Fetching the paper…
Reading the bibliography…
Batch Normalization's (BN) unique property of depending on other samples in a batch is known to cause problems in several tasks, including sequence modeling.
Temporal memory relation network for workflow recognition from surgical video
Jin, Y., Long, Y., Chen, C., Zhao, Z., Dou, Q., Heng, P.A., 2021 · 1923
Earlier work this paper cites.
Long short-term memory
Hochreiter, S., Schmidhuber, J., 1997 · 1997
Earlier work this paper cites.
Towards stabilizing batch statistics in backward propagation of batch normalization
Yan, J., Wan, R., Zhang, X., Zhang, W., Wei, Y., Sun, J., 2020 · 2001
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., et al., 2020 · 2010
Earlier work this paper cites.
Learning to recognize objects in egocentric activities, in: CVPR 2011, IEEE. pp. 3281–3288
Fathi, A., Ren, X., Rehg, J.M., 2011 · 2011
Earlier work this paper cites.
Privileged knowledge distillation for online action detection
Zhao, P., Xie, L., Zhang, Y., Wang, Y., Tian, Q., 2020 · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks, in: Pereira, F., Burges, C.J.C., Bottou, L., Weinberger, K.Q. (Eds.), Advances in Neural Information Processing Systems, Curran Associates, Inc
Krizhevsky, A., Sutskever, I., Hinton, G.E., 2012 · 2012
Earlier work this paper cites.
Combining embedded accelerometers with computer vision for recognizing food preparation activities, in: Proceedings of the 2013 ACM international joint conference on Pervasive and ubiquitous computing, pp. 729–738
Stein, S., McKenna, S.J., 2013 · 2013
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, K., Zisserman, A., 2014 · 2014
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift, in: International conference on machine learning, PMLR. pp. 448–456
Ioffe, S., Szegedy, C., 2015 · 2015
Earlier work this paper cites.
Ba, J.L., Kiros, J.R., Hinton, G.E., 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 770–778
He, K., Zhang, X., Ren, S., Sun, J., 2016 · 2016
Earlier work this paper cites.
Endorcn: recurrent convolutional networks for recognition of surgical workflow in cholecystectomy procedure video
Jin, Y., Dou, Q., Chen, H., Yu, L., Heng, P.A., 2016 · 2016
Earlier work this paper cites.
Weight normalization: A simple reparameterization to accelerate training of deep neural networks
Salimans, T., Kingma, D.P., 2016 · 2016
Earlier work this paper cites.
The tum lapchole dataset for the m2cai 2016 workflow challenge
Stauder, R., Ostler, D., Kranzfelder, M., Koller, S., Feußner, H., Navab, N., 2016 · 2016
Earlier work this paper cites.
Rethinking the inception architecture for computer vision, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 2818–2826
Szegedy, C., Vanhoucke, V., Ioffe, S., Shlens, J., Wojna, Z., 2016 · 2016
Earlier work this paper cites.
Endonet: a deep architecture for recognition tasks on laparoscopic videos
Twinanda, A.P., Shehata, S., Mutter, D., Marescaux, J., De Mathelin, M., Padoy, N., 2016 · 2016
Earlier work this paper cites.
Instance normalization: The missing ingredient for fast stylization
Ulyanov, D., Vedaldi, A., Lempitsky, V., 2016 · 2016
Earlier work this paper cites.
Deep neural networks predict remaining surgery duration from cholecystectomy videos, in: International Conference on Medical Image Computing and Computer-Assisted Intervention, Springer. pp. 586–593
Aksamentov, I., Twinanda, A.P., Mutter, D., Marescaux, J., Padoy, N., 2017 · 2017
Earlier work this paper cites.
Quo vadis, action recognition? a new model and the kinetics dataset, in: proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 6299–6308
Carreira, J., Zisserman, A., 2017 · 2017
Earlier work this paper cites.
Red: Reinforced encoder-decoder networks for action anticipation
Gao, J., Yang, Z., Nevatia, R., 2017 · 2017
Earlier work this paper cites.
Train longer, generalize better: closing the generalization gap in large batch training of neural networks
Hoffer, E., Hubara, I., Soudry, D., 2017 · 2017
Earlier work this paper cites.
Batch renormalization: Towards reducing minibatch dependence in batch-normalized models
Ioffe, S., 2017 · 2017
Earlier work this paper cites.
Sv-rcnet: workflow recognition from surgical videos using recurrent convolutional network
Jin, Y., Dou, Q., Chen, H., Yu, L., Qin, J., Fu, C.W., Heng, P.A., 2017 · 2017
Earlier work this paper cites.
Temporal convolutional networks for action segmentation and detection, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Lea, C., Flynn, M.D., Vidal, R., Reiter, A., Hager, G.D., 2017 · 2017
Earlier work this paper cites.
Surgical data science for next-generation interventions
Maier-Hein, L., Vedula, S.S., Speidel, S., Navab, N., Kikinis, R., Park, A., Eisenmann, M., Feussner, H., Forestier, G., Giannarou, S., et al., 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, Ł., Polosukhin, I., 2017 · 2017
Earlier work this paper cites.
Understanding batch normalization
Bjorck, N., Gomes, C.P., Selman, B., Weinberger, K.Q., 2018 · 2018
Earlier work this paper cites.
Rsdnet: Learning to predict remaining surgery duration from laparoscopic videos without manual annotations
Twinanda, A.P., Yengera, G., Mutter, D., Marescaux, J., Padoy, N., 2018 · 2018
Cited alongside, same era.
Group normalization, in: Proceedings of the European conference on computer vision (ECCV), pp. 3–19
Wu, Y., He, K., 2018 · 2018
Cited alongside, same era.
Yengera, G., Mutter, D., Marescaux, J., Padoy, N., 2018 · 2018
Cited alongside, same era.
Deepphase: surgical phase recognition in cataracts videos, in: International conference on medical image computing and computer-assisted intervention, Springer. pp. 265–272
Zisimopoulos, O., Flouty, E., Luengo, I., Giataganas, P., Nehme, J., Chow, A., Stoyanov, D., 2018 · 2018
Cited alongside, same era.
Lrtd: long-range temporal dependency based active learning for surgical workflow recognition
Shi, X., Jin, Y., Dou, Q., Heng, P.A., 2020 · 2020
Later among the works it cites.
Filter response normalization layer: Eliminating batch dependence in the training of deep neural networks, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 11237–11246
Singh, S., Krishnan, S., 2020 · 2020
Later among the works it cites.
Boundary-aware cascade networks for temporal action segmentation, in: European Conference on Computer Vision, Springer. pp. 34–51
Wang, Z., Gao, Z., Wang, L., Li, Z., Wu, G., 2020 · 2020
Later among the works it cites.
High-performance large-scale image recognition without normalization, in: International Conference on Machine Learning, PMLR. pp. 1059–1071
Brock, A., De, S., Smith, S.L., Simonyan, K., 2021 · 2021
Later among the works it cites.
Dynamic normalization and relay for video action recognition
Cai, D., Yao, A., Chen, Y., 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Farha, Y.A., Gall, J., 2019 · 2019
Cited alongside, same era.
Using 3d convolutional neural networks to learn spatiotemporal features for automatic surgical gesture recognition in video, in: International conference on medical image computing and computer-assisted intervention, Springer. pp. 467–475
Funke, I., Bodenstedt, S., Oehme, F., von Bechtolsheim, F., Weitz, J., Speidel, S., 2019 · 2019
Cited alongside, same era.
Future-state predicting lstm for early surgery type recognition
Kannan, S., Yengera, G., Mutter, D., Marescaux, J., Padoy, N., 2019 · 2019
Cited alongside, same era.
Time-conditioned action anticipation in one shot, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 9925–9934
Ke, Q., Fritz, M., Schiele, B., 2019 · 2019
Cited alongside, same era.
Weakly supervised convolutional lstm approach for tool tracking in laparoscopic videos
Nwoye, C.I., Mutter, D., Marescaux, J., Padoy, N., 2019 · 2019
Cited alongside, same era.
Unsupervised temporal video segmentation as an auxiliary task for predicting the remaining surgery duration, in: OR 2.0 Context-Aware Operating Theaters and Machine Learning in Clinical Neuroimaging. Springer, pp. 29–37
Rivoir, D., Bodenstedt, S., Bechtolsheim, F.v., Distler, M., Weitz, J., Speidel, S., 2019 · 2019
Cited alongside, same era.
Evalnorm: Estimating batch normalization statistics for evaluation, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 3633–3641
Singh, S., Shrivastava, A., 2019 · 2019
Cited alongside, same era.
Efficientnet: Rethinking model scaling for convolutional neural networks, in: International conference on machine learning, PMLR. pp. 6105–6114
Tan, M., Le, Q., 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Opera: Attention-regularized transformers for surgical phase recognition, in: International Conference on Medical Image Computing and Computer-Assisted Intervention, Springer. pp. 604–614
Czempiel, T., Paschali, M., Ostler, D., Kim, S.T., Busam, B., Navab, N., 2021 · 2021
Later among the works it cites.
Trans-svnet: accurate phase recognition from surgical videos via hybrid embedding aggregation transformer, in: International Conference on Medical Image Computing and Computer-Assisted Intervention, Springer. pp. 593–603
Gao, X., Jin, Y., Long, Y., Dou, Q., Heng, P.A., 2021 · 2021
Later among the works it cites.
Anticipative video transformer, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 13505–13515
Girdhar, R., Grauman, K., 2021 · 2021
Later among the works it cites.
Alleviating over-segmentation errors by detecting action boundaries, in: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pp. 2322–2331
Ishikawa, Y., Kasai, S., Aoki, Y., Kataoka, H., 2021 · 2021
Later among the works it cites.
Proxy-normalizing activations to match batch normalization while removing batch dependence
Labatie, A., Masters, D., Eaton-Rosen, Z., Luschi, C., 2021 · 2021
Later among the works it cites.
Catanet: Predicting remaining cataract surgery duration, in: International Conference on Medical Image Computing and Computer-Assisted Intervention, Springer. pp. 426–435
Marafioti, A., Hayoz, M., Gallardo, M., Márquez Neila, P., Wolf, S., Zinkernagel, M., Sznitman, R., 2021 · 2021
Later among the works it cites.
Rethinking" batch" in batchnorm
Wu, Y., Johnson, J., 2021 · 2021
Later among the works it cites.
Back to the future: Cycle encoding prediction for self-supervised video representation learning, in: The 32nd British Machine Vision Conference
Yang, X., Mirmehdi, M., Burghardt, T., 2021 · 2021
Later among the works it cites.
Asformer: Transformer for action segmentation, in: The British Machine Vision Conference (BMVC)
Yi, F., Wen, H., Jiang, T., 2021 · 2021
Later among the works it cites.
Surgical workflow anticipation using instrument interaction, in: International conference on medical image computing and computer-assisted intervention, Springer. pp. 615–625
Yuan, K., Holden, M., Gao, S., Lee, W.S., 2021 · 2021
Later among the works it cites.
Swnet: Surgical workflow recognition with deep convolutional network, in: Medical Imaging with Deep Learning, PMLR. pp. 855–869
Zhang, B., Ghanem, A., Simes, A., Choi, H., Yoo, A., Min, A., 2021 · 2021
Later among the works it cites.
Spatio-temporal causal transformer for multi-grained surgical phase recognition, in: 2022 44th Annual International Conference of the IEEE Engineering in Medicine & Biology Society (EMBC), IEEE. pp. 1663–1666
Chen, H.B., Li, Z., Fu, P., Ni, Z.L., Bian, G.B., 2022 · 2022
Closest in time.
Surgical workflow recognition: From analysis of challenges to architectural study, in: European Conference on Computer Vision, Springer. pp. 556–568
Czempiel, T., Sharghi, A., Paschali, M., Navab, N., Mohareri, O., 2022 · 2022
Closest in time.
Weakly-supervised online action segmentation in multi-view instructional videos, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 13780–13790
Ghoddoosian, R., Dwivedi, I., Agarwal, N., Choi, C., Dariush, B., 2022 · 2022
Closest in time.
An empirical study on activity recognition in long surgical videos, in: Machine Learning for Health, PMLR. pp. 356–372
He, Z., Mottaghi, A., Sharghi, A., Jamal, M.A., Mohareri, O., 2022 · 2022
Closest in time.
Patg: position-aware temporal graph networks for surgical phase recognition on laparoscopic videos
Kadkhodamohammadi, A., Luengo, I., Stoyanov, D., 2022 · 2022
Closest in time.
Surgical data science–from concepts toward clinical translation
Maier-Hein, L., Eisenmann, M., Sarikaya, D., März, K., Collins, T., Malpani, A., Fallert, J., Feussner, H., Giannarou, S., Mascagni, P., et al., 2022 · 2022
Closest in time.
Continual normalization: Rethinking batch normalization for online continual learning
Pham, Q., Liu, C., Hoi, S., 2022 · 2022
Closest in time.
Autolaparo: A new dataset of integrated multi-tasks for image-guided surgical automation in laparoscopic hysterectomy, in: International Conference on Medical Image Computing and Computer-Assisted Intervention, Springer. pp. 486–496
Wang, Z., Lu, B., Long, Y., Zhong, F., Cheung, T.H., Dou, Q., Liu, Y., 2022 · 2022
Closest in time.
Not end-to-end: Explore multi-stage architecture for online surgical phase recognition, in: Proceedings of the Asian Conference on Computer Vision, pp. 2613–2628
Yi, F., Yang, Y., Jiang, T., 2022 · 2022
Closest in time.
Large-scale surgical workflow segmentation for laparoscopic sacrocolpopexy
Zhang, Y., Bano, S., Page, A.S., Deprest, J., Stoyanov, D., Vasconcelos, F., 2022 · 2022
Closest in time.
Real-time online video detection with temporal smoothing transformers, in: European Conference on Computer Vision, Springer. pp. 485–502
Zhao, Y., Krähenbühl, P., 2022 · 2022
Closest in time.