Fetching the paper…
Reading the bibliography…
Capsule networks were proposed as an alternative approach to Convolutional Neural Networks (CNNs) for learning object-centric representations, which can be leveraged for improved generalization and sample complexity.
I. Rock, Orientation and form . Academic Press, 1973. [Online]. Available: https://books.google.co.uk/books?id=hgQEAQAAIAAJ
1973
Earlier work this paper cites.
A. P. Dempster, N. M. Laird, and D. B. Rubin, “Maximum likelihood from incomplete data via the em algorithm,” Journal of the Royal Statistical Society: Series B (Methodological) , vol. 39, no. 1, pp. 1–22, 1977
1977
Earlier work this paper cites.
G. Hinton, “Some demonstrations of the effects of structural descriptions in mental imagery,” Cognitive Science , vol. 3, no. 3, pp. 231–250, 1979
1979
Earlier work this paper cites.
M. A. Fischler and R. C. Bolles, “Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography,” Communications of the ACM , vol. 24, no. 6, pp. 381–395, 1981
1981
Earlier work this paper cites.
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel, “Backpropagation applied to handwritten zip code recognition,” Neural computation , vol. 1, no. 4, pp. 541–551, 1989
1989
Earlier work this paper cites.
Y. LeCun, B. Boser, J. Denker, D. Henderson, R. Howard, W. Hubbard, and L. Jackel, “Handwritten digit recognition with a back-propagation network,” Advances in neural information processing systems , vol. 2, 1989
1989
Earlier work this paper cites.
D. Kahneman, A. Treisman, and B. J. Gibbs, “The reviewing of object files: Object-specific integration of information,” Cognitive psychology , vol. 24, no. 2, pp. 175–219, 1992
1992
Earlier work this paper cites.
G. E. Hinton and D. Van Camp, “Keeping the neural networks simple by minimizing the description length of the weights,” in Proceedings of the sixth annual conference on Computational learning theory , 1993, pp. 5–13
1993
Earlier work this paper cites.
V. Bruce and G. W. Humphreys, “Recognizing objects and faces,” Visual cognition , vol. 1, no. 2-3, pp. 141–180, 1994
1994
Earlier work this paper cites.
M. I. Jordan, Z. Ghahramani, T. S. Jaakkola, and L. K. Saul, “An introduction to variational methods for graphical models,” Machine learning , vol. 37, no. 2, pp. 183–233, 1999
1999
Earlier work this paper cites.
H. Attias, “Inferring parameters and structure of latent variable models by variational bayes,” in Proceedings of the Fifteenth Conference on Uncertainty in Artificial Intelligence , ser. UAI’99. San Francisco, CA, USA: Morgan Kaufmann Publishers Inc., 1999, p. 21–30
1999
Earlier work this paper cites.
G. Schwarzer, “Development of face processing: The effect of face inversion,” Child development , vol. 71, no. 2, pp. 391–401, 2000
2000
Earlier work this paper cites.
A. M. Andrew, “Multiple view geometry in computer vision,” Kybernetes , 2001
2001
Earlier work this paper cites.
C. M. Bishop, Pattern recognition and machine learning . springer, 2006
2006
Earlier work this paper cites.
A. Krizhevsky, G. Hinton et al. , “Learning multiple layers of features from tiny images,” 2009
2009
Earlier work this paper cites.
G. E. Hinton, A. Krizhevsky, and S. D. Wang, “Transforming auto-encoders,” in International conference on artificial neural networks . Springer, 2011, pp. 44–51
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
K. P. Murphy, Machine learning: a probabilistic perspective . MIT press, 2012
2012
Earlier work this paper cites.
Y. Bengio, A. Courville, and P. Vincent, “Representation learning: A review and new perspectives,” IEEE transactions on pattern analysis and machine intelligence , vol. 35, no. 8, pp. 1798–1828, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
J. Feldman, “The neural binding problem (s),” Cognitive neurodynamics , vol. 7, no. 1, pp. 1–11, 2013
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep learning,” nature , vol. 521, no. 7553, pp. 436–444, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
M. Jaderberg, K. Simonyan, A. Zisserman et al. , “Spatial transformer networks,” Advances in neural information processing systems , vol. 28, pp. 2017–2025, 2015
2015
Earlier work this paper cites.
T. D. Kulkarni, W. F. Whitney, P. Kohli, and J. Tenenbaum, “Deep convolutional inverse graphics network,” in Advances in neural information processing systems , 2015, pp. 2539–2547
2015
Earlier work this paper cites.
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri, “Learning spatiotemporal features with 3d convolutional networks,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 4489–4497
2015
Earlier work this paper cites.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: Convolutional networks for biomedical image segmentation,” 2015
2015
Earlier work this paper cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 1–9
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
I. Goodfellow, Y. Bengio, and A. Courville, Deep learning . MIT press, 2016
2016
Earlier work this paper cites.
H. Pashler, Attention . Psychology Press, 2016
2016
Earlier work this paper cites.
T. Cohen and M. Welling, “Group equivariant convolutional networks,” in International conference on machine learning . PMLR, 2016, pp. 2990–2999
2016
Earlier work this paper cites.
S. Dieleman, J. De Fauw, and K. Kavukcuoglu, “Exploiting cyclic symmetry in convolutional neural networks,” in International conference on machine learning . PMLR, 2016, pp. 1889–1898
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
T. S. Cohen and M. Welling, “Steerable CNNs,” in 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Conference Track Proceedings . OpenReview.net, 2017. [Online]. Available: https://openreview.net/forum?id=rJQKYt5ll
2017
Earlier work this paper cites.
S. Sabour, N. Frosst, and G. E. Hinton, “Dynamic routing between capsules,” in Advances in neural information processing systems , 2017, pp. 3856–3866
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in neural information processing systems , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
D. E. Worrall, S. J. Garbin, D. Turmukhambetov, and G. J. Brostow, “Harmonic networks: Deep translation and rotation equivariance,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 5028–5037
2017
Earlier work this paper cites.
K. Greff, S. Van Steenkiste, and J. Schmidhuber, “Neural expectation maximization,” Advances in Neural Information Processing Systems , vol. 30, 2017
2017
Earlier work this paper cites.
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra, “Grad-cam: Visual explanations from deep networks via gradient-based localization,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 618–626
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. M. Bronstein, J. Bruna, Y. LeCun, A. Szlam, and P. Vandergheynst, “Geometric deep learning: going beyond euclidean data,” IEEE Signal Processing Magazine , vol. 34, no. 4, pp. 18–42, 2017
2017
Cited alongside, same era.
G. Hinton, S. Sabour, and N. Frosst, “Matrix capsules with em routing,” in 6th international conference on learning representations, ICLR , 2018, pp. 1–15
2018
Cited alongside, same era.
T. S. Cohen, M. Geiger, J. Köhler, and M. Welling, “Spherical CNNs,” in International Conference on Learning Representations , 2018. [Online]. Available: https://openreview.net/forum?id=Hkbd5xZRb
2018
Cited alongside, same era.
2018
Cited alongside, same era.
F. D. S. Ribeiro, G. Leontidis, and S. Kollias, “Capsule routing via variational bayes,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 04, 2020, pp. 3749–3756
2020
Later among the works it cites.
S. Hooker, “The hardware lottery,” arXiv preprint arXiv:2009.06489 , 2020
2020
Later among the works it cites.
R. Shi and L. Niu, “A brief survey on capsule network,” in 2020 IEEE/WIC/ACM International Joint Conference on Web Intelligence and Intelligent Agent Technology (WI-IAT) . IEEE, 2020, pp. 682–686
2020
Later among the works it cites.
D. Romero, E. Bekkers, J. Tomczak, and M. Hoogendoorn, “Attentive group equivariant convolutional networks,” in International Conference on Machine Learning . PMLR, 2020, pp. 8188–8199
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
K. Duarte, Y. S. Rawat, and M. Shah, “Videocapsulenet: A simplified network for action detection,” Advances in Neural Information Processing Systems , 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Z. Xinyi and L. Chen, “Capsule graph neural network,” in International conference on learning representations , 2018
2018
Cited alongside, same era.
J. E. Lenssen, M. Fey, and P. Libuschewski, “Group equivariant capsule networks,” Advances in Neural Information Processing Systems , vol. 31, 2018
2018
Cited alongside, same era.
A. Jaiswal, W. AbdAlmageed, Y. Wu, and P. Natarajan, “Capsulegan: Generative adversarial capsule network,” in Proceedings of the European conference on computer vision (ECCV) workshops , 2018, pp. 0–0
2018
Cited alongside, same era.
M. Yang, W. Zhao, J. Ye, Z. Lei, Z. Zhao, and S. Zhang, “Investigating capsule networks with dynamic routing for text classification,” in Proceedings of the 2018 conference on empirical methods in natural language processing , 2018, pp. 3110–3119
2018
Cited alongside, same era.
C. Xia, C. Zhang, X. Yan, Y. Chang, and P. S. Yu, “Zero-shot user intent detection via capsule neural networks,” Proceedings of the 2018 conference on empirical methods in natural language processing , 2018
2018
Cited alongside, same era.
F. De Sousa Ribeiro, G. Leontidis, and S. Kollias, “Introducing routing uncertainty in capsule networks,” in Advances in Neural Information Processing Systems , vol. 33, 2020, pp. 6490–6502
2020
Later among the works it cites.
Y. Qin, N. Frosst, S. Sabour, C. Raffel, G. Cottrell, and G. Hinton, “Detecting and diagnosing adversarial images with class-conditional capsule reconstructions,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=Skgy464Kvr
2020
Later among the works it cites.
F. Locatello, D. Weissenborn, T. Unterthiner, A. Mahendran, G. Heigold, J. Uszkoreit, A. Dosovitskiy, and T. Kipf, “Object-centric learning with slot attention,” Advances in Neural Information Processing Systems , vol. 33, pp. 11 525–11 538, 2020
2020
Later among the works it cites.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly et al. , “An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale,” in International Conference on Learning Representations , 2020
2020
Later among the works it cites.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in European conference on computer vision . Springer, 2020, pp. 213–229
2020
Later among the works it cites.
B. McIntosh, K. Duarte, Y. S. Rawat, and M. Shah, “Visual-textual capsule routing for text-based video segmentation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 9942–9951
2020
Later among the works it cites.
J. Gu and V. Tresp, “Interpretable graph capsule networks for object recognition,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , 2020
2020
Later among the works it cites.
Y. Zhao, T. Birdal, J. E. Lenssen, E. Menegatti, L. Guibas, and F. Tombari, “Quaternion equivariant capsule networks for 3d point clouds,” in European Conference on Computer Vision . Springer, 2020, pp. 1–19
2020
Later among the works it cites.
X. Wen, Z. Han, X. Liu, and Y.-S. Liu, “Point2spatialcapsule: Aggregating features and spatial relationships of local regions on point clouds using spatial-aware capsules,” IEEE Transactions on Image Processing , vol. 29, pp. 8855–8869, 2020
2020
Later among the works it cites.
M. Edraki, N. Rahnavard, and M. Shah, “Subspace capsule network,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 07, 2020, pp. 10 745–10 753
2020
Later among the works it cites.
D. Jung, J. Lee, J. Yi, and S. Yoon, “icaps: An interpretable classifier via disentangled capsule networks,” in European Conference on Computer Vision . Springer, 2020, pp. 314–330
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Kim, S. Jang, E. Park, and S. Choi, “Text classification using capsules,” Neurocomputing , vol. 376, pp. 214–221, 2020
2020
Later among the works it cites.
H. Lin, F. Meng, J. Su, Y. Yin, Z. Yang, Y. Ge, J. Zhou, and J. Luo, “Dynamic context-guided capsule network for multimodal machine translation,” in Proceedings of the 28th ACM International Conference on Multimedia , 2020, pp. 1320–1329
2020
Later among the works it cites.
T. Wang, A. Bezerianos, A. Cichocki, and J. Li, “Multikernel capsule network for schizophrenia identification,” IEEE transactions on Cybernetics , 2020
2020
Later among the works it cites.
P. Afshar, S. Heidarian, F. Naderkhani, A. Oikonomou, K. N. Plataniotis, and A. Mohammadi, “Covid-caps: A capsule network-based framework for identification of covid-19 cases from x-ray images,” Pattern Recognition Letters , vol. 138, pp. 638–643, 2020. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0167865520303512
2020
Later among the works it cites.
Y. Bengio, Y. Lecun, and G. Hinton, “Deep learning for ai,” Communications of the ACM , vol. 64, no. 7, pp. 58–65, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
Y. Tay, D. Bahri, D. Metzler, D.-C. Juan, Z. Zhao, and C. Zheng, “Synthesizer: Rethinking self-attention for transformer models,” in International Conference on Machine Learning . PMLR, 2021, pp. 10 183–10 192
2021
Later among the works it cites.
L. Li, B. Wang, M. Verma, Y. Nakashima, R. Kawasaki, and H. Nagahara, “SCOUTER: Slot attention-based classifier for explainable image recognition,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 1046–1055
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
A. Roy, M. Saffar, A. Vaswani, and D. Grangier, “Efficient content-based sparse attention with routing transformers,” Transactions of the Association for Computational Linguistics , vol. 9, pp. 53–68, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
Z. Xu, X. Shen, Y. Wong, and M. Kankanhalli, “Unsupervised motion representation learning with capsule autoencoders,” in Thirty-Fifth Conference on Neural Information Processing Systems , 2021
2021
Later among the works it cites.
D. Ma and X. Wu, “Capsulerrt: Relationships-aware regression tracking via capsules,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 10 948–10 957
2021
Later among the works it cites.
Y. Li, W. Zhao, E. Cambria, S. Wang, and S. Eger, “Graph routing between capsules,” Neural Networks , vol. 143, pp. 345–354, 2021
2021
Later among the works it cites.
W. Sun, A. Tagliasacchi, B. Deng, S. Sabour, S. Yazdani, G. E. Hinton, and K. M. Yi, “Canonical Capsules: Self-Supervised Capsules in Canonical Pose,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Later among the works it cites.
T. Keller and M. Welling, “Topographic vaes learn equivariant capsules,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Later among the works it cites.
A. Urooj, H. Kuehne, K. Duarte, C. Gan, N. Lobo, and M. Shah, “Found a reason for me? weakly-supervised grounded visual question answering using capsules,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 8465–8474
2021
Later among the works it cites.
Q. Cao, W. Wan, K. Wang, X. Liang, and L. Lin, “Linguistically routing capsule network for out-of-distribution visual question answering,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 1614–1623
2021
Later among the works it cites.
R. LaLonde, Z. Xu, I. Irmakci, S. Jain, and U. Bagci, “Capsules for biomedical image segmentation,” Medical image analysis , vol. 68, p. 101889, 2021
2021
Later among the works it cites.
F. Li, X. Lu, and J. Yuan, “Mha-corocapsule: Multi-head attention routing-based capsule network for covid-19 chest x-ray image classification,” IEEE Transactions on Medical Imaging , 2021
2021
Later among the works it cites.
A. Luo, E. Li, Y. Liu, X. Kang, and Z. J. Wang, “A capsule network based approach for detection of audio spoofing attacks,” in ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2021, pp. 6359–6363
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
A. Ramesh, M. Pavlov, G. Goh, S. Gray, C. Voss, A. Radford, M. Chen, and I. Sutskever, “Zero-shot text-to-image generation,” in International Conference on Machine Learning . PMLR, 2021, pp. 8821–8831
2021
Later among the works it cites.
S. Sabour, A. Tagliasacchi, S. Yazdani, G. Hinton, and D. J. Fleet, “Unsupervised part representation by flow capsules,” in International Conference on Machine Learning . PMLR, 2021, pp. 9213–9223
2021
Later among the works it cites.
2021
Later among the works it cites.
M. K. Patrick, A. F. Adekoya, A. A. Mighty, and B. Y. Edward, “Capsule networks–a survey,” Journal of King Saud University-computer and information sciences , vol. 34, no. 1, pp. 1295–1310, 2022
2022
Closest in time.