Self-attention Generative Adversarial Networks
Zhang, H., Goodfellow, I., Metaxas, D., and Odena, A · 2019
Later among the works it cites.
B-spline CNNs on Lie groups
Bekkers, E. J · 2020
Closest in time.
Probabilistic symmetries and invariant neural networks
Bloem-Reddy, B. and Teh, Y. W · 2020
Closest in time.
Language models are few-shot learners
Original
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Closest in time.
An image is worth 16x16 words: Transformers for image recognition at scale
Original
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., et al · 2020
Closest in time.
Spin-weighted spherical CNNs
Esteves, C., Makadia, A., and Daniilidis, K · 2020
Closest in time.
Generalizing convolutional neural networks for equivariance to Lie groups on arbitrary continuous data
Finzi, M., Stanton, S., Izmailov, P., and Wilson, A. G · 2020
Closest in time.
SE(3)-Transformers: 3D roto-translation equivariant attention networks
Fuchs, F. B., Worrall, D. E., Fischer, V., and Welling, M · 2020
Closest in time.
Array programming with numpy
Harris, C. R., Millman, K. J., van der Walt, S. J., Gommers, R., Virtanen, P., Cournapeau, D., Wieser, E., Taylor, J., Berg, S., Smith, N. J., et al · 2020
Closest in time.
Transformers are rnns: Fast autoregressive transformers with linear attention
Katharopoulos, A., Vyas, A., Pappas, N., and Fleuret, F · 2020
Closest in time.
Reformer: The efficient transformer
Kitaev, N., Kaiser, Ł., and Levskaya, A · 2020
Closest in time.
Fast and uncertainty-aware directional message passing for non-equilibrium molecules
Original
Klicpera, J., Giri, S., Margraf, J. T., and Günnemann, S · 2020
Closest in time.
Relevance of rotationally equivariant convolutions for predicting molecular properties
Original
Miller, B. K., Geiger, M., Smidt, T. E., and Noé, F · 2020
Closest in time.
Stabilizing Transformers for reinforcement learning
Parisotto, E., Song, H. F., Rae, J. W., Pascanu, R., Gulcehre, C., Jayakumar, S. M., Jaderberg, M., Kaufman, R. L., Clark, A., Noury, S., Botvinick, M. M., Heess, N., and Hadsell, R · 2020
Closest in time.
Co-attentive equivariant neural networks: Focusing equivariance on transformations co-occurring in data
Romero, D. W. and Hoogendoorn, M · 2020
Closest in time.
Attentive group equivariant convolutional networks
Romero, D. W., Bekkers, E. J., Tomczak, J. M., and Hoogendoorn, M · 2020
Closest in time.
Linformer: Self-attention with linear complexity
Original
Wang, S., Li, B., Khabsa, M., Fang, H., and Ma, H · 2020
Closest in time.
Big bird: Transformers for longer sequences
Zaheer, M., Guruganesh, G., Dubey, K. A., Ainslie, J., Alberti, C., Ontanon, S., Pham, P., Ravula, A., Wang, Q., Yang, L., et al · 2020
Closest in time.
Symplectic ode-net: Learning hamiltonian dynamics with control
Zhong, Y. D., Dey, B., and Chakraborty, A · 2020
Closest in time.
Group equivariant stand-alone self-attention for vision
Romero, D. W. and Cordonnier, J.-B · 2021
Closest in time.