Gradient-centralization: A new optimization technique for deep neural networks
Hongwei Yong, Jianqiang Huang, Xiansheng Hua, and Lei Zhang · 2020
Later among the works it cites.
A survey of transformers
Original
Tianyang Lin, Yuxin Wang, Xiangyang Liu, and Xipeng Qiu · 2021
Later among the works it cites.
Metaformer is actually what you need for vision
Original
Weihao Yu, Mi Luo, Pan Zhou, Chenyang Si, Yichen Zhou, Xinchao Wang, Jiashi Feng, and Shuicheng Yan · 2021
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Later among the works it cites.
Large associative memory problem in neurobiology and machine learning
Dmitry Krotov and John J. Hopfield · 2021
Later among the works it cites.
A remark on a paper of krotov and hopfield [arxiv: 2008.06996]
Original
Fei Tang and Michael Kopp · 2021
Later among the works it cites.
Hierarchical associative memory
Original
Dmitry Krotov · 2021
Later among the works it cites.
Pick and choose: a gnn-based imbalanced learning approach for fraud detection
Yang Liu, Xiang Ao, Zidi Qin, Jianfeng Chi, Jinghua Feng, Hao Yang, and Qing He · 2021
Later among the works it cites.
Hierarchical multi-view graph pooling with structure learning
Zhen Zhang, Jiajun Bu, Martin Ester, Jianfeng Zhang, Zhao Li, Chengwei Yao, Dai Huifen, Zhi Yu, and Can Wang · 2021
Later among the works it cites.
Inductive anomaly detection on attributed networks
Kaize Ding, Jundong Li, Nitin Agarwal, and Huan Liu · 2021
Later among the works it cites.
Do transformers really perform bad for graph representation?
Original
Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng, Guolin Ke, Di He, Yanming Shen, and Tie-Yan Liu · 2021
Later among the works it cites.
Graphformers: Gnn-nested transformers for representation learning on textual graph
Junhan Yang, Zheng Liu, Shitao Xiao, Chaozhuo Li, Defu Lian, Sanjay Agrawal, Amit Singh, Guangzhong Sun, and Xing Xie · 2021
Later among the works it cites.
Transformers from an optimization perspective
Original
Yongyi Yang, Zengfeng Huang, and David Wipf · 2022
Later among the works it cites.
Rethinking graph neural networks for anomaly detection
Original
Jianheng Tang, Jiajin Li, Ziqi Gao, and Jia Li · 2022
Later among the works it cites.
Benchmarking graph neural networks, 2022
Vijay Prakash Dwivedi, Chaitanya K. Joshi, Anh Tuan Luu, Thomas Laurent, Yoshua Bengio, and Xavier Bresson · 2022
Later among the works it cites.
Universal graph transformer self-attention networks
Dai Quoc Nguyen, Tu Dinh Nguyen, and Dinh Phung · 2022
Later among the works it cites.
A new perspective on the effects of spectrum in graph neural networks
Mingqi Yang, Yanming Shen, Rui Li, Heng Qi, Qiang Zhang, and Baocai Yin · 2022
Later among the works it cites.
Modern hopfield networks for graph embedding
Yuchen Liang, Dmitry Krotov, and Mohammed J Zaki · 2022
Later among the works it cites.
Global self-attention as a replacement for graph convolution
Md Shamim Hussain, Mohammed J. Zaki, and Dharmashankar Subramanian · 2022
Later among the works it cites.
A new frontier for hopfield networks
Dmitry Krotov · 2023
Closest in time.
Exphormer: Sparse transformers for graphs
Original
Hamed Shirzad, Ameya Velingker, Balaji Venkatachalam, Danica J Sutherland, and Ali Kemal Sinop · 2023
Closest in time.