Fetching the paper…
Reading the bibliography…
Convolution-based and Transformer-based vision backbone networks process images into the grid or sequence structures, respectively, which are inflexible for capturing irregular objects.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner. 1998 · 1998
Earlier work this paper cites.
Digital selection and analogue amplification coexist in a cortex-inspired silicon circuit
Richard HR Hahnloser, Rahul Sarpeshkar, Misha A Mahowald, Rodney J Douglas, and H Sebastian Seung. 2000 · 2000
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database. In 2009 IEEE conference on computer vision and pattern recognition . Ieee, 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
Spectral networks and locally connected networks on graphs
Joan Bruna, Wojciech Zaremba, Arthur Szlam, and Yann LeCun. 2013 · 2013
Earlier work this paper cites.
Microsoft coco: Common objects in context. In Computer Vision–ECCV 2014: 13th European Conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part V 13 . Springer, 740–755
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman. 2014 · 2014
Earlier work this paper cites.
Deep convolutional networks on graph-structured data
Mikael Henaff, Joan Bruna, and Yann LeCun. 2015 · 2015
Earlier work this paper cites.
Going deeper with convolutions. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1–9
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich. 2015 · 2015
Earlier work this paper cites.
Diffusion-convolutional neural networks
James Atwood and Don Towsley. 2016 · 2016
Earlier work this paper cites.
Convolutional neural networks on graphs with fast localized spectral filtering
Michaël Defferrard, Xavier Bresson, and Pierre Vandergheynst. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Bridging nonlinearities and stochastic regularizers with gaussian error linear units
Dan Hendrycks and Kevin Gimpel. 2016 · 2016
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Thomas N Kipf and Max Welling. 2016 · 2016
Earlier work this paper cites.
Structural deep network embedding. In Proceedings of the 22nd ACM SIGKDD international conference on Knowledge discovery and data mining . 1225–1234
Daixin Wang, Peng Cui, and Wenwu Zhu. 2016 · 2016
Earlier work this paper cites.
Inductive representation learning on large graphs
Will Hamilton, Zhitao Ying, and Jure Leskovec. 2017 · 2017
Earlier work this paper cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Andrew G Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam. 2017 · 2017
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton. 2017 · 2017
Earlier work this paper cites.
Geometric deep learning on graphs and manifolds using mixture model cnns. In Proceedings of the IEEE conference on computer vision and pattern recognition . 5115–5124
Federico Monti, Davide Boscaini, Jonathan Masci, Emanuele Rodola, Jan Svoboda, and Michael M Bronstein. 2017 · 2017
Earlier work this paper cites.
Petar Veličković, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Lio, and Yoshua Bengio. 2017 · 2017
Earlier work this paper cites.
Adaptive sampling towards fast graph representation learning
Wenbing Huang, Tong Zhang, Yu Rong, and Junzhou Huang. 2018 · 2018
Earlier work this paper cites.
Cayleynets: Graph convolutional neural networks with complex rational spectral filters
Ron Levie, Federico Monti, Xavier Bresson, and Michael M Bronstein. 2018 · 2018
Earlier work this paper cites.
Representation learning on graphs with jumping knowledge networks. In International conference on machine learning . PMLR, 5453–5462
Keyulu Xu, Chengtao Li, Yonglong Tian, Tomohiro Sonobe, Ken-ichi Kawarabayashi, and Stefanie Jegelka. 2018 · 2018
Earlier work this paper cites.
Mixhop: Higher-order graph convolutional architectures via sparsified neighborhood mixing. In international conference on machine learning . PMLR, 21–29
Sami Abu-El-Haija, Bryan Perozzi, Amol Kapoor, Nazanin Alipourfard, Kristina Lerman, Hrayr Harutyunyan, Greg Ver Steeg, and Aram Galstyan. 2019 · 2019
Earlier work this paper cites.
Yu Chen, Lingfei Wu, and Mohammed J Zaki. 2019b · 2019
Earlier work this paper cites.
Reinforcement learning based graph-to-sequence model for natural question generation
Yu Chen, Lingfei Wu, and Mohammed J Zaki. 2019c · 2019
Cited alongside, same era.
Diffusion improves graph learning
Johannes Gasteiger, Stefan Weißenberger, and Stephan Günnemann. 2019 · 2019
Cited alongside, same era.
Rethinking knowledge graph propagation for zero-shot learning. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 11487–11496
Michael Kampffmeyer, Yinbo Chen, Xiaodan Liang, Hao Wang, Yujia Zhang, and Eric P Xing. 2019 · 2019
Cited alongside, same era.
Gcan: Graph convolutional adversarial network for unsupervised domain adaptation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 8266–8276
Xinhong Ma, Tianzhu Zhang, and Changsheng Xu. 2019 · 2019
Cited alongside, same era.
Dropedge: Towards deep graph convolutional networks on node classification
Training data-efficient image transformers & distillation through attention. In International conference on machine learning . PMLR, 10347–10357
Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Hervé Jégou. 2021 · 2021
Later among the works it cites.
Pyramid vision transformer: A versatile backbone for dense prediction without convolutions. In Proceedings of the IEEE/CVF international conference on computer vision . 568–578
Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao. 2021 · 2021
Later among the works it cites.
Resnet strikes back: An improved training procedure in timm
Ross Wightman, Hugo Touvron, and Hervé Jégou. 2021 · 2021
Later among the works it cites.
Cvt: Introducing convolutions to vision transformers. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 22–31
Haiping Wu, Bin Xiao, Noel Codella, Mengchen Liu, Xiyang Dai, Lu Yuan, and Lei Zhang. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yu Rong, Wenbing Huang, Tingyang Xu, and Junzhou Huang. 2019 · 2019
Cited alongside, same era.
Dynamic graph cnn for learning on point clouds
Yue Wang, Yongbin Sun, Ziwei Liu, Sanjay E Sarma, Michael M Bronstein, and Justin M Solomon. 2019 · 2019
Cited alongside, same era.
Pairnorm: Tackling oversmoothing in gnns
Lingxiao Zhao and Leman Akoglu. 2019 · 2019
Cited alongside, same era.
Iterative deep graph learning for graph neural networks: Better and robust node embeddings
Yu Chen, Lingfei Wu, and Mohammed Zaki. 2020 · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
Class-wise dynamic graph convolution for semantic segmentation. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XVII 16 . Springer, 1–17
Hanzhe Hu, Deyi Ji, Weihao Gan, Shuai Bai, Wei Wu, and Junjie Yan. 2020 · 2020
Cited alongside, same era.
Progressive graph learning for open-set domain adaptation. In International Conference on Machine Learning . PMLR, 6468–6478
Yadan Luo, Zijian Wang, Zi Huang, and Mahsa Baktashmotlagh. 2020 · 2020
Cited alongside, same era.
Amr-to-text generation with graph transformer
Tianming Wang, Xiaojun Wan, and Hanqi Jin. 2020 · 2020
Cited alongside, same era.
Vitae: Vision transformer advanced by exploring intrinsic inductive bias
Yufei Xu, Qiming Zhang, Jing Zhang, and Dacheng Tao. 2021 · 2021
Later among the works it cites.
Focal self-attention for local-global interactions in vision transformers
Jianwei Yang, Chunyuan Li, Pengchuan Zhang, Xiyang Dai, Bin Xiao, Lu Yuan, and Jianfeng Gao. 2021 · 2021
Later among the works it cites.
Tokens-to-token vit: Training vision transformers from scratch on imagenet. In Proceedings of the IEEE/CVF international conference on computer vision . 558–567
Li Yuan, Yunpeng Chen, Tao Wang, Weihao Yu, Yujun Shi, Zi-Hang Jiang, Francis EH Tay, Jiashi Feng, and Shuicheng Yan. 2021 · 2021
Later among the works it cites.
Analogous to evolutionary algorithm: Designing a unified sequence model
Jiangning Zhang, Chao Xu, Jian Li, Wenzhou Chen, Yabiao Wang, Ying Tai, Shuo Chen, Chengjie Wang, Feiyue Huang, and Yong Liu. 2021 · 2021
Later among the works it cites.
GraphFPN: Graph feature pyramid network for object detection. In Proceedings of the IEEE/CVF international conference on computer vision . 2763–2772
Gangming Zhao, Weifeng Ge, and Yizhou Yu. 2021 · 2021
Later among the works it cites.
Gradinit: Learning to initialize neural networks for stable and efficient training
Chen Zhu, Renkun Ni, Zheng Xu, Kezhi Kong, W Ronny Huang, and Tom Goldstein. 2021 · 2021
Later among the works it cites.
Simple spectral graph convolution. In International conference on learning representations
Hao Zhu and Piotr Koniusz. 2021 · 2021
Later among the works it cites.
Compound domain generalization via meta-knowledge encoding. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 7119–7129
Chaoqi Chen, Jiongcheng Li, Xiaoguang Han, Xiaoqing Liu, and Yizhou Yu. 2022 · 2022
Later among the works it cites.
Scaling up your kernels to 31x31: Revisiting large kernel design in cnns. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 11963–11975
Xiaohan Ding, Xiangyu Zhang, Jungong Han, and Guiguang Ding. 2022 · 2022
Later among the works it cites.
Cswin transformer: A general vision transformer backbone with cross-shaped windows. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 12124–12134
Xiaoyi Dong, Jianmin Bao, Dongdong Chen, Weiming Zhang, Nenghai Yu, Lu Yuan, Dong Chen, and Baining Guo. 2022 · 2022
Later among the works it cites.
Vision gnn: An image is worth graph of nodes
Kai Han, Yunhe Wang, Jianyuan Guo, Yehui Tang, and Enhua Wu. 2022 · 2022
Later among the works it cites.
Uniformer: Unified transformer for efficient spatiotemporal representation learning
Kunchang Li, Yali Wang, Peng Gao, Guanglu Song, Yu Liu, Hongsheng Li, and Yu Qiao. 2022b · 2022
Later among the works it cites.
SFNet: Faster, Accurate, and Domain Agnostic Semantic Segmentation via Semantic Flow
Xiangtai Li, Jiangning Zhang, Yibo Yang, Guangliang Cheng, Kuiyuan Yang, Yu Tong, and Dacheng Tao. 2022c · 2022
Later among the works it cites.
Eatformer: Improving vision transformer inspired by evolutionary algorithm
Jiangning Zhang, Xiangtai Li, Yabiao Wang, Chengjie Wang, Yibo Yang, Yong Liu, and Dacheng Tao. 2022 · 2022
Later among the works it cites.
Reference Twice: A Simple and Unified Baseline for Few-Shot Instance Segmentation
Yue Han, Jiangning Zhang, Zhucun Xue, Chao Xu, Xintian Shen, Yabiao Wang, Chengjie Wang, Yong Liu, and Xiangtai Li. 2023 · 2023
Closest in time.
Learning convolutional neural networks for graphs. In International conference on machine learning . PMLR, 2014–2023
Mathias Niepert, Mohamed Ahmed, and Konstantin Kutzkov. 2016 · 2023
Closest in time.
CrossFormer++: A Versatile Vision Transformer Hinging on Cross-scale Attention
Wenxiao Wang, Wei Chen, Qibo Qiu, Long Chen, Boxi Wu, Binbin Lin, Xiaofei He, and Wei Liu. 2023 · 2023
Closest in time.
Towards Open Vocabulary Learning: A Survey
Jianzong Wu, Xiangtai Li, Shilin Xu Haobo Yuan, Henghui Ding, Yibo Yang, Xia Li, Jiangning Zhang, Yunhai Tong, Xudong Jiang, Bernard Ghanem, et al · 2023
Closest in time.
Rethinking Mobile Block for Efficient Neural Models
Jiangning Zhang, Xiangtai Li, Jian Li, Liang Liu, Zhucun Xue, Boshen Zhang, Zhengkai Jiang, Tianxin Huang, Yabiao Wang, and Chengjie Wang. 2023 · 2023
Closest in time.