Fetching the paper…
Reading the bibliography…
We introduce a novel cost aggregation network, dubbed Volumetric Aggregation with Transformers (VAT), to tackle the few-shot segmentation task by using both convolutions and transformers to efficiently handle high dimensional correlation maps between query and support.
Neural network studies, 1. comparison of overfitting and overtraining
Igor V. Tetko, David J. Livingstone, and Alexander I. Luik · 1995
Earlier work this paper cites.
A taxonomy and evaluation of dense two-frame stereo correspondence algorithms
Daniel Scharstein and Richard Szeliski · 2002
Earlier work this paper cites.
Object retrieval with large vocabularies and fast spatial matching
James Philbin, Ondrej Chum, Michael Isard, Josef Sivic, and Andrew Zisserman · 2007
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
Mark Everingham, Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman · 2010
Earlier work this paper cites.
Fast cost-volume filtering for visual correspondence and beyond
Asmaa Hosni, Christoph Rhemann, Michael Bleyer, Carsten Rother, and Margrit Gelautz · 2012
Earlier work this paper cites.
Simultaneous detection and segmentation
Bharath Hariharan, Pablo Arbeláez, Ross Girshick, and Jitendra Malik · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation
Jonathan Long, Evan Shelhamer, and Trevor Darrell · 2015
Earlier work this paper cites.
Learning deconvolution network for semantic segmentation
Hyeonwoo Noh, Seunghoon Hong, and Bohyung Han · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton · 2016
Earlier work this paper cites.
Proposal flow
Bumsub Ham, Minsu Cho, Cordelia Schmid, and Jean Ponce · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Optimization as a model for few-shot learning
Sachin Ravi and Hugo Larochelle · 2016
Earlier work this paper cites.
Matching networks for one shot learning
Oriol Vinyals, Charles Blundell, Timothy Lillicrap, Daan Wierstra, et al · 2016
Earlier work this paper cites.
Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L Yuille · 2017
Earlier work this paper cites.
Proposal flow: Semantic correspondences from object proposals
Bumsub Ham, Minsu Cho, Cordelia Schmid, and Jean Ponce · 2017
Earlier work this paper cites.
Fcss: Fully convolutional self-similarity for dense semantic correspondence
Seungryong Kim, Dongbo Min, Bumsub Ham, Sangryul Jeon, Stephen Lin, and Kwanghoon Sohn · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Earlier work this paper cites.
Convolutional neural network architecture for geometric matching
Ignacio Rocco, Relja Arandjelovic, and Josef Sivic · 2017
Earlier work this paper cites.
One-shot learning for semantic segmentation
Amirreza Shaban, Shray Bansal, Zhen Liu, Irfan Essa, and Byron Boots · 2017
Earlier work this paper cites.
Prototypical networks for few-shot learning
Jake Snell, Kevin Swersky, and Richard S Zemel · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Encoder-decoder with atrous separable convolution for semantic image segmentation
Liang-Chieh Chen, Yukun Zhu, George Papandreou, Florian Schroff, and Hartwig Adam · 2018
Earlier work this paper cites.
Few-shot semantic segmentation with prototype learning
Nanqing Dong and Eric P Xing · 2018
Earlier work this paper cites.
Relation networks for object detection
Han Hu, Jiayuan Gu, Zheng Zhang, Jifeng Dai, and Yichen Wei · 2018
Earlier work this paper cites.
End-to-end weakly-supervised semantic alignment
Ignacio Rocco, Relja Arandjelović, and Josef Sivic · 2018
Cited alongside, same era.
Neighbourhood consensus networks
Ignacio Rocco, Mircea Cimpoi, Relja Arandjelović, Akihiko Torii, Tomas Pajdla, and Josef Sivic · 2018
Cited alongside, same era.
Pwc-net: Cnns for optical flow using pyramid, warping, and cost volume
Deqing Sun, Xiaodong Yang, Ming-Yu Liu, and Jan Kautz · 2018
Cited alongside, same era.
Local relation networks for image recognition
Han Hu, Zheng Zhang, Zhenda Xie, and Stephen Lin · 2019
Cited alongside, same era.
Dynamic context correspondence network for semantic alignment
Shuaiyi Huang, Qiuyue Wang, Songyang Zhang, Shipeng Yan, and Xuming He · 2019
Cited alongside, same era.
Sfnet: Learning object-aware semantic correspondence
Few-shot semantic segmentation with democratic attention networks
Haochen Wang, Xudong Zhang, Yutao Hu, Yandan Yang, Xianbin Cao, and Xiantong Zhen · 2020
Later among the works it cites.
Linformer: Self-attention with linear complexity
Sinong Wang, Belinda Z Li, Madian Khabsa, Han Fang, and Hao Ma · 2020
Later among the works it cites.
Prototype mixture models for few-shot semantic segmentation
Boyu Yang, Chang Liu, Bohao Li, Jianbin Jiao, and Qixiang Ye · 2020
Later among the works it cites.
Few-shot segmentation without meta-learning: A good transductive inference is all you need?
Malik Boudiaf, Hoel Kervadec, Ziko Imtiaz Masud, Pablo Piantanida, Ismail Ben Ayed, and Jose Dolz · 2021
Closest in time.
Cats: Cost aggregation transformers for visual correspondence
Seokju Cho, Sunghwan Hong, Sangryul Jeon, Yunsung Lee, Kwanghoon Sohn, and Seungryong Kim · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Junghyup Lee, Dohyung Kim, Jean Ponce, and Bumsub Ham · 2019
Cited alongside, same era.
Hyperpixel flow: Semantic correspondence with multi-layer neural features
Juhong Min, Jongmin Lee, Jean Ponce, and Minsu Cho · 2019
Cited alongside, same era.
Spair-71k: A large-scale benchmark for semantic correspondence
Juhong Min, Jongmin Lee, Jean Ponce, and Minsu Cho · 2019
Cited alongside, same era.
Feature weighting and boosting for few-shot segmentation
Khoi Nguyen and Sinisa Todorovic · 2019
Cited alongside, same era.
Adaptive masked proxies for few-shot segmentation
Mennatullah Siam, Boris Oreshkin, and Martin Jagersand · 2019
Cited alongside, same era.
Panet: Few-shot image semantic segmentation with prototype alignment
Kaixin Wang, Jun Hao Liew, Yingtian Zou, Daquan Zhou, and Jiashi Feng · 2019
Cited alongside, same era.
Pyramid graph networks with connection attentions for region-based one-shot semantic segmentation
Chi Zhang, Guosheng Lin, Fayao Liu, Jiushuang Guo, Qingyao Wu, and Rui Yao · 2019
Cited alongside, same era.
Deep matching prior: Test-time optimization for dense correspondence
Sunghwan Hong and Seungryong Kim · 2021
Closest in time.
Adaptive prototype learning and allocation for few-shot segmentation
Gen Li, Varun Jampani, Laura Sevilla-Lara, Deqing Sun, Jonghyun Kim, and Joongkyu Kim · 2021
Closest in time.
Few-shot segmentation with optimal transport matching and message flow
Weide Liu, Chi Zhang, Henghui Ding, Tzu-Yi Hung, and Guosheng Lin · 2021
Closest in time.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Closest in time.
Soft: Softmax-free transformer with linear complexity
Jiachen Lu, Jinghan Yao, Junge Zhang, Xiatian Zhu, Hang Xu, Weiguo Gao, Chunjing Xu, Tao Xiang, and Li Zhang · 2021
Closest in time.
Simpler is better: Few-shot semantic segmentation with classifier weight transformer
Zhihe Lu, Sen He, Xiatian Zhu, Li Zhang, Yi-Zhe Song, and Tao Xiang · 2021
Closest in time.
Hypercorrelation squeeze for few-shot segmentation
Juhong Min, Dahyun Kang, and Minsu Cho · 2021
Closest in time.
Convolutional hough matching networks for robust and efficient visual correspondence
Juhong Min, Seungwook Kim, and Minsu Cho · 2021
Closest in time.
Do vision transformers see like convolutional neural networks?
Maithra Raghu, Thomas Unterthiner, Simon Kornblith, Chiyuan Zhang, and Alexey Dosovitskiy · 2021
Closest in time.
Boosting few-shot semantic segmentation with transformers
Guolei Sun, Yun Liu, Jingyun Liang, and Luc Van Gool · 2021
Closest in time.
Learning accurate dense correspondences and when to trust them
Prune Truong, Martin Danelljan, Luc Van Gool, and Radu Timofte · 2021
Closest in time.
Pyramid vision transformer: A versatile backbone for dense prediction without convolutions
Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao · 2021
Closest in time.
Fastformer: Additive attention can be all you need
Chuhan Wu, Fangzhao Wu, Tao Qi, Yongfeng Huang, and Xing Xie · 2021
Closest in time.
Learning meta-class memory for few-shot semantic segmentation
Zhonghua Wu, Xiangxi Shi, Guosheng Lin, and Jianfei Cai · 2021
Closest in time.
Early convolutions help transformers see better
Tete Xiao, Mannat Singh, Eric Mintun, Trevor Darrell, Piotr Dollár, and Ross Girshick · 2021
Closest in time.
Scale-aware graph neural network for few-shot semantic segmentation
Guo-Sen Xie, Jie Liu, Huan Xiong, and Ling Shao · 2021
Closest in time.
Few-shot semantic segmentation with cyclic memory network
Guo-Sen Xie, Huan Xiong, Jie Liu, Yazhou Yao, and Ling Shao · 2021
Closest in time.
Mining latent classes for few-shot segmentation
Lihe Yang, Wei Zhuo, Lei Qi, Yinghuan Shi, and Yang Gao · 2021
Closest in time.
Self-guided and cross-guided learning for few-shot segmentation
Bingfeng Zhang, Jimin Xiao, and Terry Qin · 2021
Closest in time.
Few-shot segmentation via cycle-consistent transformer
Gengwei Zhang, Guoliang Kang, Yunchao Wei, and Yi Yang · 2021
Closest in time.
Prototypical matching and open set rejection for zero-shot semantic segmentation
Hui Zhang and Henghui Ding · 2021
Closest in time.
Multi-scale matching networks for semantic correspondence
Dongyang Zhao, Ziyang Song, Zhenghao Ji, Gangming Zhao, Weifeng Ge, and Yizhou Yu · 2021
Closest in time.