Fetching the paper…
Reading the bibliography…
End-to-end paradigms significantly improve the accuracy of various deep-learning-based computer vision models.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Simultaneous detection and segmentation
Bharath Hariharan, Pablo Arbeláez, Ross Girshick, and Jitendra Malik · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
V-net: Fully convolutional neural networks for volumetric medical image segmentation
Fausto Milletari, Nassir Navab, and Seyed-Ahmad Ahmadi · 2016
Earlier work this paper cites.
Recurrent instance segmentation
Bernardino Romera-Paredes and Philip Hilaire Sean Torr · 2016
Earlier work this paper cites.
Semantic instance segmentation with a discriminative loss function
Bert De Brabandere, Davy Neven, and Luc Van Gool · 2017
Earlier work this paper cites.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick · 2017
Earlier work this paper cites.
Fully convolutional instance-aware semantic segmentation
Yi Li, Haozhi Qi, Jifeng Dai, Xiangyang Ji, and Yichen Wei · 2017
Earlier work this paper cites.
Feature pyramid networks for object detection
Tsung-Yi Lin, Piotr Dollár, Ross Girshick, Kaiming He, Bharath Hariharan, and Serge Belongie · 2017
Earlier work this paper cites.
Focal loss for dense object detection
Tsung-Yi Lin, Priya Goyal, Ross Girshick, Kaiming He, and Piotr Dollár · 2017
Earlier work this paper cites.
Sgn: Sequential grouping networks for instance segmentation
Shu Liu, Jiaya Jia, Sanja Fidler, and Raquel Urtasun · 2017
Earlier work this paper cites.
End-to-end instance segmentation with recurrent attention
Mengye Ren and Richard S Zemel · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Masklab: Instance segmentation by refining object detection with semantic and direction features
Liang-Chieh Chen, Alexander Hermans, George Papandreou, Florian Schroff, Peng Wang, and Hartwig Adam · 2018
Earlier work this paper cites.
Relation networks for object detection
Han Hu, Jiayuan Gu, Zheng Zhang, Jifeng Dai, and Yichen Wei · 2018
Earlier work this paper cites.
Path aggregation network for instance segmentation
Shu Liu, Lu Qi, Haifang Qin, Jianping Shi, and Jiaya Jia · 2018
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2018
Earlier work this paper cites.
Yolact: Real-time instance segmentation
Daniel Bolya, Chong Zhou, Fanyi Xiao, and Yong Jae Lee · 2019
Earlier work this paper cites.
Triply supervised decoder networks for joint detection and segmentation
Jiale Cao, Yanwei Pang, and Xuelong Li · 2019
Cited alongside, same era.
Tensormask: A foundation for dense object segmentation
Xinlei Chen, Ross Girshick, Kaiming He, and Piotr Dollár · 2019
Cited alongside, same era.
Ssap: Single-shot instance segmentation with affinity pyramid
Naiyu Gao, Yanhu Shan, Yupei Wang, Xin Zhao, Yinan Yu, Ming Yang, and Kaiqi Huang · 2019
Cited alongside, same era.
Video action transformer network
Rohit Girdhar, Joao Carreira, Carl Doersch, and Andrew Zisserman · 2019
Cited alongside, same era.
Mask scoring r-cnn
Zhaojin Huang, Lichao Huang, Yongchao Gong, Chang Huang, and Xinggang Wang · 2019
Cited alongside, same era.
Generalized intersection over union: A metric and a loss for bounding box regression
Hamid Rezatofighi, Nathan Tsoi, JunYoung Gwak, Amir Sadeghian, Ian Reid, and Silvio Savarese · 2019
Conditional convolutions for instance segmentation
Zhi Tian, Chunhua Shen, and Hao Chen · 2020
Later among the works it cites.
Training data-efficient image transformers & distillation through attention
Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Hervé Jégou · 2020
Later among the works it cites.
End-to-end object detection with fully convolutional network
Jianfeng Wang, Lin Song, Zeming Li, Hongbin Sun, Jian Sun, and Nanning Zheng · 2020
Later among the works it cites.
Solo: Segmenting objects by locations
Xinlong Wang, Tao Kong, Chunhua Shen, Yuning Jiang, and Lei Li · 2020
Later among the works it cites.
Sceneformer: Indoor scene generation with transformers
Xinpeng Wang, Chandan Yeshwanth, and Matthias Nießner · 2020
Later among the works it cites.
Solov2: Dynamic and fast instance segmentation
Xinlong Wang, Rufeng Zhang, Tao Kong, Lei Li, and Chunhua Shen · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Cyclic guidance for weakly supervised joint detection and segmentation
Yunhang Shen, Rongrong Ji, Yan Wang, Yongjian Wu, and Liujuan Cao · 2019
Cited alongside, same era.
Videobert: A joint model for video and language representation learning
Chen Sun, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid · 2019
Cited alongside, same era.
Lxmert: Learning cross-modality encoder representations from transformers
Hao Tan and Mohit Bansal · 2019
Cited alongside, same era.
Decoders matter for semantic segmentation: Data-dependent decoding enables flexible feature aggregation
Zhi Tian, Tong He, Chunhua Shen, and Youliang Yan · 2019
Cited alongside, same era.
Cross-modal self-attention network for referring image segmentation
Linwei Ye, Mrigank Rochan, Zhi Liu, and Yang Wang · 2019
Cited alongside, same era.
Objects as points
Xingyi Zhou, Dequan Wang, and Philipp Krähenbühl · 2019
Cited alongside, same era.
Later among the works it cites.
Polarmask: Single shot instance segmentation with polar representation
Enze Xie, Peize Sun, Xiaoge Song, Wenhai Wang, Xuebo Liu, Ding Liang, Chunhua Shen, and Ping Luo · 2020
Later among the works it cites.
Learning texture transformer network for image super-resolution
Fuzhi Yang, Huan Yang, Jianlong Fu, Hongtao Lu, and Baining Guo · 2020
Later among the works it cites.
Few-shot learning via embedding adaptation with set-to-set functions
Han-Jia Ye, Hexiang Hu, De-Chuan Zhan, and Fei Sha · 2020
Later among the works it cites.
Deep variational instance segmentation
Jialin Yuan, Chao Chen, and Li Fuxin · 2020
Later among the works it cites.
Mask encoding for single shot instance segmentation
Rufeng Zhang, Zhi Tian, Chunhua Shen, Mingyu You, and Youliang Yan · 2020
Later among the works it cites.
Deformable detr: Deformable transformers for end-to-end object detection
Xizhou Zhu, Weijie Su, Lewei Lu, Bin Li, Xiaogang Wang, and Jifeng Dai · 2020
Later among the works it cites.
Pre-trained image processing transformer
Hanting Chen, Yunhe Wang, Tianyu Guo, Chang Xu, Yiping Deng, Zhenhua Liu, Siwei Ma, Chunjing Xu, Chao Xu, and Wen Gao · 2021
Closest in time.
Transformers in vision: A survey
Salman Khan, Muzammal Naseer, Munawar Hayat, Syed Waqas Zamir, Fahad Shahbaz Khan, and Mubarak Shah · 2021
Closest in time.
Colorization transformer
Manoj Kumar, Dirk Weissenborn, and Nal Kalchbrenner · 2021
Closest in time.
Sparse r-cnn: End-to-end object detection with learnable proposals
Peize Sun, Rufeng Zhang, Yi Jiang, Tao Kong, Chenfeng Xu, Wei Zhan, Masayoshi Tomizuka, Lei Li, Zehuan Yuan, Changhu Wang, et al · 2021
Closest in time.
End-to-end video instance segmentation with transformers
Yuqing Wang, Zhaoliang Xu, Xinlong Wang, Chunhua Shen, Baoshan Cheng, Hao Shen, and Huaxia Xia · 2021
Closest in time.
Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers
Sixiao Zheng, Jiachen Lu, Hengshuang Zhao, Xiatian Zhu, Zekun Luo, Yabiao Wang, Yanwei Fu, Jianfeng Feng, Tao Xiang, Philip HS Torr, et al · 2021
Closest in time.