Fetching the paper…
Reading the bibliography…
The standard approach to image instance segmentation is to perform the object detection first, and then segment the object from the detection bounding-box.
Cluster analysis of multivariate data: efficiency versus interpretability of classifications
Edward W Forgy · 1965
Earlier work this paper cites.
Pedestrian detection: A benchmark
Piotr Dollár, Christian Wojek, Bernt Schiele, and Pietro Perona · 2009
Earlier work this paper cites.
Annotated facial landmarks in the wild: A large-scale, real-world database for facial landmark localization
Martin Koestinger, Paul Wohlhart, Peter M Roth, and Horst Bischof · 2011
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
Andreas Geiger, Philip Lenz, and Raquel Urtasun · 2012
Earlier work this paper cites.
Extensive facial landmark localization with coarse-to-fine convolutional network cascade
Erjin Zhou, Haoqiang Fan, Zhimin Cao, Yuning Jiang, and Qi Yin · 2013
Earlier work this paper cites.
Simultaneous detection and segmentation
Bharath Hariharan, Pablo Arbeláez, Ross Girshick, and Jitendra Malik · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Facial landmark detection by deep multi-task learning
Zhanpeng Zhang, Ping Luo, Chen Change Loy, and Xiaoou Tang · 2014
Earlier work this paper cites.
Convolutional feature masking for joint object and stuff segmentation
Jifeng Dai, Kaiming He, and Jian Sun · 2015
Earlier work this paper cites.
Fast r-cnn
Ross Girshick · 2015
Earlier work this paper cites.
Deformable part models are convolutional neural networks
Ross Girshick, Forrest Iandola, Trevor Darrell, and Jitendra Malik · 2015
Earlier work this paper cites.
Hypercolumns for object segmentation and fine-grained localization
Bharath Hariharan, Pablo Arbeláez, Ross Girshick, and Jitendra Malik · 2015
Earlier work this paper cites.
Learning to segment object candidates
Pedro O Pinheiro, Ronan Collobert, and Piotr Dollár · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Instance-sensitive fully convolutional networks
Jifeng Dai, Kaiming He, Yi Li, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
DeeperCut: A deeper, stronger, and faster multi-person pose estimation model
Eldar Insafutdinov, Leonid Pishchulin, Bjoern Andres, Mykhaylo Andriluka, and Bernt Schiele · 2016
Cited alongside, same era.
A fast face detection architecture for auto-focus in smart-phones and digital cameras
Peng Ouyang, Shouyi Yin, Chenchen Deng, Leibo Liu, and Shaojun Wei · 2016
Cited alongside, same era.
You only look once: Unified, real-time object detection
Joseph Redmon, Santosh Divvala, Ross Girshick, and Ali Farhadi · 2016
Cited alongside, same era.
Automatic portrait segmentation for image stylization
Xiaoyong Shen, Aaron Hertzmann, Jiaya Jia, Sylvain Paris, Brian Price, Eli Shechtman, and Ian Sachs · 2016
Cited alongside, same era.
Deep automatic portrait matting
Xiaoyong Shen, Tao Xin, Hongyun Gao, Zhou Chao, and Jiaya Jia · 2016
What can help pedestrian detection?
Jiayuan Mao, Tete Xiao, Yuning Jiang, and Zhimin Cao · 2017
Later among the works it cites.
Associative embedding: End-to-end learning for joint detection and grouping
Alejandro Newell, Zhiao Huang, and Jia Deng · 2017
Later among the works it cites.
Towards accurate multi-person pose estimation in the wild
George Papandreou, Tyler Zhu, Nori Kanazawa, Alexander Toshev, Jonathan Tompson, Chris Bregler, and Kevin Murphy · 2017
Later among the works it cites.
Deepcut: Object segmentation from bounding box annotations using convolutional neural networks
Martin Rajchl, Matthew CH Lee, Ozan Oktay, Konstantinos Kamnitsas, Jonathan Passerat-Palmbach, Wenjia Bai, Mellisa Damodaram, Mary A Rutherford, Joseph V Hajnal, Bernhard Kainz, et al · 2017
Later among the works it cites.
High-quality correspondence and segmentation estimation for dual-lens smart-phone portraits
Xiaoyong Shen, Hongyun Gao, Xin Tao, Chao Zhou, and Jiaya Jia · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Is faster r-cnn doing well for pedestrian detection?
Liliang Zhang, Liang Lin, Xiaodan Liang, and Kaiming He · 2016
Cited alongside, same era.
Realtime multi-person 2d pose estimation using part affinity fields
Zhe Cao, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2017
Cited alongside, same era.
Cascaded pyramid network for multi-person pose estimation
Yilun Chen, Zhicheng Wang, Yuxiang Peng, Zhiqiang Zhang, Gang Yu, and Jian Sun · 2017
Cited alongside, same era.
RMPE: Regional multi-person pose estimation
Hao-Shu Fang, Shuqin Xie, Yu-Wing Tai, and Cewu Lu · 2017
Cited alongside, same era.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick · 2017
Cited alongside, same era.
A coarse-fine network for keypoint localization
Shaoli Huang, Mingming Gong, and Dacheng Tao · 2017
Cited alongside, same era.
Subarna Tripathi, Maxwell Collins, Matthew Brown, and Serge Belongie · 2017
Later among the works it cites.
Joint head pose and facial landmark regression from depth images
Jie Wang, Juyong Zhang, Changwei Luo, and Falai Chen · 2017
Later among the works it cites.
A survey on human performance capture and animation
Shihong Xia, Lin Gao, Yu-Kun Lai, Ming-Ze Yuan, and Jinxiang Chai · 2017
Later among the works it cites.
Detectron
Ross Girshick, Ilija Radosavovic, Georgia Gkioxari, Piotr Dollár, and Kaiming He · 2018
Closest in time.
Robust sparse representation based face recognition in an adaptive weighted spatial pyramid structure
Xiao Ma, Fandong Zhang, Yuelong Li, and Jufu Feng · 2018
Closest in time.
George Papandreou, Tyler Zhu, Liang-Chieh Chen, Spyros Gidaris, Jonathan Tompson, and Kevin Murphy · 2018
Closest in time.
Crowdhuman: A benchmark for detecting human in a crowd
Shuai Shao, Zijian Zhao, Boxun Li, Tete Xiao, Gang Yu, Xiangyu Zhang, and Jian Sun · 2018
Closest in time.
Occluded pedestrian detection through guided attention in cnns
Shanshan Zhang, Jian Yang, and Bernt Schiele · 2018
Closest in time.
Real-time avatar pose transfer and motion generation using locally encoded laplacian offsets
Masoud Zadghorban Lifkooee, Celong Liu, Yongqing Liang, Yimin Zhu, and Xin Li · 2019
Closest in time.