Fetching the paper…
Reading the bibliography…
Many image-based perception tasks can be formulated as detecting, associating and tracking semantic keypoints, e.g., human body pose estimation and tracking.
N. Dinesh Reddy, M. Vo, and S. G. Narasimhan, “Carfusion: Combining point tracking and part detection for dynamic 3d reconstruction of vehicles,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 1906–1915
1915
Earlier work this paper cites.
X. Chen, H. Ma, J. Wan, B. Li, and T. Xia, “Multi-view 3d object detection network for autonomous driving,” in Proceedings of the IEEE conference on Computer Vision and Pattern Recognition , 2017, pp. 1907–1915
1915
Earlier work this paper cites.
1955
Earlier work this paper cites.
M. Kristan, A. Leonardis, J. Matas, M. Felsberg, R. Pflugfelder, L. Cehovin Zajc, T. Vojir, G. Hager, A. Lukezic, A. Eldesokey et al. , “The visual object tracking vot2017 challenge results,” in Proceedings of the IEEE international conference on computer vision workshops , 2017, pp. 1949–1972
1972
Earlier work this paper cites.
B. D. Lucas, T. Kanade et al. , “An iterative image registration technique with an application to stereo vision,” 1981
1981
Earlier work this paper cites.
Y. Nesterov, “A method of solving a convex programming problem with convergence rate o(1/k2),” in Soviet Mathematics Doklady , vol. 27, no. 2, 1983, pp. 372–376
1983
Earlier work this paper cites.
D. Ruppert, “Efficient estimations from a slowly convergent robbins-monro process,” Cornell University Operations Research and Industrial Engineering, Tech. Rep., 1988
1988
Earlier work this paper cites.
B. T. Polyak and A. B. Juditsky, “Acceleration of stochastic approximation by averaging,” SIAM journal on control and optimization , vol. 30, no. 4, pp. 838–855, 1992
1992
Earlier work this paper cites.
A. Simonelli, S. R. Bulo, L. Porzi, M. López-Antequera, and P. Kontschieder, “Disentangling monocular 3d object detection,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 1991–1999
1999
Earlier work this paper cites.
C. Dugas, Y. Bengio, F. Bélisle, C. Nadeau, and R. Garcia, “Incorporating second-order functional knowledge for better option pricing,” Advances in neural information processing systems , vol. 13, pp. 472–478, 2000
2000
Earlier work this paper cites.
K. Bernardin and R. Stiefelhagen, “Evaluating multiple object tracking performance: the clear mot metrics,” EURASIP Journal on Image and Video Processing , vol. 2008, pp. 1–10, 2008
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in Computer Vision and Pattern Recognition, 2009. CVPR 2009. IEEE Conference on . IEEE, 2009, pp. 248–255
2009
Earlier work this paper cites.
L. Bottou, “Large-scale machine learning with stochastic gradient descent,” in Proceedings of COMPSTAT’2010 . Springer, 2010, pp. 177–186
2010
Earlier work this paper cites.
A. Geiger, P. Lenz, C. Stiller, and R. Urtasun, “Vision meets robotics: The kitti dataset,” International Journal of Robotics Research (IJRR) , 2013
2013
Earlier work this paper cites.
M. Andriluka, L. Pishchulin, P. Gehler, and B. Schiele, “2d human pose estimation: New benchmark and state of the art analysis,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2014
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in European Conference on Computer Vision (ECCV) . Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
A. Toshev and C. Szegedy, “Deeppose: Human pose estimation via deep neural networks,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2014, pp. 1653–1660
2014
Earlier work this paper cites.
Y. Xiang, R. Mottaghi, and S. Savarese, “Beyond pascal: A benchmark for 3d object detection in the wild,” in Proceeding of the IEEE Winter Conference on Applications of Computer Vision (WACV) . IEEE, 2014, pp. 75–82
2014
Earlier work this paper cites.
M. Kristan, J. Matas, A. Leonardis, M. Felsberg, L. Cehovin, G. Fernandez, T. Vojir, G. Hager, G. Nebehay, and R. Pflugfelder, “The visual object tracking vot2015 challenge results,” in Proceedings of the IEEE international conference on computer vision workshops , 2015, pp. 1–23
2015
Earlier work this paper cites.
M. Loper, N. Mahmood, J. Romero, G. Pons-Moll, and M. J. Black, “Smpl: A skinned multi-person linear model,” ACM transactions on graphics (TOG) , vol. 34, no. 6, pp. 1–16, 2015
2015
Earlier work this paper cites.
M. Everingham, S. A. Eslami, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes challenge: A retrospective,” International journal of computer vision , vol. 111, no. 1, pp. 98–136, 2015
2015
Earlier work this paper cites.
X. Chen, K. Kundu, Y. Zhu, A. G. Berneshawi, H. Ma, S. Fidler, and R. Urtasun, “3d object proposals for accurate object class detection,” in Advances in Neural Information Processing Systems . Citeseer, 2015, pp. 424–432
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
S.-E. Wei, V. Ramakrishna, T. Kanade, and Y. Sheikh, “Convolutional pose machines,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 4724–4732
2016
Earlier work this paper cites.
A. Newell, K. Yang, and J. Deng, “Stacked hourglass networks for human pose estimation,” in European Conference on Computer Vision (ECCV) . Springer, 2016, pp. 483–499
2016
Earlier work this paper cites.
L. Pishchulin, E. Insafutdinov, S. Tang, B. Andres, M. Andriluka, P. V. Gehler, and B. Schiele, “Deepcut: Joint subset partition and labeling for multi person pose estimation,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 4929–4937
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 770–778
2016
Cited alongside, same era.
W. Shi, J. Caballero, F. Huszár, J. Totz, A. P. Aitken, R. Bishop, D. Rueckert, and Z. Wang, “Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 1874–1883
2016
Cited alongside, same era.
A. Crow, “How safe are self-driving cars?” Rocky Mountain Institute, 5 2017. [Online]. Available: https://rmi.org/safe-self-driving-cars/
2017
Cited alongside, same era.
K. Sun, B. Xiao, D. Liu, and J. Wang, “Deep high-resolution representation learning for human pose estimation,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 5693–5703
2019
Later among the works it cites.
J. Hwang, J. Lee, S. Park, and N. Kwak, “Pose estimator and tracker using temporal flow maps for limbs,” in 2019 International Joint Conference on Neural Networks (IJCNN) . IEEE, 2019, pp. 1–8
2019
Later among the works it cites.
Y. Raaj, H. Idrees, G. Hidalgo, and Y. Sheikh, “Efficient online multi-person 2d pose tracking with recurrent spatio-temporal affinity fields,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 4620–4628
2019
Later among the works it cites.
S. Jin, W. Liu, W. Ouyang, and C. Qian, “Multi-person articulated tracking with spatial and temporal embeddings,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 5664–5673
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Insafutdinov, M. Andriluka, L. Pishchulin, S. Tang, E. Levinkov, B. Andres, and B. Schiele, “Arttrack: Articulated multi-person tracking in the wild,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 6457–6465
2017
Cited alongside, same era.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask r-cnn,” in Computer Vision (ICCV), 2017 IEEE International Conference on . IEEE, 2017, pp. 2980–2988
2017
Cited alongside, same era.
Z. Cao, T. Simon, S.-E. Wei, and Y. Sheikh, “Realtime multi-person 2d pose estimation using part affinity fields,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 7291–7299
2017
Cited alongside, same era.
A. Newell, Z. Huang, and J. Deng, “Associative embedding: End-to-end learning for joint detection and grouping,” in Advances in Neural Information Processing Systems , 2017, pp. 2277–2287
2017
Cited alongside, same era.
G. Papandreou, T. Zhu, N. Kanazawa, A. Toshev, J. Tompson, C. Bregler, and K. Murphy, “Towards accurate multi-person pose estimation in the wild,” in Conference on Computer Vision and Pattern Recognition (CVPR) , vol. 3, no. 4, 2017, p. 6
2017
Cited alongside, same era.
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollár, “Focal loss for dense object detection,” in Proceedings of the IEEE international conference on computer vision (ICCV) , 2017, pp. 2980–2988
2017
Cited alongside, same era.
A. Kendall and Y. Gal, “What uncertainties do we need in bayesian deep learning for computer vision?” in Advances in neural information processing systems , 2017, pp. 5574–5584
2017
Cited alongside, same era.
H.-S. Fang, S. Xie, Y.-W. Tai, and C. Lu, “Rmpe: Regional multi-person pose estimation,” in International Conference on Computer Vision (ICCV) , 2017, pp. 2334–2343
2017
Cited alongside, same era.
2019
Later among the works it cites.
J. Cao, H. Tang, H.-S. Fang, X. Shen, C. Lu, and Y.-W. Tai, “Cross-domain adaptation for animal pose estimation,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2019, pp. 9498–9507
2019
Later among the works it cites.
N. D. Reddy, M. Vo, and S. G. Narasimhan, “Occlusion-net: 2d/3d occluded keypoint localization using graph networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 7326–7335
2019
Later among the works it cites.
T. D. Pereira, D. E. Aldarondo, L. Willmore, M. Kislin, S. S.-H. Wang, M. Murthy, and J. W. Shaevitz, “Fast animal pose estimation using deep neural networks,” Nature methods , vol. 16, no. 1, pp. 117–125, 2019
2019
Later among the works it cites.
S. Zuffi, A. Kanazawa, T. Berger-Wolf, and M. Black, “Three-d safari: Learning to estimate zebra pose, shape, and texture from images “in the wild”,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2019, pp. 5358–5367
2019
Later among the works it cites.
Z. Cao, G. H. Martinez, T. Simon, S.-E. Wei, and Y. A. Sheikh, “Openpose: realtime multi-person 2d pose estimation using part affinity fields,” IEEE transactions on pattern analysis and machine intelligence , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
X. Song, P. Wang, D. Zhou, R. Zhu, C. Guan, Y. Dai, H. Su, H. Li, and R. Yang, “Apollocar3d: A large 3d car instance understanding benchmark for autonomous driving,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 5452–5462
2019
Later among the works it cites.
J. Ku, A. D. Pon, and S. L. Waslander, “Monocular 3d object detection leveraging accurate proposals and shape reconstruction,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 11 867–11 876
2019
Later among the works it cites.
X. Zhou, D. Wang, and P. Krähenbühl, “Objects as points,” arXiv preprint arXiv:1904.07850 , 2019
2019
Later among the works it cites.
U. Iqbal, A. Milan, and J. Gall, “Posetrack: Joint multi-person pose estimation and tracking,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 2011–2020
2020
Later among the works it cites.
B. Cheng, B. Xiao, J. Wang, H. Shi, T. S. Huang, and L. Zhang, “Higherhrnet: Scale-aware representation learning for bottom-up human pose estimation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 5386–5395
2020
Later among the works it cites.
G. Ning, J. Pei, and H. Huang, “Lighttrack: A generic framework for online top-down human pose tracking,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops , 2020, pp. 1034–1035
2020
Later among the works it cites.
B. Biggs, O. Boyne, J. Charles, A. Fitzgibbon, and R. Cipolla, “Who left the dogs out? 3d animal reconstruction with expectation maximization in the loop,” in European Conference on Computer Vision (ECCV) . Springer, 2020, pp. 195–211
2020
Later among the works it cites.
S. Li, J. Li, H. Tang, R. Qian, and W. Lin, “Atrw: A benchmark for amur tiger re-identification in the wild,” in Proceedings of the 28th ACM International Conference on Multimedia . New York, NY, USA: Association for Computing Machinery, 2020, p. 2590–2598
2020
Later among the works it cites.
J. Mu, W. Qiu, G. D. Hager, and A. L. Yuille, “Learning from synthetic animals,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2020, pp. 12 386–12 395
2020
Later among the works it cites.
L. Ke, S. Li, Y. Sun, Y.-W. Tai, and C.-K. Tang, “Gsnet: Joint vehicle pose and shape reconstruction with geometrical and scene-aware supervision,” in European Conference on Computer Vision (ECCV) . Springer, 2020, pp. 515–532
2020
Later among the works it cites.
H. C. Sánchez, A. H. Martínez, R. I. Gonzalo, N. H. Parra, I. P. Alonso, and D. Fernandez-Llorca, “Simple baseline for vehicle pose estimation: Experimental validation,” IEEE Access , vol. 8, pp. 132 539–132 550, 2020
2020
Later among the works it cites.
M. Wang, J. Tighe, and D. Modolo, “Combining detection and tracking for human pose estimation in videos,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 11 088–11 096
2020
Later among the works it cites.
Z. Liu, Z. Wu, and R. Tóth, “Smoke: Single-stage monocular 3d object detection via keypoint estimation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops , 2020, pp. 996–997
2020
Later among the works it cites.
A. Mathis, M. Yüksekgönül, B. Rogers, M. Bethge, and M. W. Mathis, “Pretraining boosts out-of-domain robustness for pose estimation,” in Proceeding of the IEEE Winter Conference on Applications of Computer Vision (WACV) , 2021
2021
Closest in time.