Fetching the paper…
Reading the bibliography…
We present a novel approach for tracking multiple people in video.
Scale-space and edge detection using anisotropic diffusion
Pietro Perona and Jitendra Malik · 1990
Earlier work this paper cites.
Tracking people with twists and exponential maps
Christoph Bregler and Jitendra Malik · 1998
Earlier work this paper cites.
Framework for performance evaluation of face, text, and vehicle detection and tracking in video: Data, metrics, and protocol
Rangachar Kasturi, Dmitry Goldgof, Padmanabhan Soundararajan, Vasant Manohar, John Garofolo, Rachel Bowers, Matthew Boonstra, Valentina Korzhova, and Jing Zhang · 2008
Earlier work this paper cites.
Estimating human shape and pose from a single image
Peng Guan, Alexander Weiss, Alexandru O Balan, and Michael J Black · 2009
Earlier work this paper cites.
Bayesian 3D tracking from monocular video
Ernesto Brau, Jinyan Guan, Kyle Simek, Luca Del Pero, Colin Reimer Dawson, and Kobus Barnard · 2013
Earlier work this paper cites.
Human3.6M: Large scale datasets and predictive methods for 3D human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2013
Earlier work this paper cites.
2D human pose estimation: New benchmark and state of the art analysis
Mykhaylo Andriluka, Leonid Pishchulin, Peter Gehler, and Bernt Schiele · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Recurrent network models for human dynamics
Katerina Fragkiadaki, Sergey Levine, Panna Felsen, and Jitendra Malik · 2015
Earlier work this paper cites.
SMPL: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J Black · 2015
Earlier work this paper cites.
Faster R-CNN: towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2016
Earlier work this paper cites.
Performance measures and a data set for multi-target, multi-camera tracking
Ergys Ristani, Francesco Solera, Roger Zou, Rita Cucchiara, and Carlo Tomasi · 2016
Earlier work this paper cites.
View synthesis by appearance flow
Tinghui Zhou, Shubham Tulsiani, Weilun Sun, Jitendra Malik, and Alexei A Efros · 2016
Earlier work this paper cites.
RMPE: Regional multi-person pose estimation
Hao-Shu Fang, Shuqin Xie, Yu-Wing Tai, and Cewu Lu · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
PoseTrack: A benchmark for human pose estimation and tracking
Mykhaylo Andriluka, Umar Iqbal, Eldar Insafutdinov, Leonid Pishchulin, Anton Milan, Juergen Gall, and Bernt Schiele · 2018
Earlier work this paper cites.
From lifestyle VLOGs to everyday interactions
David F Fouhey, Wei-cheng Kuo, Alexei A Efros, and Jitendra Malik · 2018
Earlier work this paper cites.
Detect-and-track: Efficient pose estimation in videos
Rohit Girdhar, Georgia Gkioxari, Lorenzo Torresani, Manohar Paluri, and Du Tran · 2018
Earlier work this paper cites.
AVA: A video dataset of spatio-temporally localized atomic visual actions
Chunhui Gu, Chen Sun, David A Ross, Carl Vondrick, Caroline Pantofaru, Yeqing Li, Sudheendra Vijayanarasimhan, George Toderici, Susanna Ricco, Rahul Sukthankar, Cordelia Schmid, and Jitendra Malik · 2018
Cited alongside, same era.
End-to-end recovery of human shape and pose
Angjoo Kanazawa, Michael J Black, David W Jacobs, and Jitendra Malik · 2018
Cited alongside, same era.
Learning category-specific mesh reconstruction from image collections
Angjoo Kanazawa, Shubham Tulsiani, Alexei A Efros, and Jitendra Malik · 2018
Cited alongside, same era.
Single-shot multi-person 3D pose estimation from monocular RGB
Dushyant Mehta, Oleksandr Sotnychenko, Franziska Mueller, Weipeng Xu, Srinath Sridhar, Gerard Pons-Moll, and Christian Theobalt · 2018
Cited alongside, same era.
Sfv: Reinforcement learning of physical skills from videos
Xue Bin Peng, Angjoo Kanazawa, Jitendra Malik, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Coherent reconstruction of multiple humans from a single image
Wen Jiang, Nikos Kolotouros, Georgios Pavlakos, Xiaowei Zhou, and Kostas Daniilidis · 2020
Later among the works it cites.
VIBE: Video inference for human body pose and shape estimation
Muhammed Kocabas, Nikos Athanasiou, and Michael J Black · 2020
Later among the works it cites.
Recursive bayesian filtering for multiple human pose tracking from multiple cameras
Oh-Hun Kwon, Julian Tanke, and Juergen Gall · 2020
Later among the works it cites.
XNect: Real-time multi-person 3D motion capture with a single RGB camera
Dushyant Mehta, Oleksandr Sotnychenko, Franziska Mueller, Weipeng Xu, Mohamed Elgharib, Pascal Fua, Hans-Peter Seidel, Helge Rhodin, Gerard Pons-Moll, and Christian Theobalt · 2020
Later among the works it cites.
Human mesh recovery from multiple shots
Georgios Pavlakos, Jitendra Malik, and Angjoo Kanazawa · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bin Xiao, Haiping Wu, and Yichen Wei · 2018
Cited alongside, same era.
Pose Flow: Efficient online pose tracking
Yuliang Xiu, Jiefeng Li, Haoyu Wang, Yinghong Fang, and Cewu Lu · 2018
Cited alongside, same era.
Deep network for the integrated 3D sensing of multiple people in natural images
Andrei Zanfir, Elisabeta Marinoiu, Mihai Zanfir, Alin-Ionut Popa, and Cristian Sminchisescu · 2018
Cited alongside, same era.
Exploiting temporal context for 3D human pose estimation in the wild
Anurag Arnab, Carl Doersch, and Andrew Zisserman · 2019
Cited alongside, same era.
Tracking without bells and whistles
Philipp Bergmann, Tim Meinhardt, and Laura Leal-Taixe · 2019
Cited alongside, same era.
Learning 3D human dynamics from video
Angjoo Kanazawa, Jason Y Zhang, Panna Felsen, and Jitendra Malik · 2019
Cited alongside, same era.
Learning to reconstruct 3D human pose and shape via model-fitting in the loop
Nikos Kolotouros, Georgios Pavlakos, Michael J Black, and Kostas Daniilidis · 2019
Cited alongside, same era.
15 keypoints is all you need
Michael Snower, Asim Kadav, Farley Lai, and Hans Peter Graf · 2020
Later among the works it cites.
GNN3DMOT: Graph neural network for 3D multi-object tracking with 2D-3D multi-feature learning
Xinshuo Weng, Yongxin Wang, Yunze Man, and Kris M Kitani · 2020
Later among the works it cites.
3D human shape and pose from a single low-resolution image with self-supervised learning
Xiangyu Xu, Hao Chen, Francesc Moreno-Noguer, László A Jeni, and Fernando De la Torre · 2020
Later among the works it cites.
How to train your deep multi-object tracker
Yihong Xu, Aljosa Osep, Yutong Ban, Radu Horaud, Laura Leal-Taixé, and Xavier Alameda-Pineda · 2020
Later among the works it cites.
4D association graph for realtime multi-person motion capture using multiple video cameras
Yuxiang Zhang, Liang An, Tao Yu, Xiu Li, Kun Li, and Yebin Liu · 2020
Later among the works it cites.
Tracking objects as points
Xingyi Zhou, Vladlen Koltun, and Philipp Krähenbühl · 2020
Later among the works it cites.
MOTChallenge: A benchmark for single-camera multiple target tracking
Patrick Dendorfer, Aljosa Osep, Anton Milan, Konrad Schindler, Daniel Cremers, Ian Reid, Stefan Roth, and Laura Leal-Taixé · 2021
Closest in time.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Closest in time.
TrackFormer: Multi-object tracking with transformers
Tim Meinhardt, Alexander Kirillov, Laura Leal-Taixe, and Christoph Feichtenhofer · 2021
Closest in time.
TesseTrack: End-to-end learnable multi-person articulated 3D pose tracking
N Dinesh Reddy, Laurent Guigues, Leonid Pischulin, Jayan Eledath, and Srinivasa Narasimhan · 2021
Closest in time.
Monocular, one-stage, regression of multiple 3D people
Yu Sun, Qian Bao, Wu Liu, Yili Fu, Michael J Black, and Tao Mei · 2021
Closest in time.
3D human pose, shape and texture from low-resolution images and videos
Xiangyu Xu, Hao Chen, Francesc Moreno-Noguer, Laszlo A Jeni, and Fernando De la Torre · 2021
Closest in time.