Fetching the paper…
Reading the bibliography…
We present EgoHumans, a new multi-view multi-human video benchmark to advance the state-of-the-art of egocentric human 3D pose estimation and tracking.
Solving matching problems with linear programming
Martin Grötschel and Olaf Holland · 1985
Earlier work this paper cites.
Dynamic stereo vision
Larry Henry Matthies · 1989
Earlier work this paper cites.
A survey of augmented reality
Ronald T Azuma · 1997
Earlier work this paper cites.
Efficient variants of the icp algorithm
Szymon Rusinkiewicz and Marc Levoy · 2001
Earlier work this paper cites.
Iterative procrustes alignment with the em algorithm
Bin Luo and Edwin R Hancock · 2002
Earlier work this paper cites.
Multiple view geometry in computer vision
Richard Hartley and Andrew Zisserman · 2003
Earlier work this paper cites.
Theory of point estimation
Erich L Lehmann and George Casella · 2006
Earlier work this paper cites.
Object tracking: A survey
Alper Yilmaz, Omar Javed, and Mubarak Shah · 2006
Earlier work this paper cites.
The dynamic hungarian algorithm for the assignment problem with changing costs
G Ayorkor Mills-Tettey, Anthony Stentz, and M Bernardine Dias · 2007
Earlier work this paper cites.
Evaluating multiple object tracking performance: the clear mot metrics
Keni Bernardin and Rainer Stiefelhagen · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Pascal voc 2008 challenge
Derek Hoiem, Santosh K Divvala, and James H Hays · 2009
Earlier work this paper cites.
Clustered pose and nonlinear appearance models for human pose estimation
Sam Johnson and Mark Everingham · 2010
Earlier work this paper cites.
Humaneva: Synchronized video and motion capture dataset and baseline algorithm for evaluation of articulated human motion
Leonid Sigal, Alexandru O Balan, and Michael J Black · 2010
Earlier work this paper cites.
3dpes: 3d people dataset for surveillance and forensics
Davide Baltieri, Roberto Vezzani, and Rita Cucchiara · 2011
Earlier work this paper cites.
Augmented reality: an overview
Julie Carmigniani and Borko Furht · 2011
Earlier work this paper cites.
Understanding egocentric activities
Alireza Fathi, Ali Farhadi, and James M Rehg · 2011
Earlier work this paper cites.
Fast unsupervised ego-action learning for first-person sports videos
Kris M Kitani, Takahiro Okabe, Yoichi Sato, and Akihiro Sugimoto · 2011
Earlier work this paper cites.
Coupling eye-motion and ego-motion features for first-person activity recognition
Keisuke Ogaki, Kris M Kitani, Yusuke Sugano, and Yoichi Sato · 2012
Earlier work this paper cites.
Detecting activities of daily living in first-person camera views
Hamed Pirsiavash and Deva Ramanan · 2012
Earlier work this paper cites.
Virtual reality and virtual reality system components
Oluleke Bamodu and Xu Ming Ye · 2013
Earlier work this paper cites.
Vision meets robotics: The kitti dataset
Andreas Geiger, Philip Lenz, Christoph Stiller, and Raquel Urtasun · 2013
Earlier work this paper cites.
Teleoperation and beyond for assistive humanoid robots
Michael A Goodrich, Jacob W Crandall, and Emilia Barakova · 2013
Earlier work this paper cites.
Human3. 6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2013
Earlier work this paper cites.
Computational augmented reality eyeglasses
Andrew Maimone and Henry Fuchs · 2013
Earlier work this paper cites.
First-person activity recognition: What are they doing to me?
Michael S Ryoo and Larry Matthies · 2013
Earlier work this paper cites.
Acceptance of socially assistive humanoid robot by preschool and elementary school teachers
Marina Fridin and Mark Belokopytov · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Action and interaction recognition in first-person videos
Sanath Narayan, Mohan S Kankanhalli, and Kalpathi R Ramakrishnan · 2014
Earlier work this paper cites.
Lending a hand: Detecting hands and recognizing activities in complex egocentric interactions
Sven Bambach, Stefan Lee, David J Crandall, and Chen Yu · 2015
Earlier work this paper cites.
A survey of augmented reality
Mark Billinghurst, Adrian Clark, Gun Lee, et al · 2015
Earlier work this paper cites.
Near-online multi-target tracking with aggregated local flow descriptor
Wongun Choi · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Panoptic studio: A massively multiview system for social motion capture
Hanbyul Joo, Hao Liu, Lei Tan, Lin Gui, Bart Nabbe, Iain Matthews, Takeo Kanade, Shohei Nobuhara, and Yaser Sheikh · 2015
Earlier work this paper cites.
Smpl: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J Black · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Simple online and realtime tracking
Alex Bewley, Zongyuan Ge, Lionel Ott, Fabio Ramos, and Ben Upcroft · 2016
Earlier work this paper cites.
Simple online and realtime tracking
Alex Bewley, Zongyuan Ge, Lionel Ott, Fabio Ramos, and Ben Upcroft · 2016
Earlier work this paper cites.
Keep it smpl: Automatic estimation of 3d human pose and shape from a single image
Federica Bogo, Angjoo Kanazawa, Christoph Lassner, Peter Gehler, Javier Romero, and Michael J Black · 2016
Earlier work this paper cites.
The cityscapes dataset for semantic urban scene understanding
Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele · 2016
Earlier work this paper cites.
A survey on multiple object tracking algorithm
Litong Fan, Zhongli Wang, Baigen Cail, Chuanqi Tao, Zhiyi Zhang, Yinling Wang, Shanwen Li, Fengtian Huang, Shuangfu Fu, and Feng Zhang · 2016
Earlier work this paper cites.
Performance measures and a data set for multi-target, multi-camera tracking
Ergys Ristani, Francesco Solera, Roger Zou, Rita Cucchiara, and Carlo Tomasi · 2016
Earlier work this paper cites.
The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes
German Ros, Laura Sellart, Joanna Materzynska, David Vazquez, and Antonio M Lopez · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L Schonberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Recognizing micro-actions and reactions from paired egocentric videos
Ryo Yonetani, Kris M Kitani, and Yoichi Sato · 2016
Cited alongside, same era.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick · 2017
Cited alongside, same era.
Towards accurate marker-less human shape and pose estimation over time
Yinghao Huang, Federica Bogo, Christoph Lassner, Angjoo Kanazawa, Peter V Gehler, Javier Romero, Ijaz Akhter, and Michael J Black · 2017
Cited alongside, same era.
End-to-end recovery of human shape and pose. corr abs/1712.06584 (2017)
Angjoo Kanazawa, Michael J Black, David W Jacobs, and Jitendra Malik · 2017
Cited alongside, same era.
The kinetics human action video dataset
Will Kay, Joao Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Paul Natsev, et al · 2017
Cited alongside, same era.
Humbi: A large multiview dataset of human body expressions
Zhixuan Yu, Jae Shin Yoon, In Kyu Lee, Prashanth Venkatesh, Jaesik Park, Jihun Yu, and Hyun Soo Park · 2020
Later among the works it cites.
Srnet: Improving generalization in 3d human pose estimation with a split-and-recombine approach
Ailing Zeng, Xiao Sun, Fuyang Huang, Minhao Liu, Qiang Xu, and Stephen Lin · 2020
Later among the works it cites.
Wandering eyes: Eye movements during mind wandering in video lectures
Han Zhang, Kevin F Miller, Xin Sun, and Kai S Cortina · 2020
Later among the works it cites.
Tracking objects as points
Xingyi Zhou, Vladlen Koltun, and Philipp Krähenbühl · 2020
Later among the works it cites.
Motchallenge: A benchmark for single-camera multiple target tracking
Patrick Dendorfer, Aljosa Osep, Anton Milan, Konrad Schindler, Daniel Cremers, Ian Reid, Stefan Roth, and Laura Leal-Taixé · 2021
Later among the works it cites.
Learning to regress bodies from images using differentiable semantic rendering
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Monocular 3d human pose estimation in the wild using improved cnn supervision
Dushyant Mehta, Helge Rhodin, Dan Casas, Pascal Fua, Oleksandr Sotnychenko, Weipeng Xu, and Christian Theobalt · 2017
Cited alongside, same era.
Feasibility study of a socially assistive humanoid robot for guiding elderly individuals during walking
Chiara Piezzo and Kenji Suzuki · 2017
Cited alongside, same era.
Learning from synthetic humans
Gul Varol, Javier Romero, Xavier Martin, Naureen Mahmood, Michael J Black, Ivan Laptev, and Cordelia Schmid · 2017
Cited alongside, same era.
Simple online and realtime tracking with a deep association metric
Nicolai Wojke, Alex Bewley, and Dietrich Paulus · 2017
Cited alongside, same era.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Cited alongside, same era.
Scaling egocentric vision: The epic-kitchens dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Sanja Fidler, Antonino Furnari, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, et al · 2018
Cited alongside, same era.
Total capture: A 3d deformation model for tracking faces, hands, and bodies
Hanbyul Joo, Tomas Simon, and Yaser Sheikh · 2018
Cited alongside, same era.
Sai Kumar Dwivedi, Nikos Athanasiou, Muhammed Kocabas, and Michael J Black · 2021
Later among the works it cites.
Yolox: Exceeding yolo series in 2021
Zheng Ge, Songtao Liu, Feng Wang, Zeming Li, and Jian Sun · 2021
Later among the works it cites.
Bilevel online adaptation for out-of-domain human mesh reconstruction
Shanyan Guan, Jingwei Xu, Yunbo Wang, Bingbing Ni, and Xiaokang Yang · 2021
Later among the works it cites.
Human poseitioning system (hps): 3d human pose estimation and self-localization in large scenes from body-mounted sensors
Vladimir Guzov, Aymen Mir, Torsten Sattler, and Gerard Pons-Moll · 2021
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2021
Later among the works it cites.
Hota: A higher order metric for evaluating multi-object tracking
Luiten Jonathon, Osep Aljosa, Patrick Dendorfer, Philip Torr, Andreas Geiger, Laura Leal-Taixé, and Leibe Bastian · 2021
Later among the works it cites.
Pare: Part attention regressor for 3d human body estimation
Muhammed Kocabas, Chun-Hao P Huang, Otmar Hilliges, and Michael J Black · 2021
Later among the works it cites.
Spec: Seeing people in the wild with an estimated camera
Muhammed Kocabas, Chun-Hao P Huang, Joachim Tesch, Lea Müller, Otmar Hilliges, and Michael J Black · 2021
Later among the works it cites.
H2o: Two hands manipulating objects for first person interaction recognition
Taein Kwon, Bugra Tekin, Jan Stühmer, Federica Bogo, and Marc Pollefeys · 2021
Later among the works it cites.
Project starline: A high-fidelity telepresence system
Jason Lawrence, Dan B Goldman, Supreeth Achar, Gregory Major Blascovich, Joseph G Desloge, Tommy Fortes, Eric M Gomez, Sascha Häberling, Hugues Hoppe, Andy Huibers, et al · 2021
Later among the works it cites.
Hybrik: A hybrid analytical-neural inverse kinematics solution for 3d human pose and shape estimation
Jiefeng Li, Chao Xu, Zhicun Chen, Siyuan Bian, Lixin Yang, and Cewu Lu · 2021
Later among the works it cites.
Ai choreographer: Music conditioned 3d dance generation with aist++
Ruilong Li, Shan Yang, David A Ross, and Angjoo Kanazawa · 2021
Later among the works it cites.
Ego-exo: Transferring visual representations from third-person to first-person videos
Yanghao Li, Tushar Nagarajan, Bo Xiong, and Kristen Grauman · 2021
Later among the works it cites.
End-to-end human pose and mesh reconstruction with transformers
Kevin Lin, Lijuan Wang, and Zicheng Liu · 2021
Later among the works it cites.
Mesh graphormer
Kevin Lin, Lijuan Wang, and Zicheng Liu · 2021
Later among the works it cites.
Mixture of volumetric primitives for efficient neural rendering
Stephen Lombardi, Tomas Simon, Gabriel Schwartz, Michael Zollhoefer, Yaser Sheikh, and Jason Saragih · 2021
Later among the works it cites.
Pixel codec avatars
Shugao Ma, Tomas Simon, Jason Saragih, Dawei Wang, Yuecheng Li, Fernando De La Torre, and Yaser Sheikh · 2021
Later among the works it cites.
Quasi-dense similarity learning for multiple object tracking
Jiangmiao Pang, Linlu Qiu, Xia Li, Haofeng Chen, Qi Li, Trevor Darrell, and Fisher Yu · 2021
Later among the works it cites.
Agora: Avatars in geography optimized for regression analysis
Priyanka Patel, Chun-Hao P Huang, Joachim Tesch, David T Hoffmann, Shashank Tripathi, and Michael J Black · 2021
Later among the works it cites.
Monocular, one-stage, regression of multiple 3d people
Yu Sun, Qian Bao, Wu Liu, Yili Fu, Michael J Black, and Tao Mei · 2021
Later among the works it cites.
Encoder-decoder with multi-level attention for 3d human shape and pose estimation
Ziniu Wan, Zhengjia Li, Maoqing Tian, Jianbo Liu, Shuai Yi, and Hongsheng Li · 2021
Later among the works it cites.
Pymaf: 3d human pose and shape regression with pyramidal mesh alignment feedback loop
Hongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang, Yebin Liu, Limin Wang, and Zhenan Sun · 2021
Later among the works it cites.
Fairmot: On the fairness of detection and re-identification in multiple object tracking
Yifu Zhang, Chunyu Wang, Xinggang Wang, Wenjun Zeng, and Wenyu Liu · 2021
Later among the works it cites.
Observation-centric sort: Rethinking sort for robust multi-object tracking
Jinkun Cao, Xinshuo Weng, Rawal Khirodkar, Jiangmiao Pang, and Kris Kitani · 2022
Later among the works it cites.
Rescaling egocentric vision: collection, pipeline and challenges for epic-kitchens-100
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Antonino Furnari, Evangelos Kazakos, Jian Ma, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, et al · 2022
Later among the works it cites.
Out-of-domain human mesh reconstruction via dynamic bilevel online adaptation
Shanyan Guan, Jingwei Xu, Michelle Z He, Yunbo Wang, Bingbing Ni, and Xiaokang Yang · 2022
Later among the works it cites.
Occluded human mesh recovery
Rawal Khirodkar, Shashank Tripathi, and Kris Kitani · 2022
Later among the works it cites.
Cliff: Carrying location information in full frames into human pose and shape estimation
Zhihao Li, Jianzhuang Liu, Zhensong Zhang, Songcen Xu, and Youliang Yan · 2022
Later among the works it cites.
Aria pilot dataset
Zhaoyang Lv, Edward Miller, Jeff Meissner, Luis Pesqueira, Chris Sweeney, Jing Dong, Lingni Ma, Pratik Patel, Pierre Moulon, Kiran Somasundaram, Omkar Parkhi, Yuyang Zou, Nikhil Raina, Steve Saarinen, Yusuf M Mansour, Po-Kang Huang, Zijian Wang, Anton Troynikov, Raul Mur Artal, Daniel DeTone, Daniel Barnes, Elizabeth Argall, Andrey Lobanovskiy, David Jaeyun Kim, Philippe Bouttefroy, Julian Straub, Jakob Julian Engel, Prince Gupta, Mingfei Yan, Renzo De Nardi, and Richard Newcombe · 2022
Later among the works it cites.
Evaluation of different lidar technologies for the documentation of forgotten cultural heritage under forest environments
Miguel Ángel Maté-González, Vincenzo Di Pietra, and Marco Piras · 2022
Later among the works it cites.
Human mesh recovery from multiple shots
Georgios Pavlakos, Jitendra Malik, and Angjoo Kanazawa · 2022
Later among the works it cites.
Tracking people by predicting 3d appearance, location and pose
Jathushan Rajasegaran, Georgios Pavlakos, Angjoo Kanazawa, and Jitendra Malik · 2022
Later among the works it cites.
Dancetrack: Multi-object tracking in uniform appearance and diverse motion
Peize Sun, Jinkun Cao, Yi Jiang, Zehuan Yuan, Song Bai, Kris Kitani, and Ping Luo · 2022
Later among the works it cites.
Human-aware object placement for visual environment reconstruction
Hongwei Yi, Chun-Hao P. Huang, Dimitrios Tzionas, Muhammed Kocabas, Mohamed Hassan, Siyu Tang, Justus Thies, and Michael J. Black · 2022
Later among the works it cites.
Egobody: Human body shape and motion of interacting people from head-mounted devices
Siwei Zhang, Qianli Ma, Yan Zhang, Zhiyin Qian, Taein Kwon, Marc Pollefeys, Federica Bogo, and Siyu Tang · 2022
Later among the works it cites.
Bytetrack: Multi-object tracking by associating every detection box
Yifu Zhang, Peize Sun, Yi Jiang, Dongdong Yu, Fucheng Weng, Zehuan Yuan, Ping Luo, Wenyu Liu, and Xinggang Wang · 2022
Later among the works it cites.
Voxeltrack: Multi-person 3d human pose estimation and tracking in the wild
Yifu Zhang, Chunyu Wang, Xinggang Wang, Wenyu Liu, and Wenjun Zeng · 2022
Later among the works it cites.
Can gaze inform egocentric action recognition?
Zehua Zhang, David Crandall, Michael Proulx, Sachin Talathi, and Abhishek Sharma · 2022
Later among the works it cites.
Shihao Zou, Yuanlu Xu, Chao Li, Lingni Ma, Li Cheng, and Minh Vo · 2022
Later among the works it cites.