Fetching the paper…
Reading the bibliography…
Estimating the 2D human poses in each view is typically the first step in calibrated multi-view 3D pose estimation.
Multiple view geometry in computer vision
Alex M Andrew · 2001
Earlier work this paper cites.
Deying Kong, Haoyu Ma, and Xiaohui Xie · 2009
Earlier work this paper cites.
Latent structured models for human pose estimation
Cristian Sminchisescu Catalin Ionescu, Fuxin Li · 2011
Earlier work this paper cites.
End-to-end video instance segmentation with transformers
Yuqing Wang, Zhaoliang Xu, Xinlong Wang, Chunhua Shen, Baoshan Cheng, Hao Shen, and Huaxia Xia · 2011
Earlier work this paper cites.
Determination of external forces in alpine skiing using a differential global navigation satellite system
Matthias Gilgien, Jörg Spörri, Julien Chardonnens, Josef Kröll, and Erich Müller · 2013
Earlier work this paper cites.
The effect of different global navigation satellite system methods on positioning accuracy in elite alpine skiing
Matthias Gilgien, Jörg Spörri, Philippe Limpach, Alain Geiger, and Erich Müller · 2014
Earlier work this paper cites.
Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Joint training of a convolutional network and a graphical model for human pose estimation
Jonathan J Tompson, Arjun Jain, Yann LeCun, and Christoph Bregler · 2014
Earlier work this paper cites.
Determination of the centre of mass kinematics in alpine skiing using differential global navigation satellite systems
Matthias Gilgien, Jörg Spörri, Julien Chardonnens, Josef Kröll, Philippe Limpach, and Erich Müller · 2015
Earlier work this paper cites.
Three-dimensional body and centre of mass kinematics in alpine ski racing using differential gnss and inertial sensors
Benedikt Fasel, Jörg Spörri, Matthias Gilgien, Geo Boffi, Julien Chardonnens, Erich Müller, and Kamiar Aminian · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Stacked hourglass networks for human pose estimation
Alejandro Newell, Kaiyu Yang, and Jia Deng · 2016
Earlier work this paper cites.
Reasearch dedicated to sports injury prevention-the’sequence of prevention’on the example of alpine ski racing
Jörg Spörri · 2016
Earlier work this paper cites.
Convolutional pose machines
Shih-En Wei, Varun Ramakrishna, Takeo Kanade, and Yaser Sheikh · 2016
Earlier work this paper cites.
Joint inertial sensor orientation drift reduction for highly dynamic movements
Benedikt Fasel, Jörg Spörri, Julien Chardonnens, Josef Kröll, Erich Müller, and Kamiar Aminian · 2017
Earlier work this paper cites.
Vnect: Real-time 3d human pose estimation with a single rgb camera
Dushyant Mehta, Srinath Sridhar, Oleksandr Sotnychenko, Helge Rhodin, Mohammad Shafiei, Hans-Peter Seidel, Weipeng Xu, Dan Casas, and Christian Theobalt · 2017
Earlier work this paper cites.
Hand keypoint detection in single images using multiview bootstrapping
Tomas Simon, Hanbyul Joo, Iain Matthews, and Yaser Sheikh · 2017
Earlier work this paper cites.
Lifting from the deep: Convolutional 3d pose estimation from a single image
Denis Tome, Chris Russell, and Lourdes Agapito · 2017
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Learning monocular 3d human pose estimation from multi-view images
Helge Rhodin, Jörg Spörri, Isinsu Katircioglu, Victor Constantin, Frédéric Meyer, Erich Müller, Mathieu Salzmann, and Pascal Fua · 2018
Cited alongside, same era.
Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning
Piyush Sharma, Nan Ding, Sebastian Goodman, and Radu Soricut · 2018
Cited alongside, same era.
Non-local neural networks
Xiaolong Wang, Ross Girshick, Abhinav Gupta, and Kaiming He · 2018
Cited alongside, same era.
Epipolar transformers
Yihui He, Rui Yan, Katerina Fragkiadaki, and Shoou-I Yu · 2020
Later among the works it cites.
Unicoder-vl: A universal encoder for vision and language by cross-modal pre-training
Gen Li, Nan Duan, Yuejian Fang, Ming Gong, and Daxin Jiang · 2020
Later among the works it cites.
End-to-end human pose and mesh reconstruction with transformers
Kevin Lin, Lijuan Wang, and Zicheng Liu · 2020
Later among the works it cites.
Training data-efficient image transformers & distillation through attention
Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Hervé Jégou · 2020
Later among the works it cites.
Metafuse: A pre-trained fusion model for human pose estimation
Rongchang Xie, Chunyu Wang, and Yizhou Wang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Simple baselines for human pose estimation and tracking
Bin Xiao, Haiping Wu, and Yichen Wei · 2018
Cited alongside, same era.
3d hand shape and pose estimation from a single rgb image
Liuhao Ge, Zhou Ren, Yuncheng Li, Zehao Xue, Yingying Wang, Jianfei Cai, and Junsong Yuan · 2019
Cited alongside, same era.
Learnable triangulation of human pose
Karim Iskakov, Egor Burkov, Victor Lempitsky, and Yury Malkov · 2019
Cited alongside, same era.
Adaptive graphical model network for 2d handpose estimation
Deying Kong, Yifei Chen, Haoyu Ma, Xiangyi Yan, and Xiaohui Xie · 2019
Cited alongside, same era.
Cross view fusion for 3d human pose estimation
Haibo Qiu, Chunyu Wang, Jingdong Wang, Naiyan Wang, and Wenjun Zeng · 2019
Cited alongside, same era.
Vl-bert: Pre-training of generic visual-linguistic representations
Weijie Su, Xizhou Zhu, Yue Cao, Bin Li, Lewei Lu, Furu Wei, and Jifeng Dai · 2019
Cited alongside, same era.
Deep high-resolution representation learning for human pose estimation
Ke Sun, Bin Xiao, Dong Liu, and Jingdong Wang · 2019
Cited alongside, same era.
Sen Yang, Zhibin Quan, Mu Nie, and Wankou Yang · 2020
Later among the works it cites.
Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers
Sixiao Zheng, Jiachen Lu, Hengshuang Zhao, Xiatian Zhu, Zekun Luo, Yabiao Wang, Yanwei Fu, Jianfeng Feng, Tao Xiang, Philip HS Torr, et al · 2020
Later among the works it cites.
Deformable detr: Deformable transformers for end-to-end object detection
Xizhou Zhu, Weijie Su, Lewei Lu, Bin Li, Xiaogang Wang, and Jifeng Dai · 2020
Later among the works it cites.
Vilt: Vision-and-language transformer without convolution or region supervision
Wonjae Kim, Bokyung Son, and Ildoo Kim · 2021
Closest in time.
Pose recognition with cascade transformers
Ke Li, Shijie Wang, Xiang Zhang, Yifan Xu, Weijian Xu, and Zhuowen Tu · 2021
Closest in time.
Tfpose: Direct human pose estimation with transformers
Weian Mao, Yongtao Ge, Chunhua Shen, Zhi Tian, Xinlong Wang, and Zhibin Wang · 2021
Closest in time.
Spatial context-aware self-attention model for multi-organ segmentation
Hao Tang, Xingwei Liu, Kun Han, Xiaohui Xie, Xuming Chen, Huang Qian, Yong Liu, Shanlin Sun, and Narisu Bai · 2021
Closest in time.
Adafuse: Adaptive multiview fusion for accurate human pose estimation in the wild
Zhe Zhang, Chunyu Wang, Weichao Qiu, Wenhu Qin, and Wenjun Zeng · 2021
Closest in time.
Camera pose matters: Improving depth prediction by mitigating pose distribution bias
Yunhan Zhao, Shu Kong, and Charless Fowlkes · 2021
Closest in time.
3d human pose estimation with spatial and temporal transformers
Ce Zheng, Sijie Zhu, Matias Mendieta, Taojiannan Yang, Chen Chen, and Zhengming Ding · 2021
Closest in time.
Kaleido-bert: Vision-language pre-training on fashion domain
Mingchen Zhuge, Dehong Gao, Deng-Ping Fan, Linbo Jin, Ben Chen, Haoming Zhou, Minghui Qiu, and Ling Shao · 2021
Closest in time.
Sscap: Self-supervised co-occurrence action parsing for unsupervised temporal action segmentation
Zhe Wang, Hao Chen, Xinyu Li, Chunhui Liu, Yuanjun Xiong, Joseph Tighe, and Charless C Fowlkes · 2022
Closest in time.
After-unet: Axial fusion transformer unet for medical image segmentation
Xiangyi Yan, Hao Tang, Shanlin Sun, Haoyu Ma, Deying Kong, and Xiaohui Xie · 2022
Closest in time.