Fetching the paper…
Reading the bibliography…
We present a framework for pre-training of 3D hand pose estimation from in-the-wild hand images sharing with similar hand characteristics, dubbed SimHand.
Learning a similarity metric discriminatively, with application to face verification
S. Chopra, R. Hadsell, and Y. LeCun · 2005
Earlier work this paper cites.
Joint generative and contrastive learning for unsupervised person re-identification
H. Chen, Y. Wang, B. Lagadec, A. Dantcheva, and F. Bremond · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Facenet: A unified embedding for face recognition and clustering
F. Schroff, D. Kalenichenko, and J. Philbin · 2015
Earlier work this paper cites.
Netvlad: CNN architecture for weakly supervised place recognition
R. Arandjelovic, P. Gronat, A. Torii, T. Pajdla, and J. Sivic · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Improved deep metric learning with multi-class n-pair loss objective
K. Sohn · 2016
Earlier work this paper cites.
Deep metric learning via lifted structured feature embedding
H. Song, Y. Xiang, S. Jegelka, and S. Savarese · 2016
Earlier work this paper cites.
Large batch training of convolutional networks
Y. You, I. Gitman, and B. Ginsburg · 2017
Earlier work this paper cites.
Weakly-supervised 3D hand pose estimation from monocular RGB images
Y. Cai, L. Ge, J. Cai, and J. Yuan · 2018
Earlier work this paper cites.
3D hand shape and pose estimation from a single RGB image
L. Ge, Z. Ren, Y. Li, Z. Xue, Y. Wang, J. Cai, and J. Yuan · 2019
Earlier work this paper cites.
Mediapipe: A framework for building perception pipelines
C. Lugaresi, J. Tang, H. Nash, C. McClanahan, E. Uboweja, M. Hays, and F. Zhang et al · 2019
Earlier work this paper cites.
A2J: anchor-to-joint regression network for 3D articulated pose estimation from a single depth image
F. Xiong, B. Zhang, Y. Xiao, Z. Cao, T. Yu, J. T. Zhou, and J. Yuan · 2019
Earlier work this paper cites.
FreiHAND: A dataset for markerless capture of hand pose and shape from single RGB images
C. Zimmermann, D. Ceylan, J. Yang, B. Russell, M. J. Argus, and T. Brox · 2019
Earlier work this paper cites.
Unsupervised learning of visual features by contrasting cluster assignments
M. Caron, I. Misra, J. Mairal, P. Goyal, P. Bojanowski, and A. Joulin · 2020
Earlier work this paper cites.
A simple framework for contrastive learning of visual representations
T. Chen, S. Kornblith, M. Norouzi, and G. E. Hinton · 2020
Earlier work this paper cites.
Self-supervising fine-grained region similarities for large-scale image localization
Y. Ge, H. Wang, F. Zhu, R. Zhao, and H. Li · 2020
Earlier work this paper cites.
Bootstrap your own latent: A new approach to self-supervised learning
J. Grill, F. Strub, F. Altché, C. Tallec, P. Richemond, E. Buchatskaya, C. Doersch, B. Avila Pires, Z. Guo, M. Gheshlaghi Azar, B. Piot, K. Kavukcuoglu, R. Munos, and M. Valko · 2020
Cited alongside, same era.
Momentum contrast for unsupervised visual representation learning
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick · 2020
Cited alongside, same era.
InterHand2.6M: A dataset and baseline for 3D interacting hand pose estimation from a single RGB image
G. Moon, S.-I. Yu, H. Wen, T. Shiratori, and K. M. Lee · 2020
Cited alongside, same era.
Understanding human hands in contact at internet scale
D. Shan, J. Geng, M. Shu, and D. Fouhey · 2020
Cited alongside, same era.
Weakly supervised 3D hand pose estimation via biomechanical constraints
A. Spurr, U. Iqbal, P. Molchanov, O. Hilliges, and J. Kautz · 2020
Cited alongside, same era.
Ego4D: Around the world in 3, 000 hours of egocentric video
K. Grauman, A. Westbury, E. Byrne, Z. Chavis, A. Furnari, R. Girdhar, J. Hamburger, H. Jiang, M. Liu, X. Liu, M. Martin, T. Nagarajan, I. Radosavovic, S. K. Ramakrishnan, F. Ryan, J. Sharma, M. Wray, M.g Xu, E. Zhongcong Xu, C. Zhao, S. Bansal, D. Batra, V. Cartillier, S. Crane, T. Do, M. Doulaty, A. Erapalli, C. Feichtenhofer, A. Fragomeni, Q. Fu, C. Fuegen, A. Gebreselasie, C. Gonzalez, J. Hillis, X. Huang, Y. Huang, W. Jia, W. Khoo, J. Kolar, S. Kottur, A. Kumar, F. Landini, C. Li, Y. Li, Z. Li, K. Mangalam, R. Modhugu, J. Munro, T. Murrell, T. Nishiyasu, W. Price, P. R. Puentes, M. Ramazanova, L. Sari, K. Somasundaram, A. Southerland, Y. Sugano, R. Tao, M. Vo, Y. Wang, X. Wu, T. Yagi, Y. Zhu, P. Arbelaez, D. Crandall, D. Damen, G. M. Farinella, B. Ghanem, V. K. Ithapu, C. V. Jawahar, H. Joo, K. Kitani, H. Li, R. Newcombe, A. Oliva, H. Soo Park, J. M. Rehg, Y. Sato, J. Shi, M. Z. Shou, A. Torralba, Lo Torresani, M.i Yan, and J. Malik · 2022
Later among the works it cites.
UmeTrack: Unified multi-view end-to-end hand tracking for VR
S. Han, P.-C. Wu, Y. Zhang, B. Liu, L. Zhang, Z. Wang, W. Si, P. Zhang, Y. Cai, T. Hodan, R. Cabezas, L. Tran, M. Akbay, T.-H. Yu, C. Keskin, and R. Wang · 2022
Later among the works it cites.
Domain adaptive hand keypoint and pixel localization in the wild
T. Ohkawa, Y.-J. Li, Q. Fu, R. Furuta, K. M. Kitani, and Y. Sato · 2022
Later among the works it cites.
Handoccnet: Occlusion-robust 3d hand mesh estimation network
J. Park, Y. Oh, G. Moon, H. Choi, and K. M. Lee · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M.-Y. Wu, P.-W. Ting, Y.-H. Tang, E. T. Chou, and L.-C. Fu · 2020
Cited alongside, same era.
Monocular real-time hand shape and motion capture using multi-modal data
Y. Zhou, M. Habermann, W. Xu, I. Habibie, C. Theobalt, and F. Xu · 2020
Cited alongside, same era.
Emerging properties in self-supervised vision transformers
M. Caron, H. Touvron, I. Misra, H. Jégou, J. Mairal, P. Bojanowski, and A. Joulin · 2021
Cited alongside, same era.
DexYCB: A benchmark for capturing hand grasping of objects
Y.-W. Chao, W. Yang, Y. Xiang, P. Molchanov, A. Handa, J. Tremblay, Y. S. Narang, K. Van Wyk, U. Iqbal, S. Birchfield, J. Kautz, and D. Fox · 2021
Cited alongside, same era.
Exploring simple siamese representation learning
X. Chen and K. He · 2021
Cited alongside, same era.
Patch-netvlad: Multi-scale fusion of locally-global descriptors for place recognition
S. Hausler, S. Garg, M. Xu, M. Milford, and T. Fischer · 2021
Cited alongside, same era.
Semi-supervised 3D hand-object poses estimation with interactions in time
S. Liu, H. Jiang, J. Xu, S. Liu, and X. Wang · 2021
Cited alongside, same era.
DexMV: Imitation learning for dexterous manipulation from human videos
Y. Qin, Y.-H. Wu, S. Liu, H. Jiang, R. Yang, Y. Fu, and X. Wang · 2022
Later among the works it cites.
Assembly101: A large-scale multi-view video dataset for understanding procedural activities
F. Sener, D. Chatterjee, D. Shelepov, K. He, D. Singhania, R. Wang, and A. Yao · 2022
Later among the works it cites.
Background mixup data augmentation for hand and object-in-contact detection
K. Tango, T. Ohkawa, R. Furuta, and Y. Sato · 2022
Later among the works it cites.
Collaborative learning for hand and object reconstruction with attention-guided graph convolution
T. Tse, K. Kim, A. Leonardis, and H. Chang · 2022
Later among the works it cites.
Contrastive positive mining for unsupervised 3d action representation learning
H. Zhang, Y. Hou, W. Zhang, and W. Li · 2022
Later among the works it cites.
Tempclr: Reconstructing hands via time-coherent contrastive learning
A. Ziani, Z. Fan, M. Kocabas, S. J. Christen, and O. Hilliges · 2022
Later among the works it cites.
Weakly supervised temporal sentence grounding with uncertainty-guided self-training
Y. Huang, L. Yang, and Y. Sato · 2023
Later among the works it cites.
Generative hierarchical temporal transformer for hand action recognition and motion prediction
Y. Wen, H. Pan, T. Ohkawa, L. Yang, J. Pan, Y. Sato, T. Komura, and W. Wang · 2023
Later among the works it cites.
Hamuco: Hand pose estimation via multiview collaborative self-supervised learning
X. Zheng, C. Wen, Z. Xue, P. Ren, and J. Wang · 2023
Later among the works it cites.
Benchmarks and challenges in pose estimation for egocentric hand interactions with objects
Z. Fan, T. Ohkawa, L. Yang, N. Lin, Z. Zhou, S. Zhou, J. Liang, Z. Gao, X. Zhang, X. Zhang, F. Li, L. Zheng, F. Lu, K. A. Zeid, B. Leibe, J. On, S. Baek, A. Prakash, S. Gupta, K. He, Y. Sato, O. Hilliges, H. J. Chang, and A. Yao · 2024
Later among the works it cites.
Self-weighted contrastive learning among multiple views for mitigating representation degeneration
X. Jie, S. Chen, Y. Ren, X. Shi, H. Shen, G. Niu, and X. Zhu · 2024
Later among the works it cites.
Single-to-dual-view adaptation for egocentric 3d hand pose estimation
R. Liu, T. Ohkawa, M. Zhang, and Y. Sato · 2024
Later among the works it cites.