Fetching the paper…
Reading the bibliography…
Directly regressing the non-rigid shape and camera pose from the individual 2D frame is ill-suited to the Non-Rigid Structure-from-Motion (NRSfM) problem.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning representations by back-propagating errors,” nature , vol. 323, no. 6088, pp. 533–536, 1986
1986
Earlier work this paper cites.
P. J. Werbos, “Backpropagation through time: what it does and how to do it,” Proceedings of the IEEE , vol. 78, no. 10, pp. 1550–1560, 1990
1990
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
C. Bregler, A. Hertzmann, and H. Biermann, “Recovering non-rigid 3d shape from image streams,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2000, pp. 690–696
2000
Earlier work this paper cites.
W. Brand, “Morphable 3d models from video,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2001, pp. 456–463
2001
Earlier work this paper cites.
J. Xiao, J.-x. Chai, and T. Kanade, “A closed-form solution to non-rigid shape and motion recovery,” in Eur. Conf. Comput. Vis. , 2004, pp. 573–587
2004
Earlier work this paper cites.
M. Brand, “A direct method for 3d factorization of nonrigid motion observed in 2d,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2005, pp. 122–128
2005
Earlier work this paper cites.
I. Akhter, Y. Sheikh, S. Khan, and T. Kanade, “Nonrigid structure from motion in trajectory space,” in Adv. Neural Inform. Process. Syst. , 2008, pp. 41–48
2008
Earlier work this paper cites.
L. Torresani, A. Hertzmann, and C. Bregler, “Nonrigid structure-from-motion: Estimating shape and motion with hierarchical priors,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 30, no. 5, pp. 878–892, 2008
2008
Earlier work this paper cites.
I. Akhter, Y. Sheikh, and S. Khan, “In defense of orthonormality constraints for nonrigid structure from motion,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2009, pp. 1534–1541
2009
Earlier work this paper cites.
P. F. Gotardo and A. M. Martinez, “Computing smooth time trajectories for camera and deformable shape in structure from motion with occlusion,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 33, no. 10, pp. 2051–2065, 2011
2011
Earlier work this paper cites.
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath et al. , “Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups,” IEEE Signal processing magazine , vol. 29, no. 6, pp. 82–97, 2012
2012
Earlier work this paper cites.
G. Liu, Z. Lin, S. Yan, J. Sun, Y. Yu, and Y. Ma, “Robust recovery of subspace structures by low-rank representation,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 35, no. 1, pp. 171–184, 2012
2012
Earlier work this paper cites.
M. Lee, J. Cho, C.-H. Choi, and S. Oh, “Procrustean normal distribution for non-rigid structure from motion,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2013, pp. 1280–1287
2013
Earlier work this paper cites.
R. Garg, A. Roussos, and L. Agapito, “Dense variational reconstruction of non-rigid surfaces from monocular video,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2013, pp. 1272–1279
2013
Earlier work this paper cites.
Y. Zhu, D. Huang, F. De La Torre, and S. Lucey, “Complex non-rigid motion 3d reconstruction by union of subspaces,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2014, pp. 1542–1549
2014
Earlier work this paper cites.
Y. Dai, H. Li, and M. He, “A simple prior-free method for non-rigid structure-from-motion factorization,” Int. J. Comput. Vis. , vol. 107, no. 2, pp. 101–122, 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in Adv. Neural Inform. Process. Syst. , 2014, p. 3104–3112
2014
Earlier work this paper cites.
C. Ionescu, D. Papava, V. Olaru, and C. Sminchisescu, “Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 36, no. 7, pp. 1325–1339, 2014
2014
Earlier work this paper cites.
S. Kumar, Y. Dai, and H. Li, “Multi-body non-rigid structure-from-motion,” in In Proc. of the International Conf. on 3D Vision (3DV) , 2016, pp. 148–156
2016
Earlier work this paper cites.
A. Agudo and F. Moreno-Noguer, “Recovering pose and 3d deformable shape from multi-instance image ensembles,” in Proc. of the Asian Conf. on Computer Vision , 2016, pp. 291–307
2016
Earlier work this paper cites.
T. Simon, J. Valmadre, I. Matthews, and Y. Sheikh, “Kronecker-markov prior for dynamic 3d reconstruction,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 39, no. 11, pp. 2201–2214, 2017
2017
Cited alongside, same era.
Y. Dai, H. Deng, and M. He, “Dense non-rigid structure-from-motion made easy—a spatial-temporal smoothness based solution,” in IEEE Int. Conf. Image Process. , 2017, pp. 4532–4536
2017
Cited alongside, same era.
S. Kumar, Y. Dai, and H. Li, “Spatio-temporal union of subspaces for multi-body non-rigid structure-from-motion,” Pattern Recognition , vol. 71, pp. 428–443, 2017
2017
Cited alongside, same era.
——, “Procrustean regression: A flexible alignment-based framework for nonrigid structure estimation,” IEEE Trans. Image Process. , vol. 27, no. 1, pp. 249–264, 2017
2017
Cited alongside, same era.
V. Sidhu, E. Tretschk, V. Golyanik, A. Agudo, and C. Theobalt, “Neural dense non-rigid structure from motion with latent space constraints,” in Eur. Conf. Comput. Vis. , 2020, pp. 204–222
2020
Later among the works it cites.
C. Wang, C.-H. Lin, and S. Lucey, “Deep nrsfm++: Towards 3d reconstruction in the wild,” in In Proc. of the International Conf. on 3D Vision (3DV) , 2020, pp. 12–22
2020
Later among the works it cites.
S. Park, M. Lee, and N. Kwak, “Procrustean regression networks: Learning 3d structure of non-rigid objects from 2d annotations,” in Eur. Conf. Comput. Vis. , 2020, pp. 1–18
2020
Later among the works it cites.
M. Lewis, Y. Liu, N. Goyal, M. Ghazvininejad, A. Mohamed, O. Levy, V. Stoyanov, and L. Zettlemoyer, “Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension,” 01 2020, pp. 7871–7880
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask r-cnn,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 2961–2969
2017
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Adv. Neural Inform. Process. Syst. , 2017, pp. 5998–6008
2017
Cited alongside, same era.
P. Ji, T. Zhang, H. Li, M. Salzmann, and I. Reid, “Deep subspace clustering networks,” Adv. Neural Inform. Process. Syst. , vol. 30, pp. 24–33, 2017
2017
Cited alongside, same era.
A. Agudo and F. Moreno-Noguer, “Robust spatio-temporal clustering and reconstruction of multiple deformable bodies,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 41, no. 4, pp. 971–984, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
A. Radford, K. Narasimhan, T. Salimans, and I. Sutskever, “Improving language understanding by generative pre-training,” URL https://s3-us-west-2.amazonaws.com/openai-assets/research-covers/language-unsupervised/language_understanding_paper.pdf , 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2020
Later among the works it cites.
A. Baevski, H. Zhou, A. Mohamed, and M. Auli, “wav2vec 2.0: A framework for self-supervised learning of speech representations,” in Adv. Neural Inform. Process. Syst. , 2020, pp. 12 449–12 460
2020
Later among the works it cites.
Z. Zhang, Y. Shi, C. Yuan, B. Li, P. Wang, W. Hu, and Z. Zha, “Object relational graph with teacher-recommended learning for video captioning,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2020, pp. 13 275–13 285
2020
Later among the works it cites.
M. Kocabas, N. Athanasiou, and M. J. Black, “Vibe: Video inference for human body pose and shape estimation,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2020, pp. 5253–5263
2020
Later among the works it cites.
C. Kong and S. Lucey, “Deep non-rigid structure from motion with missing data,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 43, no. 12, pp. 4365–4377, 2021
2021
Later among the works it cites.
X. Xu and E. Dunn, “Gtt-net: Learned generalized trajectory triangulation,” in Int. Conf. Comput. Vis. , 2021, pp. 5795–5804
2021
Later among the works it cites.
C. Wang and S. Lucey, “Paul: Procrustean autoencoder for unsupervised lifting,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2021, pp. 434–443
2021
Later among the works it cites.
H. Zeng, Y. Dai, X. Yu, X. Wang, and Y. Yang, “Pr-rrn: Pairwise-regularized residual-recursive networks for non-rigid structure-from-motion,” in Int. Conf. Comput. Vis. , 2021, pp. 5600–5609
2021
Later among the works it cites.
G. Yang, D. Sun, V. Jampani, D. Vlasic, F. Cole, H. Chang, D. Ramanan, W. T. Freeman, and C. Liu, “Lasr: Learning articulated shape reconstruction from a monocular video,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2021, pp. 15 980–15 989
2021
Later among the works it cites.
G. Yang, D. Sun, V. Jampani, D. Vlasic, F. Cole, C. Liu, and D. Ramanan, “Viser: Video-specific surface embeddings for articulated 3d shape reconstruction,” Adv. Neural Inform. Process. Syst. , vol. 34, pp. 19 326–19 338, 2021
2021
Later among the works it cites.
Z. Geng, M.-H. Guo, H. Chen, X. Li, K. Wei, and Z. Lin, “Is attention better than matrix decomposition?” in Int. Conf. Learn. Represent. , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
C. Xu, S. Chen, M. Li, and Y. Zhang, “Invariant teacher and equivariant student for unsupervised 3d human pose estimation,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 35, no. 4, 2021, pp. 3013–3021
2021
Later among the works it cites.
S. Kumar and L. V. G. ;, “Organic priors in non-rigid structure from motion,” in ECCV , 2022
2022
Closest in time.
G. Yang, M. Vo, N. Neverova, D. Ramanan, A. Vedaldi, and H. Joo, “Banmo: Building animatable 3d neural models from many casual videos,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2022, pp. 2863–2873
2022
Closest in time.
2022
Closest in time.
O. Press, N. Smith, and M. Lewis, “Train short, test long: Attention with linear biases enables input length extrapolation,” in International Conference on Learning Representations , 2022. [Online]. Available: https://openreview.net/forum?id=R8sQPpGCv0
2022
Closest in time.