Fetching the paper…
Reading the bibliography…
We present HAHA - a novel approach for animatable human avatar generation from monocular input videos.
Kanopoulos, N., Vasanthavada, N., Baker, R.L.: Design of an image edge detection filter using the sobel operator. IEEE Journal of solid-state circuits 23
1988
Earlier work this paper cites.
Chambolle, A.: An algorithm for total variation minimization and applications. Journal of Mathematical imaging and vision 20
2004
Earlier work this paper cites.
Kilian, M., Mitra, N.J., Pottmann, H.: Geometric modeling in shape space. In: ACM TOG, pp. 64–es (2007)
2007
Earlier work this paper cites.
Kingma, D.P., Ba, J.: Adam: A method for stochastic optimization. CoRR abs/1412.6980
2014
Earlier work this paper cites.
Milletari, F., Navab, N., Ahmadi, S.A.: V-net: Fully convolutional neural networks for volumetric medical image segmentation. In: International conference on 3D vision (3DV). pp. 565–571. Ieee (2016)
2016
Earlier work this paper cites.
Li, T., Bolkart, T., Black, M.J., Li, H., Romero, J.: Learning a model of facial shape and expression from 4D scans. ACM TOG 36
2017
Earlier work this paper cites.
Alldieck, T., Magnor, M., Xu, W., Theobalt, C., Pons-Moll, G.: Detailed human avatars from monocular video. In: International Conference on 3D Vision (3DV). pp. 98–109. IEEE (2018)
2018
Earlier work this paper cites.
Alldieck, T., Magnor, M., Xu, W., Theobalt, C., Pons-Moll, G.: Video based reconstruction of 3d people models. In: CVPR. pp. 8387–8397 (Jun 2018). https://doi.org/10.1109/CVPR.2018.00875, CVPR Spotlight Paper
2018
Earlier work this paper cites.
Zhang, R., Isola, P., Efros, A.A., Shechtman, E., Wang, O.: The unreasonable effectiveness of deep features as a perceptual metric. In: CVPR (2018)
2018
Earlier work this paper cites.
Zhang, R., Isola, P., Efros, A.A., Shechtman, E., Wang, O.: The unreasonable effectiveness of deep features as a perceptual metric. In: CVPR. pp. 586–595 (2018)
2018
Earlier work this paper cites.
Alldieck, T., Magnor, M., Bhatnagar, B.L., Theobalt, C., Pons-Moll, G.: Learning to reconstruct people in clothing from a single rgb camera. In: CVPR. pp. 1175–1186 (2019)
2019
Earlier work this paper cites.
Gong, K., Gao, Y., Liang, X., Shen, X., Wang, M., Lin, L.: Graphonomy: Universal human parsing via graph transfer learning. In: CVPR (2019)
2019
Earlier work this paper cites.
Pavlakos, G., Choutas, V., Ghorbani, N., Bolkart, T., Osman, A.A., Tzionas, D., Black, M.J.: Expressive body capture: 3d hands, face, and body from a single image. In: CVPR. pp. 10975–10985 (2019)
2019
Earlier work this paper cites.
Saito, S., Huang, Z., Natsume, R., Morishima, S., Kanazawa, A., Li, H.: Pifu: Pixel-aligned implicit function for high-resolution clothed human digitization. In: ICCV. pp. 2304–2314 (2019)
2019
Earlier work this paper cites.
Thies, J., Zollhöfer, M., Nießner, M.: Deferred neural rendering: Image synthesis using neural textures. ACM TOG 38
2019
Earlier work this paper cites.
Kocabas, M., Athanasiou, N., Black, M.J.: Vibe: Video inference for human body pose and shape estimation. In: CVPR. pp. 5253–5263 (2020)
2020
Earlier work this paper cites.
Laine, S., Hellsten, J., Karras, T., Seol, Y., Lehtinen, J., Aila, T.: Modular primitives for high-performance differentiable rendering. ACM TOG 39
2020
Earlier work this paper cites.
Yang, L., Song, Q., Wang, Z., Hu, M., Liu, C., Xin, X., Jia, W., Xu, S.: Renovating parsing r-cnn for accurate multiple human parsing. In: ECCV. pp. 421–437. Springer (2020)
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
Grigorev, A., Iskakov, K., Ianina, A., Bashirov, R., Zakharkin, I., Vakhitov, A., Lempitsky, V.: Stylepeople: A generative model of fullbody human avatars. In: CVPR. pp. 5151–5160 (2021)
2021
Earlier work this paper cites.
He, T., Xu, Y., Saito, S., Soatto, S., Tung, T.: Arch++: Animation-ready clothed human reconstruction revisited. In: ICCV. pp. 11046–11056 (2021)
2021
Earlier work this paper cites.
Jones, B., Zhang, Y., Wong, P.N., Rintel, S.: Belonging there: Vroom-ing into the uncanny valley of xr telepresence. Proceedings of the ACM on Human-Computer Interaction 5
2021
Earlier work this paper cites.
Peng, S., Zhang, Y., Xu, Y., Wang, Q., Shuai, Q., Bao, H., Zhou, X.: Neural body: Implicit neural representations with structured latent codes for novel view synthesis of dynamic humans. In: CVPR. pp. 9054–9063 (2021)
2021
Earlier work this paper cites.
Raj, A., Tanke, J., Hays, J., Vo, M., Stoll, C., Lassner, C.: Anr: Articulated neural rendering for virtual avatars. In: CVPR. pp. 3722–3731 (2021)
2021
Cited alongside, same era.
Sun, Y., Bao, Q., Liu, W., Fu, Y., Black, M.J., Mei, T.: Monocular, one-stage, regression of multiple 3d people. In: ICCV. pp. 11179–11188 (2021)
2021
Cited alongside, same era.
Alldieck, T., Zanfir, M., Sminchisescu, C.: Photorealistic monocular 3d reconstruction of humans wearing clothing. In: CVPR. pp. 1506–1515 (2022)
2022
Cited alongside, same era.
Grassal, P.W., Prinzler, M., Leistner, T., Rother, C., Nießner, M., Thies, J.: Neural head avatars from monocular rgb videos. In: CVPR. pp. 18653–18664 (2022)
2022
Cited alongside, same era.
Jiang, T., Chen, X., Song, J., Hilliges, O.: Instantavatar: Learning avatars from monocular video in 60 seconds. CVPR pp. 16922–16932 (2022)
Introducing Apple Vision Pro: Apple’s first spatial computer. https://www.apple.com/newsroom/2023/06/introducing-apple-vision-pro/ , [Online; accessed 27-June-2024]
2024
Closest in time.
Mark zuckerberg: First interview in the metaverse. https://lexfridman.com/mark-zuckerberg-3/ , online; accessed 27-February-2024
2024
Closest in time.
Texel 3d body model dataset. https://texel.graphics/texel-3d-body-model-dataset/ , online; accessed 27-June-2024
2024
Closest in time.
Bashirov, R., Larionov, A., Ustinova, E., Sidorenko, M., Svitov, D., Zakharkin, I., Lempitsky, V.: Morf: Mobile realistic fullbody avatars from a monocular video. In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision. pp. 3545–3555 (2024)
2024
Closest in time.
Duan, Y., Wei, F., Dai, Q., He, Y., Chen, W., Chen, B.: 4d gaussian splatting: Towards efficient novel view synthesis for dynamic scenes (2024)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
Zhao, H., Zhang, J., Lai, Y.K., Zheng, Z., Xie, Y., Liu, Y., Li, K.: High-fidelity human avatars from a single rgb camera. In: CVPR. pp. 15904–15913 (2022)
2022
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Duan, H.B., Wang, M., Shi, J.C., Chen, X.C., Cao, Y.P.: Bakedavatar: Baking neural fields for real-time head avatar synthesis. ACM TOG 42
2023
Cited alongside, same era.
Işık, M., Rünz, M., Georgopoulos, M., Khakhulin, T., Starck, J., Agapito, L., Nießner, M.: Humanrf: High-fidelity neural radiance fields for humans in motion. ACM TOG 42
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Jiang, T., Chen, X., Song, J., Hilliges, O.: Instantavatar: Learning avatars from monocular video in 60 seconds. In: CVPR. pp. 16922–16932 (2023)
2023
Cited alongside, same era.
2024
Closest in time.
Hu, L., Zhang, H., Zhang, Y., Zhou, B., Liu, B., Zhang, S., Nie, L.: Gaussianavatar: Towards realistic human avatar modeling from a single video via animatable 3d gaussians. CVPR pp. 634–644 (2024)
2024
Closest in time.
Hu, S., Liu, Z.: Gauhuman: Articulated gaussian splatting from monocular human videos. CVPR pp. 20418–20431 (2024)
2024
Closest in time.
Huang, L., Bai, J., Guo, J., Li, Y., Guo, Y.: On the error analysis of 3d gaussian splatting and an optimal projection strategy (2024)
2024
Closest in time.
Jiang, Y., Tu, J., Liu, Y., Gao, X., Long, X., Wang, W., Ma, Y.: Gaussianshader: 3d gaussian splatting with shading functions for reflective surfaces. CVPR pp. 5322–5332 (2024)
2024
Closest in time.
Jiang, Y., Shen, Z., Wang, P., Su, Z., Hong, Y., Zhang, Y., Yu, J., Xu, L.: Hifi4g: High-fidelity human performance rendering via compact gaussian splatting. CVPR pp. 19734–19745 (2024)
2024
Closest in time.
Lee, B., Lee, H., Sun, X., Ali, U., Park, E.: Deblurring 3d gaussian splatting (2024)
2024
Closest in time.
Lei, J., Wang, Y., Pavlakos, G., Liu, L., Daniilidis, K.: Gart: Gaussian articulated template models. CVPR pp. 19876–19887 (2024)
2024
Closest in time.
Li, Z., Zheng, Z., Wang, L., Liu, Y.: Animatable gaussians: Learning pose-dependent gaussian maps for high-fidelity human avatar modeling. CVPR pp. 19711–19722 (2024)
2024
Closest in time.
Luiten, J., Kopanas, G., Leibe, B., Ramanan, D.: Dynamic 3d gaussians: Tracking by persistent dynamic view synthesis pp. 800–809 (2024)
2024
Closest in time.
Moreau, A., Song, J., Dhamo, H., Shaw, R., Zhou, Y., Pérez-Pellitero, E.: Human gaussian splatting: Real-time rendering of animatable avatars. In: CVPR (2024)
2024
Closest in time.
Pang, H., Zhu, H., Kortylewski, A., Theobalt, C., Habermann, M.: Ash: Animatable gaussian splats for efficient and photoreal human rendering. CVPR pp. 1165–1175 (2024)
2024
Closest in time.
Qian, S., Kirschstein, T., Schoneveld, L., Davoli, D., Giebenhain, S., Nießner, M.: Gaussianavatars: Photorealistic head avatars with rigged 3d gaussians. CVPR pp. 20299–20309 (2024)
2024
Closest in time.
Qian, Z., Wang, S., Mihajlovic, M., Geiger, A., Tang, S.: 3dgs-avatar: Animatable avatars via deformable 3d gaussian splatting. CVPR pp. 5020–5030 (2024)
2024
Closest in time.
Saito, S., Schwartz, G., Simon, T., Li, J., Nam, G.: Relightable gaussian codec avatars. CVPR pp. 130–141 (2024)
2024
Closest in time.
Waczyńska, J., Borycki, P., Tadeja, S., Tabor, J., Spurek, P.: Games: Mesh-based adapting and modification of gaussian splatting (2024)
2024
Closest in time.
Yu, Z., Chen, A., Huang, B., Sattler, T., Geiger, A.: Mip-splatting: Alias-free 3d gaussian splatting. CVPR pp. 19447–19456 (2024)
2024
Closest in time.
Zheng, S., Zhou, B., Shao, R., Liu, B., Zhang, S., Nie, L., Liu, Y.: Gps-gaussian: Generalizable pixel-wise 3d gaussian splatting for real-time human novel view synthesis. CVPR pp. 19680–19690 (2024)
2024
Closest in time.