Fetching the paper…
Reading the bibliography…
Humans inhabit a world defined by interactions -- with other humans, objects, and environments.
URL https://arxiv.org/abs/1904.10681
Ji, Y., Xu, F., Yang, Y., Shen, F., Shen, H.T., Zheng, W.S.: A large-scale varying-view rgb-d action dataset for arbitrary-view human action recognition (2019) · 1904
Earlier work this paper cites.
URL https://arxiv.org/abs/1904.09882
Ng, E., Xiang, D., Joo, H., Grauman, K.: You2me: Inferring body pose in egocentric video via first and second person interactions (2020) · 1904
Earlier work this paper cites.
URL https://arxiv.org/abs/1904.05866
Pavlakos, G., Choutas, V., Ghorbani, N., Bolkart, T., Osman, A.A.A., Tzionas, D., Black, M.J.: Expressive body capture: 3d hands, face, and body from a single image (2019) · 1904
Earlier work this paper cites.
URL https://arxiv.org/abs/1906.04158
Joo, H., Simon, T., Cikara, M., Sheikh, Y.: Towards social artificial intelligence: Nonverbal social signal prediction in a triadic interaction (2019) · 1906
Earlier work this paper cites.
URL https://arxiv.org/abs/1910.02181
Ahuja, C., Ma, S., Morency, L.P., Sheikh, Y.: To react or not to react: End-to-end visual pose forecasting for personalized avatar during dyadic conversations (2019) · 1910
Earlier work this paper cites.
In: Computer Vision and Pattern Recognition (CVPR) (2020)
Zhang, Y., Hassan, M., Neumann, H., Black, M.J., Tang, S.: Generating 3d people in scenes without people · 1912
Earlier work this paper cites.
nature 323
Rumelhart, D.E., Hinton, G.E., Williams, R.J.: Learning representations by back-propagating errors · 1986
Earlier work this paper cites.
In: 1994 Proceedings of IEEE conference on computer vision and pattern recognition, pp. 593–600. IEEE (1994)
Shi, J., et al.: Good features to track · 1994
Earlier work this paper cites.
In: Readings in human–computer interaction, pp. 5–21. Elsevier (1995)
Norman, D.A.: The psychopathology of everyday things · 1995
Earlier work this paper cites.
MIT press (1998)
Clark, A.: Being there: Putting brain, body, and world together again · 1998
Earlier work this paper cites.
Communications of the ACM 42
Badler, N.I., Palmer, M.S., Bindiganavale, R.: Animation control for real-time virtual humans · 1999
Earlier work this paper cites.
In: ACM SIGGRAPH 2003 Papers, pp. 402–408 (2003)
Arikan, O., Forsyth, D.A., O’Brien, J.F.: Motion synthesis from annotations · 2003
Earlier work this paper cites.
In: 5th IEEE-RAS International Conference on Humanoid Robots, 2005., pp. 104–109. IEEE (2005)
Naksuk, N., Lee, C.G., Rietdyk, S.: Whole-body human-to-humanoid motion transfer · 2005
Earlier work this paper cites.
In: 2006 IEEE computer society conference on computer vision and pattern recognition (CVPR’06), vol. 1, pp. 519–528. IEEE (2006)
Seitz, S.M., Curless, B., Diebel, J., Scharstein, D., Szeliski, R.: A comparison and evaluation of multi-view stereo reconstruction algorithms · 2006
Earlier work this paper cites.
In: 2009 IEEE/RSJ International Conference on Intelligent Robots and Systems, pp. 2518–2524. IEEE (2009)
Kim, S., Kim, C., You, B., Oh, S.: Stable whole-body motion generation for humanoid robots to imitate human motions · 2009
Earlier work this paper cites.
URL https://arxiv.org/abs/2010.11929
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., Uszkoreit, J., Houlsby, N.: An image is worth 16x16 words: Transformers for image recognition at scale (2021) · 2010
Earlier work this paper cites.
http://cvrc.ece.utexas.edu/SDHA2010/Human_Interaction.html (2010)
Ryoo, M.S., Aggarwal, J.K.: UT-Interaction Dataset, ICPR contest on Semantic Description of Human Activities (SDHA) · 2010
Earlier work this paper cites.
In: Computer Graphics Forum, vol. 29, pp. 2530–2554. Wiley Online Library (2010)
Van Welbergen, H., Van Basten, B.J., Egges, A., Ruttkay, Z.M., Overmars, M.H.: Real time animation of virtual humans: a trade-off between naturalness and control · 2010
Earlier work this paper cites.
In: 2011 IEEE International Conference on Computer Vision Workshops (ICCV Workshops), pp. 1264–1269 (2011)
van der Aa, N., Luo, X., Giezeman, G., Tan, R., Veltkamp, R.: Umpm benchmark: A multi-person dataset with synchronized video and motion capture data for evaluation of articulated human motion and interaction · 2011
Earlier work this paper cites.
In: CVPR 2011, pp. 1249–1256 (2011)
Liu, Y., Stoll, C., Gall, J., Seidel, H.P., Theobalt, C.: Markerless motion capture of interacting characters using multi-view image segmentation · 2011
Earlier work this paper cites.
In: 2011 10th IEEE international symposium on mixed and augmented reality, pp. 127–136. Ieee (2011)
Newcombe, R.A., Izadi, S., Hilliges, O., Molyneaux, D., Kim, D., Davison, A.J., Kohi, P., Shotton, J., Hodges, S., Fitzgibbon, A.: Kinectfusion: Real-time dense surface mapping and tracking · 2011
Earlier work this paper cites.
International Journal of Production Research 50
Kuo, C.F., Wang, M.J.J.: Motion generation and virtual simulation in a digital environment · 2012
Earlier work this paper cites.
pp. 28–35 (2012)
Yun, K., Honorio, J., Chattopadhyay, D., Berg, T., Samaras, D.: Two-person interaction detection using body-pose features and multiple instance learning · 2012
Earlier work this paper cites.
IEEE multimedia 19
Zhang, Z.: Microsoft kinect sensor and its effect · 2012
Earlier work this paper cites.
In: Proceedings of the 21st ACM international conference on Multimedia, pp. 835–838 (2013)
Eyben, F., Weninger, F., Gross, F., Schuller, B.: Recent developments in opensmile, the munich open-source multimedia feature extractor · 2013
Earlier work this paper cites.
Mathematical Problems in Engineering 2013
Hu, T., Zhu, X., Su, K.: Efficient interaction recognition through positive action representation · 2013
Earlier work this paper cites.
IEEE transactions on pattern analysis and machine intelligence 36
Ionescu, C., Papava, D., Olaru, V., Sminchisescu, C.: Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments · 2013
Earlier work this paper cites.
arXiv preprint arXiv:1312.6114 (2013)
Kingma, D.P.: Auto-encoding variational bayes · 2013
Earlier work this paper cites.
URL https://arxiv.org/abs/1406.1078
Cho, K., van Merrienboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., Bengio, Y.: Learning phrase representations using rnn encoder-decoder for statistical machine translation (2014) · 2014
Earlier work this paper cites.
In: European Conference on Computer Vision (2014)
Huang, D.A., Kitani, K.M.: Action-reaction: Forecasting the dynamics of human interaction · 2014
Earlier work this paper cites.
In: 2015 international conference on advanced robotics (ICAR), pp. 510–517. IEEE (2015)
Calli, B., Singh, A., Walsman, A., Srinivasa, S., Abbeel, P., Dollar, A.M.: The ycb object and model set: Towards common benchmarks for manipulation research · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1502.03143 (2015)
Calli, B., Walsman, A., Singh, A., Srinivasa, S., Abbeel, P., Dollar, A.M.: Benchmarking in manipulation research: The ycb object and model set and benchmarking protocols · 2015
Earlier work this paper cites.
Int J Sci Res 4
Ebrahim, M.A.B.: 3d laser scanners’ techniques overview · 2015
Earlier work this paper cites.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2015)
Elhayek, A., de Aguiar, E., Jain, A., Tompson, J., Pishchulin, L., Andriluka, M., Bregler, C., Schiele, B., Theobalt, C.: Efficient convnet-based marker-less motion capture in general scenes with a low number of cameras · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1505.04474 (2015)
Gupta, S., Malik, J.: Visual semantic role labeling · 2015
Earlier work this paper cites.
In: SciPy, pp. 18–24 (2015)
McFee, B., Raffel, C., Liang, D., Ellis, D.P., McVicar, M., Battenberg, E., Nieto, O.: librosa: Audio and music signal analysis in python · 2015
Earlier work this paper cites.
Advances in neural information processing systems 28
Sohn, K., Lee, H., Yan, X.: Learning structured output representation using deep conditional generative models · 2015
Earlier work this paper cites.
In: Proceedings of the IEEE international conference on computer vision, pp. 4489–4497 (2015)
Tran, D., Bourdev, L., Fergus, R., Torresani, L., Paluri, M.: Learning spatiotemporal features with 3d convolutional networks · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1611.02648 (2016)
Dilokthanakul, N., Mediano, P.A., Garnelo, M., Lee, M.C., Salimbeni, H., Arulkumaran, K., Shanahan, M.: Deep unsupervised clustering with gaussian mixture variational autoencoders · 2016
Earlier work this paper cites.
In: M. Chetouani, J. Cohn, A.A. Salah (eds.) Human Behavior Understanding, pp. 116–133. Springer International Publishing, Cham (2016)
van Gemeren, C., Poppe, R., Veltkamp, R.C.: Spatio-temporal detection of fine-grained dyadic human interactions · 2016
Earlier work this paper cites.
URL https://arxiv.org/abs/1612.03153
Joo, H., Simon, T., Li, X., Liu, H., Tan, L., Gui, L., Banerjee, S., Godisart, T., Nabbe, B., Matthews, I., Kanade, T., Nobuhara, S., Sheikh, Y.: Panoptic studio: A massively multiview system for social interaction capture (2016) · 2016
Earlier work this paper cites.
In: SIGGRAPH ASIA 2016 virtual reality meets physical reality: Modelling and simulating virtual humans and environments, pp. 1–4 (2016)
Lin, J., Guo, X., Shao, J., Jiang, C., Zhu, Y., Zhu, S.C.: A virtual reality platform for dynamic human-scene interaction · 2016
Earlier work this paper cites.
Big Data 4
Plappert, M., Mandery, C., Asfour, T.: The KIT motion-language dataset · 2016
Earlier work this paper cites.
ACM Transactions on Graphics (TOG) 35
Savva, M., Chang, A.X., Hanrahan, P., Fisher, M., Nießner, M.: PiGraphs: Learning Interaction Snapshots from Observations · 2016
Earlier work this paper cites.
URL https://arxiv.org/abs/1604.03692
Shu, T., Ryoo, M.S., Zhu, S.C.: Learning social affordance for human-robot interaction (2016) · 2016
Earlier work this paper cites.
In: Proceedings of European Conference on Computer Vision (ECCV) (2016)
Sridhar, S., Mueller, F., Zollhoefer, M., Casas, D., Oulasvirta, A., Theobalt, C.: Real-time joint tracking of a hand manipulating an object from rgb-d input · 2016
Earlier work this paper cites.
arXiv preprint arXiv:1609.03499 12
Van Den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A., Kavukcuoglu, K., et al.: Wavenet: A generative model for raw audio · 2016
Earlier work this paper cites.
Advances in neural information processing systems 29
Vondrick, C., Pirsiavash, H., Torralba, A.: Generating videos with scene dynamics · 2016
Earlier work this paper cites.
Advances in neural information processing systems 29
Wu, J., Zhang, C., Xue, T., Freeman, B., Tenenbaum, J.: Learning a probabilistic latent space of object shapes via 3d generative-adversarial modeling · 2016
Earlier work this paper cites.
In: Proc. Computer Vision and Pattern Recognition (CVPR), IEEE (2017)
Dai, A., Chang, A.X., Savva, M., Halber, M., Funkhouser, T., Nießner, M.: Scannet: Richly-annotated 3d reconstructions of indoor scenes · 2017
Earlier work this paper cites.
In: I. Guyon, U.V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, R. Garnett (eds.) Advances in Neural Information Processing Systems, vol. 30. Curran Associates, Inc. (2017)
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., Hochreiter, S.: Gans trained by a two time-scale update rule converge to a local nash equilibrium · 2017
Earlier work this paper cites.
Sensors 17
Merriaux, P., Dupuis, Y., Boutteau, R., Vasseur, P., Savatier, X.: A study of vicon system positioning performance · 2017
Earlier work this paper cites.
URL https://arxiv.org/abs/1612.00593
Qi, C.R., Su, H., Mo, K., Guibas, L.J.: Pointnet: Deep learning on point sets for 3d classification and segmentation (2017) · 2017
Earlier work this paper cites.
ACM Transactions on Graphics, (Proc. SIGGRAPH Asia) 36
Romero, J., Tzionas, D., Black, M.J.: Embodied hands: Modeling and capturing hands and bodies together · 2017
Earlier work this paper cites.
Advances in neural information processing systems 30
Van Den Oord, A., Vinyals, O., et al.: Neural discrete representation learning · 2017
Earlier work this paper cites.
arXiv preprint arXiv:1711.00199 (2017)
Xiang, Y., Schmidt, T., Narayanan, V., Fox, D.: Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes · 2017
Earlier work this paper cites.
In: 2018 ieee winter conference on applications of computer vision (wacv), pp. 381–389. IEEE (2018)
Chao, Y.W., Liu, Y., Liu, X., Zeng, H., Deng, J.: Learning to detect human-object interactions · 2018
Earlier work this paper cites.
URL https://arxiv.org/abs/1803.08319
Fabbri, M., Lanzi, F., Calderara, S., Palazzi, A., Vezzani, R., Cucchiara, R.: Learning to detect and track visible and occluded body joints in a virtual world (2018) · 2018
Earlier work this paper cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 409–419 (2018)
Garcia-Hernando, G., Yuan, S., Baek, S., Kim, T.K.: First-person hand action benchmark with rgb-d videos and 3d hand pose annotations · 2018
Earlier work this paper cites.
arXiv preprint arXiv:1809.04281 (2018)
Huang, C.Z.A., Vaswani, A., Uszkoreit, J., Shazeer, N., Simon, I., Hawthorne, C., Dai, A.M., Hoffman, M.D., Dinculescu, M., Eck, D.: Music transformer · 2018
Earlier work this paper cites.
In: Proceedings of the European Conference on Computer Vision (ECCV) (2018)
von Marcard, T., Henschel, R., Black, M.J., Rosenhahn, B., Pons-Moll, G.: Recovering accurate 3d human pose in the wild using imus and a moving camera · 2018
Earlier work this paper cites.
URL https://arxiv.org/abs/1712.03453
Mehta, D., Sotnychenko, O., Mueller, F., Xu, W., Sridhar, S., Pons-Moll, G., Theobalt, C.: Single-shot multi-person 3d pose estimation from monocular rgb (2018) · 2018
Earlier work this paper cites.
In: 2018 IEEE Conference on Virtual Reality and 3D User Interfaces (VR), pp. 57–64 (2018)
Mousas, C.: Performance-driven dance motion control of a virtual partner character · 2018
Earlier work this paper cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 8494–8502 (2018)
Puig, X., Ra, K., Boben, M., Li, J., Wang, T., Fidler, S., Torralba, A.: Virtualhome: Simulating household activities via programs · 2018
Earlier work this paper cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 1526–1535 (2018)
Tulyakov, S., Liu, M.Y., Yang, X., Kautz, J.: Mocogan: Decomposing motion and content for video generation · 2018
Earlier work this paper cites.
In: Intelligent Distributed Computing XII, pp. 247–258. Springer (2018)
de la Vega-Hazas, L., Calatayud, F., Iglesias, A.: Applying evolutionary computation operators for automatic human motion generation in computer animation and video games · 2018
Earlier work this paper cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 8639–8648 (2018)
Villegas, R., Yang, J., Ceylan, D., Lee, H.: Neural kinematic networks for unsupervised motion retargetting · 2018
Earlier work this paper cites.
Optics and lasers in engineering 106
Zhang, S.: High-speed 3d shape measurement with structured light methods: A review · 2018
Earlier work this paper cites.
In: Computer graphics forum, vol. 37, pp. 625–652. Wiley Online Library (2018)
Zollhöfer, M., Stotko, P., Görlitz, A., Theobalt, C., Nießner, M., Klein, R., Kolb, A.: State of the art on 3d reconstruction with rgb-d cameras · 2018
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2019)
Brahmbhatt, S., Ham, C., Kemp, C.C., Hays, J.: Contactdb: Analyzing and predicting grasp contact via thermal imaging · 2019
Earlier work this paper cites.
In: 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), pp. 2386–2393. IEEE (2019)
Brahmbhatt, S., Handa, A., Hays, J., Fox, D.: Contactgrasp: Functional multi-finger grasp synthesis from contact · 2019
Earlier work this paper cites.
Computer Vision and Image Understanding 192
Chen, Y., Tian, Y., He, M.: Monocular human pose estimation: A survey of deep learning-based methods · 2019
Earlier work this paper cites.
In: Advances in Motion Sensing and Control for Robotic Applications: Selected Papers from the Symposium on Mechatronics, Robotics, and Control (SMRC’18)-CSME International Congress 2018, May 27-30, 2018 Toronto, Canada, pp. 15–31. Springer (2019)
Furtado, J.S., Liu, H.H., Lai, G., Lacheray, H., Desouza-Coelho, J.: Comparative analysis of optitrack motion capture systems · 2019
Earlier work this paper cites.
In: International Conference on Computer Vision (2019)
Hassan, M., Choutas, V., Tzionas, D., Black, M.J.: Resolving 3D human pose ambiguities with 3D scene constraints · 2019
Earlier work this paper cites.
In: CVPR (2019)
Hasson, Y., Varol, G., Tzionas, D., Kalevatykh, I., Black, M.J., Laptev, I., Schmid, C.: Learning joint reconstruction of hands and manipulated objects · 2019
Earlier work this paper cites.
arXiv preprint arXiv:1812.04948 (2019)
Karras, T.: A style-based generator architecture for generative adversarial networks · 2019
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) (2019)
Lee, G., Deng, Z., Ma, S., Shiratori, T., Srinivasa, S.S., Sheikh, Y.: Talking with hands 16.2m: A large-scale dataset of synchronized body-finger motion and audio for conversational motion analysis and synthesis · 2019
Earlier work this paper cites.
In: IEEE Conference on Computer Vision and Pattern Recognition (2019)
Li, X., Liu, S., Kim, K., Wang, X., Yang, M.H., Kautz, J.: Putting humans in a scene: Learning affordance in 3d indoor environments · 2019
Earlier work this paper cites.
IEEE transactions on pattern analysis and machine intelligence 42
Liu, J., Shahroudy, A., Perez, M., Wang, G., Duan, L.Y., Kot, A.C.: Ntu rgb+d 120: A large-scale benchmark for 3d human activity understanding · 2019
Earlier work this paper cites.
In: International Conference on Computer Vision, pp. 5442–5451 (2019)
Mahmood, N., Ghorbani, N., Troje, N.F., Pons-Moll, G., Black, M.J.: AMASS: Archive of motion capture as surface shapes · 2019
Earlier work this paper cites.
In: 2019 IEEE International Symposium on Medical Measurements and Applications (MeMeA), pp. 1–5. IEEE (2019)
Mihcin, S., Kose, H., Cizmeciogullari, S., Ciklacandir, S., Kocak, M., Tosun, A., Akan, A.: Investigation of wearable motion capture system towards biomechanical modelling · 2019
Earlier work this paper cites.
ACM SIGGRAPH (2019)
Monszpart, A., Guerrero, P., Ceylan, D., Yumer, E., J. Mitra, N.: iMapper: Interaction-guided scene mapping from monocular videos · 2019
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp. 165–174 (2019)
Park, J.J., Florence, P., Straub, J., Newcombe, R., Lovegrove, S.: Deepsdf: Learning continuous signed distance functions for shape representation · 2019
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF international conference on computer vision, pp. 4332–4341 (2019)
Prokudin, S., Lassner, C., Romero, J.: Efficient learning on point clouds with basis point sets · 2019
Earlier work this paper cites.
IEEE Transactions on Visualization and Computer Graphics 26
Shen, Y., Yang, L., Ho, E.S.L., Shum, H.P.H.: Interaction-based human activity comparison · 2019
Earlier work this paper cites.
ACM Transactions on Graphics 38
Starke, S., Zhang, H., Komura, T., Saito, J.: Neural state machine for character-scene interactions · 2019
Earlier work this paper cites.
ACM Trans. Graph. 38
Starke, S., Zhang, H., Komura, T., Saito, J.: Neural state machine for character-scene interactions · 2019
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp. 5745–5753 (2019)
Zhou, Y., Barnes, C., Lu, J., Yang, J., Li, H.: On the continuity of rotation representations in neural networks · 2019
Earlier work this paper cites.
IEEE Robotics and Automation Letters 5
Adeli, V., Adeli, E., Reid, I., Niebles, J.C., Rezatofighi, H.: Socially and contextually aware human motion and pose forecasting · 2020
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops (2020)
Baruah, M., Banerjee, B.: A multimodal predictive agent model for human interaction generation · 2020
Earlier work this paper cites.
In: Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XIII 16, pp. 361–378. Springer (2020)
Brahmbhatt, S., Tang, C., Twigg, C.D., Kemp, C.C., Hays, J.: Contactpose: A dataset of grasps with object contact and hand pose · 2020
Earlier work this paper cites.
Advances in neural information processing systems 33
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.: Language models are few-shot learners · 2020
Earlier work this paper cites.
Cao, Z., Gao, H., Mangalam, K., Cai, Q., Vo, M., Malik, J.: Long-term human motion prediction with scene context (2020)
2020
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp. 5031–5041 (2020)
Corona, E., Pumarola, A., Alenya, G., Moreno-Noguer, F., Rogez, G.: Ganhand: Predicting human grasp affordances in multi-object scenes · 2020
Earlier work this paper cites.
In: 16th Robotics: Science and Systems, RSS 2020, Workshops (2020)
Denninger, M., Sundermeyer, M., Winkelbauer, D., Olefir, D., Hodan, T., Zidan, Y., Elbadrawy, M., Knauer, M., Katam, H., Lodhi, A.: Blenderproc: Reducing the reality gap with photorealistic rendering · 2020
Earlier work this paper cites.
arXiv preprint arXiv:2005.00341 (2020)
Dhariwal, P., Jun, H., Payne, C., Kim, J.W., Radford, A., Sutskever, I.: Jukebox: A generative model for music · 2020
Earlier work this paper cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2020)
Fieraru, M., Zanfir, M., Oneata, E., Popa, A.I., Olaru, V., Sminchisescu, C.: Three-dimensional reconstruction of human interactions · 2020
Cited alongside, same era.
Communications of the ACM 63
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., Bengio, Y.: Generative adversarial networks · 2020
Cited alongside, same era.
In: Proceedings of the 28th ACM International Conference on Multimedia, pp. 2021–2029 (2020)
Guo, C., Zuo, X., Wang, S., Zou, S., Sun, Q., Deng, A., Gong, M., Cheng, L.: Action2motion: Conditioned generation of 3d human motions · 2020
Cited alongside, same era.
IEEE transactions on pattern analysis and machine intelligence 43
Guo, Y., Wang, H., Hu, Q., Liu, H., Liu, L., Bennamoun, M.: Deep learning for 3d point clouds: A survey · 2020
Cited alongside, same era.
In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp. 3196–3206 (2020)
Hampali, S., Rad, M., Oberweger, M., Lepetit, V.: Honnotate: A method for 3d annotation of hand and object poses · 2020
In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), pp. 20609–20620 (2023)
Liu, S., Zhou, Y., Yang, J., Gupta, S., Wang, S.: Contactgen: Generative contact modeling for grasp generation · 2023
Later among the works it cites.
arXiv preprint arXiv:2312.02700 (2023)
Liu, X., Hou, H., Yang, Y., Li, Y.L., Lu, C.: Revisit human-scene interaction via space occupancy · 2023
Later among the works it cites.
Association for Computing Machinery, New York, NY, USA (2023)
Loper, M., Mahmood, N., Romero, J., Pons-Moll, G., Black, M.J.: SMPL: A Skinned Multi-Person Linear Model, 1 edn · 2023
Later among the works it cites.
URL https://arxiv.org/abs/2304.02061
Mir, A., Puig, X., Kanazawa, A., Pons-Moll, G.: Generating continual human motion in diverse 3d scenes (2023) · 2023
Later among the works it cites.
IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) (2023)
Mullen Jr, J.F., Kothandaraman, D., Bera, A., Manocha, D.: Placing human animations into 3d scenes by learning interaction- and geometry-driven keyframes · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Advances in neural information processing systems 33
Ho, J., Jain, A., Abbeel, P.: Denoising diffusion probabilistic models · 2020
Cited alongside, same era.
In: 2020 International Conference on 3D Vision (3DV 2020), pp. 333–344. IEEE, Piscataway, NJ (2020)
Karunratanakul, K., Yang, J., Zhang, Y., Black, M., Muandet, K., Tang, S.: Grasping field: Learning implicit representations for human grasps · 2020
Cited alongside, same era.
In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp. 5253–5263 (2020)
Kocabas, M., Athanasiou, N., Black, M.J.: Vibe: Video inference for human body pose and shape estimation · 2020
Cited alongside, same era.
IEEE Robotics and Automation Letters 6
Kratzer, P., Bihlmaier, S., Midlagajni, N.B., Prakash, R., Toussaint, M., Mainprice, J.: Mogaze: A dataset of full-body motions that includes workspace geometry and eye-gaze · 2020
Cited alongside, same era.
In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) (2020)
Kundu, J.N., Buckchash, H., Mandikal, P., V, R.M., Jamkhandi, A., RADHAKRISHNAN, V.B.: Cross-conditioned recurrent networks for long-term synthesis of inter-person human motion interactions · 2020
Cited alongside, same era.
In: Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XX 16, pp. 548–564. Springer (2020)
Moon, G., Yu, S.I., Wen, H., Shiratori, T., Lee, K.M.: Interhand2. 6m: A dataset and baseline for 3d interacting hand pose estimation from a single rgb image · 2020
Cited alongside, same era.
Journal of machine learning research 21
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., Liu, P.J.: Exploring the limits of transfer learning with a unified text-to-text transformer · 2020
Cited alongside, same era.
Later among the works it cites.
arXiv preprint arXiv:2312.06553 (2023)
Peng, X., Xie, Y., Wu, Z., Jampani, V., Sun, D., Jiang, H.: Hoi-diff: Text-driven synthesis of 3d human-object interactions using diffusion models · 2023
Later among the works it cites.
URL https://arxiv.org/abs/2303.01418
Shafir, Y., Tevet, G., Kapon, R., Bermano, A.H.: Human motion diffusion as a generative prior (2023) · 2023
Later among the works it cites.
Image Analysis and Stereology 42
Surmen, H.K.: Photogrammetry for 3d reconstruction of objects: Effects of geometry, texture and photographing · 2023
Later among the works it cites.
In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), pp. 15999–16009 (2023)
Tanaka, M., Fujiwara, K.: Role-aware interaction generation from textual description · 2023
Later among the works it cites.
In: 2023 IEEE/CVF International Conference on Computer Vision (ICCV), pp. 9567–9577 (2023)
Tanke, J., Zhang, L., Zhao, A., Tang, C., Cai, Y., Wang, L., Wu, P.C., Gall, J., Keskin, C.: Social diffusion: Long-term multiple human motion anticipation · 2023
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 21179–21189 (2023)
Tendulkar, P., Surís, D., Vondrick, C.: Flex: Full-body grasping without full-body grasps · 2023
Later among the works it cites.
In: The Eleventh International Conference on Learning Representations (2023)
Tevet, G., Raab, S., Gordon, B., Shafir, Y., Cohen-or, D., Bermano, A.H.: Human motion diffusion model · 2023
Later among the works it cites.
In: The Eleventh International Conference on Learning Representations (2023)
Tevet, G., Raab, S., Gordon, B., Shafir, Y., Cohen-or, D., Bermano, A.H.: Human motion diffusion model · 2023
Later among the works it cites.
arXiv preprint arXiv:2302.13971 (2023)
Touvron, H., Lavril, T., Izacard, G., Martinet, X., Lachaux, M.A., Lacroix, T., Rozière, B., Goyal, N., Hambro, E., Azhar, F., et al.: Llama: Open and efficient foundation language models · 2023
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 448–458 (2023)
Tseng, J., Castellon, R., Liu, K.: Edge: Editable dance generation from music · 2023
Later among the works it cites.
URL https://arxiv.org/abs/1706.03762
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, L., Polosukhin, I.: Attention is all you need (2023) · 2023
Later among the works it cites.
In: Thirty-seventh Conference on Neural Information Processing Systems (2023)
Wang, R., Mao, W., Li, H.: DeepsimHO: Stable pose estimation for hand-object interaction via physics simulation · 2023
Later among the works it cites.
arXiv preprint arXiv:2312.04393 (2023)
Wang, Y., Lin, J., Zeng, A., Luo, Z., Zhang, J., Zhang, L.: Physhoi: Physics-based imitation of dynamic human-object interaction · 2023
Later among the works it cites.
arXiv preprint arXiv:2310.08580 (2023)
Xie, Y., Jampani, V., Zhong, L., Sun, D., Jiang, H.: Omnicontrol: Control any joint at any time for human motion generation · 2023
Later among the works it cites.
URL https://arxiv.org/abs/2312.16051
Xu, L., Lv, X., Yan, Y., Jin, X., Wu, S., Xu, C., Liu, Y., Zhou, Y., Rao, F., Sheng, X., Liu, Y., Zeng, W., Yang, X.: Inter-x: Towards versatile human-human interaction analysis (2023) · 2023
Later among the works it cites.
In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 14928–14940 (2023)
Xu, S., Li, Z., Wang, Y.X., Gui, L.Y.: Interdiff: Generating 3d human-object interactions with physics-informed diffusion · 2023
Later among the works it cites.
In: arXiv preprint arXiv:2303.09410 (2023)
Xuan, H., Li, X., Zhang, J., Zhang, H., Liu, Y., Li, K.: Narrator: Towards natural control of human-scene interaction generation via relationship reasoning · 2023
Later among the works it cites.
ACM Computing Surveys 56
Yang, L., Zhang, Z., Song, Y., Hong, S., Xu, R., Zhao, Y., Zhang, W., Cui, B., Yang, M.H.: Diffusion models: A comprehensive survey of methods and applications · 2023
Later among the works it cites.
URL https://arxiv.org/abs/2303.15380
Yin, Y., Guo, C., Kaufmann, M., Zarate, J.J., Song, J., Hilliges, O.: Hi4d: 4d instance segmentation of close human interaction (2023) · 2023
Later among the works it cites.
In: CVPR (2023)
Zhang, J., Luo, H., Yang, H., Xu, X., Wu, Q., Shi, Y., Yu, J., Xu, L., Wang, J.: Neuraldome: A neural modeling pipeline on multi-view human-object interactions · 2023
Later among the works it cites.
In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp. 14730–14740 (2023)
Zhang, J., Zhang, Y., Cun, X., Zhang, Y., Zhao, H., Lu, H., Shen, X., Shan, Y.: Generating human motion from textual descriptions with discrete representations · 2023
Later among the works it cites.
URL https://arxiv.org/abs/2302.05543
Zhang, L., Rao, A., Agrawala, M.: Adding conditional control to text-to-image diffusion models (2023) · 2023
Later among the works it cites.
In: International conference on computer vision (ICCV) (2023)
Zhao, K., Zhang, Y., Wang, S., Beeler, T., , Tang, S.: Synthesizing diverse human motions in 3d indoor scenes · 2023
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 585–594 (2023)
Zheng, J., Zheng, Q., Fang, L., Liu, Y., Yi, L.: Cams: Canonicalized manipulation spaces for category-level functional hand-object manipulation synthesis · 2023
Later among the works it cites.
URL https://arxiv.org/abs/2307.10894
Zhu, W., Ma, X., Ro, D., Ci, H., Zhang, J., Shi, J., Gao, F., Tian, Q., Wang, Y.: Human motion generation: A survey (2023) · 2023
Later among the works it cites.
In: SIGGRAPH Asia 2024 Conference Papers, pp. 1–11 (2024)
Athanasiou, N., Cseke, A., Diomataris, M., Black, M.J., Varol, G.: Motionfix: Text-driven 3d human motion editing · 2024
Later among the works it cites.
In: International Conference on 3D Vision (3DV) (2024)
Braun, J., Christen, S., Kocabas, M., Aksan, E., Hilliges, O.: Physically plausible full-body hand-object interaction synthesis · 2024
Later among the works it cites.
In: CVPR (2024)
Cen, Z., Pi, H., Peng, S., Shen, Z., Yang, M., Shuai, Z., Bao, H., Zhou, X.: Generating human motion in 3d scenes from text descriptions · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 1577–1585 (2024)
Cha, J., Kim, J., Yoon, J.S., Baek, S.: Text2hoi: Text-guided 3d motion generation for hand-object interaction · 2024
Later among the works it cites.
arXiv preprint arXiv:2410.10790 (2024)
Chen, J., Hu, P., Chang, X., Shi, Z., Kampffmeyer, M., Liang, X.: Sitcom-crafter: A plot-driven human motion generation system in 3d scenes · 2024
Later among the works it cites.
In: SIGGRAPH Asia 2024 Conference Papers (2024)
Christen, S., Hampali, S., Sener, F., Remelli, E., Hodan, T., Sauser, E., Ma, S., Tekin, B.: Diffh2o: Diffusion-based synthesis of hand-object interactions from textual descriptions · 2024
Later among the works it cites.
Cong, P., Wang, Z., Dou, Z., Ren, Y., Yin, W., Cheng, K., Sun, Y., Long, X., Zhu, X., Ma, Y.: Laserhuman: Language-guided scene-aware human motion generation in free environment (2024)
2024
Later among the works it cites.
In: ECCV (2024)
Dai, S., Li, W., Sun, H., Huang, H., Ma, C., Huang, H., Xu, K., Hu, R.: Interfusion: Text-driven generation of 3d human-object interaction · 2024
Later among the works it cites.
In: Proc. Computer Vision and Pattern Recognition (CVPR), IEEE (2024)
Diller, C., Dai, A.: Cg-hoi: Contact-guided 3d human-object interaction generation · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2405.15763
Fan, K., Tang, J., Cao, W., Yi, R., Li, M., Gong, J., Zhang, J., Wang, Y., Wang, C., Ma, L.: Freemotion: A unified framework for number-free text-to-motion synthesis (2024) · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2405.18700
Gao, X., Yang, Y., Wu, Y., Du, S., Qi, G.J.: Multi-condition latent diffusion network for scene-aware neural human motion prediction (2024) · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2311.17057
Ghosh, A., Dabral, R., Golyanik, V., Theobalt, C., Slusallek, P.: Remos: 3d motion-conditioned reaction synthesis for two-person interactions (2024) · 2024
Later among the works it cites.
arXiv e-prints (2024)
Gong, J., Zhang, C., Liu, F., Fan, K., Zhou, Q., Tan, X., Zhang, Z., Xie, Y., Ma, L.: Diffusion implicit policy for unpaired scene-aware motion synthesis · 2024
Later among the works it cites.
In: International Conference on 3D Vision (3DV) (2024)
Guzov, V., Chibane, J., Marin, R., He, Y., Saracoglu, Y., Sattler, T., Pons-Moll, G.: Interaction replica: Tracking human–object interaction and scene changes from human motion · 2024
Later among the works it cites.
arXiv preprint arXiv:2410.10010 (2024)
Javed, M.G., Guo, C., Cheng, L., Li, X.: Intermask: 3d human interaction generation via collaborative masked modelling · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2410.03187
Jiang, N., He, Z., Wang, Z., Li, H., Chen, Y., Huang, S., Zhu, Y.: Autonomous character-scene interaction synthesis from text instruction (2024) · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 1737–1747 (2024)
Jiang, N., Zhang, Z., Li, H., Ma, X., Wang, Z., Chen, Y., Liu, T., Zhu, Y., Huang, S.: Scaling up dynamic human-scene interaction modeling · 2024
Later among the works it cites.
Kim, J., Kim, J., Na, J., Joo, H.: Parahome: Parameterizing everyday home activities towards 3d generative modeling of human-object interactions (2024)
2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 947–957 (2024)
Kulkarni, N., Rempe, D., Genova, K., Kundu, A., Johnson, J., Fouhey, D., Guibas, L.: Nifty: Neural object interaction fields for guided human motion synthesis · 2024
Later among the works it cites.
arXiv preprint arXiv:2410.13911 (2024)
Kwon, P., Joo, H.: Graspdiffusion: Synthesizing realistic whole-body hand-object interaction · 2024
Later among the works it cites.
URL http://mocap.cs.cmu.edu/
Lab, C.M.U.G.: Cmu graphics lab motion capture database (2003) · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 527–537 (2024)
Lee, J., Saito, S., Nam, G., Sung, M., Kim, T.K.: Interhandgen: Two-hand interaction generation via cascaded reverse diffusion · 2024
Later among the works it cites.
In: European Conference on Computer Vision, pp. 54–72. Springer (2024)
Li, J., Clegg, A., Mottaghi, R., Wu, J., Puig, X., Liu, C.K.: Controllable human-object interaction synthesis · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024)
Li, L., Dai, A.: GenZI: Zero-shot 3D human-scene interaction generation · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pp. 3035–3044 (2024)
Li, Q., Wang, J., Loy, C.C., Dai, B.: Task-oriented human-object interactions generation with implicit neural representations · 2024
Later among the works it cites.
International Journal of Computer Vision (2024)
Liang, H., Zhang, W., Li, W., Yu, J., Xu, L.: Intergen: Diffusion-based multi-human motion generation under complex interactions · 2024
Later among the works it cites.
In: The Twelfth International Conference on Learning Representations (2024)
Liu, X., Yi, L.: Geneoh diffusion: Towards generalizable hand-object interaction denoising via denoising diffusion · 2024
Later among the works it cites.
In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2024)
Lou, Z., Cui, Q., Wang, H., Tang, X., Zhou, H.: Multimodal sense-informed prediction of 3d human motions · 2024
Later among the works it cites.
European Conference on Computer Vision (2024)
Lv, X., Xu, L., Yan, Y., Jin, X., Xu, C., Wu, S., Liu, Y., Li, L., Bi, M., Zeng, W., et al.: Himo: A new benchmark for full-body human interacting with multiple objects · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2405.18438
Milacski, Z.A., Niinuma, K., Kawamura, R., de la Torre, F., Jeni, L.A.: Ghost: Grounded human motion generation with open vocabulary scene-and-text contexts (2024) · 2024
Later among the works it cites.
arXiv preprint arXiv:2402.06196 (2024)
Minaee, S., Mikolov, T., Nikzad, N., Chenaghlu, M., Socher, R., Amatriain, X., Gao, J.: Large language models: A survey · 2024
Later among the works it cites.
In: International Conference on 3D Vision (3DV) (2024)
Pan, L., Wang, J., Huang, B., Zhang, J., Wang, H., Tang, X., Wang, Y.: Synthesizing physically plausible human motions in 3d scenes · 2024
Later among the works it cites.
arXiv preprint arXiv:2408.16770 (2024)
Paschalidis, G., Wilschut, R., Antić, D., Taheri, O., Tzionas, D.: 3d whole-body grasp synthesis with directional controllability · 2024
Later among the works it cites.
arXiv preprint arXiv:2410.10780 (2024)
Pinyoanuntapong, E., Saleem, M.U., Karunratanakul, K., Wang, P., Xue, H., Chen, C., Guo, C., Cao, J., Ren, J., Tulyakov, S.: Controlmm: Controllable masked motion generation · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2404.09988
Ponce, P.R., Barquero, G., Palmero, C., Escalera, S., Garcia-Rodriguez, J.: in2in: Leveraging individual information to generate human interactions (2024) · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2403.14947
Qu, H., Guo, Z., Liu, J.: Gpt-connect: Interaction between text-driven human motion generator and 3d scenes in a training-free manner (2024) · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2405.18483
Shan, M., Dong, L., Han, Y., Yao, Y., Liu, T., Nwogu, I., Qi, G.J., Hill, M.: Towards open domain text-driven synthesis of multi-person motions (2024) · 2024
Later among the works it cites.
In: International Conference on 3D Vision (3DV) (2024)
Shimada, S., Mueller, F., Bednarik, J., Doosti, B., Bickel, B., Tang, D., Golyanik, V., Taylor, J., Theobalt, C., Beeler, T.: Macs: Mass conditioned 3d hand and object motion synthesis · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2403.18811
Siyao, L., Gu, T., Yang, Z., Lin, Z., Liu, Z., Ding, H., Yang, L., Loy, C.C.: Duolando: Follower gpt with off-policy reinforcement learning for dance accompaniment (2024) · 2024
Later among the works it cites.
In: International Conference on 3D Vision (3DV) (2024)
Taheri, O., Zhou, Y., Tzionas, D., Zhou, Y., Ceylan, D., Pirk, S., Black, M.J.: GRIP: Generating interaction poses using latent consistency and spatial cues · 2024
Later among the works it cites.
arXiv preprint arXiv:2404.04890 (2024)
Tang, J., Jingya, W., Ji, K., Xu, L., Yu, J., Shi, Y.: A unified diffusion framework for scene-aware human motion estimation from sparse signals · 2024
Later among the works it cites.
arXiv preprint arXiv:2410.03441 (2024)
Tevet, G., Raab, S., Cohan, S., Reda, D., Luo, Z., Peng, X.B., Bermano, A.H., van de Panne, M.: Closd: Closing the loop between simulation and diffusion for multi-task character control · 2024
Later among the works it cites.
Ugrinovic, N., Lucas, T., Baradel, F., Weinzaepfel, P., Rogez, G., Moreno-Noguer, F.: Purposer: Putting human motion generation in context (2024)
2024
Later among the works it cites.
URL https://arxiv.org/abs/2411.19921
Wang, W., Pan, L., Dou, Z., Liao, Z., Lou, Y., Yang, L., Wang, J., Komura, T.: Sims: Simulating human-scene interactions with real world script planning (2024) · 2024
Later among the works it cites.
arXiv preprint arXiv:2410.07995 (2024)
Wang, Y., Guo, C., Cheng, L., Jiang, H.: Regiongrasp: A novel task for contact region controllable hand grasp generation · 2024
Later among the works it cites.
In: European Conference on Computer Vision, pp. 467–487. Springer (2024)
Wang, Y., Wang, Z., Liu, L., Daniilidis, K.: Tram: Global trajectory and motion of 3d humans from in-the-wild videos · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024)
Wang, Z., Chen, Y., Jia, B., Li, P., Zhang, J., Zhang, J., Liu, T., Zhu, Y., Liang, W., Huang, S.: Move as you say, interact as you can: Language-guided human motion generation with scene affordance · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2311.15864
Wang, Z., Wang, J., Li, Y., Lin, D., Dai, B.: Intercontrol: Zero-shot human interaction generation by controlling every joint (2024) · 2024
Later among the works it cites.
In: The Twelfth International Conference on Learning Representations (2024)
Xiao, Z., Wang, T., Wang, J., Cao, J., Zhang, W., Dai, B., Lin, D., Pang, J.: Unified human-scene interaction via prompted chain-of-contacts · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2310.00615
Xing, C., Mao, W., Liu, M.: Scene-aware human motion forecasting via mutual distance prediction (2024) · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2403.11882
Xu, L., Zhou, Y., Yan, Y., Jin, X., Zhu, W., Rao, F., Yang, X., Zeng, W.: Regennet: Towards human action-reaction synthesis (2024) · 2024
Later among the works it cites.
The Thirty-Eighth Annual Conference on Neural Information Processing Systems (2024)
Xu, S., Wang, Z., Wang, Y.X., Gui, L.Y.: Interdreamer: Zero-shot text to 3d dynamic human-object interaction · 2024
Later among the works it cites.
European Conference on Computer Vision (2024)
Yang, J., Niu, X., Jiang, N., Zhang, R., Siyuan, H.: F-hoi: Toward fine-grained semantic-aligned 3d human-object interactions · 2024
Later among the works it cites.
In: CVPR (2024)
Ye, Y., Gupta, A., Kitani, K., Tulsiani, S.: G-hop: Generative hand-object prior for interaction reconstruction and grasp synthesis · 2024
Later among the works it cites.
Yi, H., Thies, J., Black, M.J., Peng, X.B., Rempe, D.: Generating human interaction motions in scenes with text control · 2024
Later among the works it cites.
In: 2024 International Conference on 3D Vision (3DV), pp. 235–246. IEEE (2024)
Zhang, H., Christen, S., Fan, Z., Zheng, L., Hwangbo, J., Song, J., Hilliges, O.: Artigrasp: Physically plausible synthesis of bi-manual dexterous grasping and articulation · 2024
Later among the works it cites.
URL https://arxiv.org/abs/2404.00299
Zhang, J., Zhang, J., Song, Z., Shi, Z., Zhao, C., Shi, Y., Yu, J., Xu, L., Wang, J.: Hoi-m3:capture multiple humans and objects interaction within contextual environment (2024) · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 8521–8531 (2024)
Zhang, M., Fu, Y., Ding, Z., Liu, S., Tu, Z., Wang, X.: Hoidiffusion: Generating realistic 3d hand-object interaction data · 2024
Later among the works it cites.
arxiv preprint (2024)
Zhang, X., Bhatnagar, B.L., Starke, S., Petrov, I., Guzov, V., Dhamo, H., Pérez Pellitero, E., Pons-Moll, G.: Force: Dataset and method for intuitive physics guided human-object interaction · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 729–741 (2024)
Zhao, C., Zhang, J., Du, J., Shan, Z., Wang, J., Yu, J., Wang, J., Xu, L.: I’m hoi: Inertia-aware monocular capture of 3d human-object interactions · 2024
Later among the works it cites.
In: European Conference on Computer Vision, pp. 405–421. Springer (2024)
Zhong, L., Xie, Y., Jampani, V., Sun, D., Jiang, H.: Smoodi: Stylized motion diffusion model · 2024
Later among the works it cites.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 20634–20643 (2024)
Zhou, K., Bhatnagar, B.L., Lenssen, J.E., Pons-Moll, G.: Gears: Local geometry-aware hand-object interaction synthesis · 2024
Later among the works it cites.
IEEE Transactions on Visualization and Computer Graphics (2024)
Zuo, B., Zhao, Z., Sun, W., Yuan, X., Yu, Z., Wang, Y.: Graspdiff: Grasping generation for hand-object interaction with multimodal guided diffusion · 2024
Later among the works it cites.
arXiv preprint arXiv:2501.02765 (2025)
Li, Y., Lai, Z., Bao, W., Tan, Z., Dao, A., Sui, K., Shen, J., Liu, D., Liu, H., Kong, Y.: Visual large language models for generalized and specialized applications · 2025
Closest in time.
In: International Conference on Learning Representations (ICLR) (2025)
Wang, H., Zhu, W., Miao, L., Xu, Y., Gao, F., Tian, Q., Wang, Y.: Aligning motion generation with human perceptions · 2025
Closest in time.