Fetching the paper…
Reading the bibliography…
We address the problem of action-conditioned generation of human motion sequences.
Badler, N.: Temporal scene analysis: Conceptual descriptions of object movements. In: PhD thesis, University of Toronto (1975)
1975
Earlier work this paper cites.
Badler, N.I., Phillips, C.B., Webber, B.L.: Simulating humans: Computer graphics animation and control. In: Oxford University Press (1993)
1993
Earlier work this paper cites.
Bowden, R.: Learning statistical models of human motion. In: CVPRW (2000)
2000
Earlier work this paper cites.
Herda, L., Fua, P., Plankers, R., Boulic, R., Thalmann, D.: Skeleton-based motion capture for robust reconstruction of human motion. In: Proceedings Computer Animation 2000 (2000)
2000
Earlier work this paper cites.
Galata, A., Johnson, N., Hogg, D.: Learning variable-length markov models of behavior. CVIU (2001)
2001
Earlier work this paper cites.
Agarwal, A., Triggs, B.: Recovering 3d human pose from monocular images. IEEE trans. PAMI (2005)
2005
Earlier work this paper cites.
Taylor, G.W., Hinton, G.E., Roweis, S.: Modeling human motion using binary latent variables. NeurIPS (2006)
2006
Earlier work this paper cites.
Urtasun, R., Fleet, D.J., Lawrence, N.D.: Modeling human locomotion with topologically constrained latent variable models. In: Workshop on Human Motion (2007)
2007
Earlier work this paper cites.
Jegou, H., Douze, M., Schmid, C.: Product quantization for nearest neighbor search. IEEE trans. PAMI (2010)
2010
Earlier work this paper cites.
2013
Earlier work this paper cites.
Kingma, D.P., Welling, M.: Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 (2013)
2013
Earlier work this paper cites.
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., Bengio, Y.: Generative adversarial nets. In: NeurIPS (2014)
2014
Earlier work this paper cites.
Ionescu, C., Papava, D., Olaru, V., Sminchisescu, C.: Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments. In: IEEE. trans. PAMI (2014)
2014
Earlier work this paper cites.
Rezende, D.J., Mohamed, S., Wierstra, D.: Stochastic backpropagation and approximate inference in deep, generative models. In: ICML (2014)
2014
Earlier work this paper cites.
Fragkiadaki, K., Levine, S., Felsen, P., Malik, J.: Recurrent network models for human dynamics. In: ICCV (2015)
2015
Earlier work this paper cites.
Kingma, D.P., Ba, J.: Adam: A method for stochastic optimization. In: ICLR (2015)
2015
Earlier work this paper cites.
Loper, M., Mahmood, N., Romero, J., Pons-Moll, G., Black, M.J.: Smpl: A skinned multi-person linear model. In: TOG (2015)
2015
Earlier work this paper cites.
Van den Oord, A., Kalchbrenner, N., Espeholt, L., Vinyals, O., Graves, A., et al.: Conditional image generation with pixelcnn decoders. In: NeurIPS (2016)
2016
Earlier work this paper cites.
van den Oord, A., Kalchbrenner, N., Kavukcuoglu, K.: Pixel recurrent neural networks. In: ICML (2016)
2016
Earlier work this paper cites.
Ghosh, P., Song, J., Aksan, E., Hilliges, O.: Learning human motion models for long-term predictions. In: 3DV (2017)
2017
Earlier work this paper cites.
Habibie, I., Holden, D., Schwarz, J., Yearsley, J., Komura, T.: A recurrent variational autoencoder for human motion synthesis. In: BMVC (2017)
2017
Earlier work this paper cites.
Holden, D., Komura, T., Saito, J.: Phase-functioned neural networks for character control. In: TOG (2017)
2017
Earlier work this paper cites.
Martinez, J., Black, M.J., Romero, J.: On human motion prediction using recurrent neural networks. In: CVPR (2017)
2017
Earlier work this paper cites.
Van Den Oord, A., Vinyals, O., et al.: Neural discrete representation learning. In: NeurIPS (2017)
2017
Earlier work this paper cites.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, L., Polosukhin, I.: Attention is all you need. In: NeurIPS (2017)
2017
Earlier work this paper cites.
Ahn, H., Ha, T., Choi, Y., Yoo, H., Oh, S.: Text2action: Generative adversarial synthesis from language to action. In: ICRA (2018)
2018
Cited alongside, same era.
Angela S. Lin, Lemeng Wu, R.C.K.T.Q.H.R.J.M.: Generating animated videos of human activities from natural language descriptions. In: Proceedings of the Visually Grounded Interaction and Language Workshop at NeurIPS 2018 (2018)
2018
Cited alongside, same era.
Barratt, S., Sharma, R.: A note on the inception score. arXiv preprint arXiv:1801.01973 (2018)
2018
Cited alongside, same era.
Barsoum, E., Kender, J., Liu, Z.: Hp-gan: Probabilistic 3d human motion prediction via gan. In: CVPRW (2018)
2018
Cited alongside, same era.
Chen, X., Mishra, N., Rohaninejad, M., Abbeel, P.: Pixelsnail: An improved autoregressive generative model. In: ICML (2018)
Cao, Z., Gao, H., Mangalam, K., Cai, Q.Z., Vo, M., Malik, J.: Long-term human motion prediction with scene context. In: ECCV (2020)
2020
Later among the works it cites.
Chen, M., Radford, A., Child, R., Wu, J., Jun, H., Dhariwal, P., Luan, D., Ilya, S.: Generative pretraining from pixels. In: ICML (2020)
2020
Later among the works it cites.
Guo, C., Zuo, X., Wang, S., Zou, S., Sun, Q., Deng, A., Minglun, Cheng, L.: Action2motion: Conditioned generation of 3d human motions. In: ACMMM (2020)
2020
Later among the works it cites.
Kocabas, M., Athanasiou, N., Black, M.J.: Vibe: Video inference for human body pose and shape estimation. In: CVPR (2020)
2020
Later among the works it cites.
Naeem, M.F., Oh, S.J., Uh, Y., Choi, Y., Yoo, J.: Reliable fidelity and diversity metrics for generative models. In: ICML (2020)
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Gupta, A., Johnson, J., Fei-Fei, L., Savarese, S., Alahi, A.: Social gan: Socially acceptable trajectories with generative adversarial networks. In: CVPR (2018)
2018
Cited alongside, same era.
Lin, X., Amer, M.R.: Human motion modeling using dvgans. arXiv preprint arXiv:1804.10652 (2018)
2018
Cited alongside, same era.
van den Oord, A., Oriol, V., Kavukcuoglu, K.: Neural discrete representation learning. In: ICML (2018)
2018
Cited alongside, same era.
Radford, A., Narasimhan, K., Salimans, T., Sutskever, I., et al.: Improving language understanding by generative pre-training (2018)
2018
Cited alongside, same era.
Shmelkov, K., Schmid, C., Alahari, K.: How good is my gan? In: ECCV (2018)
2018
Cited alongside, same era.
Ahuja, C., Morency, L.: Language2pose: Natural language grounded pose forecasting. In: 3DV (2019)
2019
Cited alongside, same era.
Aksan, E., Kaufmann, M., Hilliges, O.: Structured prediction helps 3d human motion modelling. In: ICCV (2019)
2019
Cited alongside, same era.
Taheri, O., Ghorbani, N., Black, M.J., Tzionas, D.: GRAB: A dataset of whole-body human grasping of objects. In: ECCV (2020)
2020
Later among the works it cites.
Weinzaepfel, P., Brégier, R., Combaluzier, H., Leroy, V., Rogez, G.: DOPE: Distillation of part experts for whole-body 3D pose estimation in the wild. In: ECCV (2020)
2020
Later among the works it cites.
Weissenborn, D., Täckström, O., Uszkoreit, J.: Scaling autoregressive video models. In: ICLR (2020)
2020
Later among the works it cites.
Yuan, Y., Kitani, K.: Dlow: Diversifying latent flows for diverse human motion prediction. In: ECCV (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
Zou, S., Zuo, X., Qian, Y., Wang, S., Xu, C., Gong, M., Cheng, L.: 3D human shape reconstruction from a polarization image. In: ECCV (2020)
2020
Later among the works it cites.
Baradel, F., Groueix, T., Weinzaepfel, P., Brégier, R., Kalantidis, Y., Rogez, G.: Leveraging mocap data for human mesh recovery. In: 3DV (2021)
2021
Later among the works it cites.
Brégier, R.: Deep regression on manifolds: a 3d rotation case study. In: 3DV (2021)
2021
Later among the works it cites.
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S.: An image is worth 16x16 words: Transformers for image recognition at scale. In: ICLR (2021)
2021
Later among the works it cites.
Esser, P., Rombach, R., Ommer, B.: Taming transformers for high-resolution image synthesis. In: CVPR (2021)
2021
Later among the works it cites.
Ghosh, A., Cheema, N., Oguz, C., Theobalt, C., Slusallek, P.: Synthesis of compositional animations from textual descriptions. In: CVPR (2021)
2021
Later among the works it cites.
2021
Later among the works it cites.
Petrovich, M., Black, M.J., Varol, G.: Action-conditioned 3d human motion synthesis with transformer vae. In: ICCV (2021)
2021
Later among the works it cites.
Punnakkal, A.R., Chandrasekaran, A., Athanasiou, N., Quiros-Ramirez, A., Black, M.J.: BABEL: Bodies, action and behavior with english labels. In: CVPR (2021)
2021
Later among the works it cites.
Rempe, D., Birdal, T., Hertzmann, A., Yang, J., Sridhar, S., Guibas, L.J.: Humor: 3d human motion model for robust pose estimation. ICCV (2021)
2021
Later among the works it cites.
2021
Later among the works it cites.
Zhang, Y., Black, M.J., Tang, S.: We are more than our joints: Predicting how 3D bodies move. In: CVPR (2021)
2021
Later among the works it cites.
Delmas, G., Weinzaepfel, P., Lucas, T., Moreno-Noguer, F., Rogez, G.: Posescript: 3d human poses from natural language. In: ECCV (2022)
2022
Closest in time.
Siyao, L., Yu, W., Gu, T., Lin, C., Wang, Q., Qian, C., Loy, C.C., Liu, Z.: Bailando: 3d dance generation by actor-critic gpt with choreographic memory. In: CVPR (2022)
2022
Closest in time.