Fetching the paper…
Reading the bibliography…
After many researchers observed fruitfulness from the recent diffusion probabilistic model, its effectiveness in image generation is actively studied these days.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
L. Sigal and M. J. Black, “Humaneva: Synchronized video and motion capture dataset for evaluation of articulated human motion,” Brown Univertsity TR , vol. 120, no. 2, 2006
2006
Earlier work this paper cites.
2013
Earlier work this paper cites.
C. Ionescu, D. Papava, V. Olaru, and C. Sminchisescu, “Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments,” Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , vol. 36, no. 7, pp. 1325–1339, jul 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
B. Zhou, X. Tang, and X. Wang, “Learning collective crowd behaviors with dynamic pedestrian-agents,” International Journal of Computer Vision (IJCV) , vol. 111, no. 1, pp. 50–68, 2015
2015
Earlier work this paper cites.
M. Loper, N. Mahmood, J. Romero, G. Pons-Moll, and M. J. Black, “Smpl: A skinned multi-person linear model,” ACM transactions on graphics (TOG) , vol. 34, no. 6, pp. 1–16, 2015
2015
Earlier work this paper cites.
K. Fragkiadaki, S. Levine, P. Felsen, and J. Malik, “Recurrent network models for human dynamics,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 4346–4354
2015
Earlier work this paper cites.
A. Jain, A. R. Zamir, S. Savarese, and A. Saxena, “Structural-rnn: Deep learning on spatio-temporal graphs,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 5308–5317
2016
Earlier work this paper cites.
J. Martinez, M. J. Black, and J. Romero, “On human motion prediction using recurrent neural networks,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 2891–2900
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems (NeurIPS) , vol. 30, 2017
2017
Earlier work this paper cites.
Y. Xu, Z. Piao, and S. Gao, “Encoding crowd interaction with deep neural network for pedestrian trajectory prediction,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 5275–5284
2018
Earlier work this paper cites.
J. Bütepage, H. Kjellström, and D. Kragic, “Anticipating many futures: Online human motion prediction and generation for human-robot interaction,” in International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 4563–4570
2018
Earlier work this paper cites.
E. Barsoum, J. Kender, and Z. Liu, “Hp-gan: Probabilistic 3d human motion prediction via gan,” in Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) , 2018, pp. 1418–1427
2018
Cited alongside, same era.
X. Yan, A. Rastogi, R. Villegas, K. Sunkavalli, E. Shechtman, S. Hadap, E. Yumer, and H. Lee, “Mt-vae: Learning motion transformations to generate multimodal human dynamics,” in European conference on computer vision (ECCV) , 2018, pp. 265–281
2018
Cited alongside, same era.
A. Rasouli, I. Kotseruba, T. Kunic, and J. K. Tsotsos, “Pie: A large-scale dataset and models for pedestrian intention estimation and trajectory prediction,” in International Conference on Computer Vision (ICCV) , 2019, pp. 6262–6271
2019
Cited alongside, same era.
N. Mahmood, N. Ghorbani, N. F. Troje, G. Pons-Moll, and M. J. Black, “Amass: Archive of motion capture as surface shapes,” in International Conference on Computer Vision (ICCV) , 2019, pp. 5442–5451
2019
Cited alongside, same era.
Y. Tashiro, J. Song, Y. Song, and S. Ermon, “Csdi: Conditional score-based diffusion models for probabilistic time series imputation,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 34, pp. 24 804–24 816, 2021
2021
Later among the works it cites.
S. Aliakbarian, F. Saleh, L. Petersson, S. Gould, and M. Salzmann, “Contextually plausible and diverse 3d human motion prediction,” in International Conference on Computer Vision (ICCV) , 2021, pp. 11 333–11 342
2021
Later among the works it cites.
V. Popov, I. Vovk, V. Gogoryan, T. Sadekova, and M. Kudinov, “Grad-tts: A diffusion probabilistic model for text-to-speech,” in International Conference on Machine Learning . PMLR, 2021, pp. 8599–8608
2021
Later among the works it cites.
T. Salimans and J. Ho, “Progressive distillation for fast sampling of diffusion models,” in International Conference on Learning Representations , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Mao, M. Liu, M. Salzmann, and H. Li, “Learning trajectory dependencies for human motion prediction,” in International Conference on Computer Vision (ICCV) , 2019, pp. 9489–9497
2019
Cited alongside, same era.
K. Kim, Y. K. Lee, H. Ahn, S. Hahn, and S. Oh, “Pedestrian intention prediction for autonomous driving using a multiple stakeholder perspective model,” in International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2020, pp. 7957–7962
2020
Cited alongside, same era.
Y. Yuan and K. Kitani, “Dlow: Diversifying latent flows for diverse human motion prediction,” in European Conference on Computer Vision (ECCV) . Springer, 2020, pp. 346–364
2020
Cited alongside, same era.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 33, pp. 6840–6851, 2020
2020
Cited alongside, same era.
W. Mao, M. Liu, and M. Salzmann, “History repeats itself: Human motion prediction via motion attention,” in European Conference on Computer Vision (ECCV) . Springer, 2020, pp. 474–489
2020
Cited alongside, same era.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial networks,” Communications of the ACM , vol. 63, no. 11, pp. 139–144, 2020
2020
Cited alongside, same era.
E. Aksan, M. Kaufmann, P. Cao, and O. Hilliges, “A spatio-temporal transformer for 3d human motion prediction,” in International Conference on 3D Vision (3DV) . IEEE, 2021, pp. 565–574
2021
Cited alongside, same era.
W. Mao, M. Liu, and M. Salzmann, “Generating smooth pose sequences for diverse human motion prediction,” in International Conference on Computer Vision (ICCV) , 2021, pp. 13 309–13 318
2021
Cited alongside, same era.
E. Valls Mascaro, S. Ma, H. Ahn, and D. Lee, “Robust human motion forcasting using transformer-based model,” in International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
C. Saharia, W. Chan, H. Chang, C. Lee, J. Ho, T. Salimans, D. Fleet, and M. Norouzi, “Palette: Image-to-image diffusion models,” in SIGGRAPH Conference Proceedings , 2022, pp. 1–10
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.