Fetching the paper…
Reading the bibliography…
We present InfiniDreamer, a novel framework for arbitrarily long human motion generation.
The kit motion-language dataset
Matthias Plappert, Christian Mandery, and Tamim Asfour · 2016
Earlier work this paper cites.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Earlier work this paper cites.
Language2pose: Natural language grounded pose forecasting
Chaitanya Ahuja and Louis-Philippe Morency · 2019
Earlier work this paper cites.
Decoupled weight decay regularization, 2019
Ilya Loshchilov and Frank Hutter · 2019
Earlier work this paper cites.
Amass: Archive of motion capture as surface shapes
Naureen Mahmood, Nima Ghorbani, Nikolaus F Troje, Gerard Pons-Moll, and Michael J Black · 2019
Earlier work this paper cites.
Action2motion: Conditioned generation of 3d human motions
Chuan Guo, Xinxin Zuo, Sen Wang, Shihao Zou, Qingyao Sun, Annan Deng, Minglun Gong, and Li Cheng · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Earlier work this paper cites.
Character controllers using motion vaes
Hung Yu Ling, Fabio Zinno, George Cheng, and Michiel Van De Panne · 2020
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis, 2020
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Earlier work this paper cites.
Motionet: 3d human motion reconstruction from monocular video with skeleton consistency
Mingyi Shi, Kfir Aberman, Andreas Aristidou, Taku Komura, Dani Lischinski, Daniel Cohen-Or, and Baoquan Chen · 2020
Earlier work this paper cites.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2020
Earlier work this paper cites.
Perpetual motion: Generating unbounded human motion
Yan Zhang, Michael J Black, and Siyu Tang · 2020
Earlier work this paper cites.
Improved denoising diffusion probabilistic models
Alexander Quinn Nichol and Prafulla Dhariwal · 2021
Earlier work this paper cites.
Action-conditioned 3d human motion synthesis with transformer vae
Mathis Petrovich, Michael J Black, and Gül Varol · 2021
Earlier work this paper cites.
Babel: Bodies, action and behavior with english labels
Abhinanda R Punnakkal, Arjun Chandrasekaran, Nikos Athanasiou, Alejandra Quiros-Ramirez, and Michael J Black · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision, 2021
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Earlier work this paper cites.
Teach: Temporal action composition for 3d humans
Nikos Athanasiou, Mathis Petrovich, Michael J Black, and Gül Varol · 2022
Earlier work this paper cites.
Generating diverse and natural 3d human motions from text
Chuan Guo, Shihao Zou, Xinxin Zuo, Sen Wang, Wei Ji, Xingyu Li, and Li Cheng · 2022
Earlier work this paper cites.
Temos: Generating diverse human motions from textual descriptions
Mathis Petrovich, Michael J Black, and Gül Varol · 2022
Earlier work this paper cites.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T Barron, and Ben Mildenhall · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Cited alongside, same era.
Motionclip: Exposing human motion generation to clip space
Guy Tevet, Brian Gordon, Amir Hertz, Amit H Bermano, and Daniel Cohen-Or · 2022
Cited alongside, same era.
Motiondiffuse: Text-driven human motion generation with diffusion model
Mingyuan Zhang, Zhongang Cai, Liang Pan, Fangzhou Hong, Xinying Guo, Lei Yang, and Ziwei Liu · 2022
Cited alongside, same era.
Mofusion: A framework for denoising-diffusion-based motion synthesis
Rishabh Dabral, Muhammad Hamza Mughal, Vladislav Golyanik, and Christian Theobalt · 2023
Cited alongside, same era.
Mamba: Linear-time sequence modeling with selective state spaces
Albert Gu and Tri Dao · 2023
Cited alongside, same era.
Whitenedcse: Whitening-based contrastive learning of sentence embeddings
Wenjie Zhuo, Yifan Sun, Xiaohan Wang, Linchao Zhu, and Yi Yang · 2023
Later among the works it cites.
Seamless human motion composition with blended positional encodings
German Barquero, Sergio Escalera, and Cristina Palmero · 2024
Closest in time.
M2d2m: Multi-motion generation from text with discrete diffusion models
Seunggeun Chi, Hyung-gun Chi, Hengbo Ma, Nakul Agarwal, Faizan Siddiqui, Karthik Ramani, and Kwonjoon Lee · 2024
Closest in time.
Evolved hierarchical masking for self-supervised learning
Zhanzhou Feng and Shiliang Zhang · 2024
Closest in time.
Momask: Generative masked modeling of 3d human motions
Chuan Guo, Yuxuan Mu, Muhammad Gohar Javed, Sen Wang, and Li Cheng · 2024
Closest in time.
Amd: Autoregressive motion diffusion
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Delta denoising score
Amir Hertz, Kfir Aberman, and Daniel Cohen-Or · 2023
Cited alongside, same era.
Dreamtime: An improved optimization strategy for text-to-3d content creation
Yukun Huang, Jianan Wang, Yukai Shi, Xianbiao Qi, Zheng-Jun Zha, and Lei Zhang · 2023
Cited alongside, same era.
3d gaussian splatting for real-time radiance field rendering
Bernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, and George Drettakis · 2023
Cited alongside, same era.
Multiact: Long-term 3d human motion generation from multiple action labels
Taeryung Lee, Gyeongsik Moon, and Kyoung Mu Lee · 2023
Cited alongside, same era.
Sequential texts driven cohesive motions synthesis with natural transitions
Shuai Li, Sisi Zhuang, Wenfeng Song, Xinyu Zhang, Hejia Chen, and Aimin Hao · 2023
Cited alongside, same era.
Breaking the limits of text-conditioned 3d motion synthesis with elaborative descriptions
Yijun Qian, Jack Urbanek, Alexander G Hauptmann, and Jungdam Won · 2023
Cited alongside, same era.
Story-to-motion: Synthesizing infinite and controllable character animation from long text
Zhongfei Qing, Zhongang Cai, Zhitao Yang, and Lei Yang · 2023
Cited alongside, same era.
Bo Han, Hao Peng, Minjing Dong, Yi Ren, Yixuan Shen, and Chang Xu · 2024
Closest in time.
Motiongpt: Human motion as a foreign language
Biao Jiang, Xin Chen, Wen Liu, Jingyi Yu, Gang Yu, and Tao Chen · 2024
Closest in time.
T2lm: Long-term 3d human motion generation from multiple sentences
Taeryung Lee, Fabien Baradel, Thomas Lucas, Kyoung Mu Lee, and Grégory Rogez · 2024
Closest in time.
Psychometry: An omnifit model for image reconstruction from human brain activity
Ruijie Quan, Wenguan Wang, Zhibo Tian, Fan Ma, and Yi Yang · 2024
Closest in time.
Fonts: Text rendering with typography and style controls
Wenda Shi, Yiren Song, Dengming Zhang, Jiaming Liu, and Xingxing Zou · 2024
Closest in time.
Knowledge-enhanced dual-stream zero-shot composed image retrieval
Yucheng Suo, Fan Ma, Linchao Zhu, and Yi Yang · 2024
Closest in time.
Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation
Zhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao, Chongxuan Li, Hang Su, and Jun Zhu · 2024
Closest in time.
Zero-shot video editing through adaptive sliding score distillation
Lianghan Zhu, Yanqi Bao, Jing Huo, Jing Wu, Yu-Kun Lai, Wenbin Li, and Yang Gao · 2024
Closest in time.
Vividdreamer: invariant score distillation for hyper-realistic text-to-3d generation
Wenjie Zhuo, Fan Ma, Hehe Fan, and Yi Yang · 2024
Closest in time.
Longanimation: Long animation generation with dynamic global-local memory
Nan Chen, Mengqi Huang, Yihao Meng, and Zhendong Mao · 2025
Closest in time.
Runze He, Bo Cheng, Yuhang Ma, Qingxiang Jia, Shanyuan Liu, Ao Ma, Xiaoyu Wu, Liebucha Wu, Dawei Leng, and Yuhui Yin · 2025
Closest in time.
Dynamicid: Zero-shot multi-id image personalization with flexible facial editability
Xirui Hu, Jiahao Wang, Hao Chen, Weizhan Zhang, Benqi Wang, Yikun Li, and Haishun Nan · 2025
Closest in time.
Magicid: Hybrid preference optimization for id-consistent and dynamic-preserved video customization
Hengjia Li, Lifan Jiang, Xi Xiao, Tianyang Wang, Hongwei Yi, Boxi Wu, and Deng Cai · 2025
Closest in time.
Autonomous llm-enhanced adversarial attack for text-to-motion
Honglei Miao, Fan Ma, Ruijie Quan, Kun Zhan, and Yi Yang · 2025
Closest in time.
Insert anything: Image insertion via in-context editing in dit
Wensong Song, Hong Jiang, Zongxing Yang, Ruijie Quan, and Yi Yang · 2025
Closest in time.