Fetching the paper…
Reading the bibliography…
Text-to-motion diffusion models can generate realistic animations from text prompts, but do not support fine-grained motion editing controls.
Motion warping
Andrew P. Witkin and Zoran Popovic. 1995 · 1995
Earlier work this paper cites.
Motion Editing with Spacetime Constraints. In
Michael Gleicher. 1997 · 1997
Earlier work this paper cites.
A Hierarchical Approach to Interactive Motion Editing for Human-like Figures. In
Jehee Lee and Sung Yong Shin. 1999 · 1999
Earlier work this paper cites.
Motion Path Editing. In
Michael Gleicher. 2001 · 2001
Earlier work this paper cites.
Iterative Training of Dynamic Skills Inspired by Human Coaching Techniques
Sehoon Ha and C. Karen Liu. 2015 · 2015
Earlier work this paper cites.
Keep it SMPL: Automatic Estimation of 3D Human Pose and Shape from a Single Image. In
Federica Bogo, Angjoo Kanazawa, Christoph Lassner, Peter Gehler, Javier Romero, and Michael J. Black. 2016 · 2016
Earlier work this paper cites.
Deep Motifs and Motion Signatures
Andreas Aristidou, Daniel Cohen-Or, Jessica K. Hodgins, Yiorgos Chrysanthou, and Ariel Shamir. 2018 · 2018
Earlier work this paper cites.
AMASS: Archive of Motion Capture as Surface Shapes. In
Naureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll, and Michael J. Black. 2019 · 2019
Earlier work this paper cites.
Unpaired Motion Style Transfer from Video to Animation
Kfir Aberman, Yijia Weng, Dani Lischinski, Daniel Cohen-Or, and Baoquan Chen. 2020 · 2020
Earlier work this paper cites.
fairmotion - Tools to load, process and visualize motion capture data
Deepak Gopinath and Jungdam Won. 2020 · 2020
Earlier work this paper cites.
Denoising Diffusion Probabilistic Models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
Hierarchical Motion Understanding via Motion Programs. In
Sumith Kulal, Jiayuan Mao, Alex Aiken, and Jiajun Wu. 2021 · 2021
Earlier work this paper cites.
AI Choreographer: Music Conditioned 3D Dance Generation with AIST++
Ruilong Li, Shan Yang, David A. Ross, and Angjoo Kanazawa. 2021 · 2021
Earlier work this paper cites.
VideoGPT: Video Generation using VQ-VAE and Transformers
Wilson Yan, Yunzhi Zhang, Pieter Abbeel, and Aravind Srinivas. 2021 · 2021
Earlier work this paper cites.
Blended Diffusion for Text-Driven Editing of Natural Images. In
Omri Avrahami, Dani Lischinski, and Ohad Fried. 2022 · 2022
Earlier work this paper cites.
Generating Diverse and Natural 3D Human Motions From Text. In
Chuan Guo, Shihao Zou, Xinxin Zuo, Sen Wang, Wei Ji, Xingyu Li, and Li Cheng. 2022 · 2022
Earlier work this paper cites.
Prompt-to-prompt image editing with cross attention control
Amir Hertz, Ron Mokady, Jay Tenenbaum, Kfir Aberman, Yael Pritch, and Daniel Cohen-Or. 2022 · 2022
Earlier work this paper cites.
Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents
Wenlong Huang, Pieter Abbeel, Deepak Pathak, and Igor Mordatch. 2022 · 2022
Cited alongside, same era.
SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations
Chenlin Meng, Yutong He, Yang Song, Jiaming Song, Jiajun Wu, Jun-Yan Zhu, and Stefano Ermon. 2022 · 2022
Cited alongside, same era.
ProtoRes: Proto-Residual Network for Pose Authoring via Learned Inverse Kinematics. In
Boris N. Oreshkin, Florent Bocquelet, Félix G. Harvey, Bay Raitt, and Dominic Laflamme. 2022 · 2022
Cited alongside, same era.
Motion In-Betweening via Two-Stage Transformers
Jia Qin, Youyi Zheng, and Kun Zhou. 2022 · 2022
Cited alongside, same era.
EDGE: Editable Dance Generation From Music
Jonathan Tseng, Rodrigo Castellon, and C Karen Liu. 2022 · 2022
Object motion guided human motion synthesis
Jiaman Li, Jiajun Wu, and C Karen Liu. 2023 · 2023
Closest in time.
Code as Policies: Language Model Programs for Embodied Control. In
Jacky Liang, Wenlong Huang, Fei Xia, Peng Xu, Karol Hausman, Brian Ichter, Pete Florence, and Andy Zeng. 2023 · 2023
Closest in time.
Generative Proxemics: A Prior for 3D Social Interaction from Images
Lea Müller, Vickie Ye, Georgios Pavlakos, Michael Black, and Angjoo Kanazawa. 2023 · 2023
Closest in time.
Trace and Pace: Controllable Pedestrian Animation via Guided Trajectory Diffusion. In
Davis Rempe, Zhengyi Luo, Xue Bin Peng, Ye Yuan, Kris Kitani, Karsten Kreis, Sanja Fidler, and Or Litany. 2023 · 2023
Closest in time.
InsActor: Instruction-driven Physics-based Characters
Jiawei Ren, Mingyuan Zhang, Cunjun Yu, Xiao Ma, Liang Pan, and Ziwei Liu. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model
Mingyuan Zhang, Zhongang Cai, Liang Pan, Fangzhou Hong, Xinying Guo, Lei Yang, and Ziwei Liu. 2022 · 2022
Cited alongside, same era.
Unpredictable Black Boxes are Terrible Interfaces
Maneesh Agrawala. 2023 · 2023
Cited alongside, same era.
InstructPix2Pix: Learning to Follow Image Editing Instructions. In
Tim Brooks, Aleksander Holynski, and Alexei A. Efros. 2023 · 2023
Cited alongside, same era.
PoseFix: Correcting 3D Human Poses with Natural Language. In
Ginger Delmas, Philippe Weinzaepfel, Francesc Moreno-Noguer, and Grégory Rogez. 2023 · 2023
Cited alongside, same era.
Motion Question Answering via Modular Motion Programs
Mark Endo, Joy Hsu, Jiaman Li, and Jiajun Wu. 2023 · 2023
Cited alongside, same era.
PoseGPT: Chatting about 3D Human Pose
Yao Feng, Jing Lin, Sai Kumar Dwivedi, Yu Sun, Priyanka Patel, and Michael J. Black. 2023 · 2023
Cited alongside, same era.
MultiModal-GPT: A Vision and Language Model for Dialogue with Humans
Tao Gong, Chengqi Lyu, Shilong Zhang, Yudong Wang, Miao Zheng, Qian Zhao, Kuikun Liu, Wenwei Zhang, Ping Luo, and Kai Chen. 2023 · 2023
Cited alongside, same era.
Vishnu Sarukkai, Linden Li, Arden Ma, Christopher Ré, and Kayvon Fatahalian. 2023 · 2023
Closest in time.
Human motion diffusion as a generative prior
Yonatan Shafir, Guy Tevet, Roy Kapon, and Amit H Bermano. 2023 · 2023
Closest in time.
Reflexion: Language Agents with Verbal Reinforcement Learning
Noah Shinn, Federico Cassano, Edward Berman, Ashwin Gopinath, Karthik Narasimhan, and Shunyu Yao. 2023 · 2023
Closest in time.
ProgPrompt: Generating Situated Robot Task Plans using Large Language Models. In
Ishika Singh, Valts Blukis, Arsalan Mousavian, Ankit Goyal, Danfei Xu, Jonathan Tremblay, Dieter Fox, Jesse Thomason, and Animesh Garg. 2023 · 2023
Closest in time.
ViperGPT: Visual Inference via Python Execution for Reasoning
Dídac Surís, Sachit Menon, and Carl Vondrick. 2023 · 2023
Closest in time.
Human Motion Diffusion Model. In
Guy Tevet, Sigal Raab, Brian Gordon, Yoni Shafir, Daniel Cohen-or, and Amit Haim Bermano. 2023 · 2023
Closest in time.
Understanding Text-driven Motion Synthesis with Keyframe Collaboration via Diffusion Models
Dong Wei, Xiaoning Sun, Huaijiang Sun, Bin Li, Sheng liang Hu, Weiqing Li, and Jian-Zhou Lu. 2023a · 2023
Closest in time.
ReAct: Synergizing Reasoning and Acting in Language Models
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao. 2023 · 2023
Closest in time.
Dance Style Transfer with Cross-modal Transformer. In
Wenjie Yin, Hang Yin, Kim Baraka, Danica Kragic, and Mårten Björkman. 2023 · 2023
Closest in time.
FineMoGen: Fine-Grained Spatio-Temporal Motion Generation and Editing
Mingyuan Zhang, Huirong Li, Zhongang Cai, Jiawei Ren, Lei Yang, and Ziwei Liu. 2023 · 2023
Closest in time.
Back to optimization: Diffusion-based zero-shot 3d human pose estimation. In
Zhongyu Jiang, Zhuoran Zhou, Lei Li, Wenhao Chai, Cheng-Yen Yang, and Jenq-Neng Hwang. 2024 · 2024
Closest in time.