Fetching the paper…
Reading the bibliography…
Numerous methods have been proposed to detect, estimate, and analyze properties of people in images, including 3D pose, shape, contact, human-object interaction, and emotion.
A 3D face model for pose and illumination invariant face recognition
P. Paysan, R. Knothe, B. Amberg, S. Romdhani, and T. Vetter · 2009
Earlier work this paper cites.
Cognitive Load Theory
John Sweller, Paul Ayres, and Slava Kalyuga · 2011
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge J. Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick · 2014
Earlier work this paper cites.
SMPL: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J. Black · 2015
Earlier work this paper cites.
Keep it SMPL: Automatic estimation of 3D human pose and shape from a single image
Federica Bogo, Angjoo Kanazawa, Christoph Lassner, Peter Gehler, Javier Romero, and Michael J. Black · 2016
Earlier work this paper cites.
Learning a model of facial shape and expression from 4D scans
Tianye Li, Timo Bolkart, Michael. J. Black, Hao Li, and Javier Romero · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Embodied hands: Modeling and capturing hands and bodies together
Javier Romero, Dimitrios Tzionas, and Michael J. Black · 2017
Earlier work this paper cites.
Mofa: Model-based deep convolutional face autoencoder for unsupervised monocular reconstruction
Ayush Tewari, Michael Zollhofer, Hyeongwoo Kim, Pablo Garrido, Florian Bernard, Patrick Perez, and Christian Theobalt · 2017
Earlier work this paper cites.
End-to-end recovery of human shape and pose
Angjoo Kanazawa, Michael J. Black, David W. Jacobs, and Jitendra Malik · 2018
Earlier work this paper cites.
Recovering accurate 3D human pose in the wild using IMUs and a moving camera
Timo von Marcard, Roberto Henschel, Michael J. Black, Bodo Rosenhahn, and Gerard Pons-Moll · 2018
Earlier work this paper cites.
Learning character-agnostic motion for motion retargeting in 2d
Kfir Aberman, Rundi Wu, Dani Lischinski, Baoquan Chen, and Daniel Cohen-Or · 2019
Earlier work this paper cites.
Accurate 3d face reconstruction with weakly-supervised learning: From single image to image set
Yu Deng, Jiaolong Yang, Sicheng Xu, Dong Chen, Yunde Jia, and Xin Tong · 2019
Earlier work this paper cites.
Learning to reconstruct 3D human pose and shape via model-fitting in the loop
Nikos Kolotouros, Georgios Pavlakos, Michael J. Black, and Kostas Daniilidis · 2019
Earlier work this paper cites.
Pyrender
Matthew Matl · 2019
Earlier work this paper cites.
Expressive body capture: 3D hands, face, and body from a single image
Georgios Pavlakos, Vasileios Choutas, Nima Ghorbani, Timo Bolkart, Ahmed A. A. Osman, Dimitrios Tzionas, and Michael J. Black · 2019
Earlier work this paper cites.
Exemplar fine-tuning for 3D human pose fitting towards in-the-wild 3D human pose estimation
Hanbyul Joo, Natalia Neverova, and Andrea Vedaldi · 2020
Earlier work this paper cites.
Jeff Rasley, Samyam Rajbhandari, Olatunji Ruwase, and Yuxiong He · 2020
Earlier work this paper cites.
GHUM & GHUML: Generative 3D human shape and articulated pose models
Hongyi Xu, Eduard Gabriel Bazavan, Andrei Zanfir, William T. Freeman, Rahul Sukthankar, and Cristian Sminchisescu · 2020
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2021
Earlier work this paper cites.
End-to-end human pose and mesh reconstruction with transformers
Kevin Lin, Lijuan Wang, and Zicheng Liu · 2021
Earlier work this paper cites.
On self-contact and human pose
Lea Muller, Ahmed AA Osman, Siyu Tang, Chun-Hao P Huang, and Michael J Black · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Frankmocap: A monocular 3d whole-body pose estimation system via regression and integration
Yu Rong, Takaaki Shiratori, and Hanbyul Joo · 2021
Cited alongside, same era.
Pymaf: 3d human pose and shape regression with pyramidal mesh alignment feedback loop
Hongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang, Yebin Liu, Limin Wang, and Zhenan Sun · 2021
Cited alongside, same era.
Langchain, 2022
Harrison Chase and LangChain Contributors · 2022
Cited alongside, same era.
Accurate 3d body shape regression using metric and semantic attributes
Vasileios Choutas, Lea Müller, Chun-Hao P Huang, Siyu Tang, Dimitrios Tzionas, and Michael J Black · 2022
Cited alongside, same era.
Emoca: Emotion driven monocular face capture and animation
Drop: Dynamics responses from human motion prior and projective dynamics
Yifeng Jiang, Jungdam Won, Yuting Ye, and C Karen Liu · 2023
Later among the works it cites.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Object motion guided human motion synthesis
Jiaman Li, Jiajun Wu, and C Karen Liu · 2023
Later among the works it cites.
One-stage 3d whole-body mesh recovery with component aware transformer
Jing Lin, Ailing Zeng, Haoqian Wang, Lei Zhang, and Yu Li · 2023
Later among the works it cites.
GPT-4 technical report
OpenAI · 2023
Later among the works it cites.
TMR: Text-to-motion retrieval using contrastive 3D human motion synthesis
Mathis Petrovich, Michael J. Black, and Gül Varol · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Radek Daněček, Michael J Black, and Timo Bolkart · 2022
Cited alongside, same era.
Posescript: 3d human poses from natural language
Ginger Delmas, Philippe Weinzaepfel, Thomas Lucas, Francesc Moreno-Noguer, and Grégory Rogez · 2022
Cited alongside, same era.
Avatarclip: Zero-shot text-driven generation and animation of 3d avatars
Fangzhou Hong, Mingyuan Zhang, Liang Pan, Zhongang Cai, Lei Yang, and Ziwei Liu · 2022
Cited alongside, same era.
CLIFF: Carrying location information in full frames into human pose and shape estimation
Zhihao Li, Jianzhuang Liu, Zhensong Zhang, Songcen Xu, and Youliang Yan · 2022
Cited alongside, same era.
Posegpt: Quantization-based 3d human motion generation and forecasting
Thomas Lucas*, Fabien Baradel*, Philippe Weinzaepfel, and Grégory Rogez · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Cited alongside, same era.
One embedder, any task: Instruction-finetuned text embeddings
Hongjin Su, Weijia Shi, Jungo Kasai, Yizhong Wang, Yushi Hu, Mari Ostendorf, Wen-tau Yih, Noah A Smith, Luke Zettlemoyer, and Tao Yu · 2022
Cited alongside, same era.
Visual chatgpt: Talking, drawing and editing with visual foundation models
Jeff Rasley, Samyam Rajbhandari, Olatunji Ruwase, and Yuxiong He · 2023
Later among the works it cites.
Tptu: Task planning and tool usage of large language model-based ai agents
Jingqing Ruan, Yihong Chen, Bin Zhang, Zhiwei Xu, Tianpeng Bao, Hangyu Mao, Ziyue Li, Xingyu Zeng, Rui Zhao, et al · 2023
Later among the works it cites.
Vipergpt: Visual inference via python execution for reasoning
Dídac Surís, Sachit Menon, and Carl Vondrick · 2023
Later among the works it cites.
Human motion diffusion model
Guy Tevet, Sigal Raab, Brian Gordon, Yoni Shafir, Daniel Cohen-or, and Amit Haim Bermano · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Amadeusgpt: a natural language interface for interactive animal behavioral analysis
Shaokai Ye, Jessy Lauer, Mu Zhou, Alexander Mathis, and Mackenzie Mathis · 2023
Later among the works it cites.
Dreamface: Progressive generation of animatable 3d faces under text guidance, 2023
Longwen Zhang, Qiwei Qiu, Hongyang Lin, Qixuan Zhang, Cheng Shi, Wei Yang, Ye Shi, Sibei Yang, Lan Xu, and Jingyi Yu · 2023
Later among the works it cites.
Retrieving multimodal information for augmented generation: A survey
Ruochen Zhao, Hailin Chen, Weishi Wang, Fangkai Jiao, Xuan Long Do, Chengwei Qin, Bosheng Ding, Xiaobao Guo, Minzhi Li, Xingxuan Li, et al · 2023
Later among the works it cites.
Anytool: Self-reflective, hierarchical agents for large-scale api calls
Yu Du, Fangyun Wei, and Hongyang Zhang · 2024
Closest in time.
ChatPose: Chatting about 3d human pose
Yao Feng, Jing Lin, Sai Kumar Dwivedi, Yu Sun, Priyanka Patel, and Michael J. Black · 2024
Closest in time.
Wham: Reconstructing world-grounded humans with accurate 3d motion
Soyong Shin, Juyong Kim, Eni Halilaj, and Michael J. Black · 2024
Closest in time.
Mllm-tool: A multimodal large language model for tool agent learning
Chenyu Wang, Weixin Luo, Qianyu Chen, Haonan Mai, Jindi Guo, Sixun Dong, Xiaohua (Michael) Xuan, Zhengxin Li, Lin Ma, and Shenghua Gao · 2024
Closest in time.
Toolplanner: A tool augmented llm for multi granularity instructions with path planning and feedback
Qinzhuo Wu, Wei Liu, Jian Luan, and Bin Wang · 2024
Closest in time.
Amadeusgpt: a natural language interface for interactive animal behavioral analysis
Shaokai Ye, Jessy Lauer, Mu Zhou, Alexander Mathis, and Mackenzie Mathis · 2024
Closest in time.
Easytool: Enhancing llm-based agents with concise tool instruction
Siyu Yuan, Kaitao Song, Jiangjie Chen, Xu Tan, Yongliang Shen, Ren Kan, Dongsheng Li, and Deqing Yang · 2024
Closest in time.
Champ: Controllable and consistent human image animation with 3d parametric guidance, 2024
Shenhao Zhu, Junming Leo Chen, Zuozhuo Dai, Yinghui Xu, Xun Cao, Yao Yao, Hao Zhu, and Siyu Zhu · 2024
Closest in time.