Fetching the paper…
Reading the bibliography…
Unsupervised learning of keypoints and landmarks has seen significant progress with the help of modern neural network architectures, but performance is yet to match the supervised counterpart, making their practicability questionable.
Distinctive Image Features from Scale-Invariant Keypoints
David G. Lowe · 2004
Earlier work this paper cites.
The Caltech-UCSD Birds-200-2011 Dataset
Catherine Wah, Steve Branson, Peter Welinder, Pietro Perona, and Serge Belongie · 2011
Earlier work this paper cites.
Human3. 6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2013
Earlier work this paper cites.
Understanding image representations by measuring their equivariance and equivalence
Karel Lenc and Andrea Vedaldi · 2015
Earlier work this paper cites.
Deep learning face attributes in the wild
Ziwei Liu, Ping Luo, Xiaogang Wang, and Xiaoou Tang · 2015
Earlier work this paper cites.
Deepfashion: Powering robust clothes recognition and retrieval with rich annotations
Ziwei Liu, Ping Luo, Shi Qiu, Xiaogang Wang, and Xiaoou Tang · 2016
Earlier work this paper cites.
Realtime multi-person 2d pose estimation using part affinity fields
Zhe Cao, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2017
Earlier work this paper cites.
Revisiting unreasonable effectiveness of data in deep learning era
Chen Sun, Abhinav Shrivastava, Saurabh Singh, and Abhinav Gupta · 2017
Earlier work this paper cites.
Unsupervised learning of object landmarks by factorized spatial embeddings
James Thewlis, Hakan Bilen, and Andrea Vedaldi · 2017
Earlier work this paper cites.
Deep feature factorization for concept discovery
Edo Collins, Radhakrishna Achanta, and Sabine Susstrunk · 2018
Earlier work this paper cites.
Superpoint: Self-supervised interest point detection and description
Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2018
Earlier work this paper cites.
Unsupervised learning of object landmarks through conditional image generation
Tomas Jakab, Ankush Gupta, Hakan Bilen, and Andrea Vedaldi · 2018
Earlier work this paper cites.
Personlab: Person pose estimation and instance segmentation with a bottom-up, part-based, geometric embedding model
George Papandreou, Tyler Zhu, Liang-Chieh Chen, Spyros Gidaris, Jonathan Tompson, and Kevin Murphy · 2018
Earlier work this paper cites.
Unsupervised discovery of object landmarks as structural representations
Yuting Zhang, Yijie Guo, Yixin Jin, Yijun Luo, Zhiyuan He, and Honglak Lee · 2018
Earlier work this paper cites.
Scops: Self-supervised co-part segmentation
Wei-Chih Hung, Varun Jampani, Sifei Liu, Pavlo Molchanov, Ming-Hsuan Yang, and Jan Kautz · 2019
Earlier work this paper cites.
Unsupervised part-based disentangling of object shape and appearance
Dominik Lorenz, Leonard Bereska, Timo Milbich, and Bjorn Ommer · 2019
Earlier work this paper cites.
Facial landmark detection: A literature survey
Yue Wu and Qiang Ji · 2019
Earlier work this paper cites.
A survey on hand pose estimation with wearable sensors and computer-vision-based methods
Weiya Chen, Chenchen Yu, Chenyu Tu, Zehua Lyu, Jing Tang, Shiqi Ou, Yan Fu, and Zhidong Xue · 2020
Earlier work this paper cites.
Self-supervised learning of interpretable keypoints from unlabelled videos
Tomas Jakab, Ankush Gupta, Hakan Bilen, and Andrea Vedaldi · 2020
Earlier work this paper cites.
Deep keypoint-based camera pose estimation with geometric constraints
You-Yi Jau, Rui Zhu, Hao Su, and Manmohan Chandraker · 2020
Earlier work this paper cites.
Label-efficient semantic segmentation with diffusion models
Dmitry Baranchuk, Ivan Rubachev, Andrey Voynov, Valentin Khrulkov, and Artem Babenko · 2021
Earlier work this paper cites.
Unsupervised part discovery from contrastive reconstruction
Subhabrata Choudhury, Iro Laina, Christian Rupprecht, and Andrea Vedaldi · 2021
Cited alongside, same era.
Latentkeypointgan: Controlling gans via latent keypoints
Xingzhe He, Bastian Wandt, and Helge Rhodin · 2021
Cited alongside, same era.
Improved denoising diffusion probabilistic models
Alexander Quinn Nichol and Prafulla Dhariwal · 2021
Cited alongside, same era.
Unsupervised human pose estimation through transforming shape templates
Luca Schmidtke, Athanasios Vlontzos, Simon Ellershaw, Anna Lukens, Tomoki Arichi, and Bernhard Kainz · 2021
Cited alongside, same era.
Motion-supervised co-part segmentation
Aliaksandr Siarohin, Subhankar Roy, Stéphane Lathuilière, Sergey Tulyakov, Elisa Ricci, and Nicu Sebe · 2021
Cited alongside, same era.
Diffusiondet: Diffusion model for object detection
Shoufa Chen, Peize Sun, Yibing Song, and Ping Luo · 2022
Zoomnas: searching for whole-body human pose estimation in the wild
Lumin Xu, Sheng Jin, Wentao Liu, Chen Qian, Wanli Ouyang, Ping Luo, and Xiaogang Wang · 2022
Later among the works it cites.
Self-supervised part segmentation via motion imitation
Yanping Zhang, Qiaokang Liang, Kunlin Zou, Zhengwei Li, Wei Sun, and Yaonan Wang · 2022
Later among the works it cites.
Synthetic data from diffusion models improves imagenet classification
Shekoofeh Azizi, Simon Kornblith, Chitwan Saharia, Mohammad Norouzi, and David J Fleet · 2023
Closest in time.
Text-to-image diffusion models are zero-shot classifiers
Kevin Clark and Priyank Jaini · 2023
Closest in time.
Unsupervised semantic correspondence using stable diffusion
Eric Hedlin, Gopal Sharma, Shweta Mahajan, Hossam Isack, Abhishek Kar, Andrea Tagliasacchi, and Kwang Moo Yi · 2023
Closest in time.
68 landmarks are efficient for 3d face alignment: what about more? 3d face alignment method applied to face recognition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Cited alongside, same era.
Alphapose: Whole-body regional multi-person pose estimation and tracking in real-time, 2022
Hao-Shu Fang, Jiefeng Li, Hongyang Tang, Chao Xu, Haoyi Zhu, Yuliang Xiu, Yong-Lu Li, and Cewu Lu · 2022
Cited alongside, same era.
Animal pose estimation: A closer look at the state-of-the-art, existing gaps and opportunities
Le Jiang, Caleb Lee, Divyang Teotia, and Sarah Ostadabbas · 2022
Cited alongside, same era.
Tusk: Task-agnostic unsupervised keypoints
Yuhe Jin, Weiwei Sun, Jan Hosang, Eduard Trulls, and Kwang Moo Yi · 2022
Cited alongside, same era.
Recent advances of monocular 2d and 3d human pose estimation: A deep learning perspective
Wu Liu, Qian Bao, Yu Sun, and Tao Mei · 2022
Cited alongside, same era.
Null-text inversion for editing real images using guided diffusion models
Ron Mokady, Amir Hertz, Kfir Aberman, Yael Pritch, and Daniel Cohen-Or · 2022
Cited alongside, same era.
Marwa Jabberi, Ali Wali, Bidyut Baran Chaudhuri, and Adel M Alimi · 2023
Closest in time.
Slime: Segment like me
Aliasghar Khani, Saeid Asgari Taghanaki, Aditya Sanghi, Ali Mahdavi Amiri, and Ghassan Hamarneh · 2023
Closest in time.
Magic3d: High-resolution text-to-3d content creation
Chen-Hsuan Lin, Jun Gao, Luming Tang, Towaki Takikawa, Xiaohui Zeng, Xun Huang, Karsten Kreis, Sanja Fidler, Ming-Yu Liu, and Tsung-Yi Lin · 2023
Closest in time.
Dynamic 3d gaussians: Tracking by persistent dynamic view synthesis
Jonathon Luiten, Georgios Kopanas, Bastian Leibe, and Deva Ramanan · 2023
Closest in time.
Diffusion hyperfeatures: Searching through time and space for semantic correspondence
Grace Luo, Lisa Dunlap, Dong Huk Park, Aleksander Holynski, and Trevor Darrell · 2023
Closest in time.
6d object position estimation from 2d images: a literature review
Giorgia Marullo, Leonardo Tanzi, Pietro Piazzolla, and Enrico Vezzetti · 2023
Closest in time.
Latent-nerf for shape-guided generation of 3d shapes and textures
Gal Metzer, Elad Richardson, Or Patashnik, Raja Giryes, and Daniel Cohen-Or · 2023
Closest in time.
Gpt-4 technical report, 2023
OpenAI · 2023
Closest in time.
Emergent correspondence from image diffusion
Luming Tang, Menglin Jia, Qianqian Wang, Cheng Perng Phoo, and Bharath Hariharan · 2023
Closest in time.
Diffuse, attend, and segment: Unsupervised zero-shot segmentation using stable diffusion
Junjiao Tian, Lavisha Aggarwal, Andrea Colaco, Zsolt Kira, and Mar Gonzalez-Franco · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
Diffumask: Synthesizing images with pixel-level annotations for semantic segmentation using diffusion models
Weijia Wu, Yuzhong Zhao, Mike Zheng Shou, Hong Zhou, and Chunhua Shen · 2023
Closest in time.
From text to mask: Localizing entities using the attention of text-to-image diffusion models
Changming Xiao, Qi Yang, Feng Zhou, and Changshui Zhang · 2023
Closest in time.
Open-vocabulary panoptic segmentation with text-to-image diffusion models
Jiarui Xu, Sifei Liu, Arash Vahdat, Wonmin Byeon, Xiaolong Wang, and Shalini De Mello · 2023
Closest in time.
A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence
Junyi Zhang, Charles Herrmann, Junhwa Hur, Luisa Polania Cabrera, Varun Jampani, Deqing Sun, and Ming-Hsuan Yang · 2023
Closest in time.
Deep learning-based human pose estimation: A survey
Ce Zheng, Wenhan Wu, Chen Chen, Taojiannan Yang, Sijie Zhu, Ju Shen, Nasser Kehtarnavaz, and Mubarak Shah · 2023
Closest in time.