Fetching the paper…
Reading the bibliography…
Zero-shot 6D object pose estimation involves the detection of novel objects with their 6D poses in cluttered scenes, presenting significant challenges for model generalizability.
Shapenet: An information-rich 3d model repository
Angel X Chang, Thomas Funkhouser, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, et al · 2015
Earlier work this paper cites.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick · 2017
Earlier work this paper cites.
An efficient algebraic solution to the perspective-three-point problem
Tong Ke and Stergios I Roumeliotis · 2017
Earlier work this paper cites.
Pointnet++: Deep hierarchical feature learning on point sets in a metric space
Charles Ruizhongtai Qi, Li Yi, Hao Su, and Leonidas J Guibas · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes
Yu Xiang, Tanner Schmidt, Venkatraman Narayanan, and Dieter Fox · 2017
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Earlier work this paper cites.
Pvn3d: A deep point-wise 3d keypoints voting network for 6dof pose estimation
Yisheng He, Wei Sun, Haibin Huang, Jianran Liu, Haoqiang Fan, and Jian Sun · 2020
Earlier work this paper cites.
Transformers are rnns: Fast autoregressive transformers with linear attention
Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, and François Fleuret · 2020
Earlier work this paper cites.
Cosypose: Consistent multi-view multi-object 6d pose estimation
Yann Labbé, Justin Carpentier, Mathieu Aubry, and Josef Sivic · 2020
Earlier work this paper cites.
Shape prior deformation for categorical 6d object pose and size estimation
Meng Tian, Marcelo H Ang, and Gim Hee Lee · 2020
Earlier work this paper cites.
Sgpa: Structure-guided prior adaptation for category-level 6d object pose estimation
Kai Chen and Qi Dou · 2021
Earlier work this paper cites.
Ffb6d: A full flow bidirectional fusion network for 6d pose estimation
Yisheng He, Haibin Huang, Haoqiang Fan, Qifeng Chen, and Jian Sun · 2021
Earlier work this paper cites.
Predator: Registration of 3d point clouds with low overlap
Shengyu Huang, Zan Gojcic, Mikhail Usvyatsov, Andreas Wieser, and Konrad Schindler · 2021
Earlier work this paper cites.
Zephyr: Zero-shot pose hypothesis rating
Brian Okorn, Qiao Gu, Martial Hebert, and David Held · 2021
Cited alongside, same era.
Loftr: Detector-free local feature matching with transformers
Jiaming Sun, Zehong Shen, Yuang Wang, Hujun Bao, and Xiaowei Zhou · 2021
Cited alongside, same era.
Gdr-net: Geometry-guided direct regression network for monocular 6d object pose estimation
Gu Wang, Fabian Manhardt, Federico Tombari, and Xiangyang Ji · 2021
Cited alongside, same era.
Ove6d: Object viewpoint encoding for depth-based 6d object pose estimation
Dingding Cai, Janne Heikkilä, and Esa Rahtu · 2022
Cited alongside, same era.
Google scanned objects: A high-quality dataset of 3d scanned household items
Laura Downs, Anthony Francis, Nate Koenig, Brandon Kinman, Ryan Hickman, Krista Reymann, Thomas B McHugh, and Vincent Vanhoucke · 2022
Cited alongside, same era.
Zero-shot category-level object pose estimation
Pope: 6-dof promptable pose estimation of any object, in any scene, with one reference
Zhiwen Fan, Panwang Pan, Peihao Wang, Yifan Jiang, Dejia Xu, Hanwen Jiang, and Zhangyang Wang · 2023
Closest in time.
Scalable mask annotation for video text spotting
Haibin He, Jing Zhang, Mengyang Xu, Juhua Liu, Bo Du, and Dacheng Tao · 2023
Closest in time.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Closest in time.
Vi-net: Boosting category-level 6d object pose estimation via learning decoupled rotations on the spherical representations
Jiehong Lin, Zewei Wei, Yabin Zhang, and Kui Jia · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Walter Goodwin, Sagar Vaze, Ioannis Havoutis, and Ingmar Posner · 2022
Cited alongside, same era.
Surfemb: Dense and continuous correspondence distributions for object pose estimation with learnt surface embeddings
Rasmus Laurvig Haugaard and Anders Glent Buch · 2022
Cited alongside, same era.
Megapose: 6d pose estimation of novel objects via render & compare
Yann Labbé, Lucas Manuelli, Arsalan Mousavian, Stephen Tyree, Stan Birchfield, Jonathan Tremblay, Justin Carpentier, Mathieu Aubry, Dieter Fox, and Josef Sivic · 2022
Cited alongside, same era.
Category-level 6d object pose and size estimation using self-supervised deep prior deformation networks
Jiehong Lin, Zewei Wei, Changxing Ding, and Kui Jia · 2022
Cited alongside, same era.
Gen6d: Generalizable model-free 6-dof object pose estimation from rgb images
Yuan Liu, Yilin Wen, Sida Peng, Cheng Lin, Xiaoxiao Long, Taku Komura, and Wenping Wang · 2022
Cited alongside, same era.
Templates for 3d object pose estimation revisited: Generalization to new objects and robustness to occlusions
Van Nguyen Nguyen, Yinlin Hu, Yang Xiao, Mathieu Salzmann, and Vincent Lepetit · 2022
Cited alongside, same era.
Geometric transformer for fast and robust point cloud registration
Zheng Qin, Hao Yu, Changjian Wang, Yulan Guo, Yuxing Peng, and Kai Xu · 2022
Cited alongside, same era.
Jun Ma and Bo Wang · 2023
Closest in time.
Segment anything model for medical image analysis: an experimental study
Maciej A Mazurowski, Haoyu Dong, Hanxue Gu, Jichen Yang, Nicholas Konz, and Yixin Zhang · 2023
Closest in time.
Panwang Pan, Zhiwen Fan, Brandon Y Feng, Peihao Wang, Chenxin Li, and Zhangyang Wang · 2023
Closest in time.
Anything-3d: Towards single-view anything reconstruction in the wild
Qiuhong Shen, Xingyi Yang, and Xinchao Wang · 2023
Closest in time.
Bop challenge 2022 on detection, segmentation and pose estimation of specific rigid objects
Martin Sundermeyer, Tomáš Hodaň, Yann Labbe, Gu Wang, Eric Brachmann, Bertram Drost, Carsten Rother, and Jiří Matas · 2023
Closest in time.
Can sam segment anything? when sam meets camouflaged object detection
Lv Tang, Haoke Xiao, and Bo Li · 2023
Closest in time.
Edit everything: A text-guided generative system for images editing
Defeng Xie, Ruichen Wang, Jian Ma, Chen Chen, Haonan Lu, Dong Yang, Fobo Shi, and Xiaodong Lin · 2023
Closest in time.
Inpaint anything: Segment anything meets image inpainting
Tao Yu, Runseng Feng, Ruoyu Feng, Jinming Liu, Xin Jin, Wenjun Zeng, and Zhibo Chen · 2023
Closest in time.
Xu Zhao, Wenchao Ding, Yongqi An, Yinglong Du, Tao Yu, Min Li, Ming Tang, and Jinqiao Wang · 2023
Closest in time.
Gigapose: Fast and robust novel object pose estimation via one correspondence
Van Nguyen Nguyen, Thibault Groueix, Mathieu Salzmann, and Vincent Lepetit · 2024
Closest in time.