Fetching the paper…
Reading the bibliography…
We seek to extract a temporally consistent 6D pose trajectory of a manipulated object from an Internet instructional video.
Graspit!: A Versatile Simulator for Robotic Grasping
Andrew T Miller and Peter K Allen · 2004
Earlier work this paper cites.
EPnP: An Accurate O(n) Solution to the PnP Problem
Vincent Lepetit, Francesc Moreno-Noguer, and Pascal Fua · 2009
Earlier work this paper cites.
Learning 6D Object Pose Estimation using 3D Object Coordinates
Eric Brachmann, Alexander Krull, Frank Michel, Stefan Gumhold, Jamie Shotton, and Carsten Rother · 2014
Earlier work this paper cites.
Semantic Pose using Deep Networks Trained on Synthetic RGB-D
Jeremie Papon and Markus Schoeler · 2015
Earlier work this paper cites.
Pose-RCNN: Joint Object Detection and Pose Estimation Using 3D Object Proposals
Markus Braun, Qing Rao, Yikang Wang, and Fabian Flohr · 2016
Earlier work this paper cites.
Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2016
Earlier work this paper cites.
Structure-from-Motion Revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Mask R-CNN
Kaiming He, Georgia Gkioxari, Piotr Dollar, and Ross Girshick · 2017
Earlier work this paper cites.
T-LESS: An RGB-D Dataset for 6D Pose Estimation of Texture-less Objects
Tomáš Hodan, Pavel Haluza, Štepán Obdržálek, Jiri Matas, Manolis Lourakis, and Xenophon Zabulis · 2017
Earlier work this paper cites.
Time-Contrastive Networks: Self-Supervised Learning from Video
Pierre Sermanet, Corey Lynch, Yevgen Chebotar, Jasmine Hsu, Eric Jang, Stefan Schaal, Sergey Levine, and Google Brain · 2018
Earlier work this paper cites.
A micro Lie theory for state estimation in robotics
Joan Sola, Jeremie Deray, and Dinesh Atchuthan · 2018
Earlier work this paper cites.
PoseCNN: A Convolutional Neural Network for 6D Object Pose Estimation in Cluttered Scenes
Yu Xiang, Tanner Schmidt, Venkatraman Narayanan, and Dieter Fox · 2018
Earlier work this paper cites.
The Pinocchio C++ library – A fast and flexible implementation of rigid body dynamics algorithms and their analytical derivatives
Justin Carpentier, Guilhem Saurel, Gabriele Buondonno, Joseph Mirabel, Florent Lamiraux, Olivier Stasse, and Nicolas Mansard · 2019
Earlier work this paper cites.
Learning joint reconstruction of hands and manipulated objects
Yana Hasson, Gul Varol, Dimitrios Tzionas, Igor Kalevatykh, Michael J Black, Ivan Laptev, and Cordelia Schmid · 2019
Earlier work this paper cites.
HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips
Antoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi, Ivan Laptev, and Josef Sivic · 2019
Earlier work this paper cites.
Normalized Object Coordinate Space for Category-Level 6D Object Pose and Size Estimation
He Wang, Srinath Sridhar, Jingwei Huang, Julien Valentin, Shuran Song, and Leonidas J Guibas · 2019
Earlier work this paper cites.
GraspNet-1Billion: A Large-Scale Benchmark for General Object Grasping
Hao-Shu Fang, Chenxi Wang, Minghao Gou, and Cewu Lu · 2020
Earlier work this paper cites.
BOP Challenge 2020 on 6D Object Localization
Tomas Hodan, Martin Sundermeyer, Bertram Drost, Yann Labbé, Eric Brachmann, Frank Michel, Carsten Rother, and Jiri Matas · 2020
Earlier work this paper cites.
CosyPose: Consistent multi-view multi-object 6d pose estimation
Yann Labbé, Justin Carpentier, Mathieu Aubry, and Josef Sivic · 2020
Earlier work this paper cites.
DeepIM: Deep Iterative Matching for 6D Pose Estimation
Yi Li, Gu Wang, Xiangyang Ji, Yu Xiang, and Dieter Fox · 2020
Earlier work this paper cites.
Multiview Neural Surface Reconstruction by Disentangling Geometry and Appearance
Lior Yariv, Yoni Kasten, Dror Moran, Meirav Galun, Matan Atzmon, Basri Ronen, and Yaron Lipman · 2020
Earlier work this paper cites.
Perceiving 3D Human-Object Spatial Arrangements from a Single Image in the Wild
Jason Y Zhang, Sam Pepose, Hanbyul Joo, Deva Ramanan, Jitendra Malik, and Angjoo Kanazawa · 2020
Earlier work this paper cites.
Reconstructing Hand-Object Interactions in the Wild
Zhe Cao, Ilija Radosavovic, Angjoo Kanazawa, and Jitendra Malik · 2021
Earlier work this paper cites.
Multi-view Fusion for Multi-level Robotic Scene Understanding
Yunzhi Lin, Jonathan Tremblay, Stephen Tyree, Patricio A. Vela, and Stan Birchfield · 2021
Earlier work this paper cites.
Learning Object Manipulation Skills via Approximate State Estimation from Real Videos
Vladimír Petrík, Makarand Tapaswi, Ivan Laptev, and Josef Sivic · 2021
Earlier work this paper cites.
GDR-Net: Geometry-Guided Direct Regression Network for Monocular 6D Object Pose Estimation
Gu Wang, Fabian Manhardt, Federico Tombari, and Xiangyang Ji · 2021
Cited alongside, same era.
BundleTrack: 6D Pose Tracking for Novel Objects without Instance or Category-Level 3D Models
Bowen Wen and Kostas Bekris · 2021
Cited alongside, same era.
Learning to Manipulate Tools by Aligning Simulation to Video Demonstration
Kateryna Zorina, Justin Carpentier, Josef Sivic, and Vladimír Petrík · 2021
Cited alongside, same era.
Super-Fibonacci Spirals: Fast, Low-Discrepancy Sampling of SO(3)
Marc Alexa · 2022
Cited alongside, same era.
Deep ViT Features as Dense Visual Descriptors
Shir Amir, Yossi Gandelsman, Shai Bagon, and Tali Dekel · 2022
Cited alongside, same era.
Rescaling Egocentric Vision: Collection, Pipeline and Challenges for EPIC-KITCHENS-100
Segment Anything in High Quality
Lei Ke, Mingqiao Ye, Martin Danelljan, Yifan Liu, Yu-Wing Tai, Chi-Keung Tang, and Fisher Yu · 2023
Later among the works it cites.
Segment Anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C. Berg, Wan-Yen Lo, Piotr Dollar, and Ross Girshick · 2023
Later among the works it cites.
Are These the Same Apple? Comparing Images Based on Object Intrinsics
Klemen Kotar, Stephen Tian, Hong-Xing Yu, Daniel L.K. Yamins, and Jiajun Wu · 2023
Later among the works it cites.
Zero-1-to-3: Zero-shot One Image to 3D Object
Ruoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov, Sergey Zakharov, and Carl Vondrick · 2023
Later among the works it cites.
CNOS: A Strong Baseline for CAD-based Novel Object Segmentation
Van Nguyen Nguyen, Thibault Groueix, Georgy Ponimatkin, Vincent Lepetit, and Tomas Hodan · 2023
Later among the works it cites.
BundleSDF: Neural 6-DoF Tracking and 3D Reconstruction of Unknown Objects
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Antonino Furnari, Jian Ma, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, and Michael Wray · 2022
Cited alongside, same era.
Google Scanned Objects: A High-Quality Dataset of 3D Scanned Household Items
Laura Downs, Anthony Francis, Nate Koenig, Brandon Kinman, Ryan Hickman, Krista Reymann, Thomas B McHugh, and Vincent Vanhoucke · 2022
Cited alongside, same era.
ShAPO: Implicit Representations for Multi-Object Shape, Appearance, and Pose Optimization
Muhammad Zubair Irshad, Sergey Zakharov, Rares Ambrus, Thomas Kollar, Zsolt Kira, and Adrien Gaidon · 2022
Cited alongside, same era.
MegaPose: 6D Pose Estimation of Novel Objects via Render & Compare
Yann Labbé, Lucas Manuelli, Arsalan Mousavian, Stephen Tyree, Stan Birchfield, Jonathan Tremblay, Justin Carpentier, Mathieu Aubry, Dieter Fox, and Josef Sivic · 2022
Cited alongside, same era.
Estimating 3D Motion and Forces of Human-Object Interactions from Internet Videos
Zongmian Li, Jiri Sedlar, Justin Carpentier, Ivan Laptev, Nicolas Mansard, and Josef Sivic · 2022
Cited alongside, same era.
Templates for 3D Object Pose Estimation Revisited: Generalization to New Objects and Robustness to Occlusions
Van Nguyen Nguyen, Yinlin Hu, Yang Xiao, Mathieu Salzmann, and Vincent Lepetit · 2022
Cited alongside, same era.
Learning to Imitate Object Interactions from Internet Videos
Austin Patel, Andrew Wang, Ilija Radosavovic, and Jitendra Malik · 2022
Cited alongside, same era.
Bowen Wen, Jonathan Tremblay, Valts Blukis, Stephen Tyree, Thomas Müller, Alex Evans, Dieter Fox, Jan Kautz, and Stan Birchfield · 2023
Later among the works it cites.
Xu Zhao, Wenchao Ding, Yongqi An, Yinglong Du, Tao Yu, Min Li, Ming Tang, and Jinqiao Wang · 2023
Later among the works it cites.
Gemma 2: Improving Open Language Models at a Practical Size
Gemma Team · 2024
Later among the works it cites.
BOP Challenge 2023 on Detection, Segmentation and Pose Estimation of Seen and Unseen Rigid Objects
Tomas Hodan, Martin Sundermeyer, Yann Labbe, Van Nguyen Nguyen, Gu Wang, Eric Brachmann, Bertram Drost, Vincent Lepetit, Carsten Rother, and Jiri Matas · 2024
Later among the works it cites.
CoTracker: It is Better to Track Together
Nikita Karaev, Ignacio Rocco, Benjamin Graham, Natalia Neverova, Andrea Vedaldi, and Christian Rupprecht · 2024
Later among the works it cites.
Llama Team · 2024
Later among the works it cites.
Wonder3D: Single Image to 3D using Cross-Domain Diffusion
Xiaoxiao Long, Yuan-Chen Guo, Cheng Lin, Yuan Liu, Zhiyang Dou, Lingjie Liu, Yuexin Ma, Song-Hai Zhang, Marc Habermann, Christian Theobalt, and Wenping Wang · 2024
Later among the works it cites.
Adapting Pre-Trained Vision Models for Novel Instance Detection and Segmentation
Yangxiao Lu, Yunhui Guo, Nicholas Ruozzi, Yu Xiang, et al · 2024
Later among the works it cites.
Towards Generalist Robot Learning from Internet Video: A Survey
Robert McCarthy, Daniel C.H. Tan, Dominik Schmidt, Fernando Acero, Nathan Herr, Yilun Du, Thomas G Thuruthel, and Zhibin Li · 2024
Later among the works it cites.
GigaPose: Fast and Robust Novel Object Pose Estimation via One Correspondence
Van Nguyen Nguyen, Thibault Groueix, Mathieu Salzmann, and Vincent Lepetit · 2024
Later among the works it cites.
GPT-4 Technical Report, 2024
OpenAI · 2024
Later among the works it cites.
DINOv2: Learning robust visual features without supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy V. Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, Mido Assran, Nicolas Ballas, Wojciech Galuba, Russell Howes, Po-Yao Huang, Shang-Wen Li, Ishan Misra, Michael Rabbat, Vasu Sharma, Gabriel Synnaeve, Hu Xu, Herve Jegou, Julien Mairal, Patrick Labatut, Armand Joulin, and Piotr Bojanowski · 2024
Later among the works it cites.
FoundPose: Unseen Object Pose Estimation with Foundation Features
Evin Pınar Örnek, Yann Labbé, Bugra Tekin, Lingni Ma, Cem Keskin, Christian Forster, and Tomas Hodan · 2024
Later among the works it cites.
SAM 2: Segment Anything in Images and Videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloe Rolland, Laura Gustafson, Eric Mintun, Junting Pan, Kalyan Vasudev Alwala, Nicolas Carion, Chao-Yuan Wu, Ross Girshick, Piotr Dollár, and Christoph Feichtenhofer · 2024
Later among the works it cites.
CRISP: Object Pose and Shape Estimation with Test-Time Adaptation
Jingnan Shi, Rajat Talak, Harry Zhang, David Jin, and Luca Carlone · 2024
Later among the works it cites.
Splatter Image: Ultra-Fast Single-View 3D Reconstruction
Stanislaw Szymanowicz, Christian Rupprecht, and Andrea Vedaldi · 2024
Later among the works it cites.
DUSt3R: Geometric 3D Vision Made Easy
Shuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii, and Jerome Revaud · 2024
Later among the works it cites.
FoundationPose: Unified 6D Pose Estimation and Tracking of Novel Objects
Bowen Wen, Wei Yang, Jan Kautz, and Stan Birchfield · 2024
Later among the works it cites.
Reconstructing Hand-Held Objects in 3D
Jane Wu, Georgios Pavlakos, Georgia Gkioxari, and Jitendra Malik · 2024
Later among the works it cites.