Fetching the paper…
Reading the bibliography…
While object reconstruction has made great strides in recent years, current methods typically require densely captured images and/or known camera poses, and generalize poorly to novel object categories.
Distinctive image features from scale-invariant keypoints
David G Lowe · 2004
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, Alan Conrad Bovik, Hamid R. Sheikh, and Eero P. Simoncelli · 2004
Earlier work this paper cites.
A comparison and evaluation of multi-view stereo reconstruction algorithms
Steven M. Seitz, Brian Curless, James Diebel, Daniel Scharstein, and Richard Szeliski · 2006
Earlier work this paper cites.
Parallel tracking and mapping on a camera phone
Georg S. W. Klein and David William Murray · 2009
Earlier work this paper cites.
Beyond pascal: A benchmark for 3d object detection in the wild
Yu Xiang, Roozbeh Mottaghi, and Silvio Savarese · 2014
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
Angel X. Chang, Thomas A. Funkhouser, Leonidas J. Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, Jianxiong Xiao, L. Yi, and Fisher Yu · 2015
Earlier work this paper cites.
Delving deeper into convolutional networks for learning video representations
Nicolas Ballas, L. Yao, Christopher Joseph Pal, and Aaron C. Courville · 2016
Earlier work this paper cites.
3d-r2n2: A unified approach for single and multi-view 3d object reconstruction
Christopher Bongsoo Choy, Danfei Xu, JunYoung Gwak, Kevin Chen, and Silvio Savarese · 2016
Earlier work this paper cites.
Learning a predictable and generative vector representation for objects
Rohit Girdhar, David F. Fouhey, Mikel D. Rodriguez, and Abhinav Kumar Gupta · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, X. Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Perceptual losses for real-time style transfer and super-resolution
Justin Johnson, Alexandre Alahi, and Li Fei-Fei · 2016
Earlier work this paper cites.
Openmvg: Open multiple view geometry
Pierre Moulon, Pascal Monasse, Romuald Perrot, and Renaud Marlet · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L. Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Perspective transformer nets: Learning single-view 3d object reconstruction without 3d supervision
Xinchen Yan, Jimei Yang, Ersin Yumer, Yijie Guo, and Honglak Lee · 2016
Earlier work this paper cites.
Learning a multi-view stereo machine
Abhishek Kar, Christian Häne, and Jitendra Malik · 2017
Earlier work this paper cites.
Feature pyramid networks for object detection
Tsung-Yi Lin, Piotr Dollár, Ross B. Girshick, Kaiming He, Bharath Hariharan, and Serge J. Belongie · 2017
Earlier work this paper cites.
Octnet: Learning deep 3d representations at high resolutions
Gernot Riegler, Ali O. Ulusoy, and Andreas Geiger · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam M. Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
A survey of augmented, virtual, and mixed reality for cultural heritage
Mafkereseb Kassahun Bekele, Roberto Pierdicca, Emanuele Frontoni, Eva Savina Malinverni, and James Gain · 2018
Earlier work this paper cites.
Codeslam - learning a compact, optimisable representation for dense visual slam
Michael Bloesch, Jan Czarnowski, Ronald Clark, Stefan Leutenegger, and Andrew J. Davison · 2018
Earlier work this paper cites.
Learning category-specific mesh reconstruction from image collections
Angjoo Kanazawa, Shubham Tulsiani, Alexei A. Efros, and Jitendra Malik · 2018
Earlier work this paper cites.
Factoring shape, pose, and layout from the 2d image of a 3d scene
Shubham Tulsiani, Saurabh Gupta, David F. Fouhey, Alexei A. Efros, and Jitendra Malik · 2018
Earlier work this paper cites.
Learning 6-dof grasping interaction via deep geometry-aware 3d representations
Xinchen Yan, Jasmined Hsu, Mohammad Khansari, Yunfei Bai, Arkanath Pathak, Abhinav Gupta, James Davidson, and Honglak Lee · 2018
Earlier work this paper cites.
Unet++: A nested u-net architecture for medical image segmentation
Zongwei Zhou, Md Mahfuzur Rahman Siddiquee, Nima Tajbakhsh, and Jianming Liang · 2018
Earlier work this paper cites.
3d-relnet: Joint object and relational network for 3d prediction
Nilesh Kulkarni, Ishan Misra, Shubham Tulsiani, and Abhinav Kumar Gupta · 2019
Cited alongside, same era.
Habitat: A Platform for Embodied AI Research
Manolis Savva*, Abhishek Kadian*, Oleksandr Maksymets*, Yili Zhao, Erik Wijmans, Bhavana Jain, Julian Straub, Jia Liu, Vladlen Koltun, Jitendra Malik, Devi Parikh, and Dhruv Batra · 2019
Cited alongside, same era.
Occupancy networks: Learning 3d reconstruction in function space
Lars M. Mescheder, Michael Oechsle, Michael Niemeyer, Sebastian Nowozin, and Andreas Geiger · 2019
Cited alongside, same era.
Hologan: Unsupervised learning of 3d representations from natural images
Thu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt, and Yong-Liang Yang · 2019
Cited alongside, same era.
Deepsdf: Learning continuous signed distance functions for shape representation
Jeong Joon Park, Peter R. Florence, Julian Straub, Richard A. Newcombe, and S. Lovegrove · 2019
Cited alongside, same era.
Barf: Bundle-adjusting neural radiance fields
Chen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, and Simon Lucey · 2021
Later among the works it cites.
Nerf in the wild: Neural radiance fields for unconstrained photo collections
Ricardo Martin-Brualla, Noha Radwan, Mehdi S. M. Sajjadi, Jonathan T. Barron, Alexey Dosovitskiy, and Daniel Duckworth · 2021
Later among the works it cites.
Giraffe: Representing scenes as compositional generative neural feature fields
Michael Niemeyer and Andreas Geiger · 2021
Later among the works it cites.
Unisurf: Unifying neural implicit surfaces and radiance fields for multi-view reconstruction
Michael Oechsle, Songyou Peng, and Andreas Geiger · 2021
Later among the works it cites.
Nerfies: Deformable neural radiance fields
Keunhong Park, U. Sinha, Jonathan T. Barron, Sofien Bouaziz, Dan B. Goldman, Steven M. Seitz, and Ricardo Martin-Brualla · 2021
Later among the works it cites.
D-nerf: Neural radiance fields for dynamic scenes
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
From photo to 3d to mixed reality: A complete workflow for cultural heritage visualisation and experience
Hafizur Rahaman, Erik Champion, and Mafkereseb Bekele · 2019
Cited alongside, same era.
Scene representation networks: Continuous 3d-structure-aware neural scene representations
Vincent Sitzmann, Michael Zollhoefer, and Gordon Wetzstein · 2019
Cited alongside, same era.
Ba-net: Dense bundle adjustment network
Chengzhou Tang and Ping Tan · 2019
Cited alongside, same era.
Multi-view supervision for single-view reconstruction via differentiable ray consistency
Shubham Tulsiani, Tinghui Zhou, Alyosha A. Efros, and Jitendra Malik · 2019
Cited alongside, same era.
Learning spatial common sense with geometry-aware recurrent networks
Hsiao-Yu Fish Tung, Ricson Cheng, and Katerina Fragkiadaki · 2019
Cited alongside, same era.
Normalized object coordinate space for category-level 6d object pose and size estimation
He Wang, Srinath Sridhar, Jingwei Huang, Julien P. C. Valentin, Shuran Song, and Leonidas J. Guibas · 2019
Cited alongside, same era.
Pointrend: Image segmentation as rendering
Alexander Kirillov, Yuxin Wu, Kaiming He, and Ross B. Girshick · 2020
Cited alongside, same era.
Albert Pumarola, Enric Corona, Gerard Pons-Moll, and Francesc Moreno-Noguer · 2021
Later among the works it cites.
Common objects in 3d: Large-scale learning and evaluation of real-life 3d category reconstruction
Jeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone, Patrick Labatut, and David Novotný · 2021
Later among the works it cites.
Habitat 2.0: Training home assistants to rearrange their habitat
Andrew Szot, Alex Clegg, Eric Undersander, Erik Wijmans, Yili Zhao, John Turner, Noah Maestre, Mustafa Mukadam, Devendra Chaplot, Oleksandr Maksymets, Aaron Gokaslan, Vladimir Vondrus, Sameer Dharur, Franziska Meier, Wojciech Galuba, Angel Chang, Zsolt Kira, Vladlen Koltun, Jitendra Malik, Manolis Savva, and Dhruv Batra · 2021
Later among the works it cites.
Neus: Learning neural implicit surfaces by volume rendering for multi-view reconstruction
Peng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt, Taku Komura, and Wenping Wang · 2021
Later among the works it cites.
Ibrnet: Learning multi-view image-based rendering
Qianqian Wang, Zhicheng Wang, Kyle Genova, Pratul P. Srinivasan, Howard Zhou, Jonathan T. Barron, Ricardo Martin-Brualla, Noah Snavely, and Thomas A. Funkhouser · 2021
Later among the works it cites.
Shelf-supervised mesh prediction in the wild
Yufei Ye, Shubham Tulsiani, and Abhinav Kumar Gupta · 2021
Later among the works it cites.
pixelnerf: Neural radiance fields from one or few images
Alex Yu, Vickie Ye, Matthew Tancik, and Angjoo Kanazawa · 2021
Later among the works it cites.
Ners: Neural reflectance surfaces for sparse-view 3d reconstruction in the wild
Jason Y. Zhang, Gengshan Yang, Shubham Tulsiani, and Deva Ramanan · 2021
Later among the works it cites.
Google scanned objects: A high-quality dataset of 3d scanned household items
Laura Downs, Anthony Francis, Nate Koenig, Brandon Kinman, Ryan Michael Hickman, Krista Reymann, Thomas Barlow McHugh, and Vincent Vanhoucke · 2022
Closest in time.
Kubric: A scalable dataset generator
Klaus Greff, Francois Belletti, Lucas Beyer, Carl Doersch, Yilun Du, Daniel Duckworth, David Fleet, Dan Gnanapragasam, Florian Golemo, Charles Herrmann, Thomas Kipf, Abhijit Kundu, Dmitry Lagun, Issam H. Laradji, Hsueh-Ti Liu, Henning Meyer, Yishu Miao, Derek Nowrouzezahrai, Cengiz Oztireli, Etienne Pot, Noha Radwan, Daniel Rebain, Sara Sabour, Mehdi S. M. Sajjadi, Matan Sela, Vincent Sitzmann, Austin Stone, Deqing Sun, Suhani Vora, Ziyu Wang, Tianhao Wu, Kwang Moo Yi, Fangcheng Zhong, and Andrea Tagliasacchi · 2022
Closest in time.
Gen6d: Generalizable model-free 6-dof object pose estimation from rgb images
Yuan Liu, Yilin Wen, Sida Peng, Chu-Hsing Lin, Xiaoxiao Long, Taku Komura, and Wenping Wang · 2022
Closest in time.
Regnerf: Regularizing neural radiance fields for view synthesis from sparse inputs
Michael Niemeyer, Jonathan T. Barron, Ben Mildenhall, Mehdi S. M. Sajjadi, Andreas Geiger, and Noha Radwan · 2022
Closest in time.
The 8-point algorithm as an inductive bias for relative pose prediction by vits
Chris Rockwell, Justin Johnson, and David F. Fouhey · 2022
Closest in time.
Scene Representation Transformer: Geometry-Free Novel View Synthesis Through Set-Latent Scene Representations
Mehdi S. M. Sajjadi, Henning Meyer, Etienne Pot, Urs Bergmann, Klaus Greff, Noha Radwan, Suhani Vora, Mario Lucic, Daniel Duckworth, Alexey Dosovitskiy, Jakob Uszkoreit, Thomas Funkhouser, and Andrea Tagliasacchi · 2022
Closest in time.
Fvor: Robust joint shape and pose optimization for few-view object reconstruction
Zhenpei Yang, Zhile Ren, Miguel Angel Bautista, Zaiwei Zhang, Qi Shan, and Qixing Huang · 2022
Closest in time.
RelPose: Predicting probabilistic relative rotation for single objects in the wild
Jason Y. Zhang, Deva Ramanan, and Shubham Tulsiani · 2022
Closest in time.
Nerfusion: Fusing radiance fields for large-scale scene reconstruction
Xiaoshuai Zhang, Sai Bi, Kalyan Sunkavalli, Hao Su, and Zexiang Xu · 2022
Closest in time.
Nice-slam: Neural implicit scalable encoding for slam
Zihan Zhu, Songyou Peng, Viktor Larsson, Weiwei Xu, Hujun Bao, Zhaopeng Cui, Martin R. Oswald, and Marc Pollefeys · 2022
Closest in time.
Tong Wu, Jiarui Zhang, Xiao Fu, Yuxin Wang, Jiawei Ren, Liang Pan, Wayne Wu, Lei Yang, Jiaqi Wang, Chen Qian, Dahua Lin, and Ziwei Liu · 2023
Closest in time.