Fetching the paper…
Reading the bibliography…
For many fundamental scene understanding tasks, it is difficult or impossible to obtain per-pixel ground truth labels from real images.
Efficiently approximating the minimum-volume bounding box of a point set in three dimensions
Gill Barequet and Sariel Har-Peled · 2001
Earlier work this paper cites.
Level of Detail for 3D Graphics
David Luebke, Martin Reddy, Jonathan D. Cohen, Amitabh Varshney, Benjamin Watson, and Robert Huebner · 2002
Earlier work this paper cites.
Computing the diameter of a point set
Grégoire Malandain and Jean-Daniel Boissonnat · 2002
Earlier work this paper cites.
Efficient graph-based image segmentation
Pedro F. Felzenszwalb and Daniel P. Huttenlocher · 2004
Earlier work this paper cites.
Streaming meshes
Martin Isenburg and Peter Lindstrom · 2005
Earlier work this paper cites.
Active Learning
Burr Settles · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from RGBD images
Nathan Silberman, Pushmeet Kohli, Derek Hoiem, and Rob Fergus · 2012
Earlier work this paper cites.
Indoor semantic segmentation using depth information
Camille Couprie, Clement Farabet, Laurent Najman, and Yann LeCun · 2013
Earlier work this paper cites.
Vision meets robotics: The KITTI dataset
Andreas Geiger, Philip Lenz, Christoph Stiller, and Raquel Urtasun · 2013
Earlier work this paper cites.
Perceptual organization and recognition of indoor scenes from RGB-D images
Saurabh Gupta, Pablo Arbelaez, and Jitendra Malik · 2013
Earlier work this paper cites.
Parsing IKEA objects: Fine pose estimation
Joseph J. Lim, Hamed Pirsiavash, and Antonio Torralba · 2013
Earlier work this paper cites.
SUN3D: A database of big spaces reconstructed using SfM and object labels
Jianxiong Xiao, Andrew Owens, and Antonio Torralba · 2013
Earlier work this paper cites.
Intrinsic images in the wild
Sean Bell, Kavita Bala, and Noah Snavely · 2014
Earlier work this paper cites.
A benchmark for RGB-D visual odometry, 3D reconstruction and SLAM
Ankur Handa, Thomas Whelan, John McDonald, and Andrew J. Davison · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, Lubomir Bourdev, Ross Girshick, James Hays, Pietro Perona, Deva Ramanan, C. Lawrence Zitnick, and Piotr Dollár · 2014
Earlier work this paper cites.
Beyond PASCAL: A benchmark for 3D object detection in the wild
Yu Xiang, Roozbeh Mottaghi, and Silvio Savarese · 2014
Earlier work this paper cites.
Shape, illumination, and reflectance from shading
Jonathan T. Barron and Jitendra Malik · 2015
Earlier work this paper cites.
ShapeNet: An information-rich 3D model repository
Angel X. Chang, Thomas Funkhouser, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, Jianxiong Xiao, Li Yi, and Fisher Yu · 2015
Earlier work this paper cites.
U-Net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
ImageNet large scale visual recognition challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei · 2015
Earlier work this paper cites.
SUN RGB-D: A RGB-D scene understanding benchmark suite
Shuran Song, Samuel P. Lichtenberg, and Jianxiong Xiao · 2015
Earlier work this paper cites.
3D semantic parsing of large-scale indoor spaces
Iro Armeni, Ozan Sener, Amir R. Zamir, Helen Jiang, Ioannis Brilakis, Martin Fischer, and Silvio Savarese · 2016
Earlier work this paper cites.
The Cityscapes dataset for semantic urban scene understanding
Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele · 2016
Earlier work this paper cites.
Virtual worlds as proxy for multi-object tracking analysis
Adrien Gaidon, Qiao Wang, Yohann Cabon, and Eleonora Vig · 2016
Earlier work this paper cites.
Understanding real world indoor scenes with synthetic data
Ankur Handa, Viorica Patraucean, Vijay Badrinarayanan, Simon Stent, and Roberto Cipolla · 2016
Earlier work this paper cites.
SceneNet: An annotated model generator for indoor scene understanding
Ankur Handa, Viorica Patraucean, Simon Stent, and Roberto Cipolla · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
SceneNN: A scene meshes dataset with aNNotations
Binh-Son Hua, Quang-Hieu Pham, Duc Thanh Nguyen, Minh-Khoi Tran, Lap-Fai Yu, and Sai-Kit Yeung · 2016
Cited alongside, same era.
Amodal instance segmentation
Ke Li and Jitendra Malik · 2016
Cited alongside, same era.
Playing for data: Ground truth from computer games
Stephan Richter, Vibhav Vineet, Stefan Roth, and Vladlen Koltun · 2016
Cited alongside, same era.
The SYNTHIA Dataset: A large collection of synthetic images for semantic segmentation of urban scenes
German Ros, Laura Sellart, Joanna Materzynska, David Vazquez, and Antonio M. Lopez · 2016
Cited alongside, same era.
A large scale database for 3D object recognition
Yu Xiang, Wonhui Kim, Wei Chen, Jingwei Ji, Christopher Choy, Hao Su, Roozbeh Mottaghi, Leonidas Guibas, and Silvio Savarese · 2016
Cited alongside, same era.
Joint 2D-3D-semantic data for indoor scene understanding
Iro Armeni, Alexander Sax, Amir R. Zamir, and Silvio Savarese · 2017
Configurable 3D scene synthesis and 2D image rendering with per-pixel ground truth using stochastic grammars
Chenfanfu Jiang, Siyuan Qi, Yixin Zhu, Siyuan Huang, Jenny Lin, Lap-Fai Yu, Demetri Terzopoulos, and Song-Chun Zhu · 2018
Later among the works it cites.
Free supervision from video games
Philipp Krähenbühl · 2018
Later among the works it cites.
InteriorNet: Mega-scale multi-sensor photo-realistic indoor scenes dataset
Wenbin Li, Sajad Saeedi, John McCormac, Ronald Clark, Dimos Tzoumanikas, Qing Ye, Yuzhong Huang, Rui Tang, and Stefan Leutenegger · 2018
Later among the works it cites.
CGIntrinsics: Better intrinsic image decomposition through physically-based rendering
Zhengqi Li and Noah Snavely · 2018
Later among the works it cites.
Effective use of synthetic data for urban scene semantic segmentation
Fatemeh Sadat Saleh, Mohammad Sadegh Aliakbarian, Mathieu Salzmann, Lars Petersson, and Jose M. Alvarez · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Intrinsic decompositions for image editing
Nicolas Bonneel, Balazs Kovacs, Sylvain Paris, and Kavita Bala · 2017
Cited alongside, same era.
Matterport3D: Learning from RGB-D data in indoor environments
Angel Chang, Angela Dai, Thomas Funkhouser, Maciej Halber, Matthias Niessner, Manolis Savva, Shuran Song, Andy Zeng, and Yinda Zhang · 2017
Cited alongside, same era.
ScanNet: Richly-annotated 3D reconstructions of indoor scenes
Angela Dai, Angel X. Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Niessner · 2017
Cited alongside, same era.
CARLA: An open urban driving simulator
Alexey Dosovitskiy, German Ros, Felipe Codevilla, Antonio Lopez, and Vladlen Koltun · 2017
Cited alongside, same era.
Learning where to look: Data-driven viewpoint set selection for 3D scenes
Kyle Genova, Manolis Savva, Angel X. Chang, and Thomas Funkhouser · 2017
Cited alongside, same era.
AI2-THOR: An interactive 3D environment for visual AI
Eric Kolve, Roozbeh Mottaghi, Winson Han, Eli VanderBilt, Luca Weihs, Alvaro Herrasti, Daniel Gordon, Yuke Zhu, Abhinav Gupta, and Ali Farhadi · 2017
Cited alongside, same era.
Synscapes: A photorealistic synthetic dataset for street scene parsing
Magnus Wrenninge and Jonas Unger · 2018
Later among the works it cites.
Building generalizable agents with a realistic and rich 3D environment
Yi Wu, Yuxin Wu, Georgia Gkioxari, and Yuandong Tian · 2018
Later among the works it cites.
Gibson Env: Real-world perception for embodied agents
Fei Xia, Amir R. Zamir, Zhi-Yang He, Alexander Sax, Jitendra Malik, and Silvio Savarese · 2018
Later among the works it cites.
Mesh R-CNN
Georgia Gkioxari, Jitendra Malik, and Justin Johnson · 2019
Later among the works it cites.
Precise synthetic image and LiDAR (PreSIL) dataset for autonomous vehicle perception
Braden Hurl, Krzysztof Czarnecki, and Steven Waslander · 2019
Later among the works it cites.
ProcSy: Procedural synthetic dataset generation towards influence factor studies of semantic segmentation networks
Samin Khan, Buu Phan, Rick Salay, and Krzysztof Czarnecki · 2019
Later among the works it cites.
Furnishing your room by what you see: An end-to-end furniture set retrieval framework with rich annotated benchmark dataset
Bingyuan Liu, Jiantao Zhang, Xiaoting Zhang, Wei Zhang, Chuanhui Yu, and Yuan Zhou · 2019
Later among the works it cites.
Synthetic data for deep learning
Sergey I. Nikolenko · 2019
Later among the works it cites.
Habitat: A platform for embodied AI research
Manolis Savva, Abhishek Kadian, Oleksandr Maksymets, Yili Zhao, Erik Wijmans, Bhavana Jain, Julian Straub, Jia Liu, Vladlen Koltun, Jitendra Malik, Devi Parikh, and Dhruv Batra · 2019
Later among the works it cites.
Neural inverse rendering of an indoor scene from a single image
Soumyadip Sengupta, Jinwei Gu, Kihwan Kim, Guilin Liu, David W. Jacobs, and Jan Kautz · 2019
Later among the works it cites.
Megatron-LM: Training multi-billion parameter language models using model parallelism
Mohammad Shoeybi, Mostofa Patwary, Raul Puri, Patrick LeGresley, Jared Casper, and Bryan Catanzaro · 2019
Later among the works it cites.
Which tasks should be learned together in multi-task learning?
Trevor Standley, Amir R. Zamir, Dawn Chen, Leonidas Guibas, Jitendra Malik, and Silvio Savarese · 2019
Later among the works it cites.
The Replica Dataset: A digital replica of indoor spaces
Julian Straub, Thomas Whelan Lingni Ma, Yufan Chen, Erik Wijmans, Simon Green, Jakob J. Engel, Raul Mur-Artal, Carl Ren, Shobhit Verma, Anton Clarkson, Mingfei Yan, Brian Budge, Yajie Yan, Xiaqing Pan, June Yon, Yuyang Zou, Kimberly Leon, Nigel Carter, Jesus Briales, Tyler Gillingham, Elias Mueggler, Luis Pesqueira, Manolis Savva, Dhruv Batra, Hauke M. Strasdat, Renzo De Nardi, Michael Goesele, Steven Lovegrove, and Richard Newcombe · 2019
Later among the works it cites.
DIODE: A Dense Indoor and Outdoor DEpth Dataset
Igor Vasiljevic, Nick Kolkin, Shanyi Zhang, Ruotian Luo, Haochen Wang, Falcon Z. Dai, Andrea F. Daniele, Mohammadreza Mostajabi, Steven Basart, Matthew R. Walter, and Gregory Shakhnarovich · 2019
Later among the works it cites.
IRS: A large synthetic indoor robotics stereo dataset for disparity and surface normal estimation
Qiang Wang, Shizhen Zheng, Qingsong Yan, Fei Deng, Kaiyong Zhao, and Xiaowen Chu · 2019
Later among the works it cites.
Structured3D: A large photo-realistic dataset for structured 3D modeling
Jia Zheng, Junfei Zhang, Jing Li, Rui Tang, Shenghua Gao, and Zihan Zhou · 2019
Later among the works it cites.
Semantic understanding of scenes through ADE20K dataset
Bolei Zhou, Hang Zhao, Xavier Puig, Tete Xiao, Sanja Fidler, Adela Barriuso, and Antonio Torralba · 2019
Later among the works it cites.
3D-FUTURE: 3D furniture shape with TextURE
Huan Fu, Rongfei Jia, Lin Gao, Mingming Gong, Binqiang Zhao, Steve Maybank, and Dacheng Tao · 2020
Closest in time.
Geometric structure based and regularized depth estimation from 360 degree indoor imagery
Lei Jin, Yanyu Xu, Jia Zheng, Junfei Zhang, Rui Tang, Shugong Xu, Jingyi Yu, and Shenghua Gao · 2020
Closest in time.
Inverse rendering for complex indoor scenes: Shape, spatially-varying lighting and SVBRDF from a single image
Zhengqin Li, Mohammad Shafiei, Ravi Ramamoorthi, Kalyan Sunkavalli, and Manmohan Chandraker · 2020
Closest in time.
OpenRooms: An end-to-end open framework for photorealistic indoor scene datasets
Zhengqin Li, Ting-Wei Yu, Shen Sang, Sarah Wang, Sai Bi, Zexiang Xu, Hong-Xing Yu, Kalyan Sunkavalli, Miloš Hašan, Ravi Ramamoorthi, and Manmohan Chandraker · 2020
Closest in time.
TartanAir: A dataset to push the limits of visual SLAM
Wenshan Wang, Delong Zhu, Xiangwei Wang, Yaoyu Hu, Yuheng Qiu, Chen Wang, Yafei Hu, Ashish Kapoor, and Sebastian Scherer · 2020
Closest in time.