Fetching the paper…
Reading the bibliography…
Visual 2.5D perception involves understanding the semantics and geometry of a scene through reasoning about object relationships with respect to the viewer in an environment.
Human vision and cognitive science
BJ Rogers RJ Watt · 1989
Earlier work this paper cites.
Model-based recognition of 3d objects from single images
Isaac Weiss and Manjit Ray · 2001
Earlier work this paper cites.
Learning depth from single monocular images
Ashutosh Saxena, Sung H Chung, and Andrew Y Ng · 2006
Earlier work this paper cites.
Labelme: a database and web-based tool for image annotation
Bryan C Russell, Antonio Torralba, Kevin P Murphy, and William T Freeman · 2008
Earlier work this paper cites.
Make3d: Learning 3d scene structure from a single still image
Ashutosh Saxena, Min Sun, and Andrew Y Ng · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Observing human-object interactions: Using spatial and functional compatibility for recognition
Abhinav Gupta, Aniruddha Kembhavi, and Larry S Davis · 2009
Earlier work this paper cites.
Single image depth estimation from predicted semantic labels
Beyang Liu, Stephen Gould, and Daphne Koller · 2010
Earlier work this paper cites.
Grouplet: A structured image representation for recognizing human and object interactions
Bangpeng Yao and Li Fei-Fei · 2010
Earlier work this paper cites.
Learning person-object interactions for action recognition in still images
Vincent Delaitre, Josef Sivic, and Ivan Laptev · 2011
Earlier work this paper cites.
A segmentation-aware object detection model with occlusion handling
Tianshi Gao, Benjamin Packer, and Daphne Koller · 2011
Earlier work this paper cites.
Recovering occlusion boundaries from an image
Derek Hoiem, Alexei A Efros, and Martial Hebert · 2011
Earlier work this paper cites.
Recognition using visual phrases
Mohammad Amin Sadeghi and Ali Farhadi · 2011
Earlier work this paper cites.
Occlusion boundary detection and figure/ground assignment from optical flow
Patrik Sundberg, Thomas Brox, Michael Maire, Pablo Arbeláez, and Jitendra Malik · 2011
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
Andreas Geiger, Philip Lenz, and Raquel Urtasun · 2012
Earlier work this paper cites.
A learning-based framework for depth ordering
Zhaoyin Jia, Andrew Gallagher, Yao-Jen Chang, and Tsuhan Chen · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
Nathan Silberman, Derek Hoiem, Pushmeet Kohli, and Rob Fergus · 2012
Earlier work this paper cites.
Occlusion reasoning for object detectionunder arbitrary viewpoint
Edward Hsiao and Martial Hebert · 2014
Earlier work this paper cites.
Pulling things out of perspective
Lubor Ladicky, Jianbo Shi, and Marc Pollefeys · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Scene parsing with object instances and occlusion ordering
Joseph Tighe, Marc Niethammer, and Svetlana Lazebnik · 2014
Earlier work this paper cites.
Beyond pascal: A benchmark for 3d object detection in the wild
Yu Xiang, Roozbeh Mottaghi, and Silvio Savarese · 2014
Earlier work this paper cites.
Hico: A benchmark for recognizing human-object interactions in images
Yu-Wei Chao, Zhan Wang, Yugeng He, Jiaxuan Wang, and Jia Deng · 2015
Earlier work this paper cites.
Multi-instance object segmentation with occlusion handling
Yi-Ting Chen, Xiaokai Liu, and Ming-Hsuan Yang · 2015
Earlier work this paper cites.
Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture
David Eigen and Rob Fergus · 2015
Earlier work this paper cites.
Direction matters: Depth estimation with a surface normal classifier
Christian Hane, Lubor Ladicky, and Marc Pollefeys · 2015
Earlier work this paper cites.
Amodal completion and size constancy in natural scenes
Abhishek Kar, Shubham Tulsiani, Joao Carreira, and Jitendra Malik · 2015
Earlier work this paper cites.
Depth and surface normal estimation from monocular images using regression on deep features and hierarchical crfs
Bo Li, Chunhua Shen, Yuchao Dai, Anton Van Den Hengel, and Mingyi He · 2015
Earlier work this paper cites.
Deep convolutional neural fields for depth estimation from a single image
Fayao Liu, Chunhua Shen, and Guosheng Lin · 2015
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Cited alongside, same era.
Scene intrinsics and depth from a single image
Evan Shelhamer, Jonathan T Barron, and Trevor Darrell · 2015
Cited alongside, same era.
Sun rgb-d: A rgb-d scene understanding benchmark suite
Shuran Song, Samuel P Lichtenberg, and Jianxiong Xiao · 2015
Cited alongside, same era.
Monocular object instance segmentation and depth ordering with cnns
Ziyu Zhang, Alexander G Schwing, Sanja Fidler, and Raquel Urtasun · 2015
Cited alongside, same era.
Single-image depth perception in the wild
Weifeng Chen, Zhao Fu, Dawei Yang, and Jia Deng · 2016
Compositional learning for human object interaction
Keizo Kato, Yin Li, and Abhinav Gupta · 2018
Later among the works it cites.
Alina Kuznetsova, Hassan Rom, Neil Alldrin, Jasper Uijlings, Ivan Krasin, Jordi Pont-Tuset, Shahab Kamali, Stefan Popov, Matteo Malloci, Tom Duerig, et al · 2018
Later among the works it cites.
Attend and interact: Higher-order object interactions for video understanding
Chih-Yao Ma, Asim Kadav, Iain Melvin, Zsolt Kira, Ghassan AlRegib, and Hans Peter Graf · 2018
Later among the works it cites.
Learning human-object interactions by graph parsing neural networks
Siyuan Qi, Wenguan Wang, Baoxiong Jia, Jianbing Shen, and Song-Chun Zhu · 2018
Later among the works it cites.
Linknet: Relational embedding for scene graph
Sanghyun Woo, Dahun Kim, Donghyeon Cho, and In So Kweon · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Monocular 3d object detection for autonomous driving
Xiaozhi Chen, Kaustav Kundu, Ziyu Zhang, Huimin Ma, Sanja Fidler, and Raquel Urtasun · 2016
Cited alongside, same era.
Amodal instance segmentation
Ke Li and Jitendra Malik · 2016
Cited alongside, same era.
Visual relationship detection with language priors
Cewu Lu, Ranjay Krishna, Michael Bernstein, and Li Fei-Fei · 2016
Cited alongside, same era.
Inception-v4, inception-resnet and the impact of residual connections on learning
Christian Szegedy, Sergey Ioffe, Vincent Vanhoucke, and Alex Alemi · 2016
Cited alongside, same era.
Detecting visual relationships with deep relational networks
Bo Dai, Yuqi Zhang, and Dahua Lin · 2017
Cited alongside, same era.
Amodal detection of 3d objects: Inferring 3d bounding boxes from 2d ones in rgb-depth images
Zhuo Deng and Longin Jan Latecki · 2017
Cited alongside, same era.
Monocular relative depth perception with web stereo data supervision
Ke Xian, Chunhua Shen, Zhiguo Cao, Hao Lu, Yang Xiao, Ruibo Li, and Zhenbo Luo · 2018
Later among the works it cites.
Graph r-cnn for scene graph generation
Jianwei Yang, Jiasen Lu, Stefan Lee, Dhruv Batra, and Devi Parikh · 2018
Later among the works it cites.
Exploring visual relationship for image captioning
Ting Yao, Yingwei Pan, Yehao Li, and Tao Mei · 2018
Later among the works it cites.
Zoom-net: Mining deep feature interactions for visual relationship recognition
Guojun Yin, Lu Sheng, Bin Liu, Nenghai Yu, Xiaogang Wang, Jing Shao, and Chen Change Loy · 2018
Later among the works it cites.
Occlusion-aware r-cnn: detecting pedestrians in a crowd
Shifeng Zhang, Longyin Wen, Xiao Bian, Zhen Lei, and Stan Z Li · 2018
Later among the works it cites.
Sail-vos: Semantic amodal instance level video object segmentation-a synthetic dataset and baselines
Yuan-Ting Hu, Hong-Shuo Chen, Kexin Hui, Jia-Bin Huang, and Alexander G Schwing · 2019
Later among the works it cites.
Perspectivenet: 3d object detection from a single rgb image via perspective points
Siyuan Huang, Yixin Chen, Tao Yuan, Siyuan Qi, Yixin Zhu, and Song-Chun Zhu · 2019
Later among the works it cites.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
Katrin Lasinger, René Ranftl, Konrad Schindler, and Vladlen Koltun · 2019
Later among the works it cites.
Gs3d: An efficient 3d object detection framework for autonomous driving
Buyu Li, Wanli Ouyang, Lu Sheng, Xingyu Zeng, and Xiaogang Wang · 2019
Later among the works it cites.
Deep fitting degree scoring network for monocular 3d object detection
Lijie Liu, Jiwen Lu, Chunjing Xu, Qi Tian, and Jie Zhou · 2019
Later among the works it cites.
Occlusion-shared and feature-separated network for occlusion relationship reasoning
Rui Lu, Feng Xue, Menghan Zhou, Anlong Ming, and Yu Zhou · 2019
Later among the works it cites.
Accurate monocular 3d object detection via color-embedded 3d reconstruction for autonomous driving
Xinzhu Ma, Zhihui Wang, Haojie Li, Pengbo Zhang, Wanli Ouyang, and Xin Fan · 2019
Later among the works it cites.
Amodal instance segmentation with kins dataset
Lu Qi, Li Jiang, Shu Liu, Xiaoyong Shen, and Jiaya Jia · 2019
Later among the works it cites.
Spatialsense: An adversarially crowdsourced benchmark for spatial relation recognition
Kaiyu Yang, Olga Russakovsky, and Jia Deng · 2019
Later among the works it cites.
Large-scale visual relationship understanding
Ji Zhang, Yannis Kalantidis, Marcus Rohrbach, Manohar Paluri, Ahmed Elgammal, and Mohamed Elhoseiny · 2019
Later among the works it cites.
Kinematic 3d object detection in monocular video
Garrick Brazil, Gerard Pons-Moll, Xiaoming Liu, and Bernt Schiele · 2020
Later among the works it cites.
Oasis: A large-scale dataset for single image 3d in the wild
Weifeng Chen, Shengyi Qian, David Fan, Noriyuki Kojima, Max Hamilton, and Jia Deng · 2020
Later among the works it cites.
Rel3d: A minimally contrastive benchmark for grounding spatial relations in 3d
Ankit Goyal, Kaiyu Yang, Dawei Yang, and Jia Deng · 2020
Later among the works it cites.
Peek-a-boo: Occlusion reasoning in indoor scenes with plane representations
Ziyu Jiang, Buyu Liu, Samuel Schulter, Zhangyang Wang, and Manmohan Chandraker · 2020
Later among the works it cites.
The open images dataset v4
Alina Kuznetsova, Hassan Rom, Neil Alldrin, Jasper Uijlings, Ivan Krasin, Jordi Pont-Tuset, Shahab Kamali, Stefan Popov, Matteo Malloci, Alexander Kolesnikov, et al · 2020
Later among the works it cites.
Structure-guided ranking loss for single image depth prediction
Ke Xian, Jianming Zhang, Oliver Wang, Long Mai, Zhe Lin, and Zhiguo Cao · 2020
Later among the works it cites.
Self-supervised scene de-occlusion
Xiaohang Zhan, Xingang Pan, Bo Dai, Ziwei Liu, Dahua Lin, and Chen Change Loy · 2020
Later among the works it cites.