Fetching the paper…
Reading the bibliography…
We study the problem of synthesizing immersive 3D indoor scenes from one or more images.
Modeling and rendering architecture from photographs: A hybrid geometry-and image-based approach
Debevec, P. E.; Taylor, C. J.; and Malik, J. 1996 · 1996
Earlier work this paper cites.
Novel view synthesis in tensor space
Avidan, S.; and Shashua, A. 1997 · 1997
Earlier work this paper cites.
High-quality video view interpolation using a layered representation
Zitnick, C. L.; Kang, S. B.; Uyttendaele, M.; Winder, S.; and Szeliski, R. 2004 · 2004
Earlier work this paper cites.
Efficient reductions for imitation learning
Ross, S.; and Bagnell, D. 2010 · 2010
Earlier work this paper cites.
Indoor Segmentation and Support Inference from RGBD Images
Nathan Silberman, P. K., Derek Hoiem; and Fergus, R. 2012 · 2012
Earlier work this paper cites.
Generative adversarial networks
Goodfellow, I. J.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
ORB-SLAM: A Versatile and Accurate Monocular SLAM System
Mur-Artal, R.; Montiel, J. M. M.; and Tardós, J. D. 2015 · 2015
Earlier work this paper cites.
End to end learning for self-driving cars
Bojarski, M.; Del Testa, D.; Dworakowski, D.; Firner, B.; Flepp, B.; Goyal, P.; Jackel, L. D.; Monfort, M.; Muller, U.; Zhang, J.; et al. 2016 · 2016
Earlier work this paper cites.
DeepStereo: Learning to Predict New Views From the World’s Imagery
Flynn, J.; Neulander, I.; Philbin, J.; and Snavely, N. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Perceptual losses for real-time style transfer and super-resolution
Johnson, J.; Alahi, A.; and Fei-Fei, L. 2016 · 2016
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Mnih, V.; Badia, A. P.; Mirza, M.; Graves, A.; Lillicrap, T.; Harley, T.; Silver, D.; and Kavukcuoglu, K. 2016 · 2016
Earlier work this paper cites.
Structure-from-Motion Revisited
Schönberger, J. L.; and Frahm, J.-M. 2016 · 2016
Earlier work this paper cites.
Data augmentation generative adversarial networks
Antoniou, A.; Storkey, A.; and Edwards, H. 2017 · 2017
Earlier work this paper cites.
Matterport3D: Learning from RGB-D Data in Indoor Environments
Chang, A.; Dai, A.; Funkhouser, T.; Halber, M.; Niessner, M.; Savva, M.; Song, S.; Zeng, A.; and Zhang, Y. 2017 · 2017
Earlier work this paper cites.
Deep visual foresight for planning robot motion
Finn, C.; and Levine, S. 2017 · 2017
Earlier work this paper cites.
GANs trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M.; Ramsauer, H.; Unterthiner, T.; Nessler, B.; and Hochreiter, S. 2017 · 2017
Earlier work this paper cites.
Image-to-Image Translation with Conditional Adversarial Networks
Isola, P.; Zhu, J.-Y.; Zhou, T.; and Efros, A. A. 2017 · 2017
Earlier work this paper cites.
Learning a multi-view stereo machine
Kar, A.; Häne, C.; and Malik, J. 2017 · 2017
Earlier work this paper cites.
Geometric gan
Lim, J. H.; and Ye, J. C. 2017 · 2017
Earlier work this paper cites.
Pixelcnn++: Improving the pixelcnn with discretized logistic mixture likelihood and other modifications
Salimans, T.; Karpathy, A.; Chen, X.; and Kingma, D. P. 2017 · 2017
Earlier work this paper cites.
End-to-end driving via conditional imitation learning
Codevilla, F.; Müller, M.; López, A.; Koltun, V.; and Dosovitskiy, A. 2018 · 2018
Earlier work this paper cites.
Stochastic video generation with a learned prior
Denton, E.; and Fergus, R. 2018 · 2018
Earlier work this paper cites.
World models
Ha, D.; and Schmidhuber, J. 2018 · 2018
Cited alongside, same era.
Single-image Tomography: 3D Volumes from 2D Cranial X-Rays
Henzler, P.; Rasche, V.; Ropinski, T.; and Ritschel, T. 2018 · 2018
Cited alongside, same era.
Rednet: Residual encoder-decoder network for indoor rgb-d semantic segmentation
Jiang, J.; Zheng, L.; Luo, F.; and Zhang, Z. 2018 · 2018
Cited alongside, same era.
Megadepth: Learning single-view depth prediction from internet photos
Li, Z.; and Snavely, N. 2018 · 2018
Cited alongside, same era.
Image inpainting for irregular holes using partial convolutions
Liu, G.; Reda, F. A.; Shih, K. J.; Wang, T.-C.; Tao, A.; and Catanzaro, B. 2018 · 2018
Cited alongside, same era.
Gibson env: Real-world perception for embodied agents
Xia, F.; Zamir, A. R.; He, Z.; Sax, A.; Malik, J.; and Savarese, S. 2018 · 2018
Cited alongside, same era.
Nerf: Representing scenes as neural radiance fields for view synthesis
Mildenhall, B.; Srinivasan, P. P.; Tancik, M.; Barron, J. T.; Ramamoorthi, R.; and Ng, R. 2020 · 2020
Later among the works it cites.
3D Photography Using Context-Aware Layered Depth Inpainting
Shih, M.-L.; Su, S.-Y.; Kopf, J.; and Huang, J.-B. 2020 · 2020
Later among the works it cites.
Deep novel view synthesis from colored 3d point clouds
Song, Z.; Chen, W.; Campbell, D.; and Li, H. 2020 · 2020
Later among the works it cites.
Single-View View Synthesis with Multiplane Images
Tucker, R.; and Snavely, N. 2020 · 2020
Later among the works it cites.
Synsin: End-to-end view synthesis from a single image
Wiles, O.; Gkioxari, G.; Szeliski, R.; and Johnson, J. 2020 · 2020
Later among the works it cites.
Diagnosing the Environment Bias in Vision-and-Language Navigation
Zhang, Y.; Tan, H.; and Bansal, M. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Peeking behind objects: Layered depth prediction from a single image
Dhamo, H.; Tateno, K.; Laina, I.; Navab, N.; and Tombari, F. 2019 · 2019
Cited alongside, same era.
Deepview: View synthesis with learned gradient descent
Flynn, J.; Broxton, M.; Debevec, P.; DuVall, M.; Fyffe, G.; Overbeck, R.; Snavely, N.; and Tucker, R. 2019 · 2019
Cited alongside, same era.
Tactical Rewind: Self-Correction via Backtracking in Vision-and-Language Navigation
Liyiming Ke, Y. B. A. H. Z. G. J. L. J. G. Y. C. S. S., Xiujun Li. 2019 · 2019
Cited alongside, same era.
Local light field fusion: Practical view synthesis with prescriptive sampling guidelines
Mildenhall, B.; Srinivasan, P. P.; Ortiz-Cayon, R.; Kalantari, N. K.; Ramamoorthi, R.; Ng, R.; and Kar, A. 2019 · 2019
Cited alongside, same era.
Semantic image synthesis with spatially-adaptive normalization
Park, T.; Liu, M.-Y.; Wang, T.-C.; and Zhu, J.-Y. 2019 · 2019
Cited alongside, same era.
Generating diverse high-fidelity images with vq-vae2
Razavi, A.; van den Oord, A.; and Vinyals, O. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance Fields
Barron, J. T.; Mildenhall, B.; Tancik, M.; Hedman, P.; Martin-Brualla, R.; and Srinivasan, P. P. 2021 · 2021
Later among the works it cites.
History Aware multimodal Transformer for Vision-and-Language Navigation
Chen, S.; Guhur, P.-L.; Schmid, C.; and Laptev, I. 2021 · 2021
Later among the works it cites.
Semantics-aware multi-modal domain translation: From lidar point clouds to panoramic color images
Cortinhal, T.; Kurnaz, F.; and Aksoy, E. E. 2021 · 2021
Later among the works it cites.
Masked autoencoders are scalable vision learners
He, K.; Chen, X.; Xie, S.; Li, Y.; Dollár, P.; and Girshick, R. 2021 · 2021
Later among the works it cites.
A recurrent vision-and-language BERT for navigation
Hong, Y.; Wu, Q.; Qi, Y.; Rodriguez-Opazo, C.; and Gould, S. 2021 · 2021
Later among the works it cites.
Worldsheet: Wrapping the world in a 3d sheet for view synthesis from a single image
Hu, R.; Ravi, N.; Berg, A. C.; and Pathak, D. 2021 · 2021
Later among the works it cites.
MURAL: Multimodal, Multitask Retrieval Across Languages
Jain, A.; Guo, M.; Srinivasan, K.; Chen, T.; Kudugunta, S.; Jia, C.; Yang, Y.; and Baldridge, J. 2021 · 2021
Later among the works it cites.
Putting nerf on a diet: Semantically consistent few-shot view synthesis
Jain, A.; Tancik, M.; and Abbeel, P. 2021 · 2021
Later among the works it cites.
MannequinChallenge: Learning the Depths of Moving People by Watching Frozen People
Li, Z.; Dekel, T.; Cole, F.; Tucker, R.; Snavely, N.; Liu, C.; and Freeman, W. T. 2021 · 2021
Later among the works it cites.
Infinite nature: Perpetual view generation of natural scenes from a single image
Liu, A.; Tucker, R.; Jampani, V.; Makadia, A.; Snavely, N.; and Kanazawa, A. 2021 · 2021
Later among the works it cites.
NeRF in the Wild: Neural Radiance Fields for Unconstrained Photo Collections
Martin-Brualla, R.; Radwan, N.; Sajjadi, M. S. M.; Barron, J. T.; Dosovitskiy, A.; and Duckworth, D. 2021 · 2021
Later among the works it cites.
Vision Transformers for Dense Prediction
Ranftl, R.; Bochkovskiy, A.; and Koltun, V. 2021 · 2021
Later among the works it cites.
Pixelsynth: Generating a 3d-consistent experience from a single image
Rockwell, C.; Fouhey, D. F.; and Johnson, J. 2021 · 2021
Later among the works it cites.
Geometry-Free View Synthesis: Transformers and no 3D Priors
Rombach, R.; Esser, P.; and Ommer, B. 2021 · 2021
Later among the works it cites.
Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single Image
Ren, X.; and Wang, X. 2022 · 2022
Closest in time.
Block-NeRF: Scalable Large Scene Neural View Synthesis
Tancik, M.; Casser, V.; Yan, X.; Pradhan, S.; Mildenhall, B.; Srinivasan, P. P.; Barron, J. T.; and Kretzschmar, H. 2022 · 2022
Closest in time.