Fetching the paper…
Reading the bibliography…
Stereo video synthesis from a monocular input is a demanding task in the fields of spatial computing and virtual reality.
Large occlusion stereo
Aaron Bobick and Stephen Intille · 1999
Earlier work this paper cites.
View Synthesis Using Stereo Vision
D. Scharstein · 1999
Earlier work this paper cites.
Detecting binocular half-occlusions: empirical comparisons of five approaches
G. Egnal and R.P. Wildes · 2002
Earlier work this paper cites.
Image distortions in stereoscopic video systems
Andrew Woods, Tom Docherty, and Rolf Koch · 2002
Earlier work this paper cites.
Multi-view stereo reconstruction of dense shape and complex appearance
Hailin Jin, Stefano Soatto, and Anthony J Yezzi · 2005
Earlier work this paper cites.
A comparison and evaluation of multi-view stereo reconstruction algorithms
S.M. Seitz, B. Curless, J. Diebel, D. Scharstein, and R. Szeliski · 2006
Earlier work this paper cites.
Stereoscopic video synthesis from a monocular video
Guofeng Zhang, Wei Hua, Xueying Qin, Tien-Tsin Wong, and Hujun Bao · 2007
Earlier work this paper cites.
An overview of free view-point depth-image-based rendering (dibr)
Wenxiu Sun, Lingfeng Xu, Oscar C Au, Sung Him Chui, and Chun Wing Kwok · 2010
Earlier work this paper cites.
Occlusion refinement for stereo video using optical flow
Dmitry Akimov, Alexey Shestov, Alexander Voronov, and Dmitriy Vatolin · 2012
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L Schonberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Low-cost 360 stereo photography and video capture
Kevin Matzen, Michael F. Cohen, Bryce Evans, Johannes Kopf, and Richard Szeliski · 2017
Earlier work this paper cites.
A survey of structure from motion*
Onur Özyeşil, Vladislav Voroninski, Ronen Basri, and Amit Singer · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Learning blind video temporal consistency
Wei-Sheng Lai, Jia-Bin Huang, Oliver Wang, Eli Shechtman, Ersin Yumer, and Ming-Hsuan Yang · 2018
Earlier work this paper cites.
Stereo magnification: Learning view synthesis using multiplane images
Tinghui Zhou, Richard Tucker, John Flynn, Graham Fyffe, and Noah Snavely · 2018
Earlier work this paper cites.
A review of stereo-photogrammetry method for 3-d reconstruction in computer vision
Phuong Ngoc Binh Do and Quoc Chi Nguyen · 2019
Earlier work this paper cites.
FVD: A new metric for video generation
Thomas Unterthiner, Sjoerd van Steenkiste, Karol Kurach, Raphaël Marinier, Marcin Michalski, and Sylvain Gelly · 2019
Earlier work this paper cites.
Photo wake-up: 3d character animation from a single photo
Chung-Yi Weng, Brian Curless, and Ira Kemelmacher-Shlizerman · 2019
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis, 2020
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Earlier work this paper cites.
Softmax splatting for video frame interpolation, 2020
Simon Niklaus and Feng Liu · 2020
Earlier work this paper cites.
3d photography using context-aware layered depth inpainting
Meng-Li Shih, Shih-Yang Su, Johannes Kopf, and Jia-Bin Huang · 2020
Earlier work this paper cites.
Raft: Recurrent all-pairs field transforms for optical flow, 2020
Zachary Teed and Jia Deng · 2020
Earlier work this paper cites.
Single-view view synthesis with multiplane images
Richard Tucker and Noah Snavely · 2020
Earlier work this paper cites.
Synsin: End-to-end view synthesis from a single image
Olivia Wiles, Georgia Gkioxari, Richard Szeliski, and Justin Johnson · 2020
Earlier work this paper cites.
Dynamic view synthesis from dynamic monocular video, 2021
Chen Gao, Ayush Saraf, Johannes Kopf, and Jia-Bin Huang · 2021
Cited alongside, same era.
Slide: Single image 3d photography with soft layering and depth-aware inpainting
Varun Jampani, Huiwen Chang, Kyle Sargent, Abhishek Kar, Richard Tucker, Michael Krainin, Dominik Kaeser, William T Freeman, David Salesin, Brian Curless, et al · 2021
Cited alongside, same era.
Infinite nature: Perpetual view generation of natural scenes from a single image
Andrew Liu, Richard Tucker, Varun Jampani, Ameesh Makadia, Noah Snavely, and Angjoo Kanazawa · 2021
Cited alongside, same era.
A survey of image labelling for computer vision applications
Christoph Sager, Christian Janiesch, and Patrick Zschech · 2021
Cited alongside, same era.
Self-supervised visibility learning for novel view synthesis.*
Yujiao Shi, Hongdong Li, and Xin Yu · 2021
Cited alongside, same era.
Text2live: Text-driven layered image and video editing
Sinmpi: Novel view synthesis from a single image with expanded multiplane images, 2023
Guo Pu, Peng-Shuai Wang, and Zhouhui Lian · 2023
Later among the works it cites.
Fatezero: Fusing attentions for zero-shot text-based video editing
Chenyang Qi, Xiaodong Cun, Yong Zhang, Chenyang Lei, Xintao Wang, Ying Shan, and Qifeng Chen · 2023
Later among the works it cites.
Train stable diffusion for inpainting, 2023
Lorenzo Stacchio · 2023
Later among the works it cites.
Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation
Jay Zhangjie Wu, Yixiao Ge, Xintao Wang, Stan Weixian Lei, Yuchao Gu, Yufei Shi, Wynne Hsu, Ying Shan, Xiaohu Qie, and Mike Zheng Shou · 2023
Later among the works it cites.
Long-term photometric consistent novel view synthesis with diffusion models, 2023
Jason J. Yu, Fereshteh Forghani, Konstantinos G. Derpanis, and Marcus A. Brubaker · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Omer Bar-Tal, Dolev Ofri-Amar, Rafail Fridman, Yoni Kasten, and Tali Dekel · 2022
Cited alongside, same era.
Single-view view synthesis in the wild with learned adaptive multiplane images
Yuxuan Han, Ruicheng Wang, and Jiaolong Yang · 2022
Cited alongside, same era.
Prompt-to-prompt image editing with cross attention control, 2022
Amir Hertz, Ron Mokady, Jay Tenenbaum, Kfir Aberman, Yael Pritch, and Daniel Cohen-Or · 2022
Cited alongside, same era.
Video diffusion models
Jonathan Ho, Tim Salimans, Alexey Gritsenko, William Chan, Mohammad Norouzi, and David J Fleet · 2022
Cited alongside, same era.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
René Ranftl, Katrin Lasinger, David Hafner, Konrad Schindler, and Vladlen Koltun · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Cited alongside, same era.
Make-a-video: Text-to-video generation without text-video data
Uriel Singer, Adam Polyak, Thomas Hayes, Xi Yin, Jie An, Songyang Zhang, Qiyuan Hu, Harry Yang, Oron Ashual, Oran Gafni, et al · 2022
Cited alongside, same era.
Min Zhao, Rongzhen Wang, Fan Bao, Chongxuan Li, and Jun Zhu · 2023
Later among the works it cites.
visionos 2 brings new spatial computing experiences to apple vision pro
Andrea Schubert · 2024
Closest in time.
Novel view synthesis with view-dependent effects from a single image
Juan Luis Gonzalez Bello and Munchurl Kim · 2024
Closest in time.
Video generation models as world simulators
Tim Brooks, Bill Peebles, Connor Holmes, Will DePue, Yufei Guo, Li Jing, David Schnurr, Joe Taylor, Troy Luhman, Eric Luhman, Clarence Ng, Ricky Wang, and Aditya Ramesh · 2024
Closest in time.
Mvsplat: Efficient 3d gaussian splatting from sparse multi-view images
Yuedong Chen, Haofei Xu, Chuanxia Zheng, Bohan Zhuang, Marc Pollefeys, Andreas Geiger, Tat-Jen Cham, and Jianfei Cai · 2024
Closest in time.
Animatediff: Animate your personalized text-to-image diffusion models without specific tuning
Yuwei Guo, Ceyuan Yang, Anyi Rao, Zhengyang Liang, Yaohui Wang, Yu Qiao, Maneesh Agrawala, Dahua Lin, and Bo Dai · 2024
Closest in time.
Unifying correspondence, pose and nerf for pose-free novel view synthesis from stereo pairs, 2024
Sunghwan Hong, Jaewoo Jung, Heeseong Shin, Jiaolong Yang, Seungryong Kim, and Chong Luo · 2024
Closest in time.
Depthcrafter: Generating consistent long depth sequences for open-world videos
Wenbo Hu, Xiangjun Gao, Xiaoyu Li, Sijie Zhao, Xiaodong Cun, Yong Zhang, Long Quan, and Ying Shan · 2024
Closest in time.
Nvist: In the wild new view synthesis from a single image with transformers, 2024
Wonbong Jang and Lourdes Agapito · 2024
Closest in time.
Ocai: Improving optical flow estimation by occlusion and consistency aware interpolation, 2024
Jisoo Jeong, Hong Cai, Risheek Garrepalli, Jamie Menjay Lin, Munawar Hayat, and Fatih Porikli · 2024
Closest in time.
Repurposing diffusion-based image generators for monocular depth estimation, 2024
Bingxin Ke, Anton Obukhov, Shengyu Huang, Nando Metzger, Rodrigo Caye Daudt, and Konrad Schindler · 2024
Closest in time.
Wonderland: Navigating 3d scenes from a single image
Hanwen Liang, Junli Cao, Vidit Goel, Guocheng Qian, Sergei Korolev, Demetri Terzopoulos, Konstantinos Plataniotis, Sergey Tulyakov, and Jian Ren · 2024
Closest in time.
Sora: A review on background, technology, limitations, and opportunities of large vision models, 2024
Yixin Liu, Kai Zhang, Yuan Li, Zhiling Yan, Chujie Gao, Ruoxi Chen, Zhengqing Yuan, Yue Huang, Hanchi Sun, Jianfeng Gao, Lifang He, and Lichao Sun · 2024
Closest in time.
Multidiff: Consistent novel view synthesis from a single image
Norman Müller, Katja Schwarz, Barbara Rössle, Lorenzo Porzi, Samuel Rota Bulò, Matthias Nießner, and Peter Kontschieder · 2024
Closest in time.
What is spatial video on iphone 15 pro and vision pro
Onee · 2024
Closest in time.
Generative camera dolly: Extreme monocular dynamic novel view synthesis
Basile Van Hoorick, Rundi Wu, Ege Ozguroglu, Kyle Sargent, Ruoshi Liu, Pavel Tokmakov, Achal Dave, Changxi Zheng, and Carl Vondrick · 2024
Closest in time.
Videocomposer: Compositional video synthesis with motion controllability
Xiang Wang, Hangjie Yuan, Shiwei Zhang, Dayou Chen, Jiuniu Wang, Yingya Zhang, Yujun Shen, Deli Zhao, and Jingren Zhou · 2024
Closest in time.
Depth anything: Unleashing the power of large-scale unlabeled data, 2024
Lihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu, Jiashi Feng, and Hengshuang Zhao · 2024
Closest in time.
Nvs-solver: Video diffusion model as zero-shot novel view synthesizer
Meng You, Zhiyu Zhu, Hui Liu, and Junhui Hou · 2024
Closest in time.