Fetching the paper…
Reading the bibliography…
Recently, Transformers were shown to enhance the performance of multi-view stereo by enabling long-range feature interaction.
Large scale multi-view stereopsis evaluation
Rasmus Jensen, Anders Dahl, George Vogiatzis, Engin Tola, and Henrik Aanæs · 2014
Earlier work this paper cites.
Massively parallel multiview stereopsis by surface normal diffusion
Silvano Galliani, Katrin Lasinger, and Konrad Schindler · 2015
Earlier work this paper cites.
Pixelwise View Selection for Unstructured Multi-View Stereo
Johannes Lutz Schönberger, Enliang Zheng, Marc Pollefeys, and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Tanks and temples: Benchmarking large-scale scene reconstruction
Arno Knapitsch, Jaesik Park, Qian-Yi Zhou, and Vladlen Koltun · 2017
Earlier work this paper cites.
Feature pyramid networks for object detection
Tsung-Yi Lin, Piotr Dollár, Ross Girshick, Kaiming He, Bharath Hariharan, and Serge Belongie · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Mvsnet: Depth inference for unstructured multi-view stereo
Yao Yao, Zixin Luo, Shiwei Li, Tian Fang, and Long Quan · 2018
Earlier work this paper cites.
Multi-scale geometric consistency guided multi-view stereo
Qingshan Xu and Wenbing Tao · 2019
Earlier work this paper cites.
Recurrent mvsnet for high-resolution multi-view stereo depth inference
Yao Yao, Zixin Luo, Shiwei Li, Tianwei Shen, Tian Fang, and Long Quan · 2019
Earlier work this paper cites.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Earlier work this paper cites.
Deep stereo using adaptive thin volume representation with uncertainty awareness
Shuo Cheng, Zexiang Xu, Shilin Zhu, Zhuwen Li, Li Erran Li, Ravi Ramamoorthi, and Hao Su · 2020
Earlier work this paper cites.
Cascade cost volume for high-resolution multi-view stereo and stereo matching
Xiaodong Gu, Zhiwen Fan, Siyu Zhu, Zuozhuo Dai, Feitong Tan, and Ping Tan · 2020
Earlier work this paper cites.
Superglue: Learning feature matching with graph neural networks
Paul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2020
Earlier work this paper cites.
Dense hybrid recurrent multi-view stereo net with dynamic consistency checking
Jianfeng Yan, Zizhuang Wei, Hongwei Yi, Mingyu Ding, Runze Zhang, Yisong Chen, Guoping Wang, and Yu-Wing Tai · 2020
Cited alongside, same era.
Cost volume pyramid based depth inference for multi-view stereo
Jiayu Yang, Wei Mao, Jose M. Alvarez, and Miaomiao Liu · 2020
Cited alongside, same era.
Blendedmvs: A large-scale dataset for generalized multi-view stereo networks
Yao Yao, Zixin Luo, Shiwei Li, Jingyang Zhang, Yufan Ren, Lei Zhou, Tian Fang, and Long Quan · 2020
Cited alongside, same era.
Pyramid multi-view stereo net with self-adaptive view aggregation
Hongwei Yi, Zizhuang Wei, Mingyu Ding, Runze Zhang, Yisong Chen, Guoping Wang, and Yu-Wing Tai · 2020
Cited alongside, same era.
Visibility-aware multi-view stereo network
Jingyang Zhang, Yao Yao, Shiwei Li, Zixin Luo, and Tian Fang · 2020
Cited alongside, same era.
Transmvsnet: Global context-aware multi-view stereo network with transformers
Ze Liu, Jia Ning, Yue Cao, Yixuan Wei, Zheng Zhang, Stephen Lin, and Han Hu · 2021
Later among the works it cites.
Epp-mvsnet: Epipolar-assembling based depth prediction for multi-view stereo
Xinjun Ma, Yue Gong, Qirui Wang, Jingwei Huang, Lei Chen, and Fan Yu · 2021
Later among the works it cites.
3d object detection with pointformer
Xuran Pan, Zhuofan Xia, Shiji Song, Li Erran Li, and Gao Huang · 2021
Later among the works it cites.
Vision transformers for dense prediction
René Ranftl, Alexey Bochkovskiy, and Vladlen Koltun · 2021
Later among the works it cites.
Loftr: Detector-free local feature matching with transformers
Jiaming Sun, Zehong Shen, Yuang Wang, Hujun Bao, and Xiaowei Zhou · 2021
Later among the works it cites.
Patchmatchnet: Learned multi-view patchmatch stereo
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yikang Ding, Wentao Yuan, Qingtian Zhu, Haotian Zhang, Xiangyue Liu, Yuanjiang Wang, and Xiao Liu · 2021
Cited alongside, same era.
Cswin transformer: A general vision transformer backbone with cross-shaped windows, 2021
Xiaoyi Dong, Jianmin Bao, Dongdong Chen, Weiming Zhang, Nenghai Yu, Lu Yuan, Dong Chen, and Baining Guo · 2021
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2021
Cited alongside, same era.
Deepvideomvs: Multi-view stereo on video with recurrent spatio-temporal fusion
Arda Duzceker, Silvano Galliani, Christoph Vogel, Pablo Speciale, Mihai Dusmanu, and Marc Pollefeys · 2021
Cited alongside, same era.
Revisiting stereo depth estimation from a sequence-to-sequence perspective with transformers
Zhaoshuo Li, Xingtong Liu, Nathan Drenkow, Andy Ding, Francis X Creighton, Russell H Taylor, and Mathias Unberath · 2021
Cited alongside, same era.
Swinir: Image restoration using swin transformer
Jingyun Liang, Jiezhang Cao, Guolei Sun, Kai Zhang, Luc Van Gool, and Radu Timofte · 2021
Cited alongside, same era.
Ds-transunet: Dual swin transformer u-net for medical image segmentation
Ai-Jun Lin, Bingzhi Chen, Jiayu Xu, Zheng Zhang, and Guangming Lu · 2021
Cited alongside, same era.
Fangjinhua Wang, Silvano Galliani, Christoph Vogel, Pablo Speciale, and Marc Pollefeys · 2021
Later among the works it cites.
Aa-rmvsnet: Adaptive aggregation recurrent multi-view stereo network
Zizhuang Wei, Qingtian Zhu, Chen Min, Yisong Chen, and Guoping Wang · 2021
Later among the works it cites.
Non-local recurrent regularization networks for multi-view stereo
Qingshan Xu, Martin R Oswald, Wenbing Tao, Marc Pollefeys, and Zhaopeng Cui · 2021
Later among the works it cites.
Mvs2d: Efficient multi-view stereo via attention-driven 2d convolutions
Zhenpei Yang, Zhile Ren, Qi Shan, and Qixing Huang · 2021
Later among the works it cites.
Generalized binary search network for highly-efficient multi-view stereo
Zhenxing Mi, Chang Di, and Dan Xu · 2022
Closest in time.
Rethinking depth estimation for multi-view stereo: A unified representation
Rui Peng, Rongjie Wang, Zhenyu Wang, Yawen Lai, and Ronggang Wang · 2022
Closest in time.
Itermvs: Iterative probability estimation for efficient multi-view stereo, 2022
Fangjinhua Wang, Silvano Galliani, Christoph Vogel, and Marc Pollefeys · 2022
Closest in time.