Huynh L, Nguyen-Ha P, Matas J, Rahtu E, Heikkilä J (2020) Guiding monocular depth estimation using depth-attention volume. In: European Conference on Computer Vision (ECCV), pp 581–597
2020
Later among the works it cites.
Johnston A, Carneiro G (2020) Self-supervised monocular trained depth estimation using self-attention and discrete disparity volume. In: Computer Vision and Pattern Recognition (CVPR), pp 4756–4765
2020
Later among the works it cites.
Kolesnikov A, Beyer L, Zhai X, Puigcerver J, Yung J, Gelly S, Houlsby N (2020) Big transfer (bit): General visual representation learning. In: European Conference on Computer Vision (ECCV), pp 491–507
2020
Later among the works it cites.
Xu D, Alameda-Pineda X, Ouyang W, Ricci E, Wang X, Sebe N (2020) Probabilistic graph attention network with conditional kernels for pixel-wise prediction. Transactions on Pattern Analysis and Machine Intelligence
2020
Later among the works it cites.
Bhat SF, Alhashim I, Wonka P (2021) Adabins: Depth estimation using adaptive bins. In: Computer Vision and Pattern Recognition (CVPR), pp 4009–4018
2021
Later among the works it cites.
Dosovitskiy A, Beyer L, Kolesnikov A, Weissenborn D, Zhai X, Unterthiner T, Dehghani M, Minderer M, Heigold G, Gelly S, et al. (2021) An image is worth 16x16 words: Transformers for image recognition at scale. In: International Conference on Learning Representations (ICLR)
2021
Later among the works it cites.
Ji P, Li R, Bhanu B, Xu Y (2021) Monoindoor: Towards good practice of self-supervised monocular depth estimation for indoor environments. In: International Conference on Computer Vision (ICCV), pp 12787–12796
2021
Later among the works it cites.
Lee S, Lee J, Kim B, Yi E, Kim J (2021) Patch-wise attention network for monocular depth estimation. In: Proceedings of the AAAI Conference on Artificial Intelligence, vol 35, pp 1873–1881
2021
Later among the works it cites.
Liu Z, Lin Y, Cao Y, Hu H, Wei Y, Zhang Z, Lin S, Guo B (2021) Swin transformer: Hierarchical vision transformer using shifted windows. International Conference on Computer Vision (ICCV)
2021
Later among the works it cites.
Qiao S, Zhu Y, Adam H, Yuille A, Chen LC (2021) Vip-deeplab: Learning visual perception with depth-aware video panoptic segmentation. In: Computer Vision and Pattern Recognition (CVPR), pp 3997–4008
2021
Later among the works it cites.
Ranftl R, Bochkovskiy A, Koltun V (2021) Vision transformers for dense prediction. In: International Conference on Computer Vision (ICCV), pp 12179–12188
2021
Later among the works it cites.
Yang G, Tang H, Ding M, Sebe N, Ricci E (2021) Transformers solve the limited receptive field for monocular depth prediction. In: International Conference on Computer Vision (ICCV)
2021
Later among the works it cites.
Zheng S, Lu J, Zhao H, Zhu X, Luo Z, Wang Y, Fu Y, Feng J, Xiang T, Torr PH, et al. (2021) Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers. In: Computer Vision and Pattern Recognition (CVPR), pp 6881–6890
2021
Later among the works it cites.
Zhu X, Su W, Lu L, Li B, Wang X, Dai J (2021) Deformable detr: Deformable transformers for end-to-end object detection. In: International Conference on Learning Representations (ICLR)
2021
Later among the works it cites.