Fetching the paper…
Reading the bibliography…
Stereo matching has become a key technique for 3D environment perception in intelligent vehicles.
Q. Su and S. Ji, “ChiTransformer: Towards reliable stereo from cues,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 1939–1949
1949
Earlier work this paper cites.
D. Scharstein and C. Pal, “Learning conditional random fields for stereo,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2007, pp. 1–8
2007
Earlier work this paper cites.
H. Hirschmuller and D. Scharstein, “Evaluation of cost functions for stereo matching,” in 2007 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2007, pp. 1–8
2007
Earlier work this paper cites.
J. Deng et al. , “ImageNet: A large-scale hierarchical image database,” in 2009 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2009, pp. 248–255
2009
Earlier work this paper cites.
A. Geiger et al. , “Are we ready for autonomous driving? the kitti vision benchmark suite,” in 2012 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2012, pp. 3354–3361
2012
Earlier work this paper cites.
D. Eigen et al. , “Depth map prediction from a single image using a multi-scale deep network,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 27, 2014
2014
Earlier work this paper cites.
D. Scharstein et al. , “High-resolution stereo datasets with subpixel-accurate ground truth,” in Pattern Recognition: 36th German Conference (GCPR) . Springer, 2014, pp. 31–42
2014
Earlier work this paper cites.
M. Menze and A. Geiger, “Object scene flow for autonomous vehicles,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 3061–3070
2015
Earlier work this paper cites.
K. He et al. , “Deep residual learning for image recognition,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 770–778
2016
Earlier work this paper cites.
N. Mayer et al. , “A large dataset to train convolutional networks for disparity, optical flow, and scene flow estimation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 4040–4048
2016
Earlier work this paper cites.
T. Schops et al. , “A multi-view stereo benchmark with high-resolution images and multi-camera videos,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 3260–3269
2017
Earlier work this paper cites.
M. Sandler and pthers, “MobileNetV2: Inverted residuals and linear bottlenecks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 4510–4520
2018
Earlier work this paper cites.
J. Hu et al. , “Squeeze-and-excitation networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 7132–7141
2018
Earlier work this paper cites.
X. Guo et al. , “Group-wise correlation stereo network,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 3273–3282
2019
Earlier work this paper cites.
R. Fan et al. , “SNE-RoadSeg: Incorporating surface normal information into semantic segmentation for accurate freespace detection,” in European Conference on Computer Vision (ECCV) . Springer, 2020, pp. 340–356
2020
Earlier work this paper cites.
J.-R. Chang et al. , “Attention-aware feature aggregation for real-time stereo matching on edge devices,” in Proceedings of the Asian Conference on Computer Vision (ACCV) , 2020
2020
Earlier work this paper cites.
X. Cheng et al. , “Hierarchical neural architecture search for deep stereo matching,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 33, pp. 22 158–22 169, 2020
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
R. Ranftl et al. , “Vision Transformers for dense prediction,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2021, pp. 12 179–12 188
2021
Earlier work this paper cites.
L. Lipson et al. , “RAFT-Stereo: Multilevel recurrent field transforms for stereo matching,” in 2021 International Conference on 3D Vision (3DV) . IEEE, 2021, pp. 218–227
2021
Cited alongside, same era.
Z. Shen et al. , “CFNet: Cascade and fused cost volume for robust stereo matching,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 13 906–13 915
2021
Cited alongside, same era.
Z. Xie et al. , “Propagate yourself: Exploring pixel-level consistency for unsupervised visual representation learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 16 684–16 693
2021
Cited alongside, same era.
X. Wang et al. , “Dense contrastive learning for self-supervised visual pre-training,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 3024–3033
2021
Cited alongside, same era.
A. Kirillov et al. , “Segment anything,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 4015–4026
2023
Later among the works it cites.
2023
Later among the works it cites.
G. Xu et al. , “Iterative geometry encoding volume for stereo matching,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 21 919–21 928
2023
Later among the works it cites.
H. Xu et al. , “Unifying flow, stereo and depth estimation,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 45, no. 11, pp. 13 941–13 958, 2023
2023
Later among the works it cites.
Z. Chen et al. , “Vision Transformer adapter for dense predictions,” in International Conference on Learning Representations (ICLR) , 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Sun et al. , “LoFTR: Detector-free local feature matching with Transformers,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 8922–8931
2021
Cited alongside, same era.
A. Dosovitskiy et al. , “An image is worth 16x16 words: Transformers for image recognition at scale,” in International Conference on Learning Representations (ICLR) , 2021
2021
Cited alongside, same era.
N. Park and S. Kim, “How do vision Transformers work?” in International Conference on Learning Representations (ICLR) , 2021
2021
Cited alongside, same era.
V. Tankovich et al. , “HITnet: Hierarchical iterative tile refinement network for real-time stereo matching,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 14 362–14 372
2021
Cited alongside, same era.
Y. Li et al. , “Exploring plain vision Transformer backbones for object detection,” in European Conference on Computer Vision (ECCV) . Springer, 2022, pp. 280–296
2022
Cited alongside, same era.
J. Li et al. , “Practical stereo matching via cascaded recurrent network with adaptive correlation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 16 263–16 272
2022
Cited alongside, same era.
W. Guo et al. , “Context-enhanced stereo Transformer,” in European Conference on Computer Vision (ECCV) . Springer, 2022, pp. 263–279
2022
Cited alongside, same era.
P. Weinzaepfel et al. , “CroCo: Self-supervised pre-training for 3D vision tasks by cross-view completion,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 35, pp. 3502–3516, 2022
2022
Cited alongside, same era.
2023
Later among the works it cites.
P. Weinzaepfel et al. , “CroCo v2: Improved cross-view completion pre-training for stereo matching and optical flow,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 17 969–17 980
2023
Later among the works it cites.
Y. Quan et al. , “Centralized feature pyramid for object detection,” IEEE Transactions on Image Processing , vol. 32, pp. 4341–4354, 2023
2023
Later among the works it cites.
G. Xu et al. , “Accurate and efficient stereo matching via attention concatenation volume,” IEEE Transactions on Pattern Analysis and Machine Intelligence , pp. 1–13, 2023
2023
Later among the works it cites.
Y. Fang et al. , “Unleashing vanilla vision Transformer with masked image modeling for object detection,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 6244–6253
2023
Later among the works it cites.
G. Wang et al. , “Cross-level attentive feature aggregation for change detection,” IEEE Transactions on Circuits and Systems for Video Technology , 2023, doi: 10.1109/TCSVT.2023.3344092
2023
Later among the works it cites.
W. Liu et al. , “Learning to upsample by learning to sample,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 6027–6037
2023
Later among the works it cites.
Y. Liu et al. , “The devil is in the upsampling: Architectural decisions made simpler for denoising with deep image prior,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 12 408–12 417
2023
Later among the works it cites.
Z. Shen et al. , “Digging into uncertainty-based pseudo-label for robust stereo matching,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 45, no. 12, pp. 14 301–14 320, 2023
2023
Later among the works it cites.
O.-H. Kwon and E. Zell, “Image-coupled volume propagation for stereo matching,” in 2023 IEEE International Conference on Image Processing (ICIP) . IEEE, 2023, pp. 2510–2514
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2024
Closest in time.
Z. Liu et al. , “Global occlusion-aware Transformer for robust stereo matching,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , 2024, pp. 3535–3544
2024
Closest in time.