Fetching the paper…
Reading the bibliography…
With the advancement of collaborative perception, the role of aerial-ground collaborative perception, a crucial component, is becoming increasingly important.
A. Torralba and A. Oliva, “Depth estimation from image structure,” IEEE Transactions on pattern analysis and machine intelligence , vol. 24, no. 9, pp. 1226–1238, 2002
2002
Earlier work this paper cites.
P. Krähenbühl and V. Koltun, “Efficient inference in fully connected crfs with gaussian edge potentials,” Neural Information Processing Systems,Neural Information Processing Systems , Dec 2011
2011
Earlier work this paper cites.
F. Liu, C. Shen, and G. Lin, “Deep convolutional neural fields for depth estimation from a single image,” in 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , Jun 2015. [Online]. Available: http://dx.doi.org/10.1109/cvpr.2015.7299152
2015
Earlier work this paper cites.
S. Zheng, S. Jayasumana, B. Romera-Paredes, V. Vineet, Z. Su, D. Du, C. Huang, and P. H. Torr, “Conditional random fields as recurrent neural networks,” pp. 1529–1537, 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
Y. Cao, Z. Wu, and C. Shen, “Estimating depth from monocular images as classification using deep fully convolutional residual networks,” IEEE Transactions on Circuits and Systems for Video Technology , p. 3174–3182, Nov 2018. [Online]. Available: http://dx.doi.org/10.1109/tcsvt.2017.2740321
2017
Earlier work this paper cites.
T.-Y. Lin, P. Dollár, R. Girshick, K. He, B. Hariharan, and S. Belongie, “Feature pyramid networks for object detection,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 2117–2125
2017
Earlier work this paper cites.
P. Zhu, L. Wen, D. Du, X. Bian, H. Ling, Q. Hu, Q. Nie, H. Cheng, C. Liu, X. Liu, et al. , “Visdrone-det2018: The vision meets drone object detection in image challenge results,” in Proceedings of the European Conference on Computer Vision (ECCV) Workshops , 2018, pp. 0–0
2018
Earlier work this paper cites.
D. Du, P. Zhu, L. Wen, X. Bian, H. Lin, Q. Hu, T. Peng, J. Zheng, X. Wang, Y. Zhang, et al. , “Visdrone-det2019: The vision meets drone object detection in image challenge results,” in Proceedings of the IEEE/CVF international conference on computer vision workshops , 2019, pp. 0–0
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y.-C. Liu, J. Tian, C.-Y. Ma, N. Glaser, C.-W. Kuo, and Z. Kira, “Who2com: Collaborative perception via learnable handshake communication,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2020, pp. 6876–6883
2020
Earlier work this paper cites.
T.-H. Wang, S. Manivasagam, M. Liang, B. Yang, W. Zeng, and R. Urtasun, “V2vnet: Vehicle-to-vehicle communication for joint perception and prediction,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part II 16 . Springer, 2020, pp. 605–621
2020
Earlier work this paper cites.
Y.-C. Liu, J. Tian, N. Glaser, and Z. Kira, “When2com: Multi-agent perception via communication graph grouping,” in 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2020, pp. 4105–4114
2020
Earlier work this paper cites.
X. Shi, Z. Chen, and T.-K. Kim, “Distance-normalized unified representation for monocular 3d object detection,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXIX 16 . Springer, 2020, pp. 91–107
2020
Earlier work this paper cites.
J. Philion and S. Fidler, “Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XIV 16 . Springer, 2020, pp. 194–210
2020
Earlier work this paper cites.
N. Vadivelu, M. Ren, J. Tu, J. Wang, and R. Urtasun, “Learning to communicate and correct pose errors,” in Conference on Robot Learning . PMLR, 2021, pp. 1195–1210
2021
Earlier work this paper cites.
S. Lan, Z. Yu, C. Choy, S. Radhakrishnan, G. Liu, Y. Zhu, L. S. Davis, and A. Anandkumar, “Discobox: Weakly supervised instance segmentation and semantic correspondence from box supervision,” in 2021 IEEE/CVF International Conference on Computer Vision (ICCV) , Oct 2021. [Online]. Available: http://dx.doi.org/10.1109/iccv48922.2021.00339
2021
Earlier work this paper cites.
Y. Cao, Z. He, L. Wang, W. Wang, Y. Yuan, D. Zhang, J. Zhang, P. Zhu, L. Van Gool, J. Han, et al. , “Visdrone-det2021: The vision meets drone object detection challenge results,” in Proceedings of the IEEE/CVF International conference on computer vision , 2021, pp. 2847–2854
2021
Earlier work this paper cites.
Y. Ming, X. Meng, C. Fan, and H. Yu, “Deep learning for monocular depth estimation: A review,” Neurocomputing , vol. 438, pp. 14–33, 2021
2021
Cited alongside, same era.
T. Wang, X. Zhu, J. Pang, and D. Lin, “Fcos3d: Fully convolutional one-stage monocular 3d object detection,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 913–922
2021
Cited alongside, same era.
——, “Probabilistic and geometric depth: Detecting objects in perspective,” 5th Annual Conference on Robot Learning,5th Annual Conference on Robot Learning , Jun 2021
2021
Cited alongside, same era.
C. Reading, A. Harakeh, J. Chae, and S. L. Waslander, “Categorical depth distribution network for monocular 3d object detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 8555–8564
2021
Cited alongside, same era.
R. Xu, J. Li, X. Dong, H. Yu, and J. Ma, “Bridging the domain gap for multi-agent perception,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 6035–6042
2023
Later among the works it cites.
S. Lan, X. Yang, Z. Yu, Z. Wu, J. M. Alvarez, and A. Anandkumar, “Vision transformers are good mask auto-labelers,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 23 745–23 755
2023
Later among the works it cites.
J. Suo, T. Wang, X. Zhang, H. Chen, W. Zhou, and W. Shi, “Hit-uav: A high-altitude infrared thermal dataset for unmanned aerial vehicle-based object detection,” Scientific Data , vol. 10, no. 1, p. 227, 2023
2023
Later among the works it cites.
A. Dutta, S. Das, J. Nielsen, R. Chakraborty, and M. Shah, “Multiview aerial visual recognition (mavrec): Can multi-view improve aerial visual perception?” 2023
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Yin, X. Zhou, and P. Krahenbuhl, “Center-based 3d object detection and tracking,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 11 784–11 793
2021
Cited alongside, same era.
Y. Hu, S. Fang, Z. Lei, Y. Zhong, and S. Chen, “Where2comm: Communication-efficient collaborative perception via spatial confidence maps,” Advances in neural information processing systems , vol. 35, pp. 4874–4886, 2022
2022
Cited alongside, same era.
C. Chen, Z. Liao, Y. Ju, C. He, K. Yu, and S. Wan, “Hierarchical domain-based multicontroller deployment strategy in sdn-enabled space–air–ground integrated network,” IEEE Transactions on Aerospace and Electronic Systems , vol. 58, no. 6, pp. 4864–4879, 2022
2022
Cited alongside, same era.
J. Cui, H. Qiu, D. Chen, P. Stone, and Y. Zhu, “Coopernaut: End-to-end driving with cooperative perception for networked vehicles,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 17 252–17 262
2022
Cited alongside, same era.
W. Yuan, X. Gu, Z. Dai, S. Zhu, and P. Tan, “Newcrfs: Neural window fully-connected crfs for monocular depth estimation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2022
2022
Cited alongside, same era.
Y. Sun, B. Cao, P. Zhu, and Q. Hu, “Drone-based rgb-infrared cross-modality vehicle detection via uncertainty-aware learning,” IEEE Transactions on Circuits and Systems for Video Technology , vol. 32, no. 10, pp. 6700–6713, 2022
2022
Cited alongside, same era.
H. Yu, Y. Luo, M. Shu, Y. Huo, Z. Yang, Y. Shi, Z. Guo, H. Li, X. Hu, J. Yuan, and Z. Nie, “Dair-v2x: A large-scale dataset for vehicle-infrastructure cooperative 3d object detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2022, pp. 21 361–21 370
2022
Cited alongside, same era.
Y. Li, D. Ma, Z. An, Z. Wang, Y. Zhong, S. Chen, and C. Feng, “V2x-sim: Multi-agent collaborative perception dataset and benchmark for autonomous driving,” IEEE Robotics and Automation Letters , vol. 7, no. 4, pp. 10 914–10 921, 2022
2022
Cited alongside, same era.
H. Xiang, R. Xu, and J. Ma, “Hm-vit: Hetero-modal vehicle-to-vehicle cooperative perception with vision transformer,” in 2023 IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 284–295
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Yu, Y. Tang, E. Xie, J. Mao, P. Luo, and Z. Nie, “Flow-based feature fusion for vehicle-infrastructure cooperative 3d object detection,” in Advances in Neural Information Processing Systems , 2023
2023
Later among the works it cites.
Z. Chen, Y. Shi, and J. Jia, “Transiff: An instance-level feature fusion framework for vehicle-infrastructure cooperative 3d detection with transformers,” in 2023 IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 18 159–18 168
2023
Later among the works it cites.
Y. Hu, Y. Lu, R. Xu, W. Xie, S. Chen, and Y. Wang, “Collaboration helps camera overtake lidar in 3d detection,” in 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 9243–9252
2023
Later among the works it cites.
X. Yang, Z. Ma, Z. Ji, and Z. Ren, “Gedepth: Ground embedding for monocular depth estimation,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 12 719–12 727
2023
Later among the works it cites.
Y. Li, Z. Ge, G. Yu, J. Yang, Z. Wang, Y. Shi, J. Sun, and Z. Li, “Bevdepth: Acquisition of reliable depth for multi-view 3d object detection,” vol. 37, no. 2, pp. 1477–1485, 2023
2023
Later among the works it cites.
Z. Sun, Y. Liu, L. Zhang, and F. Deng, “Agcg: Air–ground collaboration geolocation based on visual servo with uncalibrated cameras,” IEEE Transactions on Industrial Electronics , 2024
2024
Closest in time.
2024
Closest in time.
W. Su, L. Chen, Y. Bai, X. Lin, G. Li, Z. Qu, and P. Zhou, “What makes good collaborative views? contrastive mutual information maximization for multi-agent perception,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 16, 2024, pp. 17 550–17 558
2024
Closest in time.
X. Li, J. Yin, W. Li, C. Xu, R. Yang, and J. Shen, “Di-v2x: Learning domain-invariant representation for vehicle-infrastructure collaborative 3d object detection,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 4, 2024, pp. 3208–3215
2024
Closest in time.
S. Wei, Y. Wei, Y. Hu, Y. Lu, Y. Zhong, S. Chen, and Y. Zhang, “Asynchrony-robust collaborative perception via bird’s eye view flow,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Z. Wang, S. Fan, X. Huo, T. Xu, Y. Wang, J. Liu, Y. Chen, and Y.-Q. Zhang, “Emiff: Enhanced multi-scale image feature fusion for vehicle-infrastructure cooperative 3d object detection,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) , 2024
2024
Closest in time.