Fetching the paper…
Reading the bibliography…
Object navigation in open-world environments remains a formidable and pervasive challenge for robotic systems, particularly when it comes to executing long-horizon tasks that require both open-world object detection and high-level task planning.
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi, “You only look once: Unified, real-time object detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 779–788
2016
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster R-CNN: Towards real-time object detection with region proposal networks,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, no. 6, pp. 1137–1149, 2016
2016
Earlier work this paper cites.
F. Zhong, W. Weichao Qiu, T. Yan, A. Yuille, and Y. Wang, “Gym-UnrealCV: Realistic virtual worlds for visual reinforcement learning,” Web Page, 2017. [Online]. Available: https://github.com/unrealcv/gym-unrealcv
2017
Earlier work this paper cites.
G. Bhat, M. Danelljan, L. V. Gool, and R. Timofte, “Learning discriminative model prediction for tracking,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 6182–6191
2019
Earlier work this paper cites.
W. Luo, P. Sun, F. Zhong, W. Liu, T. Zhang, and Y. Wang, “End-to-end active object tracking and its real-world deployment via reinforcement learning,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 42, no. 6, pp. 1317–1332, 2019
2019
Earlier work this paper cites.
F. Zhong, P. Sun, W. Luo, T. Yan, and Y. Wang, “AD-VAT: An asymmetric dueling mechanism for learning visual active tracking,” in International Conference on Learning Representations , 2019
2019
Earlier work this paper cites.
——, “AD-VAT+: An asymmetric dueling mechanism for learning and understanding visual active tracking,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 43, no. 5, pp. 1467–1482, 2019
2019
Earlier work this paper cites.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” pp. 213–229, 2020
2020
Earlier work this paper cites.
M. Caron, H. Touvron, I. Misra, H. Jégou, J. Mairal, P. Bojanowski, and A. Joulin, “Emerging properties in self-supervised vision transformers,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 9650–9660
2021
Earlier work this paper cites.
——, “Towards distraction-robust active visual tracking,” in International Conference on Machine Learning , 2021, pp. 12 782–12 792
2021
Earlier work this paper cites.
P. Jiang, D. Ergu, F. Liu, Y. Cai, and B. Ma, “A review of yolo algorithm developments,” Procedia Computer Science , vol. 199, pp. 1066–1073, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Cited alongside, same era.
S. Liu, Z. Zeng, T. Ren, F. Li, H. Zhang, J. Yang, Q. Jiang, C. Li, J. Yang, H. Su et al. , “Grounding DINO: Marrying DINO with grounded pre-training for open-set object detection,” in European Conference on Computer Vision , 2024, pp. 38–55
2024
Later among the works it cites.
2024
Later among the works it cites.
G. Jocher and J. Qiu, “Ultralytics YOLO11,” 2024. [Online]. Available: https://github.com/ultralytics/ultralytics
2024
Later among the works it cites.
F. Zhong, K. Wu, H. Ci, C. Wang, and H. Chen, “Empowering embodied visual tracking with visual foundation models and offline RL,” in European Conference on Computer Vision , 2024, pp. 139–155
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Liang, W. Huang, F. Xia, P. Xu, K. Hausman, B. Ichter, P. Florence, and A. Zeng, “Code as Policies: Language model programs for embodied control,” in 2023 IEEE International Conference on Robotics and Automation , 2023, pp. 9493–9500
2023
Cited alongside, same era.
2023
Cited alongside, same era.
F. Zhong, X. Bi, Y. Zhang, W. Zhang, and Y. Wang, “RSPT: reconstruct surroundings and predict trajectory for generalizable active object tracking,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 37, no. 3, 2023, pp. 3705–3714
2023
Cited alongside, same era.
A. Grattafiori, A. Dubey, A. Jauhri, A. Pandey, A. Kadian, A. Al-Dahle, A. Letman, A. Mathur, A. Schelten, A. Vaughan et al. , “The LLaMA 3 herd of models,” arXiv e-prints , pp. arXiv–2407, 2024
2024
Cited alongside, same era.
2024
Cited alongside, same era.
Q. Zhang, P. Cui, D. Yan, J. Sun, Y. Duan, G. Han, W. Zhao, W. Zhang, Y. Guo, A. Zhang et al. , “Whole-body humanoid robot locomotion with human reference,” in 2024 IEEE/RSJ International Conference on Intelligent Robots and Systems , 2024, pp. 11 225–11 231
2024
Cited alongside, same era.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.