Fetching the paper…
Reading the bibliography…
Pretrained visual-language models have extensive world knowledge and are widely used in visual and language navigation (VLN).
A. X. Chang, A. Dai, T. A. Funkhouser, M. Halber, M. Nießner, et al
2017
Earlier work this paper cites.
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Embodied question answering,” in Proc. CVPR
2018
Earlier work this paper cites.
P. Anderson, A. X. Chang, D. S. Chaplot, et al
2018
Earlier work this paper cites.
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, et al
2018
Earlier work this paper cites.
J. Devlin, M. Chang, et al
2019
Earlier work this paper cites.
F. Petroni, T. Rocktäschel, S. Riedel, P. S. H. Lewis, A. Bakhtin, Y. Wu, and A. H. Miller, “Language models as knowledge bases?,” in EMNLP/IJCNLP (1)
2019
Earlier work this paper cites.
X. Li, C. Li, Q. Xia, Y. Bisk, A. Celikyilmaz, et al
2019
Earlier work this paper cites.
H. Tan, L. Yu, et al
2019
Earlier work this paper cites.
X. Wang, Q. Huang, A. Celikyilmaz, J. Gao, et al
2019
Earlier work this paper cites.
C. Ma, J. Lu, Z. Wu, G. AlRegib, Z. Kira, et al
2019
Earlier work this paper cites.
L. Ke, X. Li, et al
2019
Earlier work this paper cites.
Y. Qi, Q. Wu, P. Anderson, et al
2020
Earlier work this paper cites.
A. Majumdar, A. Shrivastava, et al
2020
Earlier work this paper cites.
W. Hao, C. Li, X. Li, L. Carin, et al
2020
Earlier work this paper cites.
T. B. Brown, B. Mann, et al
2020
Cited alongside, same era.
W. Hao, C. Li, X. Li, L. Carin, et al
2020
Cited alongside, same era.
X. Li, X. Yin, Li, et al
2020
Cited alongside, same era.
Y. Hong, C. R. Opazo, Q. Wu, and S. Gould, “Sub-instruction aware vision-and-language navigation,” in EMNLP (1)
2020
Cited alongside, same era.
F. Zhu, Y. Zhu, X. Chang, et al
2020
Cited alongside, same era.
P.-L. Guhur, M. Tapaswi, S. Chen, et al
2021
Cited alongside, same era.
C. Liu, F. Zhu, X. Chang, et al
2021
Later among the works it cites.
Y. Qi, Z. Pan, Y. Hong, Q. Wu, et al
2021
Later among the works it cites.
D. An, Y. Qi, Q. Wu, et al
2021
Later among the works it cites.
B. Lin, Y. Zhu, Z. Chen, et al
2022
Later among the works it cites.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,” Int. J. Comput. Vis
2022
Later among the works it cites.
M. Jia, L. Tang, B. Chen, C. Cardie, S. J. Belongie, B. Hariharan, and S. Lim, “Visual prompt tuning,” in Computer Vision - ECCV 2022 - 17th European Conference, Tel Aviv, Israel, October 23-27, 2022, Proceedings, Part XXXIII
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
B. Lester, R. Al-Rfou, and N. Constant, “The power of scale for parameter-efficient prompt tuning,” in EMNLP (1)
2021
Cited alongside, same era.
A. Radford, J. W. Kim, et al
2021
Cited alongside, same era.
Y. Yao, A. Zhang, Z. Liu, et al
2021
Cited alongside, same era.
Y. Hong, Q. Wu, et al
2021
Cited alongside, same era.
X. Liu, Y. Zheng, Z. Du, M. Ding, Y. Qian, Z. Yang, and J. Tang, “GPT understands, too,” CoRR
2021
Cited alongside, same era.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,” International Journal of Computer Vision
2022
Later among the works it cites.
M. Li, L. Chen, Y. Duan, Z. Hu, J. Feng, J. Zhou, and J. Lu, “Bridge-prompt: Towards ordinal action understanding in instructional videos,” in Proc. CVPR
2022
Later among the works it cites.
J. Chen, C. Gao, E. Meng, et al
2022
Later among the works it cites.
X. Liang, F. Zhu, L. Li, H. Xu, and X. Liang, “Visual-language navigation pretraining via prompt-based environmental self-exploration,” in ACL (1)
2022
Later among the works it cites.
T. Liu, Y. Hu, W. Wu, Y. Wang, K. Xu, and Q. Yin, “Dap: Domain-aware prompt learning for vision-and-language navigation,” 2023
2023
Closest in time.
X. Liu, S. Huang, Y. Kang, H. Chen, and D. Wang, “VGDiffZero: Text-to-image diffusion models can be zero-shot visual grounders,” 2023
2023
Closest in time.
Z. Zhang, S. Qi, Z. Zhou, et al
2023
Closest in time.