Fetching the paper…
Reading the bibliography…
Most image captioning models following an autoregressive manner suffer from significant inference latency.
Hochreiter, S., Schmidhuber, J.: Long short-term memory. Neural Comput. (1997)
1997
Earlier work this paper cites.
Papineni, K., Roukos, S., Ward, T., Zhu, W.J.: Bleu: a method for automatic evaluation of machine translation. In: Proc. of ACL (2002)
2002
Earlier work this paper cites.
Lin, C.Y.: Rouge: A package for automatic evaluation of summaries. In: Text summarization branches out (2004)
2004
Earlier work this paper cites.
Banerjee, S., Lavie, A.: Meteor: An automatic metric for mt evaluation with improved correlation with human judgments. In: Proc. of ACL workshop (2005)
2005
Earlier work this paper cites.
2015
Earlier work this paper cites.
Hinton, G.E., Vinyals, O., Dean, J.: Distilling the knowledge in a neural network. CoRR (2015)
2015
Earlier work this paper cites.
Karpathy, A., Fei-Fei, L.: Deep visual-semantic alignments for generating image descriptions. In: Proc. of CVPR (2015)
2015
Earlier work this paper cites.
Ren, S., He, K., Girshick, R.B., Sun, J.: Faster R-CNN: towards real-time object detection with region proposal networks. In: Proc. of NeurIPS (2015)
2015
Earlier work this paper cites.
Vedantam, R., Lawrence Zitnick, C., Parikh, D.: Cider: Consensus-based image description evaluation. In: Proc. of CVPR (2015)
2015
Earlier work this paper cites.
Vinyals, O., Toshev, A., Bengio, S., Erhan, D.: Show and tell: A neural image caption generator. In: Proc. of CVPR (2015)
2015
Earlier work this paper cites.
Anderson, P., Fernando, B., Johnson, M., Gould, S.: Spice: Semantic propositional image caption evaluation. In: Proc. of ECCV (2016)
2016
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. In: Proc. of CVPR (2016)
2016
Cited alongside, same era.
Krishna, R., Zhu, Y., Groth, O., Johnson, J., Hata, K., Kravitz, J., Chen, S., Kalantidis, Y., Li, L., Shamma, D.A., Bernstein, M.S., Fei-Fei, L.: Visual genome: Connecting language and vision using crowdsourced dense image annotations. Int. J. Comput. Vis. (2017)
2017
Cited alongside, same era.
Rennie, S.J., Marcheret, E., Mroueh, Y., Ross, J., Goel, V.: Self-critical sequence training for image captioning. In: Proc. of CVPR (2017)
2017
Cited alongside, same era.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, Ł., Polosukhin, I.: Attention is all you need. Proc. of NeurIPS (2017)
2017
Cited alongside, same era.
Cornia, M., Stefanini, M., Baraldi, L., Cucchiara, R.: Meshed-memory transformer for image captioning. In: Proc. of CVPR (2020)
2020
Later among the works it cites.
Fei, Z.: Iterative back modification for faster image captioning. In: Proc. of ACM MM (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
Luo, R.: A better variant of self-critical sequence training. CoRR (2020)
2020
Later among the works it cites.
Zhang, Y., Zhang, Y., Qi, P., Manning, C.D., Langlotz, C.P.: Biomedical and clinical english model packages in the stanza python NLP library. CoRR (2020)
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Kaiser, L., Bengio, S., Roy, A., Vaswani, A., Parmar, N., Uszkoreit, J., Shazeer, N.: Fast decoding in sequence models using discrete latent variables. In: Proc. of ICML (2018)
2018
Cited alongside, same era.
Zhang, Y., Xiang, T., Hospedales, T.M., Lu, H.: Deep mutual learning. In: Proc. of CVPR (2018)
2018
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Huang, L., Wang, W., Chen, J., Wei, X.Y.: Attention on attention for image captioning. In: Proc. of ICCV (2019)
2019
Cited alongside, same era.
Fei, Z.: Partially non-autoregressive image captioning. In: Proc. of AAAI (2021)
2021
Later among the works it cites.
Song, Z., Zhou, X., Dong, L., Tan, J., Guo, L.: Direction relation transformer for image captioning. In: Proc. of ACM MM (2021)
2021
Later among the works it cites.
Yan, X., Fei, Z., Li, Z., Wang, S., Huang, Q., Tian, Q.: Semi-autoregressive image captioning. In: Proc. of ACM MM (2021)
2021
Later among the works it cites.
Zhou, Y., Zhang, Y., Hu, Z., Wang, M.: Semi-autoregressive transformer for image captioning. In: Proc. of ICCV (2021)
2021
Later among the works it cites.
Li, Y., Pan, Y., Yao, T., Mei, T.: Comprehending and ordering semantics for image captioning. In: Proc. of CVPR (2022)
2022
Later among the works it cites.