Fetching the paper…
Reading the bibliography…
Image captioning implies automatically generating textual descriptions of images based only on the visual input.
Panofsky, E.: Studies in iconology. humanistic themes in the art of the renaissance, new york. New York: Harper and Row (1972)
1972
Earlier work this paper cites.
Couprie, L.D.: Iconclass: an iconographic classification system. Art Libraries Journal 8
1983
Earlier work this paper cites.
Papineni, K., Roukos, S., Ward, T., Zhu, W.J.: Bleu: a method for automatic evaluation of machine translation. In: Proceedings of the 40th annual meeting of the Association for Computational Linguistics. pp. 311–318 (2002)
2002
Earlier work this paper cites.
Lin, C.Y.: Rouge: A package for automatic evaluation of summaries. In: Text summarization branches out. pp. 74–81 (2004)
2004
Earlier work this paper cites.
Crowley, E.J., Zisserman, A.: In search of art. In: European Conference on Computer Vision. pp. 54–70. Springer (2014)
2014
Earlier work this paper cites.
Denkowski, M., Lavie, A.: Meteor universal: Language specific translation evaluation for any target language. In: Proceedings of the ninth workshop on statistical machine translation. pp. 376–380 (2014)
2014
Earlier work this paper cites.
Lin, T.Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., Zitnick, C.L.: Microsoft coco: Common objects in context. In: European conference on computer vision. pp. 740–755. Springer (2014)
2014
Earlier work this paper cites.
Young, P., Lai, A., Hodosh, M., Hockenmaier, J.: From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions. Transactions of the Association for Computational Linguistics 2
2014
Earlier work this paper cites.
Ren, S., He, K., Girshick, R., Sun, J.: Faster r-cnn: Towards real-time object detection with region proposal networks. In: Advances in neural information processing systems. pp. 91–99 (2015)
2015
Earlier work this paper cites.
Vedantam, R., Lawrence Zitnick, C., Parikh, D.: Cider: Consensus-based image description evaluation. In: Proceedings of the IEEE conference on computer vision and pattern recognition. pp. 4566–4575 (2015)
2015
Earlier work this paper cites.
Vinyals, O., Toshev, A., Bengio, S., Erhan, D.: Show and tell: A neural image caption generator. In: Proceedings of the IEEE conference on computer vision and pattern recognition. pp. 3156–3164 (2015)
2015
Earlier work this paper cites.
Seguin, B., Striolo, C., Kaplan, F., et al.: Visual link retrieval in a database of paintings. In: European Conference on Computer Vision. pp. 753–767. Springer (2016)
2016
Earlier work this paper cites.
Hayn-Leichsenring, G.U., Lehmann, T., Redies, C.: Subjective ratings of beauty and aesthetics: correlations with statistical image properties in western oil paintings. i-Perception 8
2017
Earlier work this paper cites.
Krishna, R., Zhu, Y., Groth, O., Johnson, J., Hata, K., Kravitz, J., Chen, S., Kalantidis, Y., Li, L.J., Shamma, D.A., et al.: Visual genome: Connecting language and vision using crowdsourced dense image annotations. International journal of computer vision 123
2017
Earlier work this paper cites.
Baraldi, L., Cornia, M., Grana, C., Cucchiara, R.: Aligning text and document illustrations: towards visually explainable digital humanities. In: 2018 24th International Conference on Pattern Recognition (ICPR). pp. 1097–1102. IEEE (2018)
2018
Cited alongside, same era.
Cetinic, E., Lipic, T., Grgic, S.: Fine-tuning convolutional neural networks for fine art classification. Expert Systems with Applications 114
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Elgammal, A., Liu, B., Kim, D., Elhoseiny, M., Mazzone, M.: The shape of art history in the eyes of the machine. In: 32nd AAAI Conference on Artificial Intelligence, AAAI 2018. pp. 2183–2191. AAAI press (2018)
2018
Cited alongside, same era.
Sheng, S., Moens, M.F.: Generating captions for images of ancient artworks. In: Proceedings of the 27th ACM International Conference on Multimedia. pp. 2478–2486 (2019)
2019
Later among the works it cites.
Stefanini, M., Cornia, M., Baraldi, L., Corsini, M., Cucchiara, R.: Artpedia: A new visual-semantic dataset with visual and contextual sentences in the artistic domain. In: International Conference on Image Analysis and Processing. pp. 729–740. Springer (2019)
2019
Later among the works it cites.
2019
Later among the works it cites.
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Garcia, N., Vogiatzis, G.: How to read paintings: semantic art understanding with multi-modal retrieval. In: Proceedings of the European Conference on Computer Vision (ECCV). pp. 0–0 (2018)
2018
Cited alongside, same era.
Sharma, P., Ding, N., Goodman, S., Soricut, R.: Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning. In: Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). pp. 2556–2565 (2018)
2018
Cited alongside, same era.
Strezoski, G., Worring, M.: Omniart: a large-scale artistic benchmark. ACM Transactions on Multimedia Computing, Communications, and Applications (TOMM) 14
2018
Cited alongside, same era.
Cetinic, E., Lipic, T., Grgic, S.: A deep learning perspective on beauty, sentiment, and remembrance of art. IEEE Access 7
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Jenicek, T., Chum, O.: Linking art through human poses. In: 2019 International Conference on Document Analysis and Recognition (ICDAR). pp. 1338–1345. IEEE (2019)
2019
Cited alongside, same era.
Lu, J., Batra, D., Parikh, D., Lee, S.: Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks. In: Advances in Neural Information Processing Systems. pp. 13–23 (2019)
2019
Cited alongside, same era.
Madhu, P., Kosti, R., Mührenberg, L., Bell, P., Maier, A., Christlein, V.: Recognizing characters in art history using deep learning. In: Proceedings of the 1st Workshop on Structuring and Understanding of Multimedia heritAge Contents. pp. 15–22 (2019)
2019
Cited alongside, same era.
Castellano, G., Vessio, G.: Towards a tool for visual link retrieval and knowledge discovery in painting datasets. In: Italian Research Conference on Digital Libraries. pp. 105–110. Springer (2020)
2020
Later among the works it cites.
Cetinic, E., Lipic, T., Grgic, S.: Learning the principles of art history with convolutional neural networks. Pattern Recognition Letters 129
2020
Later among the works it cites.
Cornia, M., Stefanini, M., Baraldi, L., Corsini, M., Cucchiara, R.: Explaining digital humanities by aligning images and textual descriptions. Pattern Recognition Letters 129
2020
Later among the works it cites.
Deng, Y., Tang, F., Dong, W., Ma, C., Huang, F., Deussen, O., Xu, C.: Exploring the representativity of art paintings. IEEE Transactions on Multimedia (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
Gupta, J., Madhu, P., Kosti, R., Bell, P., Maier, A., Christlein, V.: Towards image caption generation for art historical data. AI methods for digital heritage, Workshop at KI2020 43rd German Conference on Artificial Intelligence (2020)
2020
Later among the works it cites.
Posthumus, E.: Brill iconclass ai test set (2020)
2020
Later among the works it cites.
Sargentis, G., Dimitriadis, P., Koutsoyiannis, D., et al.: Aesthetical issues of leonardo da vinci’s and pablo picasso’s paintings with stochastic evaluation. Heritage 3
2020
Later among the works it cites.
2020
Later among the works it cites.
Zhou, L., Palangi, H., Zhang, L., Hu, H., Corso, J.J., Gao, J.: Unified vision-language pre-training for image captioning and vqa. In: AAAI. pp. 13041–13049 (2020)
2020
Later among the works it cites.