Fetching the paper…
Reading the bibliography…
In this paper we introduce the FooDI-ML dataset.
Food-101–mining discriminative components with random forests
Lukas Bossard, Matthieu Guillaumin, and Luc Van Gool · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Peter Young, Alice Lai, Micah Hodosh, and Julia Hockenmaier · 2014
Earlier work this paper cites.
Deep residual learning for image recognition, 2015
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Deep-based ingredient recognition for cooking recipe retrieval
Jingjing Chen and Chong-Wah Ngo · 2016
Earlier work this paper cites.
Multimodal pivots for image caption translation
Julian Hitschler, Shigehiko Schamoni, and Stefan Riezler · 2016
Earlier work this paper cites.
Bag of tricks for efficient text classification
Armand Joulin, Edouard Grave, Piotr Bojanowski, and Tomas Mikolov · 2016
Earlier work this paper cites.
Cross-lingual image caption generation
Takashi Miyazaki and Nobuyuki Shimizu · 2016
Earlier work this paper cites.
Yfcc100m: The new data in multimedia research
Bart Thomee, David A Shamma, Gerald Friedland, Benjamin Elizalde, Karl Ni, Douglas Poland, Damian Borth, and Li-Jia Li · 2016
Earlier work this paper cites.
Chinesefoodnet: A large-scale image dataset for chinese food recognition
Xin Chen, Yu Zhu, Hua Zhou, Liang Diao, and Dongyan Wang · 2017
Earlier work this paper cites.
Learning cross-modal embeddings for cooking recipes and food images
Amaia Salvador, Nicholas Hynes, Yusuf Aytar, Javier Marin, Ferda Ofli, Ingmar Weber, and Antonio Torralba · 2017
Cited alongside, same era.
Revisiting unreasonable effectiveness of data in deep learning era
Chen Sun, Abhinav Shrivastava, Saurabh Singh, and Abhinav Gupta · 2017
Cited alongside, same era.
Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning
Piyush Sharma, Nan Ding, Sebastian Goodman, and Radu Soricut · 2018
Cited alongside, same era.
Didec: The dutch image description and eye-tracking corpus
Emiel van Miltenburg, Akos Kádár, Ruud Koolen, and Emiel Krahmer · 2018
Cited alongside, same era.
Regularized uncertainty-based multi-task learning model for food analysis
Eduardo Aguilar, Marc Bolaños, and Petia Radeva · 2019
Cited alongside, same era.
Large scale datasets for image and video captioning in italian
Billion-scale semi-supervised learning for image classification
I Zeki Yalniz, Hervé Jégou, Kan Chen, Manohar Paluri, and Dhruv Mahajan · 2019
Later among the works it cites.
Self-attention generative adversarial networks
Han Zhang, Ian Goodfellow, Dimitris Metaxas, and Augustus Odena · 2019
Later among the works it cites.
Ms-coco-es: Spanish coco captions, Oct 2020
Carlos Garcia · 2020
Later among the works it cites.
The open images dataset v4
Alina Kuznetsova, Hassan Rom, Neil Alldrin, Jasper Uijlings, Ivan Krasin, Jordi Pont-Tuset, Shahab Kamali, Stefan Popov, Matteo Malloci, Alexander Kolesnikov, et al · 2020
Later among the works it cites.
Uit-viic: A dataset for the first evaluation on vietnamese image captioning
Quan Hoang Lam, Quang Duy Le, Van Kiet Nguyen, and Ngan Luu-Thuy Nguyen · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Scaiella Antonio, Danilo Croce, and Roberto Basili · 2019
Cited alongside, same era.
Foodx-251: A dataset for fine-grained food classification
Parneet Kaur, Karan Sikka, Weijun Wang, Serge Belongie, and Ajay Divakaran · 2019
Cited alongside, same era.
Coco-cn for cross-lingual image tagging, captioning, and retrieval
Xirong Li, Chaoxi Xu, Xiaoxu Wang, Weiyu Lan, Zhengxiong Jia, Gang Yang, and Jieping Xu · 2019
Cited alongside, same era.
Recipe1m+: A dataset for learning cross-modal embeddings for cooking recipes and food images
Javier Marin, Aritro Biswas, Ferda Ofli, Nicholas Hynes, Amaia Salvador, Yusuf Aytar, Ingmar Weber, and Antonio Torralba · 2019
Cited alongside, same era.
Camp: Cross-modal adaptive message passing for text-image retrieval
Zihao Wang, Xihui Liu, Hongsheng Li, Lu Sheng, Junjie Yan, Xiaogang Wang, and Jing Shao · 2019
Cited alongside, same era.
Weiqing Min, Linhu Liu, Zhiling Wang, Zhengdong Luo, Xiaoming Wei, Xiaolin Wei, and Shuqiang Jiang · 2020
Later among the works it cites.
Foodi-ml-dataset/notebooks at main · glovo/foodi-ml-dataset, 2021
Glovo App · 2021
Closest in time.
Learning transferable visual models from natural language supervision, 2021
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Closest in time.
Using deep learning for ranking in dish search, https://bytes.swiggy.com/using-deep-learning-for-ranking-in-dish-search-4df2772dddce, Jul 2021
Ramkishore Saravanan · 2021
Closest in time.
Wit: Wikipedia-based image text dataset for multimodal multilingual machine learning
Krishna Srinivasan, Karthik Raman, Jiecao Chen, Michael Bendersky, and Marc Najork · 2021
Closest in time.