Fetching the paper…
Reading the bibliography…
In this paper, we introduce Recipe1M+, a new large-scale, structured corpus of over one million cooking recipes and 13 million food images.
L. van der Maaten and G. Hinton, “Visualizing data using t-sne,” Journal of Machine Learning Research , vol. 9, pp. 2579–2605, 2008
2008
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in NIPS , 2012
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
B. Zhou, A. Lapedriza, J. Xiao, A. Torralba, and A. Oliva, “Learning deep features for scene recognition using places database,” in Advances in neural information processing systems , 2014, pp. 487–495
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
L. Bossard, M. Guillaumin, and L. Van Gool, “Food-101–mining discriminative components with random forests,” in European Conference on Computer Vision . Springer, 2014, pp. 446–461
2014
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in NIPS , 2014, pp. 3104–3112
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein et al. , “Imagenet large scale visual recognition challenge,” International Journal of Computer Vision , vol. 115, no. 3, pp. 211–252, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. Myers, N. Johnston, V. Rathod, A. Korattikara, A. Gorban, N. Silberman, S. Guadarrama, G. Papandreou, J. Huang, and K. Murphy, “Im2calories: Towards an automated mobile vision food diary,” in ICCV , 2015, pp. 1233–1241
2015
Earlier work this paper cites.
Y. Kawano and K. Yanai, “Foodcam: A real-time food recognition system on a smartphone,” Multimedia Tools and Applications , vol. 74, no. 14, pp. 5263–5287, 2015
2015
Earlier work this paper cites.
R. Xu, L. Herranz, S. Jiang, S. Wang, X. Song, and R. Jain, “Geolocalized modeling for dish recognition,” IEEE Trans. Multimedia , vol. 17, no. 8, pp. 1187–1199, 2015
2015
Cited alongside, same era.
X. Wang, D. Kumar, N. Thome, M. Cord, and F. Precioso, “Recipe recognition with large multimodal food dataset,” in ICME Workshops , 2015, pp. 1–6
2015
Cited alongside, same era.
US Department of Agriculture, Agricultural Research Service, Nutrient Data Laboratory, “Usda national nutrient database for standard reference, release 27,” May 2015. [Online]. Available: http://www.ars.usda.gov/ba/bhnrc/ndl
2015
Cited alongside, same era.
R. Kiros, Y. Zhu, R. Salakhutdinov, R. Zemel, A. Torralba, R. Urtasun, and S. Fidler, “Skip-thought vectors,” in NIPS , 2015, pp. 3294–3302
2015
Cited alongside, same era.
F. Ofli, Y. Aytar, I. Weber, R. Hammouri, and A. Torralba, “Is saki #delicious? the food perception gap on instagram and its relation to health,” in Proceedings of the 26th International Conference on World Wide Web . International World Wide Web Conferences Steering Committee, 2017
2017
Later among the works it cites.
A. Salvador, N. Hynes, Y. Aytar, J. Marin, F. Ofli, I. Weber, and A. Torralba, “Learning cross-modal embeddings for cooking recipes and food images,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , July 2017
2017
Later among the works it cites.
W. Min, S. Jiang, S. Wang, J. Sang, and S. Mei, “A delicious recipe analysis framework for exploring multi-modal recipes with various attributes,” in Proceedings of the 2017 ACM on Multimedia Conference , ser. MM ’17. New York, NY, USA: ACM, 2017, pp. 402–410. [Online]. Available: http://doi.acm.org/10.1145/3123266.3123272
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan, “Show and tell: A neural image caption generator,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 3156–3164
2015
Cited alongside, same era.
B. Zhou, A. Khosla, A. Lapedriza, A. Oliva, and A. Torralba, “Object detectors emerge in deep scene cnns,” International Conference on Learning Representations , 2015
2015
Cited alongside, same era.
V. R. K. Garimella, A. Alfayad, and I. Weber, “Social media image analysis for public health,” in CHI , 2016, pp. 5543–5547
2016
Cited alongside, same era.
Y. Mejova, S. Abbar, and H. Haddadi, “Fetishizing food in digital age: #foodporn around the world,” in ICWSM , 2016, pp. 250–258
2016
Cited alongside, same era.
C. Liu, Y. Cao, Y. Luo, G. Chen, V. Vokkarane, and Y. Ma, “Deepfood: Deep learning-based food image recognition for computer-aided dietary assessment,” in International Conference on Smart Homes and Health Telematics . Springer, 2016, pp. 37–48
2016
Cited alongside, same era.
T. Kusmierczyk, C. Trattner, and K. Norvag, “Understanding and predicting online food recipe production patterns,” in HyperText , 2016
2016
Cited alongside, same era.
C.-w. N. Jing-jing Chen, “Deep-based ingredient recognition for cooking recipe retrival,” ACM Multimedia , 2016
2016
Cited alongside, same era.
J.-j. Chen, C.-W. Ngo, and T.-S. Chua, “Cross-modal recipe retrieval with rich food attributes,” in Proceedings of the 2017 ACM on Multimedia Conference , ser. MM ’17. New York, NY, USA: ACM, 2017, pp. 1771–1779. [Online]. Available: http://doi.acm.org/10.1145/3123266.3123428
2017
Later among the works it cites.
Y. Aytar, L. Castrejon, C. Vondrick, H. Pirsiavash, and A. Torralba, “Cross-modal scene networks,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 40, no. 10, pp. 2303–2314, 2018. [Online]. Available: https://doi.org/10.1109/TPAMI.2017.2753232
2017
Later among the works it cites.
2017
Later among the works it cites.
2018
Closest in time.
M. Chang, L. V. Guillain, H. Jung, V. M. Hare, J. Kim, and M. Agrawala, “Recipescape: An interactive tool for analyzing cooking instructions at scale,” in Proceedings of the 2018 CHI Conference on Human Factors in Computing Systems , ser. CHI ’18. New York, NY, USA: ACM, 2018, pp. 451:1–451:12. [Online]. Available: http://doi.acm.org/10.1145/3173574.3174025
2018
Closest in time.
M. Engilberge, L. Chevallier, P. Pérez, and M. Cord, “Finding beans in burgers: Deep semantic-visual embedding with localization,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , June 2018
2018
Closest in time.
J.-J. Chen, C.-W. Ngo, F.-L. Feng, and T.-S. Chua, “Deep understanding of cooking procedure for cross-modal recipe retrieval,” in Proceedings of the 26th ACM International Conference on Multimedia , ser. MM ’18. New York, NY, USA: ACM, 2018, pp. 1020–1028. [Online]. Available: http://doi.acm.org/10.1145/3240508.3240627
2018
Closest in time.
M. Carvalho, R. Cadène, D. Picard, L. Soulier, N. Thome, and M. Cord, “Cross-modal retrieval in the cooking context: Learning semantic text-image embeddings,” in Proceedings of the 41st International ACM SIGIR Conference on Research and Development in Information Retrieval , ser. SIGIR ’18. New York, NY, USA: ACM, 2018
2018
Closest in time.
B. Zhou, A. Lapedriza, A. Khosla, A. Oliva, and A. Torralba, “Places: A 10 Million Image Database for Scene Recognition,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 40, no. 6, pp. 1452–1464, Apr. 2018
2018
Closest in time.