RECIPE RECOGNITION WITH LARGE MULTIMODAL FOOD DATASET

Abstract : This paper deals with automatic systems for image recipe recognition. For this purpose, we compare and evaluate leading vision-based and text-based technologies on a new very large multimodal dataset (UPMC Food-101) containing about 100,000 recipes for a total of 101 food categories. Each item in this dataset is represented by one image plus textual information. We present deep experiments of recipe recognition on our dataset using visual, textual information and fusion. Additionally, we present experiments with text-based embedding technology to represent any food word in a semantical continuous space. We also compare our dataset features with a twin dataset provided by ETHZ university: we revisit their data collection protocols and carry out transfer learning schemes to highlight similarities and differences between both datasets. Finally, we propose a real application for daily users to identify recipes. This application is a web search engine that allows any mobile device to send a query image and retrieve the most relevant recipes in our dataset.
Type de document :
Communication dans un congrès
IEEE International Conference on Multimedia & Expo (ICME), workshop CEA, Jun 2015, Turin, Italy. <10.1109/ICMEW.2015.7169757>
Liste complète des métadonnées

https://hal.archives-ouvertes.fr/hal-01196959
Contributeur : Xin Wang <>
Soumis le : lundi 14 septembre 2015 - 13:11:40
Dernière modification le : mardi 6 décembre 2016 - 17:22:54
Document(s) archivé(s) le : mardi 29 décembre 2015 - 00:10:30

Fichier

CEA_ICME2015.pdf
Fichiers produits par l'(les) auteur(s)

Identifiants

Collections

UNICE | UPMC | LIP6 | I3S

Citation

Xin Wang, Devinder Kumar, Nicolas Thome, Matthieu Cord, Frederic Precioso. RECIPE RECOGNITION WITH LARGE MULTIMODAL FOOD DATASET. IEEE International Conference on Multimedia & Expo (ICME), workshop CEA, Jun 2015, Turin, Italy. <10.1109/ICMEW.2015.7169757>. <hal-01196959>

Partager

Métriques

Consultations de
la notice

136

Téléchargements du document

238