MetaFood3D: Large 3D Food Object Dataset with Nutrition Values

Chen, Yuhao; He, Jiangpeng; Czarnecki, Chris; Vinod, Gautham; Mahmud, Talha Ibn; Raghavan, Siddeshwar; Ma, Jinge; Mao, Dayou; Nair, Saeejith; Xi, Pengcheng; Wong, Alexander; Delp, Edward; Zhu, Fengqing

Abstract:Food computing is both important and challenging in computer vision (CV). It significantly contributes to the development of CV algorithms due to its frequent presence in datasets across various applications, ranging from classification and instance segmentation to 3D reconstruction. The polymorphic shapes and textures of food, coupled with high variation in forms and vast multimodal information, including language descriptions and nutritional data, make food computing a complex and demanding task for modern CV algorithms. 3D food modeling is a new frontier for addressing food-related problems, due to its inherent capability to deal with random camera views and its straightforward representation for calculating food portion size. However, the primary hurdle in the development of algorithms for food object analysis is the lack of nutrition values in existing 3D datasets. Moreover, in the broader field of 3D research, there is a critical need for domain-specific test datasets. To bridge the gap between general 3D vision and food computing research, we propose MetaFood3D. This dataset consists of 637 meticulously labeled 3D food objects across 108 categories, featuring detailed nutrition information, weight, and food codes linked to a comprehensive nutrition database. The dataset emphasizes intra-class diversity and includes rich modalities such as textured mesh files, RGB-D videos, and segmentation masks. Experimental results demonstrate our dataset's significant potential for improving algorithm performance, highlight the challenging gap between video captures and 3D scanned data, and show the strength of the MetaFood3D dataset in high-quality data generation, simulation, and augmentation.

Comments:	Dataset is coming soon
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2409.01966 [cs.CV]
	(or arXiv:2409.01966v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2409.01966

Computer Science > Computer Vision and Pattern Recognition

Title:MetaFood3D: Large 3D Food Object Dataset with Nutrition Values

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators