Zero-Shot Ingredient Recognition by Multi-Relational Graph Convolutional Network
Jingjing Chen, Liangming Pan, Zhipeng Wei, Xiang Wang, Chong-Wah Ngo, Tat-Seng Chua
Abstract
Recognizing ingredients for a given dish image is at the core of automatic dietary assessment, attracting increasing attention from both industry and academia. Nevertheless, the task is challenging due to the difficulty of collecting and labeling sufficient training data. On one hand, there are hundred thousands of food ingredients in the world, ranging from the common to rare. Collecting training samples for all of the ingredient categories is difficult. On the other hand, as the ingredient appearances exhibit huge visual variance during the food preparation, it requires to collect the training samples under different cooking and cutting methods for robust recognition. Since obtaining sufficient fully annotated training data is not easy, a more practical way of scaling up the recognition is to develop models that are capable of recognizing unseen ingredients. Therefore, in this paper, we target the problem of ingredient recognition with zero training samples. More specifically, we introduce multi-relational GCN (graph convolutional network) that integrates ingredient hierarchy, attribute as well as co-occurrence for zero-shot ingredient recognition. Extensive experiments on both Chinese and Japanese food datasets are performed to demonstrate the superior performance of multi-relational GCN and shed light on zero-shot ingredients recognition.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- Should Graph Convolution Trust Neighbors? A Simple Causal Inference MethodFuli Feng, Weiran Huang, Xiangnan He, Xin Xin et al.SIGIR 2021 · 66 citations
- Boosting the Transferability of Video Adversarial Examples via Temporal TranslationZhipeng Wei, Jingjing Chen, Zuxuan Wu, Yu-Gang JiangAAAI 2022 · 48 citations
- Cross-Modal Transferable Adversarial Attacks from Images to VideosZhipeng Wei, Jingjing Chen, Zuxuan Wu, Yu-Gang JiangCVPR 2022 · 45 citations
- Disentangled Ontology Embedding for Zero-shot LearningYuxia Geng, Jiaoyan Chen, Wen Zhang, Yajing Xu et al.KDD 2022 · 22 citations
- Multi-modal Cooking Workflow Construction for Food RecipesLiangming Pan, Jingjing Chen, Jianlong Wu, Shaoteng Liu et al.ACM MM 2020 · 20 citations
Related papers
- DSDGF-Nutri: A Decoupled Self-Distillation Network with Gating Fusion For Food Nutritional AssessmentSujuan Hou, Zhihui Feng, Hao Xiong, Weiqing Min et al.ACM MM 2025 · 2 citations
- SeeDS: Semantic Separable Diffusion Synthesizer for Zero-shot Food DetectionPengfei Zhou, Weiqing Min, Yang Zhang, Jiajun Song et al.ACM MM 2023 · 11 citations
- Ingredients-Guided and Nutrients-Prompted Network for Food Nutrition EstimationDonglin Zhang, Boyuan Ma, Xiaojun Wu, Josef KittlerACM MM 2025 · 1 citation
- View-GCN: View-Based Graph Convolutional Network for 3D Shape AnalysisXin Wei, Ruixuan Yu, Jian SunCVPR 2020
- RECIPTOR: An Effective Pretrained Model for Recipe Representation LearningDiya Li, Mohammed J. ZakiKDD 2020 · 28 citations
