Sep 2026· Journal of Food Science· Vol 91 9, pp.
e71414
· 0 citations· 6 references
Medicine
TL;DR
A domain-adaptive open-vocabulary food object detection framework that enables the identification of both existing and previously unseen food categories without category-specific box annotations for new items and provides an assistive localization module for downstream dietary assessment, sorting review, and quality-inspection workflows.
Abstract
The rapid diversification of food products, frequent packaging updates, and seasonal variations pose significant challenges for vision-based food inspection and dietary monitoring in real-world food systems. To address the need for scalable and flexible food analysis, this study proposes a domain-adaptive open-vocabulary food object detection (OVFD) framework that enables the identification of both existing and previously unseen food categories without category-specific box annotations for new items. By leveraging open-vocabulary representations, the framework allows dynamic expansion of detectable food categories as new items emerge, supporting continuous adaptation to evolving products, recipes, and dietary patterns while reducing manual annotation requirements. The proposed method integrates a dynamic prompt distribution network (DPDN) to improve region-text alignment, along with a density-aware multiple instance learning (DA-MIL) strategy that uses weakly supervised image-level tags to improve generalization and suppress background-driven detections. To reduce semantic leakage, Food2K labels are filtered by normalized matching, synonym auditing, and overlap-based rules before training. The framework was evaluated on multiple benchmark food datasets under open-vocabulary settings, demonstrating consistent improvements in detection performance for both Base (seen) and Novel (unseen) food categories, with Novel-category gains also observed under AP75 (average precision at intersection-over-union = 0.75) and COCO-style AP. Repeated-run reporting, error analysis, and limited external evaluation support stability and practical plausibility. By enabling dynamic category expansion with reduced annotation overhead, the proposed approach provides an assistive localization module for downstream dietary assessment, sorting review, and quality-inspection workflows, while task-specific validation and human verification remain necessary before operational food-safety or production-line use. PRACTICAL APPLICATIONS: This study provides an OVFD approach that can be adapted to new food items with reduced category-specific box annotation. Bounding-box localization can support upstream perception for portion estimation, ingredient-level dietary logging, assisted sorting review, and quality-inspection workflows by identifying where food items are located before secondary human or automated assessment. The present evidence supports assistive use only; human verification and validation under pilot-plant or real production conditions remain necessary before operational use. The detector is therefore intended to provide spatial evidence for subsequent review, not to replace validated food-inspection, quality-assessment, or production-control procedures.
An automated system to detect expiration dates using deep learning techniques is proposed and contributes to improving food security and reducing waste and supports intelligent automation in food inventory management with the help of semantic analysis on the basis of intensive learning of the product label.
Roshni Bhave, Vibhakti Bagade, V. Kamble et al.· International journal of com...· 0 citations
A zero-shot annotation framework that integrates OWLv2, Google’s second-generation open-vocabulary vision model, with large language models to enable multilingual, natural language-driven fruit recognition in smart agriculture, providing scalable solutions for automated annotation, real-time monitoring, and large-scale...
Ying-Dong Qin, Hao-Yu Song, Jing-Yi Li et al.· INMATEH Agricultural Enginee...· 0 citations
Automated inspection of prepared meals is important for improving food safety, quality assurance, and production efficiency. However, vision-based inspection remains challenging because boxed meals contain multiple adjacent food items with irregular shapes, varying portion sizes, and visually similar appearances, while...
Hong-Dar Lin, Guan-Ming Chen, Chou-Hsien Lin· Italian National Conference...· 0 citations
Modern smart kitchen automation requires reliable vision-based tools to provide user-advisory decision support during domestic culinary processes. However, standard deep learning models utilizing closed-set Softmax classifiers typically misclassify unknown or Out-of-Distribution (OOD) kitchen objects with high confiden...
This review systematically compares CNNs, RNNs/LSTMs, Transformers, GNNs, GANs, and hybrid architectures, as well as transfer learning, self-supervised learning, contrastive learning, few-shot learning, lightweight networks, edge computing, and multimodal fusion.
Wei-Hao Wang, Zhi-Dan Jiang, Si-Si Yang et al.· Foods· 0 citations
A structured literature review of CNN-based food image classification studies published between 2020 and 2025 identifies persistent gaps in statistical validation, class imbalance treatment, explainability, reproducibility, public code availability and deployment-oriented evaluation.
Luka Leskovec, A. R. Borges, Fernanda Brito Correia et al.· Signal, Image and Video Proc...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.