Hiring in life sciences? Share your open positions with our professional community. Read more Close

Advertisement

Visual Food Object Detection via Open-Vocabulary Learning.

Created on 05 Sep 2026

Authors

Shoulong Liu, Guorui Sheng, Shaojie Zhou, Bolin Yang, Weiqing Min, Shuqiang Jiang

Published in

Journal of food science. Volume 91. Issue 9. Pages e71414.

Abstract

The rapid diversification of food products, frequent packaging updates, and seasonal variations pose significant challenges for vision-based food inspection and dietary monitoring in real-world food systems. To address the need for scalable and flexible food analysis, this study proposes a domain-adaptive open-vocabulary food object detection (OVFD) framework that enables the identification of both existing and previously unseen food categories without category-specific box annotations for new items. By leveraging open-vocabulary representations, the framework allows dynamic expansion of detectable food categories as new items emerge, supporting continuous adaptation to evolving products, recipes, and dietary patterns while reducing manual annotation requirements. The proposed method integrates a dynamic prompt distribution network (DPDN) to improve region-text alignment, along with a density-aware multiple instance learning (DA-MIL) strategy that uses weakly supervised image-level tags to improve generalization and suppress background-driven detections. To reduce semantic leakage, Food2K labels are filtered by normalized matching, synonym auditing, and overlap-based rules before training. The framework was evaluated on multiple benchmark food datasets under open-vocabulary settings, demonstrating consistent improvements in detection performance for both Base (seen) and Novel (unseen) food categories, with Novel-category gains also observed under AP75 (average precision at intersection-over-union = 0.75) and COCO-style AP. Repeated-run reporting, error analysis, and limited external evaluation support stability and practical plausibility. By enabling dynamic category expansion with reduced annotation overhead, the proposed approach provides an assistive localization module for downstream dietary assessment, sorting review, and quality-inspection workflows, while task-specific validation and human verification remain necessary before operational food-safety or production-line use. PRACTICAL APPLICATIONS: This study provides an OVFD approach that can be adapted to new food items with reduced category-specific box annotation. Bounding-box localization can support upstream perception for portion estimation, ingredient-level dietary logging, assisted sorting review, and quality-inspection workflows by identifying where food items are located before secondary human or automated assessment. The present evidence supports assistive use only; human verification and validation under pilot-plant or real production conditions remain necessary before operational use. The detector is therefore intended to provide spatial evidence for subsequent review, not to replace validated food-inspection, quality-assessment, or production-control procedures.

PMID:
42696467
Bibliographic data and abstract were imported from PubMed on 05 Sep 2026.

Read full publication at:
Please sign in to see all details.

Advertisement

Stats

  • Community rating n/a 0 votes
  • Reviewers' rating n/a 0 votes
  • Your rating

1-terrible, 9-excellent. How would you rate this publication? Sign in in to submit your rating.

  • Recommendations n/a n/a positive of 0 vote(s)
  • Views 3
  • Comments 0

Recommended by

  • No recommendations yet.

Post a comment

You need to be signed in to post comments. You can sign in here.

Comments

There are no comments yet.

Advertisement