paper-with-me

Papers

DishSeg24k: A Large-Scale Benchmark for Food Segmentation with Stochastic Expert Decoding

2026-07-25 · Yilin Wang, Haochen Shi, Guanyu Chen, Weiqing Min, Jinkai Zheng, Chenggang Yan, Shuqiang Jiang arxiv

Food segmentation is essential for applications such as intelligent catering, dietary assessment, and recommendation. However, existing benchmarks fail to capture the complexity of real-world dining scenes. The challenges of dense inter-dish overlap, fine-grained class similarity, and extreme long-tail class distributions exceed the fidelity of current datasets. To fill this gap, we introduce \textbf{DishSeg24k}, a large-scale dish-level segmentation benchmark with 24,096 images, 112,281 instances, and 278 fine-grained categories in real-world dining environments. Based on DishSeg24k, we further propose \textbf{Food Expert-Adaptive Segmentation Transformers (FEAST)} to address these challenges. FEAST models query-based decoding as a Markov Decision Process (MDP), where each decoder layer update is treated as a sequential decision step that explores uncertainty along dish boundaries. We further redesign the decoder with a reinforcement learning (RL)-guided Mixture-of-Experts (MoE) module, in which a dual-critic decoupled optimization scheme separates task-oriented query refinement from structure-aware expert routing. This design promotes expert specialization and prevents expert collapse under long-tail category distributions. Finally, extensive experiments on DishSeg24k demonstrate the state-of-the-art performance of FEAST, which outperforms previous methods by {+3.21\%} mIoU, {+3.68\%} mDice, and {+4.00\%} mAcc, respectively. We further validate the effectiveness of FEAST on FoodSeg103. The dataset and code will be publicly released.

📄 PDF Abstract BibTeX arXiv:2607.23070

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

A Large-Scale Benchmark for Food Image Segmentation

2021-05-12 · Xiongwei Wu, Xin Fu, Ying Liu, Ee-Peng Lim 외

Food image segmentation is a critical and indispensible task for developing health-related applications such as estimating food calories and nutrients. Existing food image segmentation models are underperforming due to t…

Image SegmentationSegmentationSemantic Segmentation

BenchSeg: A Large-Scale Dataset and Benchmark for Multi-View Food Video Segmentation

2026-01-12 · Ahmad AlMughrabi, Guillermo Rivo, Carlos Jiménez-Farfán, Umair Haroon 외 arxiv

Food image segmentation is a critical task for dietary analysis, enabling accurate estimation of food volume and nutrients. However, current methods suffer from limited multi-view data and poor generalization to new view…

Image SegmentationVideo Segmentation

Large Scale Visual Food Recognition

2021-03-30 · Weiqing Min, Zhiling Wang, Yuxin Liu, Mengjiang Luo 외

Food recognition plays an important role in food choice and intake, which is essential to the health and well-being of humans. It is thus of importance to the computer vision community, and can further support many food-…

Fine-Grained Visual RecognitionFood RecognitionImage RetrievalRepresentation Learning+1

FoodLMM: A Versatile Food Assistant using Large Multi-modal Model

2023-12-22 · Yuehao Yin, Huiyan Qi, Bin Zhu, Jingjing Chen 외

Large Multi-modal Models (LMMs) have made impressive progress in many vision-language tasks. Nevertheless, the performance of general LMMs in specific domains is still far from satisfactory. This paper proposes FoodLMM, …

Food RecognitionMulti-Task LearningNutritionReasoning Segmentation+2

Incremental Learning on Food Instance Segmentation

2023-06-28 · Huu-Thanh Nguyen, Yu Cao, Chong-Wah Ngo, Wing-Kwong Chan

Food instance segmentation is essential to estimate the serving size of dishes in a food image. The recent cutting-edge techniques for instance segmentation are deep learning networks with impressive segmentation quality…

Incremental LearningInstance SegmentationSegmentationSemantic Segmentation