paper-with-me

홈 › Papers

Foodfusion: A Novel Approach for Food Image Composition via Diffusion Models

2024-08-26 · Chaohua Shi, Xuan Wang, Si Shi, Xule Wang, Mingrui Zhu, Nannan Wang, Xinbo Gao

Food image composition requires the use of existing dish images and background images to synthesize a natural new image, while diffusion models have made significant advancements in image generation, enabling the construction of end-to-end architectures that yield promising results. However, existing diffusion models face challenges in processing and fusing information from multiple images and lack access to high-quality publicly available datasets, which prevents the application of diffusion models in food image composition. In this paper, we introduce a large-scale, high-quality food image composite dataset, FC22k, which comprises 22,000 foreground, background, and ground truth ternary image pairs. Additionally, we propose a novel food image composition method, Foodfusion, which leverages the capabilities of the pre-trained diffusion models and incorporates a Fusion Module for processing and integrating foreground and background information. This fused information aligns the foreground features with the background structure by merging the global structural information at the cross-attention layer of the denoising UNet. To further enhance the content and structure of the background, we also integrate a Content-Structure Control Module. Extensive experiments demonstrate the effectiveness and scalability of our proposed method.

📄 PDF Abstract BibTeX arXiv:2408.14135

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingImage Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

FoodFusion: A Latent Diffusion Model for Realistic Food Image Generation

2023-12-06 · Olivia Markham, Yuhao Chen, Chi-en Amy Tai, Alexander Wong

Current state-of-the-art image generation models such as Latent Diffusion Models (LDMs) have demonstrated the capacity to produce visually striking food-related images. However, these generated images often exhibit an ar…

DiversityImage Generation

Training-Free Text-to-Image Compositional Food Generation via Prompt Grafting

2026-01-25 · Xinyue Pan, Yuhao Chen, Fengqing Zhu arxiv

Real-world meal images often contain multiple food items, making reliable compositional food image generation important for applications such as image-based dietary assessment, where multi-food data augmentation is neede…

Data AugmentationImage Generation

Understanding the Limitations of Diffusion Concept Algebra Through Food

2024-06-05 · E. Zhixuan Zeng, Yuhao Chen, Alexander Wong

Image generation techniques, particularly latent diffusion models, have exploded in popularity in recent years. Many techniques have been developed to manipulate and clarify the semantic concepts these large-scale models…

DiversityImage Generation

UMDFood: Vision-language models boost food composition compilation

2023-05-18 · Peihua Ma, Yixin Wu, Ning Yu, Yang Zhang 외

Nutrition information is crucial in precision nutrition and the food industry. The current food composition compilation paradigm relies on laborious and experience-dependent methods. However, these methods struggle to ke…

Language ModellingNutrition

Diffusion Model with Clustering-based Conditioning for Food Image Generation

2023-09-01 · Yue Han, Jiangpeng He, Mridul Gupta, Edward J. Delp 외

Image-based dietary assessment serves as an efficient and accurate solution for recording and analyzing nutrition intake using eating occasion images as input. Deep learning-based techniques are commonly used to perform …

ClusteringData AugmentationImage GenerationNutrition