paper-with-me

Papers

Dif4FF: Leveraging Multimodal Diffusion Models and Graph Neural Networks for Accurate New Fashion Product Performance Forecasting

2024-12-07 · Andrea Avogaro, Luigi Capogrosso, Franco Fummi, Marco Cristani

In the fast-fashion industry, overproduction and unsold inventory create significant environmental problems. Precise sales forecasts for unreleased items could drastically improve the efficiency and profits of industries. However, predicting the success of entirely new styles is difficult due to the absence of past data and ever-changing trends. Specifically, currently used deterministic models struggle with domain shifts when encountering items outside their training data. The recently proposed diffusion models address this issue using a continuous-time diffusion process. Specifically, these models enable us to predict the sales of new items, mitigating the domain shift challenges encountered by deterministic models. As a result, this paper proposes Dif4FF, a novel two-stage pipeline for New Fashion Product Performance Forecasting (NFPPF) that leverages the power of diffusion models conditioned on multimodal data related to specific clothes. Dif4FF first utilizes a multimodal score-based diffusion model to forecast multiple sales trajectories for various garments over time. The forecasts are refined using a powerful Graph Convolutional Network (GCN) architecture. By leveraging the GCN's capability to capture long-range dependencies within both the temporal and spatial data and seeking the optimal solution between these two dimensions, Dif4FF offers the most accurate and efficient forecasting system available in the literature for predicting the sales of new items. We tested Dif4FF on VISUELLE, the de facto standard for NFPPF, achieving new state-of-the-art results.

📄 PDF Abstract BibTeX arXiv:2412.05566

Code (1)

andreaavo9/Dif4FF 공식 구현

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DPDEdit: Detail-Preserved Diffusion Models for Multimodal Fashion Image Editing

2024-09-02 · Xiaolong Wang, Zhi-Qi Cheng, Jue Wang, Xiaojiang Peng

Fashion image editing is a crucial tool for designers to convey their creative ideas by visualizing design concepts interactively. Current fashion image editing techniques, though advanced with multimodal prompts and pow…

Image GenerationLanguage ModellingLarge Language ModelMultimodal fashion image editing+1

FashionSD-X: Multimodal Fashion Garment Synthesis using Latent Diffusion

2024-04-26 · Abhishek Kumar Singh, Ioannis Patras

The rapid evolution of the fashion industry increasingly intersects with technological advancements, particularly through the integration of generative AI. This study introduces a novel generative pipeline designed to tr…

Virtual Try-on

UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation

2024-08-21 · Xiangyu Zhao, Yuehan Zhang, Wenlong Zhang, Xiao-Ming Wu

The fashion domain encompasses a variety of real-world multimodal tasks, including multimodal retrieval and multimodal generation. The rapid advancements in artificial intelligence generated content, particularly in tech…

Image GenerationImage RetrievalImage to textLanguage Modeling+4

Multimodal Garment Designer: Human-Centric Latent Diffusion Models for Fashion Image Editing

2023-04-04 · ICCV 2023 1 · Alberto Baldrati, Davide Morelli, Giuseppe Cartella, Marcella Cornia 외

Fashion illustration is used by designers to communicate their vision and to bring the design idea from conceptualization to realization, showing how clothes interact with the human body. In this context, computer vision…

Multimodal fashion image editing

Multimodal-Conditioned Latent Diffusion Models for Fashion Image Editing

2024-03-21 · Alberto Baldrati, Davide Morelli, Marcella Cornia, Marco Bertini 외

Fashion illustration is a crucial medium for designers to convey their creative vision and transform design concepts into tangible representations that showcase the interplay between clothing and the human body. In the c…

DenoisingVirtual Try-on