Product Title Refinement via Multi-Modal Generative Adversarial Learning
Nowadays, an increasing number of customers are in favor of using E-commerce Apps to browse and purchase products. Since merchants are usually inclined to employ redundant and over-informative product titles to attract customers' attention, it is of great importance to concisely display short product titles on limited screen of cell phones. Previous researchers mainly consider textual information of long product titles and lack of human-like view during training and evaluation procedure. In this paper, we propose a Multi-Modal Generative Adversarial Network (MM-GAN) for short product title generation, which innovatively incorporates image information, attribute tags from the product and the textual information from original long titles. MM-GAN treats short titles generation as a reinforcement learning process, where the generated titles are evaluated by the discriminator in a human-like view.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeGenerative Adversarial Networkreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Multi-Modal Generative Adversarial Network for Short Product Title Generation in Mobile E-Commerce
Nowadays, more and more customers browse and purchase products in favor of using mobile E-Commerce Apps such as Taobao and Amazon. Since merchants are usually inclined to describe redundant and over-informative product t…
AttributeGenerative Adversarial NetworkReinforcement LearningMultimodal Prompt Learning for Product Title Generation with Extremely Limited Labels
Generating an informative and attractive title for the product is a crucial task for e-commerce. Most existing works follow the standard multimodal natural language generation approaches, e.g., image captioning, and empl…
Image CaptioningPrompt LearningText GenerationMutual Query Network for Multi-Modal Product Image Segmentation
Product image segmentation is vital in e-commerce. Most existing methods extract the product image foreground only based on the visual modality, making it difficult to distinguish irrelevant products. As product titles c…
Image SegmentationSegmentationSemantic SegmentationMAKE: Vision-Language Pre-training based Product Retrieval in Taobao Search
Taobao Search consists of two phases: the retrieval phase and the ranking phase. Given a user query, the retrieval phase returns a subset of candidate products for the following ranking phase. Recently, the paradigm of p…
RetrievalMulti-Label Product Categorization Using Multi-Modal Fusion Models
In this study, we investigated multi-modal approaches using images, descriptions, and titles to categorize e-commerce products on Amazon. Specifically, we examined late fusion models, where the modalities are fused at th…
Multi-Label ClassificationMUlTI-LABEL-ClASSIFICATIONProduct Categorization