paper-with-me

Papers

Learning Implicit Features with Flow Infused Attention for Realistic Virtual Try-On

2024-12-16 · Delong Zhang, Qiwei Huang, Yuanliu liu, Yang Sun, Wei-Shi Zheng, Pengfei Xiong, Wei zhang

Image-based virtual try-on is challenging since the generated image should fit the garment to model images in various poses and keep the characteristics and details of the garment simultaneously. A popular research stream warps the garment image firstly to reduce the burden of the generation stage, which relies highly on the performance of the warping module. Other methods without explicit warping often lack sufficient guidance to fit the garment to the model images. In this paper, we propose FIA-VTON, which leverages the implicit warp feature by adopting a Flow Infused Attention module on virtual try-on. The dense warp flow map is projected as indirect guidance attention to enhance the feature map warping in the generation process implicitly, which is less sensitive to the warping estimation accuracy than an explicit warp of the garment image. To further enhance implicit warp guidance, we incorporate high-level spatial attention to complement the dense warp. Experimental results on the VTON-HD and DressCode dataset significantly outperform state-of-the-art methods, demonstrating that FIA-VTON is effective and robust for virtual try-on.

📄 PDF Abstract BibTeX arXiv:2412.11435

Code (0)

등록된 구현이 없습니다.

Tasks

Virtual Try-on

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

KSAT: Knowledge-infused Self Attention Transformer -- Integrating Multiple Domain-Specific Contexts

2022-10-09 · Kaushik Roy, Yuxin Zi, Vignesh Narayanan, Manas Gaur 외

Domain-specific language understanding requires integrating multiple pieces of relevant contextual information. For example, we see both suicide and depression-related behavior (multiple contexts) in the text ``I have a …

Specificity

Attention-Guided Flow-Matching for Sparse 3D Geological Generation

2026-04-07 · Zhixiang Lu, Mengqi Han, Peixin Guo, Tianming Bai 외 arxiv

Constructing high-resolution 3D geological models from sparse 1D borehole and 2D surface data is a highly ill-posed inverse problem. Traditional heuristic and implicit modeling methods fundamentally fail to capture non-l…

KW-ATTN: Knowledge Infused Attention for Accurate and Interpretable Text Classification

2021-06-01 · NAACL (DeeLIO) 2021 6 · Hyeju Jang, Seojin Bang, Wen Xiao, Giuseppe Carenini 외

Text classification has wide-ranging applications in various domains. While neural network approaches have drastically advanced performance in text classification, they tend to be powered by a large amount of training da…

Classificationtext-classificationText Classification

Audio-Infused Automatic Image Colorization by Exploiting Audio Scene Semantics

2024-01-24 · Pengcheng Zhao, Yanxiang Chen, Yang Zhao, Zhao Zhang

Automatic image colorization is inherently an ill-posed problem with uncertainty, which requires an accurate semantic understanding of scenes to estimate reasonable colors for grayscale images. Although recent interactio…

ColorizationImage Colorization

The Context of Crash Occurrence: A Complexity-Infused Approach Integrating Semantic, Contextual, and Kinematic Features

2024-11-26 · Meng Wang, Zach Noonan, Pnina Gershon, Bruce Mehler 외

Understanding the context of crash occurrence in complex driving environments is essential for improving traffic safety and advancing automated driving. Previous studies have used statistical models and deep learning to …

Autonomous DrivingLanguage ModelingLanguage ModellingLarge Language Model