paper-with-me

홈 › Papers

OmniTry: Virtual Try-On Anything without Masks

2025-08-19 · Yutong Feng, Linlin Zhang, Hengyuan Cao, Yiming Chen, Xiaoduan Feng, Jian Cao, Yuxiong Wu, Bin Wang arxiv

Virtual Try-ON (VTON) is a practical and widely-applied task, for which most of existing works focus on clothes. This paper presents OmniTry, a unified framework that extends VTON beyond garment to encompass any wearable objects, e.g., jewelries and accessories, with mask-free setting for more practical application. When extending to various types of objects, data curation is challenging for obtaining paired images, i.e., the object image and the corresponding try-on result. To tackle this problem, we propose a two-staged pipeline: For the first stage, we leverage large-scale unpaired images, i.e., portraits with any wearable items, to train the model for mask-free localization. Specifically, we repurpose the inpainting model to automatically draw objects in suitable positions given an empty mask. For the second stage, the model is further fine-tuned with paired images to transfer the consistency of object appearance. We observed that the model after the first stage shows quick convergence even with few paired samples. OmniTry is evaluated on a comprehensive benchmark consisting of 12 common classes of wearable objects, with both in-shop and in-the-wild images. Experimental results suggest that OmniTry shows better performance on both object localization and ID-preservation compared with existing methods. The code, model weights, and evaluation benchmark of OmniTry will be made publicly available at https://omnitry.github.io/.

📄 PDF Abstract BibTeX arXiv:2508.13632

Code (0)

등록된 구현이 없습니다.

Tasks

Object LocalizationVirtual Try-on

Similar Papers 제목 키워드 기반

OmniTryOn: Video Try-On Anything at Once!

2026-06-07 · Changliang Xia, Chengyou Jia, Minnan Luo, Zhuohang Dang 외 arxiv

Although video virtual try-on (VVT) has achieved significant progress, existing methods still exhibit two fundamental limitations: first, they are restricted to single-garment transfer, rendering simultaneous multi-objec…

Virtual Try-on

Has Anything Changed? 3D Change Detection by 2D Segmentation Masks

2023-12-02 · Aikaterini Adam, Konstantinos Karantzalos, Lazaros Grammatikopoulos, Torsten Sattler

As capturing devices become common, 3D scans of interior spaces are acquired on a daily basis. Through scene comparison over time, information about objects in the scene and their changes is inferred. This information is…

Change DetectionObject DiscoverySegmentation

SAM3D: Segment Anything in 3D Scenes

2023-06-06 · Yunhan Yang, Xiaoyang Wu, Tong He, Hengshuang Zhao 외

In this work, we propose SAM3D, a novel framework that is able to predict masks in 3D point clouds by leveraging the Segment-Anything Model (SAM) in RGB images without further training or finetuning. For a point cloud of…

Segmentation

Diffuse Attend and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion

2024-01-01 · CVPR 2024 1 · Junjiao Tian, Lavisha Aggarwal, Andrea Colaco, Zsolt Kira 외

Producing quality segmentation masks for images is a fundamental problem in computer vision. Recent research has explored large-scale supervised training to enable zero-shot transfer segmentation on virtually any ima…

SegmentationvalidZero Shot Segmentation

Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion

2023-08-23 · Junjiao Tian, Lavisha Aggarwal, Andrea Colaco, Zsolt Kira 외

Producing quality segmentation masks for images is a fundamental problem in computer vision. Recent research has explored large-scale supervised training to enable zero-shot segmentation on virtually any image style and …

SegmentationSemantic SegmentationvalidZero Shot Segmentation