paper-with-me

홈 › Papers

ShineOn: Illuminating Design Choices for Practical Video-based Virtual Clothing Try-on

2020-12-18 · Gaurav Kuppa, Andrew Jong, Vera Liu, Ziwei Liu, Teng-Sheng Moh

Virtual try-on has garnered interest as a neural rendering benchmark task to evaluate complex object transfer and scene composition. Recent works in virtual clothing try-on feature a plethora of possible architectural and data representation choices. However, they present little clarity on quantifying the isolated visual effect of each choice, nor do they specify the hyperparameter details that are key to experimental reproduction. Our work, ShineOn, approaches the try-on task from a bottom-up approach and aims to shine light on the visual and quantitative effects of each experiment. We build a series of scientific experiments to isolate effective design choices in video synthesis for virtual clothing try-on. Specifically, we investigate the effect of different pose annotations, self-attention layer placement, and activation functions on the quantitative and qualitative performance of video virtual try-on. We find that DensePose annotations not only enhance face details but also decrease memory usage and training time. Next, we find that attention layers improve face and neck quality. Finally, we show that GELU and ReLU activation functions are the most effective in our experiments despite the appeal of newer activations such as Swish and Sine. We will release a well-organized code base, hyperparameters, and model checkpoints to support the reproducibility of our results. We expect our extensive experiments and code to greatly inform future design choices in video virtual try-on. Our code may be accessed at https://github.com/andrewjong/ShineOn-Virtual-Tryon.

📄 PDF Abstract BibTeX arXiv:2012.10495

Code (1)

andrewjong/ShineOn-Virtual-Tryon 공식 구현 pytorch

Tasks

Neural RenderingVirtual Try-on

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
(FiLe@Against@Claim)How do I file a claim against Expedia? How do I file a claim against Expedia? How Do I File a Claim Against Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Fast Help &…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…

Similar Papers 제목 키워드 기반

SmolVLM: Redefining small and efficient multimodal models

2025-04-07 · Andrés Marafioti, Orr Zohar, Miquel Farré, Merve Noyan 외

Large Vision-Language Models (VLMs) deliver exceptional performance but require significant computational resources, limiting their deployment on mobile and edge devices. Smaller VLMs typically mirror design choices of l…

GPU

Holographic Imaging with XL-MIMO and RIS: Illumination and Reflection Design

2023-12-18 · Giulia Torcolacci, Anna Guerra, Haiyang Zhang, Francesco Guidi 외

This paper addresses a near-field imaging problem utilizing extremely large-scale multiple-input multiple-output (XL-MIMO) antennas and reconfigurable intelligent surfaces (RISs) already in place for wireless communicati…

Mech-Elites: Illuminating the Mechanic Space of GVGAI

2020-02-11 · M Charity, Michael Cerny Green, Ahmed Khalifa, Julian Togelius

This paper introduces a fully automatic method of mechanic illumination for general video game level generation. Using the Constrained MAP-Elites algorithm and the GVG-AI framework, this system generates the simplest til…

Video-Oasis: Rethinking Evaluation of Video Understanding

2026-07-02 · Geuntaek Lim, Sungjune Park, Jaeyun Lee, Inwoong Lee 외 hf

The inherent complexity of video understanding makes it difficult to determine whether Video-LLM benchmark performance stems from visual perception, linguistic reasoning, or knowledge priors. While many benchmarks have e…

FlowC2S: Flowing from Current to Succeeding Frames for Fast and Memory-Efficient Video Continuation

2026-04-19 · Hovhannes Margaryan, Quentin Bammey, Christian Sandor arxiv

This paper introduces a novel methodology for generating fast and memory-efficient video continuations. Our method, dubbed FlowC2S, fine-tunes a pre-trained text-to-video flow model to learn a vector field between the cu…