paper-with-me

Papers

Auto-Retoucher(ART) - A framework for Background Replacement and Image Editing

2019-01-13 · Yunxuan Xiao, Yikai Li, Yuwei Wu, LiZhen Zhu

Replacing the background and simultaneously adjusting foreground objects is a challenging task in image editing. Current techniques for generating such images relies heavily on user interactions with image editing softwares, which is a tedious job for professional retouchers. To reduce their workload, some exciting progress has been made on generating images with a given background. However, these models can neither adjust the position and scale of the foreground objects, nor guarantee the semantic consistency between foreground and background. To overcome these limitations, we propose a framework -- ART(Auto-Retoucher), to generate images with sufficient semantic and spatial consistency. Images are first processed by semantic matting and scene parsing modules, then a multi-task verifier model will give two confidence scores for the current background and position setting. We demonstrate that our jointly optimized verifier model successfully improves the visual consistency, and our ART framework performs well on images with the human body as foregrounds.

📄 PDF Abstract BibTeX arXiv:1901.03954

Code (1)

woshiyyya/Auto-Retoucher-pytorch pytorch

Tasks

Image MattingPositionScene Parsing

Similar Papers 제목 키워드 기반

StyleRetoucher: Generalized Portrait Image Retouching with GAN Priors

2023-12-22 · Wanchao Su, Can Wang, Chen Liu, Hangzhou Han 외

Creating fine-retouched portrait images is tedious and time-consuming even for professional artists. There exist automatic retouching methods, but they either suffer from over-smoothing artifacts or lack generalization a…

feature selectionImage Retouching

Agentic Retoucher for Text-To-Image Generation

2026-01-05 · Shaocheng Shen, Jianfeng Liang, Chunlei Cai, Cong Geng 외 arxiv

Text-to-image (T2I) diffusion models such as SDXL and FLUX have achieved impressive photorealism, yet small-scale distortions remain pervasive in limbs, face, text and so on. Existing refinement approaches either perform…

Text-to-Image Generation

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation

2026-08-06 · Yaole Wang, Xiaoyu Chen, Xin Ma, Yang Ding 외 arxiv

Video identity replacement seeks to transfer the identities of one or more subjects while preserving the motion, expressions, and temporal structure of a driving video. Existing methods largely target single-person setti…

Semantic correspondenceVideo Generation

VRetouchEr: Learning Cross-frame Feature Interdependence with Imperfection Flow for Face Retouching in Videos

2024-01-01 · CVPR 2024 1 · Wen Xue, Le Jiang, Lianxin Xie, Si Wu 외

Face Video Retouching is a complex task that often requires labor-intensive manual editing. Conventional image retouching methods perform less satisfactorily in terms of generalization performance and stability when …

Image Retouching

Closed-form detector for solid sub-pixel targets in multivariate t-distributed background clutter

2018-04-05 · James Theiler, Beate Zimmer, Amanda Ziemann

The generalized likelihood ratio test (GLRT) is used to derive a detector for solid sub-pixel targets in hyperspectral imagery. A closed-form solution is obtained that optimizes the replacement target model when the back…

Form