paper-with-me

홈 › Papers

InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem

2025-12-05 · Yeobin Hong, Suhyeon Lee, Hyungjin Chung, Jong Chul Ye arxiv

Recent approaches in controllable novel view video generation often rely on fine-tuning pre-trained Video Diffusion Models (VDMs). This dominant paradigm is computationally expensive and frequently suffers from catastrophic forgetting of the model's original generative priors. To address this challenge, here we propose InverseCrafter, a VDM training-free framework that reformulates novel view video generation as an inpainting-based inverse problem in the latent space, eliminating the need for any annotated 4D training data. The core of our method is to establish operator equivalence by employing a lightweight latent mask encoder to define a latent-domain masking operation via a continuous, multi-channel representation. This principled representation faithfully models the forward process in the latent domain, enabling efficient, backpropagation-free solvers while bypassing the costly bottleneck of repeated VAE operations. InverseCrafter achieves high-fidelity, spatio-temporally coherent novel view synthesis with near-zero additional inference overhead and excels at general-purpose video inpainting and editing by fully preserving the pre-trained VDM's generative capabilities.

📄 PDF Abstract BibTeX arXiv:2512.05672

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View SynthesisVideo GenerationVideo Inpainting

Similar Papers 제목 키워드 기반

Domain Generalized Recaptured Screen Image Identification Using SWIN Transformer

2024-07-24 · Preeti Mehta, Aman Sagar, Suchi Kumari

An increasing number of classification approaches have been developed to address the issue of image rebroadcast and recapturing, a standard attack strategy in insurance frauds, face spoofing, and video piracy. However, m…

Data AugmentationDomain Generalization

ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning

2024-11-07 · CVPR 2025 1 · David Junhao Zhang, Roni Paiss, Shiran Zada, Nikhil Karnad 외

Recently, breakthroughs in video modeling have allowed for controllable camera trajectories in generated videos. However, these methods cannot be directly applied to user-provided videos that are not generated by a video…

Learning Feature Disentanglement and Dynamic Fusion for Recaptured Image Forensic

2022-06-13 · Shuyu Miao, Lin Zheng, Hong Jin

Image recapture seriously breaks the fairness of artificial intelligent (AI) systems, which deceives the system by recapturing others' images. Most of the existing recapture models can only address a single pattern of re…

DisentanglementFairness

MToFNet: Object Anti-Spoofing with Mobile Time-of-Flight Data

2021-10-06 · Yonghyun Jeong, Doyeon Kim, Jaehyeon Lee, Minki Hong 외

In online markets, sellers can maliciously recapture others' images on display screens to utilize as spoof images, which can be challenging to distinguish in human eyes. To prevent such harm, we propose an anti-spoofing …

VISION-XL: High Definition Video Inverse Problem Solver using Latent Image Diffusion Models

2024-11-29 · Taesung Kwon, Jong Chul Ye

In this paper, we propose a novel framework for solving high-definition video inverse problems using latent image diffusion models. Building on recent advancements in spatio-temporal optimization for video inverse proble…

DeblurringGPUSuper-ResolutionVideo Reconstruction