Semantic-Aware Dynamic Parameter for Video Inpainting Transformer
Recent learning-based video inpainting approaches have achieved considerable progress. However, they still cannot fully utilize semantic information within the video frames and predict improper scene layout, failing to restore clear object boundaries for mixed scenes. To mitigate this problem, we introduce a new transformer-based video inpainting technique that can exploit semantic information within the input and considerably improve reconstruction quality. In this study, we use the mixture-of-experts scheme and train multiple experts to handle mixed scenes, including various semantics. We leverage these multiple experts and produce locally (token-wise) different network parameters to achieve semantic-aware inpainting results. Extensive experiments on YouTube-VOS and DAVIS benchmark datasets demonstrate that, compared with existing conventional video inpainting approaches, the proposed method has superior performance in synthesizing visually pleasing videos with much clearer semantic structures and textures.
Code (0)
등록된 구현이 없습니다.
Tasks
Mixture-of-ExpertsVideo InpaintingMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
DreamDance: Animating Character Art via Inpainting Stable Gaussian Worlds
This paper presents DreamDance, a novel character art animation framework capable of producing stable, consistent character and scene motion conditioned on precise camera trajectories. To achieve this, we re-formulate th…
Image InpaintingVideo GenerationVideo InpaintingLearning Semantic-Aware Dynamics for Video Prediction
We propose an architecture and training scheme to predict video frames by explicitly modeling dis-occlusions and capturing the evolution of semantically consistent regions in the video. The scene layout (semantic map) an…
Optical Flow EstimationPredictionVideo PredictionVideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control
Video inpainting, which aims to restore corrupted video content, has experienced substantial progress. Despite these advances, existing methods, whether propagating unmasked region pixels through optical flow and recepti…
Image InpaintingOptical Flow EstimationText-to-Video EditingVideo Editing+1Short-Term and Long-Term Context Aggregation Network for Video Inpainting
Video inpainting aims to restore missing regions of a video and has many applications such as video editing and object removal. However, existing methods either suffer from inaccurate short-term context aggregation or ra…
Video EditingVideo InpaintingMagicRoad: Semantic-Aware 3D Road Surface Reconstruction via Obstacle Inpainting
Road surface reconstruction is essential for autonomous driving, supporting centimeter-accurate lane perception and high-definition mapping in complex urban environments.While recent methods based on mesh rendering or 3D…
Autonomous DrivingVideo Inpainting