paper-with-me

홈 › Papers

Blind Bitstream-corrupted Video Recovery via Metadata-guided Diffusion Model

2025-01-01 · CVPR 2025 1 · Shuyun Wang, Hu Zhang, Xin Shen, Dadong Wang, Xin Yu

Bitstream-corrupted video recovery aims to fill in realistic video content due to bitstream corruption during video storage or transmission. Most existing methods typically assume that the predefined masks of the corrupted regions are known in advance. However, manually annotating these masks is laborious and time-consuming, limiting the applicability of existing methods in real-world scenarios. Therefore, we expect to relax this assumption by defining a new blind video recovery setting where the recovery of corrupted regions does not rely on predefined masks. There are two significant challenges in this setting: (i) without predefined masks, how accurately can a model identify the regions requiring recovery? (ii) how to recover contents from extensive and irregular regions, especially when large portions of frames are severely degraded? To address these challenges, we introduce a Metadata-Guided Diffusion Model, dubbed M-GDM. To enable a diffusion model focusing on the corrupted regions, we leverage intrinsic video metadata as a corruption indicator and design a dual-stream metadata encoder. This encoder first embeds the motion vectors and frame types of a video separately and then merges them into a unified metadata representation. The metadata representation will interact with the corrupted latent feature through cross-attention mechanisms at each diffusion step. Meanwhile, to preserve the intact regions, we propose a prior-driven mask predictor that generates pseudo masks for the corrupted regions by leveraging the metadata prior and diffusion prior. These pseudo masks enable the separation and recombination of intact and recovered regions through hard masking. However, imperfections in pseudo mask predictions and hard masking processes often result in boundary artifacts. Thus, we introduce a post-refinement module that refines the hard-masked outputs, enhancing the consistency between intact and recovered regions. Extensive experiment results validate the effectiveness of our method and demonstrate its superiority in the blind video recovery task.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Blind Bitstream-corrupted Video Recovery via Metadata-guided Diffusion Model

2026-04-15 · Shuyun Wang, Hu Zhang, Xin Shen, Dadong Wang 외 arxiv

Bitstream-corrupted video recovery aims to restore realistic content degraded during video storage or transmission. Existing methods typically assume that predefined masks of corrupted regions are available, but manually…

Towards Blind Bitstream-corrupted Video Recovery via a Visual Foundation Model-driven Framework

2025-07-30 · Tianyi Liu, Kejun Wu, Chen Cai, Yi Wang 외 arxiv

Video signals are vulnerable in multimedia communication and storage systems, as even slight bitstream-domain corruption can lead to significant pixel-domain degradation. To recover faithful spatio-temporal content from …

Bitstream-Corrupted Video Recovery: A Novel Benchmark Dataset and Method

2023-09-25 · NeurIPS 2023 11 · Tianyi Liu, Kejun Wu, Yi Wang, Wenyang Liu 외

The past decade has witnessed great strides in video recovery by specialist technologies, like video inpainting, completion, and error concealment. However, they typically simulate the missing content by manual-designed …

Video Inpainting

NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results

2026-04-08 · Wenbin Zou, Tianyi Liu, Kejun Wu, Huiping Zhuang 외 arxiv

This paper reports on the NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration (BSCVR). The challenge aims to advance research on recovering visually coherent videos from corrupted bitstreams, whose decoding oft…

Video Restoration

Leveraging Bitstream Metadata for Fast, Accurate, Generalized Compressed Video Quality Enhancement

2022-01-31 · Max Ehrlich, Jon Barker, Namitha Padmanabhan, Larry Davis 외

Video compression is a central feature of the modern internet powering technologies from social media to video conferencing. While video compression continues to mature, for many compression settings, quality loss is sti…

QuantizationVideo Compression