paper-with-me

Papers

V2A-Mark: Versatile Deep Visual-Audio Watermarking for Manipulation Localization and Copyright Protection

2024-04-25 · Xuanyu Zhang, Youmin Xu, Runyi Li, Jiwen Yu, Weiqi Li, Zhipei Xu, Jian Zhang

AI-generated video has revolutionized short video production, filmmaking, and personalized media, making video local editing an essential tool. However, this progress also blurs the line between reality and fiction, posing challenges in multimedia forensics. To solve this urgent issue, V2A-Mark is proposed to address the limitations of current video tampering forensics, such as poor generalizability, singular function, and single modality focus. Combining the fragility of video-into-video steganography with deep robust watermarking, our method can embed invisible visual-audio localization watermarks and copyright watermarks into the original video frames and audio, enabling precise manipulation localization and copyright protection. We also design a temporal alignment and fusion module and degradation prompt learning to enhance the localization accuracy and decoding robustness. Meanwhile, we introduce a sample-level audio localization method and a cross-modal copyright extraction mechanism to couple the information of audio and video frames. The effectiveness of V2A-Mark has been verified on a visual-audio tampering dataset, emphasizing its superiority in localization precision and copyright accuracy, crucial for the sustainable development of video editing in the AIGC video era.

📄 PDF Abstract BibTeX arXiv:2404.16824

Code (1)

zhipeixu/fakeshield pytorch

Tasks

Prompt LearningVideo Editing

Similar Papers 제목 키워드 기반

High-Fidelity Face Content Recovery via Tamper-Resilient Versatile Watermarking

2026-03-25 · Peipeng Yu, Jinfeng Xie, Chengfu Ou, Xiaoyu Zhou 외 arxiv

The proliferation of AIGC-driven face manipulation and deepfakes poses severe threats to media provenance, integrity, and copyright protection. Existing versatile watermarking systems typically rely on embedding explicit…

OmniGuard: Hybrid Manipulation Localization via Augmented Versatile Deep Image Watermarking

2024-12-02 · CVPR 2025 1 · Xuanyu Zhang, Zecheng Tang, Zhipei Xu, Runyi Li 외

With the rapid growth of generative AI and its widespread application in image editing, new risks have emerged regarding the authenticity and integrity of digital content. Existing versatile watermarking approaches suffe…

LAVA: Layered Audio-Visual Anti-tampering Watermarking for Robust Deepfake Detection and Localization

2026-04-27 · Bokang Zeng, Zheng Gao, Xiaoyu Li, Xiaoyan Feng 외 arxiv

Proactive watermarking offers a promising approach for deepfake tamper detection and localization in short-form videos. However, existing methods often decouple audio and visual evidence and assume that watermark signals…

DeepFake Detection

Proactive Detection of Voice Cloning with Localized Watermarking

2024-01-30 · Robin San Roman, Pierre Fernandez, Alexandre Défossez, Teddy Furon 외

In the rapidly evolving field of speech generative models, there is a pressing need to ensure audio authenticity against the risks of voice cloning. We present AudioSeal, the first audio watermarking technique designed s…

Voice Cloning

Hide&Seek: Remove Image Watermarks with Negligible Cost via Pixel-wise Reconstruction

2026-03-01 · Huajie Chen, Tianqing Zhu, Hailin Yang, Yuchen Zhong 외 arxiv

Watermarking has emerged as a key defense against the misuse of machine-generated images (MGIs). Yet the robustness of these protections remains underexplored. To reveal the limits of SOTA proactive image watermarking de…