paper-with-me

홈 › Papers

SPFM-Net: Semantic-Prior-Guided Frequency-Constrained Mamba for Invisible Watermark Attack

2026-07-30 · Chunpeng Wang, Yanan Shi, Zhiqiu Xia, Jidong Yang, Suo Gao, Qi Li arxiv

Existing watermark attacks typically rely on predefined signal-processing operations or locally constrained restoration networks, making it difficult to capture the long-range dependencies of globally distributed watermark signals and resulting in an unfavorable trade-off between removal effectiveness and visual fidelity. In this paper, we propose SPFM-Net, a semantic-prior-guided and frequency-constrained Mamba framework for invisible watermark attack. SPFM-Net first employs high-ratio masking to disrupt the spatial coherence of invisible watermark signals, and then utilizes a partially fine-tuned pretrained Masked Autoencoder to reconstruct semantically consistent image from sparse observations while suppressing watermark-related information. A Multi-scale Residual Frequency Feature Interaction module subsequently aggregates watermark-related residual features across multiple receptive fields, while adaptively suppressing responses from watermark-irrelevant regions. To further capture the long-range dependencies of globally distributed watermark signals, a lightweight Mamba-based Global State-space Feature Modeling (GSFM) unit is introduced to separate watermark-related features from natural image content and suppress the remaining watermark traces. In addition, SPFM-Net is optimized using a multi-level objective that jointly imposes spatial-, frequency-, and edge-domain constraints, enabling effective watermark suppression while preserving perceptual quality. Extensive experiments on representative spatial-domain, transform-domain, orthogonal moment-based, and deep learning-based watermarking schemes demonstrate that SPFM-Net achieves a favorable trade-off between watermark attack effectiveness and perceptual fidelity.

📄 PDF Abstract BibTeX arXiv:2607.27811

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Training Flow Matching Models with Reliable Labels via Self-Purification

2025-09-23 · Hyeongju Kim, Yechan Yu, June Young Yi, Juheon Lee arxiv

Training datasets are inherently imperfect, often containing mislabeled samples due to human annotation errors, limitations of tagging models, and other sources of noise. Such label contamination can significantly degrad…

Pose-invariant face recognition via feature-space pose frontalization

2025-05-22 · Nikolay Stanishev, Yuhang Lu, Touradj Ebrahimi

Pose-invariant face recognition has become a challenging problem for modern AI-based face recognition systems. It aims at matching a profile face captured in the wild with a frontal face registered in a database. Existin…

Face RecognitionRobust Face Recognition

UniCSG: Unified High-Fidelity Content-Constrained Style-Driven Generation via Staged Semantic and Frequency Disentanglement

2026-04-20 · Jingwei Yang, Ruoxi Wu, Wei Shen, Meng Li 외 arxiv

Style transfer must match a target style while preserving content semantics. DiT-based diffusion models often suffer from content-style entanglement, leading to reference-content leakage and unstable generation. We prese…

Style Transfer

CFSR: Geometry-Conditioned Shadow Removal via Physical Disentanglement

2026-04-20 · Pan Wang, Yihao Hu, Xiujin Liu, Hang Wang arxiv

Traditional shadow removal networks often treat image restoration as an unconstrained mapping, lacking the physical interpretability required to balance localized texture recovery with global illumination consistency. To…

Image RestorationShadow Removal

Robust TTS Training via Self-Purifying Flow Matching for the WildSpoof 2026 TTS Track

2025-12-19 · June Young Yi, Hyeongju Kim, Juheon Lee arxiv

This paper presents a lightweight text-to-speech (TTS) system developed for the WildSpoof Challenge TTS Track. Our approach fine-tunes the recently released open-weight TTS model, \textit{Supertonic}\footnote{\url{https:…