paper-with-me

Papers

SPOT: Selective Prompt Projection via Total Variation for Inference-Only Safe Text-to-Image Generation

2026-01-31 · Minhyuk Lee, Hyekyung Yoon, Myungjoo Kang arxiv

Text-to-Image (T2I) diffusion models enable high quality open ended synthesis, but practical use requires suppressing unsafe generations while preserving behavior on benign prompts. We study this tension relative to the frozen generator, using its prompt conditioned distribution as the preservation reference. Since T2I safety is commonly evaluated by bounded risk scores on generated images, total variation (TV) bounds how much expected risk can change from this reference. We call this fixed reference constraint the Safety-Prompt Alignment Tradeoff (SPAT): reducing expected unsafety requires prompt conditioned distributional deviation. To make this deviation selective and adjustable, we define the tau safe set as prompts whose reference risk is at most tau, and cast intervention as projection toward nearby prompts in this set. We propose Selective Prompt prOjecTion (SPOT), an inference time framework that approximates this projection without retraining the generator or learning a category specific rewriter. SPOT uses an LLM to rank candidate rewrites and a safeguard VLM to accept generated images under the same tau. Across four datasets and three diffusion backbones, SPOT achieves relative inappropriate (IP) score reductions from 14.2% to 44.4% over strong safety alignment baselines while keeping benign prompt behavior close to the fixed reference.

📄 PDF Abstract BibTeX arXiv:2602.00616

Code (0)

등록된 구현이 없습니다.

Tasks

Text-to-Image Generation

Similar Papers 제목 키워드 기반

Oracle inequalities for square root analysis estimators with application to total variation penalties

2019-02-28 · Francesco Ortelli, Sara van de Geer

Through the direct study of the analysis estimator we derive oracle inequalities with fast and slow rates by adapting the arguments involving projections by Dalalyan, Hebiri and Lederer (2017). We then extend the theory …

Differentiable Jitter Correction using Deep Learning-based Image Quality Metric for Phase-Contrast Micro-CT

2026-08-27 · Junan Chen, Yiting Jia, Joscha Maier, Dominik John 외 arxiv

This paper proposes a fully differentiable jitter correction method for X-ray phase-contrast micro computed tomography using a deep learning-based image quality metric that estimates and compensates per-projection rigid …

Faster PET Reconstruction with Non-Smooth Priors by Randomization and Preconditioning

2018-08-21 · Matthias J. Ehrhardt, Pawel Markiewicz, Carola-Bibiane Schönlieb

Uncompressed clinical data from modern positron emission tomography (PET) scanners are very large, exceeding 350 million data points (projection bins). The last decades have seen tremendous advancements in mathematical i…

Vector Ontologies as an LLM world view extraction method

2025-06-16 · Kaspar Rothenfusser, Bekk Blando

Large Language Models (LLMs) possess intricate internal representations of the world, yet these latent structures are notoriously difficult to interpret or repurpose beyond the original prediction task. Building on our e…

SpotDiff: Spotting and Disentangling Interference in Feature Space for Subject-Preserving Image Generation

2025-10-07 · Yongzhi Li, Saining Zhang, Yibing Chen, Boying Li 외 arxiv

Personalized image generation aims to faithfully preserve a reference subject's identity while adapting to diverse text prompts. Existing optimization-based methods ensure high fidelity but are computationally expensive,…

Personalized Image Generation