paper-with-me

홈 › Papers

PR-MaGIC: Prompt Refinement Via Mask Decoder Gradient Flow For In-Context Segmentation

2026-04-13 · Minjae Lee, Sungwoo Hur, Soojin Hwang, Won Hwa Kim arxiv

Visual Foundation Models (VFMs) such as the Segment Anything Model (SAM) have significantly advanced broad use of image segmentation. However, SAM and its variants necessitate substantial manual effort for prompt generation and additional training for specific applications. Recent approaches address these limitations by integrating SAM into in-context (one/few shot) segmentation, enabling auto-prompting through semantic alignment between query and support images. Despite these efforts, they still generate sub-optimal prompts that degrade segmentation quality due to visual inconsistencies between support and query images. To tackle this limitation, we introduce PR-MaGIC (Prompt Refinement via Mask Decoder Gradient Flow for In-Context Segmentation), a training-free test-time framework that refines prompts via gradient flow derived from SAM's mask decoder. PR-MaGIC seamlessly integrates into in-context segmentation frameworks, being theoretically grounded yet practically stabilized through a simple top-1 selection strategy that ensures robust performance across samples. Extensive evaluations demonstrate that PR-MaGIC consistently improves segmentation quality across various benchmarks, effectively mitigating inadequate prompts without requiring additional training or architectural modifications.

📄 PDF Abstract BibTeX arXiv:2604.12113

Code (0)

등록된 구현이 없습니다.

Tasks

Image Segmentation

Similar Papers 제목 키워드 기반

Anomagic: Crossmodal Prompt-driven Zero-shot Anomaly Generation

2025-11-13 · Yuxin Jiang, Wei Luo, Hui Zhang, Qiyu Chen 외 arxiv

We propose Anomagic, a zero-shot anomaly generation method that produces semantically coherent anomalies without requiring any exemplar anomalies. By unifying both visual and textual cues through a crossmodal prompt enco…

Anomaly Detection

MagicComp: Training-free Dual-Phase Refinement for Compositional Video Generation

2025-03-18 · Hongyu Zhang, Yufan Deng, Shenghai Yuan, Peng Jin 외

Text-to-video (T2V) generation has made significant strides with diffusion models. However, existing methods still struggle with accurately binding attributes, determining spatial relationships, and capturing complex act…

DenoisingVideo Generation

MAGIC: Few-Shot Mask-Guided Anomaly Inpainting with Prompt Perturbation, Spatially Adaptive Guidance, and Context Awareness

2025-07-03 · JaeHyuck Choi, MinJun Kim, Je Hyeong Hong arxiv

Few-shot anomaly generation is a key challenge in industrial quality control. Although diffusion models are promising, existing methods struggle: global prompt-guided approaches corrupt normal regions, and existing inpai…

MAGIC: Mask-Guided Image Synthesis by Inverting a Quasi-Robust Classifier

2022-09-23 · Mozhdeh Rouhsedaghat, Masoud Monajatipoor, C. -C. Jay Kuo, Iacopo Masi

We offer a method for one-shot mask-guided image synthesis that allows controlling manipulations of a single image by inverting a quasi-robust classifier equipped with strong regularizers. Our proposed method, entitled M…

Image Generation

PromptMagician: Interactive Prompt Engineering for Text-to-Image Creation

2023-07-18 · Yingchaojie Feng, Xingbo Wang, Kam Kwai Wong, Sijia Wang 외

Generative text-to-image models have gained great popularity among the public for their powerful capability to generate high-quality images based on natural language prompts. However, developing effective prompts for des…

Prompt Engineering