paper-with-me

Papers

Leveraging Image Editing Foundation Models for Data-Efficient CT Metal Artifact Reduction

2026-04-07 · Ahmet Rasim Emirdagi, Süleyman Aslan, Mısra Yavuz, Görkay Aydemir, Yunus Bilge Kurt, Nasrin Rahimi, Burak Can Biner, M. Akın Yılmaz arxiv

Metal artifacts from high-attenuation implants severely degrade CT image quality, obscuring critical anatomical structures and posing a challenge for standard deep learning methods that require extensive paired training data. We propose a paradigm shift: reframing artifact reduction as an in-context reasoning task by adapting a general-purpose vision-language diffusion foundation model via parameter-efficient Low-Rank Adaptation (LoRA). By leveraging rich visual priors, our approach achieves effective artifact suppression with only 16 to 128 paired training examples reducing data requirements by two orders of magnitude. Crucially, we demonstrate that domain adaptation is essential for hallucination mitigation; without it, foundation models interpret streak artifacts as erroneous natural objects (e.g., waffles or petri dishes). To ground the restoration, we propose a multi-reference conditioning strategy where clean anatomical exemplars from unrelated subjects are provided alongside the corrupted input, enabling the model to exploit category-specific context to infer uncorrupted anatomy. Extensive evaluation on the AAPM CT-MAR benchmark demonstrates that our method achieves state-of-the-art performance on perceptual and radiological-feature metrics . This work establishes that foundation models, when appropriately adapted, offer a scalable alternative for interpretable, data-efficient medical image reconstruction. Code is available at https://github.com/ahmetemirdagi/CT-EditMAR.

📄 PDF Abstract BibTeX arXiv:2604.05934

Code (0)

등록된 구현이 없습니다.

Tasks

Image ReconstructionDomain AdaptationImage Editing

Similar Papers 제목 키워드 기반

HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing

2024-04-15 · Mude Hui, Siwei Yang, Bingchen Zhao, Yichun Shi 외

This study introduces HQ-Edit, a high-quality instruction-based image editing dataset with around 200,000 edits. Unlike prior approaches relying on attribute guidance or human feedback on building datasets, we devise a s…

Attribute

ShapeUP: Scalable Image-Conditioned 3D Editing

2026-02-05 · Inbar Gat, Dana Cohen-Bar, Guy Levy, Elad Richardson 외 arxiv

Recent advancements in 3D foundation models have enabled the generation of high-fidelity assets, yet precise 3D manipulation remains a significant challenge. Existing 3D editing frameworks often face a difficult trade-of…

EditCast3D: Single-Frame-Guided 3D Editing with Video Propagation and View Selection

2025-10-11 · Huaizhi Qu, Ruichen Zhang, Shuqing Luo, Luchao Qi 외 arxiv

Recent advances in foundation models have driven remarkable progress in image editing, yet their extension to 3D editing remains underexplored. A natural approach is to replace the image editing modules in existing workf…

3D ReconstructionVideo GenerationImage Editing

Pico-Banana-400K: A Large-Scale Dataset for Text-Guided Image Editing

2025-10-22 · Yusu Qian, Eli Bocek-Rivele, Liangchen Song, Jialing Tong 외 arxiv

Recent advances in multimodal models have demonstrated remarkable text-guided image editing capabilities, with systems like GPT-4o and Nano-Banana setting new benchmarks. However, the research community's progress remain…

Image Editing

VicEdit: Learning to Edit Videos from Visual In-Context Examples

2026-08-17 · Yuji Wang, Teng Hu, Yuheng Chen, Ran Yi 외 arxiv

Despite progress in instruction-based video editing, unimodal textual instructions inherently struggle to convey fine-grained textures and complex dynamics. To bridge this perceptual gap, we propose Visual In-context Edi…