paper-with-me

홈 › Papers

MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering

2026-03-07 · Trong-Thang Pham, Loc Nguyen, Anh Nguyen, Hien Nguyen, Ngan Le arxiv

Generative diffusion models are increasingly used for medical imaging data augmentation, but text prompting cannot produce causal training data. Re-prompting rerolls the entire generation trajectory, altering anatomy, texture, and background. Inversion-based editing methods introduce reconstruction error that causes structural drift. We propose MedSteer, a training-free activation-steering framework for endoscopic synthesis. MedSteer identifies a pathology vector for each contrastive prompt pair in the cross-attention layers of a diffusion transformer. At inference time, it steers image activations along this vector, generating counterfactual pairs from scratch where the only difference is the steered concept. All other structure is preserved by construction. We evaluate MedSteer across three experiments on Kvasir v3 and HyperKvasir. On counterfactual generation across three clinical concept pairs, MedSteer achieves flip rates of 0.800, 0.925, and 0.950, outperforming the best inversion-based baseline in both concept flip rate and structural preservation. On dye disentanglement, MedSteer achieves 75% dye removal against 20% (PnP) and 10% (h-Edit). On downstream polyp detection, augmenting with MedSteer counterfactual pairs achieves ViT AUC of 0.9755 versus 0.9083 for quantity-matched re-prompting, confirming that counterfactual structure drives the gain. Code is at link https://github.com/phamtrongthang123/medsteer

📄 PDF Abstract BibTeX arXiv:2603.07066

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentation

Similar Papers 제목 키워드 기반

ExtraGS: Enhancing Endoscopic View Extrapolation via Diffusion-Guided 3D Gaussian Splatting

2026-07-14 · Cheng-Tai Hsieh, Jiwei Shan, Han Fang, Jianshu Hu 외 arxiv

Robot-assisted minimally invasive surgery (MIS) critically depends on reliable endoscopic perception for navigation and safety. However, conventional endoscopes provide only a limited field of view, leaving large portion…

Novel View Synthesis

Plug-and-Play Model-Agnostic Counterfactual Policy Synthesis for Deep Reinforcement Learning based Recommendation

2022-08-10 · Siyu Wang, Xiaocong Chen, Lina Yao, Sally Cripps 외

Recent advances in recommender systems have proved the potential of Reinforcement Learning (RL) to handle the dynamic evolution processes between users and recommender systems. However, learning to train an optimal RL ag…

counterfactualData AugmentationDeep Reinforcement LearningRecommendation Systems+2

FLex: Joint Pose and Dynamic Radiance Fields Optimization for Stereo Endoscopic Videos

2024-03-18 · Florian Philipp Stilz, Mert Asim Karaoglu, Felix Tristram, Nassir Navab 외

Reconstruction of endoscopic scenes is an important asset for various medical applications, from post-surgery analysis to educational training. Neural rendering has recently shown promising results in endoscopic reconstr…

Neural RenderingNovel View Synthesis

A Deep Learning Based 6 Degree-of-Freedom Localization Method for Endoscopic Capsule Robots

2017-05-15 · Mehmet Turan, Yasin Almalioglu, Ender Konukoglu, Metin Sitti

We present a robust deep learning based 6 degrees-of-freedom (DoF) localization system for endoscopic capsule robots. Our system mainly focuses on localization of endoscopic capsule robots inside the GI tract using only …

CPUTranslation

Seeing Through Smoke: Surgical Desmoking for Improved Visual Perception

2026-03-26 · Jingpei Lu, Fengyi Jiang, Xiaorui Zhang, Lingbo Jin 외 arxiv

Minimally invasive and robot-assisted surgery relies heavily on endoscopic imaging, yet surgical smoke produced by electrocautery and vessel-sealing instruments can severely degrade visual perception and hinder vision-ba…

Synthetic Data GenerationStereo Depth EstimationImage Reconstruction