paper-with-me

홈 › Papers

Ladder Fine-tuning approach for SAM integrating complementary network

2023-06-22 · Shurong Chai, Rahul Kumar Jain, Shiyu Teng, Jiaqing Liu, Yinhao Li, Tomoko Tateyama, Yen-Wei Chen

Recently, foundation models have been introduced demonstrating various tasks in the field of computer vision. These models such as Segment Anything Model (SAM) are generalized models trained using huge datasets. Currently, ongoing research focuses on exploring the effective utilization of these generalized models for specific domains, such as medical imaging. However, in medical imaging, the lack of training samples due to privacy concerns and other factors presents a major challenge for applying these generalized models to medical image segmentation task. To address this issue, the effective fine tuning of these models is crucial to ensure their optimal utilization. In this study, we propose to combine a complementary Convolutional Neural Network (CNN) along with the standard SAM network for medical image segmentation. To reduce the burden of fine tuning large foundation model and implement cost-efficient trainnig scheme, we focus only on fine-tuning the additional CNN network and SAM decoder part. This strategy significantly reduces trainnig time and achieves competitive results on publicly available dataset. The code is available at https://github.com/11yxk/SAM-LST.

📄 PDF Abstract BibTeX arXiv:2306.12737

Code (1)

11yxk/sam-lst 공식 구현 pytorch

Tasks

DecoderImage SegmentationMedical Image SegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

SAM 설명 없음
Focus 설명 없음

Similar Papers 제목 키워드 기반

Ladder Up, Memory Down: Low-Cost Fine-Tuning With Side Nets

2025-12-16 · Estelle Zheng, Nathan Cerisara, Sébastien Warichet, Emmanuel Helbert 외 arxiv

Fine-tuning large language models (LLMs) is often limited by the memory available on commodity GPUs. Parameter-efficient fine-tuning (PEFT) methods such as QLoRA reduce the number of trainable parameters, yet still incur…

parameter-efficient fine-tuningNatural Language Understanding

LC3Net: Ladder context correlation complementary network for salient object detection

2021-10-21 · Xian Fang, Jinchao Zhu, Xiuli Shao, Hongpeng Wang

Currently, existing salient object detection methods based on convolutional neural networks commonly resort to constructing discriminative networks to aggregate high level and low level features. However, contextual info…

DecoderDiversityobject-detectionObject Detection+1

Ladder: A Model-Agnostic Framework Boosting LLM-based Machine Translation to the Next Level

2024-06-22 · Zhaopeng Feng, Ruizhe Chen, Yan Zhang, Zijie Meng 외

General-purpose Large Language Models (LLMs) like GPT-4 have achieved remarkable advancements in machine translation (MT) by leveraging extensive web content. On the other hand, translation-specific LLMs are built by pre…

Machine TranslationTranslation

TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning

2025-08-11 · Junzhe Xu, Yuyang Yin, Xi Chen arxiv

This paper introduces TBAC-UniImage, a novel unified model for multimodal understanding and generation. We achieve this by deeply integrating a pre-trained Diffusion Model, acting as a generative ladder, with a Multimoda…

LST: Ladder Side-Tuning for Parameter and Memory Efficient Transfer Learning

2022-06-13 · Yi-Lin Sung, Jaemin Cho, Mohit Bansal

Fine-tuning large pre-trained models on downstream tasks has been adopted in a variety of domains recently. However, it is costly to update the entire parameter set of large pre-trained models. Although recently proposed…

Transfer LearningVisual Question Answering (VQA)