paper-with-me

Papers

SAM2 for Image and Video Segmentation: A Comprehensive Survey

2025-03-17 · Zhang Jiaxing, Tang Hao

Despite significant advances in deep learning for image and video segmentation, existing models continue to face challenges in cross-domain adaptability and generalization. Image and video segmentation are fundamental tasks in computer vision with wide-ranging applications in healthcare, agriculture, industrial inspection, and autonomous driving. With the advent of large-scale foundation models, SAM2 - an improved version of SAM (Segment Anything Model)has been optimized for segmentation tasks, demonstrating enhanced performance in complex scenarios. However, SAM2's adaptability and limitations in specific domains require further investigation. This paper systematically analyzes the application of SAM2 in image and video segmentation and evaluates its performance in various fields. We begin by introducing the foundational concepts of image segmentation, categorizing foundation models, and exploring the technical characteristics of SAM and SAM2. Subsequently, we delve into SAM2's applications in static image and video segmentation, emphasizing its performance in specialized areas such as medical imaging and the challenges of cross-domain adaptability. As part of our research, we reviewed over 200 related papers to provide a comprehensive analysis of the topic. Finally, the paper highlights the strengths and weaknesses of SAM2 in segmentation tasks, identifies the technical challenges it faces, and proposes future development directions. This review provides valuable insights and practical recommendations for optimizing and applying SAM2 in real-world scenarios.

📄 PDF Abstract BibTeX arXiv:2503.12781

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingImage SegmentationSegmentationSemantic SegmentationSurveyVideo SegmentationVideo Semantic Segmentation

Methods 이 논문이 사용한 방법론

SAM 설명 없음

Similar Papers 제목 키워드 기반

A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects

2025-06-16 · Guohuan Xie, Syed Ariff Syed Hesham, Wenya Guo, Bing Li 외

Video Scene Parsing (VSP) has emerged as a cornerstone in computer vision, facilitating the simultaneous segmentation, recognition, and tracking of diverse visual entities in dynamic scenes. In this survey, we present a …

BenchmarkingInstance SegmentationOpen-Vocabulary Video SegmentationPanoptic Segmentation+8

Reasoning Segmentation for Images and Videos: A Survey

2025-05-24 · Yiqing Shen, Chenjia Li, Fei Xiong, Jeong-O Jeong 외

Reasoning Segmentation (RS) aims to delineate objects based on implicit text queries, the interpretation of which requires reasoning and knowledge integration. Unlike the traditional formulation of segmentation problems …

Reasoning SegmentationSurvey

A Survey on Deep Learning Technique for Video Segmentation

2021-07-02 · Tianfei Zhou, Fatih Porikli, David Crandall, Luc van Gool 외

Video segmentation -- partitioning video frames into multiple segments or objects -- plays a critical role in a broad range of practical applications, from enhancing visual effects in movie, to understanding scenes in au…

Autonomous DrivingDeep LearningScene UnderstandingSegmentation+4

A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt Engineering

2023-05-12 · Chaoning Zhang, Joseph Cho, Fachrina Dewi Puspitasari, Sheng Zheng 외

The Segment Anything Model (SAM), developed by Meta AI Research, represents a significant breakthrough in computer vision, offering a robust framework for image and video segmentation. This survey provides a comprehensiv…

Edge DetectionmodelPrompt EngineeringSurvey+2

Image Segmentation Using Deep Learning: A Survey

2020-01-15 · Shervin Minaee, Yuri Boykov, Fatih Porikli, Antonio Plaza 외

Image segmentation is a key topic in image processing and computer vision with applications such as scene understanding, medical image analysis, robotic perception, video surveillance, augmented reality, and image compre…

DecoderDeep LearningImage CompressionImage Segmentation+5