paper-with-me

Papers

A Comprehensive Survey on Segment Anything Model for Vision and Beyond

2023-05-14 · Chunhui Zhang, Li Liu, Yawen Cui, Guanjie Huang, Weilin Lin, Yiqian Yang, Yuehong Hu

Artificial intelligence (AI) is evolving towards artificial general intelligence, which refers to the ability of an AI system to perform a wide range of tasks and exhibit a level of intelligence similar to that of a human being. This is in contrast to narrow or specialized AI, which is designed to perform specific tasks with a high degree of efficiency. Therefore, it is urgent to design a general class of models, which we term foundation models, trained on broad data that can be adapted to various downstream tasks. The recently proposed segment anything model (SAM) has made significant progress in breaking the boundaries of segmentation, greatly promoting the development of foundation models for computer vision. To fully comprehend SAM, we conduct a survey study. As the first to comprehensively review the progress of segmenting anything task for vision and beyond based on the foundation model of SAM, this work focuses on its applications to various tasks and data types by discussing its historical development, recent progress, and profound impact on broad applications. We first introduce the background and terminology for foundation models including SAM, as well as state-of-the-art methods contemporaneous with SAM that are significant for segmenting anything task. Then, we analyze and summarize the advantages and limitations of SAM across various image processing applications, including software scenes, real-world scenes, and complex scenes. Importantly, many insights are drawn to guide future research to develop more versatile foundation models and improve the architecture of SAM. We also summarize massive other amazing applications of SAM in vision and beyond. Finally, we maintain a continuously updated paper list and an open-source project summary for foundation model SAM at \href{https://github.com/liliu-avril/Awesome-Segment-Anything}{\color{magenta}{here}}.

📄 PDF Abstract BibTeX arXiv:2305.08196

Code (1)

liliu-avril/Awesome-Segment-Anything 공식 구현 paddle

Methods 이 논문이 사용한 방법론

SAM 설명 없음

Similar Papers 제목 키워드 기반

A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt Engineering

2023-05-12 · Chaoning Zhang, Joseph Cho, Fachrina Dewi Puspitasari, Sheng Zheng 외

The Segment Anything Model (SAM), developed by Meta AI Research, represents a significant breakthrough in computer vision, offering a robust framework for image and video segmentation. This survey provides a comprehensiv…

Edge DetectionmodelPrompt EngineeringSurvey+2

Segment Anything for Videos: A Systematic Survey

2024-07-31 · Chunhui Zhang, Yawen Cui, Weilin Lin, Guanjie Huang 외

The recent wave of foundation models has witnessed tremendous success in computer vision (CV) and beyond, with the segment anything model (SAM) having sparked a passion for exploring task-agnostic visual foundation model…

Image SegmentationRobot Manipulation GeneralizationSemantic SegmentationSurvey+4

Segment Anything for Satellite Imagery: A Strong Baseline and a Regional Dataset for Automatic Field Delineation

2025-06-19 · Carmelo Scribano, Elena Govi, Paolo Bertellini, Simone Parisi 외

Accurate mapping of agricultural field boundaries is essential for the efficient operation of agriculture. Automatic extraction from high-resolution satellite imagery, supported by computer vision techniques, can avoid c…

Beyond Pixel-Wise Supervision for Medical Image Segmentation: From Traditional Models to Foundation Models

2024-04-20 · Yuyan Shi, Jialu Ma, Jin Yang, Shasha Wang 외

Medical image segmentation plays an important role in many image-guided clinical approaches. However, existing segmentation algorithms mostly rely on the availability of fully annotated images with pixel-wise annotations…

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation

Segment Anything for Video: A Comprehensive Review of Video Object Segmentation and Tracking from Past to Future

2025-07-30 · Guoping Xu, Jayaram K. Udupa, Yajun Yu, Hua-Chieh Shao 외 arxiv

Video Object Segmentation and Tracking (VOST) presents a complex yet critical challenge in computer vision, requiring robust integration of segmentation and tracking across temporally dynamic frames. Traditional methods …

Video Object SegmentationComputational EfficiencyDomain Generalization