VOS
2000년 도입 · 논문 117편에서 사용
VOS is a type of video object segmentation model consisting of two network components. The target appearance model consists of a light-weight module, which is learned during the inference stage using fast optimization techniques to predict a coarse but robust target segmentation. The segmentation model is exclusively trained offline, designed to process the coarse scores into high quality segmentation masks.
출처: Learning Fast and Robust Target Models for Video Object Segmentation
소개 논문: Learning Fast and Robust Target Models for Video Object Segmentation
Video Object Segmentation Models · Computer Vision