paper-with-me

Video Object Segmentation

13개 벤치마크 · 논문 608편 · 이 태스크의 논문 보기 →

Benchmarks

DAVIS 2016

결과 48개

DAVIS 2017 (val)

결과 34개

YouTube-VOS 2018

결과 34개

DAVIS 2017 (test-dev)

결과 20개

YouTube-VOS 2019

결과 20개

DAVIS 2017

결과 10개

M$^3$-VOS

결과 8개

DAVIS-2017 (test-dev)

결과 4개

FBMS

결과 4개

YouTube

결과 4개

FBMS-59

결과 2개

MOSE

결과 2개

SegTrack-v2

결과 2개

Most implemented

One-Shot Video Object Segmentation

2016-11-16 · 구현 8개

Papers

MLLM-Assisted Audio VOS: A 3rd Place Report for the MeViS-Audio Track, 8th LSVOS Challenge

2026-08-24 · Liangtao Shi, Jinxia Xie, Xiantao Hu, Ting Liu arxiv

In this technical report, we present a training-free framework for audio-guided video object segmentation, which integrates Multimodal Large Language Models (MLLMs) with SAM-based segmentation models. We decompose the ta…

Video Object SegmentationMultimodal ReasoningVideo Segmentation

SAM2Dual: Training-Free, Dual Memory for Long-Term Video Object Segmentation

2026-08-19 · JeongRae Kim, Changwon Lim arxiv

Long-term video object segmentation (VOS) remains challenging due to error accumulation under extended occlusions, re-appearance, and scene changes. Although SAM2 provides strong zero-shot performance, its streaming memo…

Video Object Segmentation

RRTrack: Robust and Recoverable Object 6D Pose Tracking for Dynamic Scenes

2026-07-26 · Junyue Li, Ye Zheng, Yifan Chen, Zhe Sun 외 arxiv

Robust object 6D pose tracking is critical for robotic systems operating in dynamic and occluded scenes. Per-frame estimators are accurate but computationally expensive, while current trackers struggle with fast motion a…

Video Object SegmentationPose Tracking

REMIND: RE-Identification with Memory for INDoor Navigation

2026-07-10 · Pablo Diaz-Pereda, Alejandro Rodriguez-Ramos, David Perez-Saura, Pascual Campoy arxiv

Mobile robots operating indoors must re-identify previously observed objects after long temporal gaps, significant viewpoint changes, and severe illumination variations. This remains a challenging problem: multi-object t…

Video Object SegmentationVehicle Re-IdentificationMulti-Object Tracking

SAM-MT: Real-Time Interactive Multi-Target Video Segmentation

2026-07-09 · Ruiqi Shen, Chang Liu, Henghui Ding arxiv

Modern Video Object Segmentation (VOS) involves tracking and segmenting user-specified targets. While recent approaches have achieved remarkable performance in single-target scenarios, extending them to multi-target sett…

Video Object SegmentationVideo Segmentation

`Attention-Guided Cross-Temporal Clustering for Self-Supervised Video Object Segmentation

2026-07-08 · Waqas Arshid, Mohammad Awrangjeb, Alan Wee-Chung Liew, Yongsheng Gao arxiv

Video object segmentation (VOS) is a fundamental task in video understanding, requiring accurate delineation and consistent tracking of objects across frames. While supervised methods achieve strong performance, they rel…

Video Object SegmentationSelf-Supervised Learning

전체 608편 보기 →