paper-with-me

Referring Video Object Segmentation 벤치마크

Referring Video Object Segmentation on MeViS

33개 결과 · ⬇ CSV · JSON

J&F

29.3 35.4 41.5 47.6 53.7 2021-11 2026-09 MTTR — 30.0 (2021-11-29) MTTR — 30.0 (2021-11-29) ReferFormer — 31.0 (2022-01-03) ReferFormer — 31.0 (2022-01-03) LBDT — 29.3 (2022-06-08) LBDT — 29.3 (2022-06-08) VLT+TC — 35.5 (2022-10-28) VLT+TC — 35.5 (2022-10-28) LMPM — 37.2 (2023-08-16) LMPM — 37.2 (2023-08-16) HTR — 42.7 (2024-03-28) HTR — 42.7 (2024-03-28) DsHmp — 46.4 (2024-04-04) DsHmp — 46.4 (2024-04-04) SAMWISE — 48.3 (2024-11-26) SAMWISE — 48.3 (2024-11-26) DsHmp + MTCM — 47.6 (2025-01-09) DsHmp + MTCM — 47.6 (2025-01-09) VRS-HQ (Chat-UniVi-13B) — 50.9 (2025-01-15) VRS-HQ (Chat-UniVi-13B) — 50.9 (2025-01-15) InternVideo2.5 — 32.0 (2025-01-21) InternVideo2.5 — 32.0 (2025-01-21) MPG-SAM 2 — 53.7 (2025-01-23) MPG-SAM 2 — 53.7 (2025-01-23) ReferDINO (Swin-B) — 49.3 (2025-01-24) ReferDINO (Swin-B) — 49.3 (2025-01-24) FindTrack — 53.2 (2025-03-05) FindTrack — 53.2 (2025-03-05) GLUS — 51.3 (2025-04-10) GLUS — 51.3 (2025-04-10) MTTR — 30.0 (2021-11-29) ReferFormer — 31.0 (2022-01-03) VLT+TC — 35.5 (2022-10-28) LMPM — 37.2 (2023-08-16) HTR — 42.7 (2024-03-28) DsHmp — 46.4 (2024-04-04) SAMWISE — 48.3 (2024-11-26) VRS-HQ (Chat-UniVi-13B) — 50.9 (2025-01-15) MPG-SAM 2 — 53.7 (2025-01-23)
RankModel J&FJF PaperCodeYear
1 MPG-SAM 2 53.750.756.7 MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation rongfu-dsb/MPG-SAM2 2025
2 FindTrack 53.250.555.9 Find First, Track Next: Decoupling Identification and Propagation in Referring Video Object Segmentation suhwan-cho/FindTrack 2025
3 GLUS 51.348.554.2 GLUS: Global-Local Reasoning Unified into A Single Large Language Model for Video Segmentation GLUS-video/GLUS 2025
4 VRS-HQ (Chat-UniVi-13B) 50.94853.7 The Devil is in Temporal Token: High Quality Video Reasoning Segmentation sitonggong/vrs-hq 2025
5 ReferDINO (Swin-B) 49.344.753.9 ReferDINO: Referring Video Object Segmentation with Visual Grounding Foundations 2025
6 SAMWISE 48.345.451.2 SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation claudiacuttano/samwise 2024
7 DsHmp + MTCM 47.644.151.1 Multi-Context Temporal Consistent Modeling for Referring Video Object Segmentation choi58/mtcm 2025
8 DsHmp 46.44349.8 Decoupling Static and Hierarchical Motion Perception for Referring Video Segmentation heshuting555/dshmp 2024
9 HTR 42.739.945.5 Temporally Consistent Referring Video Object Segmentation with Hybrid Memory bo-miao/HTR 2024
10 LMPM 37.234.240.2 MeViS: A Large-scale Benchmark for Video Segmentation with Motion Expressions henghuiding/MeViS 2023
11 VLT+TC 35.533.637.3 VLT: Vision-Language Transformer and Query Generation for Referring Segmentation henghuiding/Vision-Language-Transformer 2022
12 InternVideo2.5 32 InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling opengvlab/internvideo 2025
13 ReferFormer 31.029.832.2 Language as Queries for Referring Video Object Segmentation wjn922/referformer 2022
14 MTTR 30.028.831.2 End-to-End Referring Video Object Segmentation with Multimodal Transformers mttr2021/MTTR · JerryX1110/awesome-rvos 2021
15 LBDT 29.327.830.8 Language-Bridged Spatial-Temporal Interaction for Referring Video Object Segmentation dzh19990407/lbdt 2022
16 URVOS 27.825.729.9 URVOS: Unified Referring Video Object Segmentation Network with a Large-Scale Benchmark skynbe/Refer-Youtube-VOS
17 MPG-SAM 2 53.750.756.7 MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation rongfu-dsb/MPG-SAM2 2025
18 FindTrack 53.250.555.9 Find First, Track Next: Decoupling Identification and Propagation in Referring Video Object Segmentation suhwan-cho/FindTrack 2025
19 GLUS 51.348.554.2 GLUS: Global-Local Reasoning Unified into A Single Large Language Model for Video Segmentation GLUS-video/GLUS 2025
20 VRS-HQ (Chat-UniVi-13B) 50.94853.7 The Devil is in Temporal Token: High Quality Video Reasoning Segmentation sitonggong/vrs-hq 2025
1–20 / 33 다음 → 페이지당 10 20 50 100