paper-with-me

Referring Expression Segmentation

22개 벤치마크 · 논문 164편 · 이 태스크의 논문 보기 →

Benchmarks

RefCoCo val

결과 74개

RefCOCO+ val

결과 66개

RefCOCO+ test B

결과 60개

RefCOCO+ testA

결과 60개

A2D Sentences

결과 54개

RefCOCOg-val

결과 46개

J-HMDB

결과 42개

DAVIS 2017 (val)

결과 36개

RefCOCOg-test

결과 36개

RefCOCO testA

결과 26개

RefCOCO testB

결과 26개

PhraseCut

결과 13개

RefCOCO

결과 8개

ReferIt

결과 6개

Refer-YouTube-VOS

결과 4개

A2Dre test

결과 2개

CLEVR-Ref+

결과 2개

G-Ref test A

결과 2개

G-Ref test B

결과 2개

G-Ref val

결과 2개

Most implemented

Papers

DRAgent: Discriminative Reasoning Agent for Referring Expression Segmentation

2026-08-24 · Yujie Qi, Luyan Zhang arxiv

Referring Expression Segmentation (RES) aims to generate a pixel-level mask for the object specified by a language expression. Recent methods based on multimodal large language models (MLLMs) often rely on one-pass coord…

Referring Expression SegmentationVisual Localization

Falcon Perception-HD: High Density Perception via Reinforcement Learning

2026-08-19 · Sofian Chaybouti, Yasser Dahou, Ngoc Dung Huynh, Reda Alami 외 arxiv

Autoregressive perception models trained to localize visual entities under the open-vocabulary setting are mostly trained using Supervised fine-tuning (SFT) with maximum likelihood, yet it optimizes a proxy objective (pe…

Referring Expression SegmentationReinforcement Learning

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry

2026-07-20 · Dingyun Zhang, Lixue Gong, Wei Liu hf

In line with the prevailing direction of vision research, we explore the integration of both generation and editing capabilities for video and image modalities within a single model. Current approaches to collecting vide…

Referring Expression SegmentationImage Editing

FlowSeg: Dynamic Semantic Guidance for LLM-Conditioned Segmentation

2026-05-28 · Zekang Zhang, Guangyu Gao, Youyun Tang, ChengJing Wu 외 arxiv

LLM-conditioned segmentation has recently advanced rapidly by coupling large language models with iterative mask generation frameworks. However, we identify a persistent failure mode in current propose-then-select pipeli…

Referring Expression Segmentation

Learning to Label: A Reinforced Self-Evolving Framework for Semi-supervised Referring Expression Segmentation

2026-05-27 · Runlong Cao, Ying Zang, Chuanwei Zhou, Tianrun Chen 외 arxiv

Semi-supervised referring expression segmentation (SS-RES) aims to achieve precise pixel-level language grounding under limited annotation, yet suffers from limited supervision and unreliable pseudo-labels when exploitin…

Referring Expression Segmentation

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding

2026-05-08 · Yuan Yao, Qiushi Yang, Humen Zhong, Jiangning Wei 외 arxiv

Open-world referring segmentation requires grounding unconstrained language expressions to precise pixel-level regions. Existing multimodal large language models (MLLMs) exhibit strong open-world visual grounding, but th…

Referring Expression SegmentationVisual Grounding

전체 164편 보기 →