paper-with-me

홈 › Papers

A Unified Mutual Supervision Framework for Referring Expression Segmentation and Generation

2022-11-15 · Shijia Huang, Feng Li, Hao Zhang, Shilong Liu, Lei Zhang, LiWei Wang

Reference Expression Segmentation (RES) and Reference Expression Generation (REG) are mutually inverse tasks that can be naturally jointly trained. Though recent work has explored such joint training, the mechanism of how RES and REG can benefit each other is still unclear. In this paper, we propose a unified mutual supervision framework that enables two tasks to improve each other. Our mutual supervision contains two directions. On the one hand, Disambiguation Supervision leverages the expression unambiguity measurement provided by RES to enhance the language generation of REG. On the other hand, Generation Supervision uses expressions automatically generated by REG to scale up the training of RES. Such unified mutual supervision effectively improves two tasks by solving their bottleneck problems. Extensive experiments show that our approach significantly outperforms all existing methods on REG and RES tasks under the same setting, and detailed ablation studies demonstrate the effectiveness of all components in our framework.

📄 PDF Abstract BibTeX arXiv:2211.07919

Code (0)

등록된 구현이 없습니다.

Tasks

Reference Expression GenerationReferring ExpressionReferring Expression SegmentationText Generation

Similar Papers 제목 키워드 기반

A Joint Speaker-Listener-Reinforcer Model for Referring Expressions

2016-12-30 · CVPR 2017 7 · Licheng Yu, Hao Tan, Mohit Bansal, Tamara L. Berg

Referring expressions are natural language constructions used to identify particular objects within a scene. In this paper, we propose a unified framework for the tasks of referring expression comprehension and generatio…

Referring ExpressionReferring Expression Comprehension

MUTATT: Visual-Textual Mutual Guidance for Referring Expression Comprehension

2020-03-18 · Shuai Wang, Fan Lyu, Wei Feng, Song Wang

Referring expression comprehension (REC) aims to localize a text-related region in a given image by a referring expression in natural language. Existing methods focus on how to build convincing visual and language repres…

Referring ExpressionReferring Expression Comprehension

Referring Image Segmentation Using Text Supervision

2023-08-28 · ICCV 2023 1 · Fang Liu, Yuhao Liu, Yuqiu Kong, Ke Xu 외

Existing Referring Image Segmentation (RIS) methods typically require expensive pixel-level or box-level annotations for supervision. In this paper, we observe that the referring texts used in RIS already provide suffici…

Image SegmentationObject LocalizationReferring Expression SegmentationSegmentation+2

Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions

2025-10-26 · Kai Ye, Bowen Liu, Jianghang Lin, Jiayi Ji 외 arxiv

Referring Remote Sensing Image Segmentation (RRSIS) aims to segment instances in remote sensing images according to referring expressions. Unlike Referring Image Segmentation on general images, acquiring high-quality ref…

Referring ExpressionImage Segmentation

When redundancy is useful: A Bayesian approach to 'overinformative' referring expressions

2019-03-19 · Judith Degen, Robert D. Hawkins, Caroline Graf, Elisa Kreiss 외

Referring is one of the most basic and prevalent uses of language. How do speakers choose from the wealth of referring expressions at their disposal? Rational theories of language use have come under attack for decades f…

InformativenessSpecificity