Generalized Referring Expression Segmentation
1개 벤치마크 · 논문 18편 · 이 태스크의 논문 보기 →
Benchmarks
gRefCOCO
Most implemented
CoHD: A Counting-Aware Hierarchical Decoding Framework for Generalized Referring Expression Segmentation
GRES: Generalized Referring Expression Segmentation
DeRIS: Decoupling Perception and Cognition for Enhanced Referring Image Segmentation through Loopback Synergy
Bring Adaptive Binding Prototypes to Generalized Referring Expression Segmentation
PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model
GSVA: Generalized Segmentation via Multimodal Large Language Models
Papers
GREx: Generalized Referring Expression Segmentation, Comprehension, and Generation
Referring Expression Segmentation (RES) and Comprehension (REC) respectively segment and detect the object described by an expression, while Referring Expression Generation (REG) generates an expression for the selected …
Generalized Referring Expression SegmentationReferring expression generationGeneralized Referring Expression Segmentation on Aerial Photos
Referring expression segmentation is a fundamental task in computer vision that integrates natural language understanding with precise visual localization of target regions. Considering aerial imagery (e.g., modern aeria…
Generalized Referring Expression SegmentationNatural Language UnderstandingSemantic SegmentationVisual LocalizationLatent Expression Generation for Referring Image Segmentation and Grounding
Visual grounding tasks, such as referring image segmentation (RIS) and referring expression comprehension (REC), aim to localize a target object based on a given textual description. The target object in an image can be …
Generalized Referring Expression SegmentationContrastive LearningImage SegmentationVisual GroundingDeRIS: Decoupling Perception and Cognition for Enhanced Referring Image Segmentation through Loopback Synergy
Referring Image Segmentation (RIS) is a challenging task that aims to segment objects in an image based on natural language expressions. While prior studies have predominantly concentrated on improving vision-language in…
Data AugmentationGeneralized Referring Expression SegmentationImage SegmentationReading Comprehension+2Refer to Anything with Vision-Language Prompts
Recent image segmentation models have advanced to segment images into high-quality masks for visual entities, and yet they cannot provide comprehensive semantic understanding for complex queries based on both language an…
BenchmarkingGeneralized Referring Expression SegmentationImage SegmentationReferring Expression+3Hierarchical Alignment-enhanced Adaptive Grounding Network for Generalized Referring Expression Comprehension
In this work, we address the challenging task of Generalized Referring Expression Comprehension (GREC). Compared to the classic Referring Expression Comprehension (REC) that focuses on single-target expressions, GREC ext…
Generalized Referring Expression ComprehensionGeneralized Referring Expression SegmentationObject CountingPhrase Grounding+3