paper-with-me

홈 › Papers

HRSeg: High-Resolution Visual Perception and Enhancement for Reasoning Segmentation

2025-07-17 · Weihuang Lin, Yiwei Ma, Xiaoshuai Sun, Shuting He, Jiayi Ji, Liujuan Cao, Rongrong Ji

The reasoning segmentation task involves segmenting objects within an image by interpreting implicit user instructions, which may encompass subtleties such as contextual cues and open-world knowledge. Despite significant advancements made by existing approaches, they remain constrained by low perceptual resolution, as visual encoders are typically pre-trained at lower resolutions. Furthermore, simply interpolating the positional embeddings of visual encoders to enhance perceptual resolution yields only marginal performance improvements while incurring substantial computational costs. To address this, we propose HRSeg, an efficient model with high-resolution fine-grained perception. It features two key innovations: High-Resolution Perception (HRP) and High-Resolution Enhancement (HRE). The HRP module processes high-resolution images through cropping, integrating local and global features for multi-granularity quality. The HRE module enhances mask features by integrating fine-grained information from high-resolution images, refining their alignment with text features for precise segmentation. Extensive ablation studies validate the effectiveness of our modules, while comprehensive experiments on multiple benchmark datasets demonstrate HRSeg's superior performance.

📄 PDF Abstract BibTeX arXiv:2507.12883

Code (0)

등록된 구현이 없습니다.

Tasks

Reasoning SegmentationWorld Knowledge

Similar Papers 제목 키워드 기반

RHRSegNet: Relighting High-Resolution Night-Time Semantic Segmentation

2024-07-08 · Sarah Elmahdy, Rodaina Hebishy, Ali Hamdi

Night time semantic segmentation is a crucial task in computer vision, focusing on accurately classifying and segmenting objects in low-light conditions. Unlike daytime techniques, which often perform worse in nighttime …

Autonomous DrivingScene SegmentationSegmentationSemantic Segmentation

REHRSeg: Unleashing the Power of Self-Supervised Super-Resolution for Resource-Efficient 3D MRI Segmentation

2024-10-14 · Zhiyun Song, Yinjie Zhao, Xiaomin Li, Manman Fei 외

High-resolution (HR) 3D magnetic resonance imaging (MRI) can provide detailed anatomical structural information, enabling precise segmentation of regions of interest for various medical image analysis tasks. Due to the h…

Knowledge DistillationMedical Image AnalysisMRI segmentationSegmentation+1

Real-time High-Resolution Neural Network with Semantic Guidance for Crack Segmentation

2023-07-01 · Yongshang Li, Ronggui Ma, Han Liu, Gaoli Cheng

Deep learning plays an important role in crack segmentation, but most work utilize off-the-shelf or improved models that have not been specifically developed for this task. High-resolution convolution neural networks tha…

Crack SegmentationSegmentation

LaVida Drive: Vision-Text Interaction VLM for Autonomous Driving with Token Selection, Recovery and Enhancement

2024-11-20 · Siwen Jiao, Yangyi Fang, Baoyun Peng, Wangqun Chen 외

Recent advancements in Visual Language Models (VLMs) have made them crucial for visual question answering (VQA) in autonomous driving, enabling natural human-vehicle interactions. However, existing methods often struggle…

Autonomous DrivingComputational EfficiencyQuestion AnsweringVisual Question Answering+1

CURVE: CLIP-Utilized Reinforcement Learning for Visual Image Enhancement via Simple Image Processing

2025-05-29 · Yuka Ogino, Takahiro Toizumi, Atsushi Ito

Low-Light Image Enhancement (LLIE) is crucial for improving both human perception and computer vision tasks. This paper addresses two challenges in zero-reference LLIE: obtaining perceptually 'good' images using the Cont…

Computational EfficiencyImage EnhancementLow-Light Image Enhancementreinforcement-learning+1