paper-with-me

Papers

UniVector: Unified Vector Extraction via Instance-Geometry Interaction

2025-10-15 · Yinglong Yan, Jun Yue, Shaobo Xia, Hanmeng Sun, Tianxu Ying, Chengcheng Wu, Sifan Lan, Min He, Pedram Ghamisi, Leyuan Fang arxiv

Vector extraction retrieves structured vector geometry from raster images, offering high-fidelity representation and broad applicability. Existing methods, however, are usually tailored to a single vector type (e.g., polygons, polylines, line segments), requiring separate models for different structures. This stems from treating instance attributes (category, structure) and geometric attributes (point coordinates, connections) independently, limiting the ability to capture complex structures. Inspired by the human brain's simultaneous use of semantic and spatial interactions in visual perception, we propose UniVector, a unified VE framework that leverages instance-geometry interaction to extract multiple vector types within a single model. UniVector encodes vectors as structured queries containing both instance- and geometry-level information, and iteratively updates them through an interaction module for cross-level context exchange. A dynamic shape constraint further refines global structures and key points. To benchmark multi-structure scenarios, we introduce the Multi-Vector dataset with diverse polygons, polylines, and line segments. Experiments show UniVector sets a new state of the art on both single- and multi-structure VE tasks. Code and dataset will be released at https://github.com/yyyyll0ss/UniVector.

📄 PDF Abstract BibTeX arXiv:2510.13234

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Vector Map as Language: Toward Unified Remote Sensing Vector Mapping

2026-06-09 · Yinglong Yan, Yunkai Yang, Haoyi Wang, Wei Fu 외 arxiv

Remote sensing vector mapping aims to generate structured maps of geospatial entities, such as buildings, roads, and water bodies, from remote sensing imagery. In practice, vector maps usually contain multiple category l…

Reinforcement LearningText Generation

IRIS-SLAM: Unified Geo-Instance Representations for Robust Semantic Localization and Mapping

2026-02-21 · Tingyang Xiao, Liu Liu, Wei Feng, Zhengyu Zou 외 arxiv

Geometry foundation models have significantly advanced dense geometric SLAM, yet existing systems often lack deep semantic understanding and robust loop closure capabilities. Meanwhile, contemporary semantic mapping appr…

Semantic SLAM

DTCLMapper: Dual Temporal Consistent Learning for Vectorized HD Map Construction

2024-05-09 · Siyu Li, Jiacheng Lin, Hao Shi, Jiaming Zhang 외

Temporal information plays a pivotal role in Bird's-Eye-View (BEV) driving scene understanding, which can alleviate the visual information sparsity. However, the indiscriminate temporal fusion method will cause the barri…

Contrastive LearningScene UnderstandingSelf-Supervised Learning

EA3D: Online Open-World 3D Object Extraction from Streaming Videos

2025-10-29 · Xiaoyu Zhou, Jingqi Wang, Yuang Jia, Yongtao Wang 외 arxiv

Current 3D scene understanding methods are limited by offline-collected multi-view data or pre-constructed 3D geometry. In this paper, we present ExtractAnything3D (EA3D), a unified online framework for open-world 3D obj…

Instance SegmentationScene Understanding3D ReconstructionVisual Odometry

Multi-Sensor 3D Object Box Refinement for Autonomous Driving

2019-09-11 · Peiliang Li, Si-Qi Liu, Shaojie Shen

We propose a 3D object detection system with multi-sensor refinement in the context of autonomous driving. In our framework, the monocular camera serves as the fundamental sensor for 2D object proposal and initial 3D bou…

3D Object DetectionAutonomous DrivingObjectobject-detection+1