paper-with-me

홈 › Papers

Large Scale Joint Semantic Re-Localisation and Scene Understanding via Globally Unique Instance Coordinate Regression

2019-09-23 · Ignas Budvytis, Marvin Teichmann, Tomas Vojir, Roberto Cipolla

In this work we present a novel approach to joint semantic localisation and scene understanding. Our work is motivated by the need for localisation algorithms which not only predict 6-DoF camera pose but also simultaneously recognise surrounding objects and estimate 3D geometry. Such capabilities are crucial for computer vision guided systems which interact with the environment: autonomous driving, augmented reality and robotics. In particular, we propose a two step procedure. During the first step we train a convolutional neural network to jointly predict per-pixel globally unique instance labels and corresponding local coordinates for each instance of a static object (e.g. a building). During the second step we obtain scene coordinates by combining object center coordinates and local coordinates and use them to perform 6-DoF camera pose estimation. We evaluate our approach on real world (CamVid-360) and artificial (SceneCity) autonomous driving datasets. We obtain smaller mean distance and angular errors than state-of-the-art 6-DoF pose estimation algorithms based on direct pose regression and pose estimation from scene coordinates on all datasets. Our contributions include: (i) a novel formulation of scene coordinate regression as two separate tasks of object instance recognition and local coordinate regression and a demonstration that our proposed solution allows to predict accurate 3D geometry of static objects and estimate 6-DoF pose of camera on (ii) maps larger by several orders of magnitude than previously attempted by scene coordinate regression methods, as well as on (iii) lightweight, approximate 3D maps built from 3D primitives such as building-aligned cuboids.

📄 PDF Abstract BibTeX arXiv:1909.10239

Code (0)

등록된 구현이 없습니다.

Tasks

3D geometryAutonomous DrivingCamera Pose EstimationPose EstimationregressionScene Understanding

Similar Papers 제목 키워드 기반

That's My Point: Compact Object-centric LiDAR Pose Estimation for Large-scale Outdoor Localisation

2024-03-07 · Georgi Pramatarov, Matthew Gadd, Paul Newman, Daniele De Martini

This paper is about 3D pose estimation on LiDAR scans with extremely minimal storage requirements to enable scalable mapping and localisation. We achieve this by clustering all points of segmented scans into semantic obj…

3D Pose EstimationPose Estimation

Introspection in Learned Semantic Scene Graph Localisation

2025-10-08 · Manshika Charvi Bissessur, Efimia Panagiotaki, Daniele De Martini arxiv

This work investigates how semantics influence localisation performance and robustness in a learned self-supervised, contrastive semantic localisation framework. After training a localisation network on both original and…

SCLARO: A Dataset for Grounded Scenario-Level Scene Understanding and ScenarioCLIP for Benchmarking

2025-11-25 · Advik Sinha, Saurabh Atreya, Aashutosh A, Sk Aziz Ali 외 arxiv

In the paradigm of computer vision-based precise real-world scene understanding, joint reasoning in terms of contextual understanding about the objects present in a scene, their inter-object relations, and the action bei…

Knowledge DistillationGraph ClassificationScene UnderstandingObject Detection

HOTFLoc++: End-to-End Hierarchical LiDAR Place Recognition, Re-Ranking, and 6-DoF Metric Localisation in Forests

2025-11-12 · Ethan Griffiths, Maryam Haghighat, Simon Denman, Clinton Fookes 외 arxiv

This article presents HOTFLoc++, an end-to-end hierarchical framework for LiDAR place recognition, re-ranking, and 6-DoF metric localisation in forests. Leveraging an octree-based transformer, our approach extracts featu…

Point Clouds

On-the-Fly Adaptation of Regression Forests for Online Camera Relocalisation

2017-02-09 · CVPR 2017 7 · Tommaso Cavallari, Stuart Golodetz, Nicholas A. Lord, Julien Valentin 외

Camera relocalisation is an important problem in computer vision, with applications in simultaneous localisation and mapping, virtual/augmented reality and navigation. Common techniques either match the current image aga…

Camera Relocalizationregression