paper-with-me

Papers

Lifting GIS Maps into Strong Geometric Context for Scene Understanding

2015-07-14 · Raúl Díaz, Minhaeng Lee, Jochen Schubert, Charless C. Fowlkes

Contextual information can have a substantial impact on the performance of visual tasks such as semantic segmentation, object detection, and geometric estimation. Data stored in Geographic Information Systems (GIS) offers a rich source of contextual information that has been largely untapped by computer vision. We propose to leverage such information for scene understanding by combining GIS resources with large sets of unorganized photographs using Structure from Motion (SfM) techniques. We present a pipeline to quickly generate strong 3D geometric priors from 2D GIS data using SfM models aligned with minimal user input. Given an image resectioned against this model, we generate robust predictions of depth, surface normals, and semantic labels. We show that the precision of the predicted geometry is substantially more accurate other single-image depth estimation methods. We then demonstrate the utility of these contextual constraints for re-scoring pedestrian detections, and use these GIS contextual features alongside object detection score maps to improve a CRF-based semantic segmentation framework, boosting accuracy over baseline models.

📄 PDF Abstract BibTeX arXiv:1507.03698

Code (0)

등록된 구현이 없습니다.

Tasks

Depth Estimationobject-detectionObject DetectionScene UnderstandingSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

FisheyeGaussianLift: BEV Feature Lifting for Surround-View Fisheye Camera Perception

2025-11-21 · Shubham Sonarghare, Prasad Deshpande, Ciaran Hogan, Deepika-Rani Kaliappan-Mahalingam 외 arxiv

Accurate BEV semantic segmentation from fisheye imagery remains challenging due to extreme non-linear distortion, occlusion, and depth ambiguity inherent to wide-angle projections. We present a distortion-aware BEV segme…

Semantic SegmentationBEV Segmentation

E3DGS: Unified Geometric-Photometric Equivariance for 3D Gaussian Splatting via Color-as-Geometry Embedding

2026-07-17 · Chankyo Kim, Maani Ghaffari arxiv

3D Gaussian Splatting (3DGS) captures scenes by coupling explicit geometry (position, covariance) with view-dependent photometry (Spherical Harmonics). However, building $\mathrm{SE}(3)$-equivariant architectures on thes…

Coalgebraic Fuzzy geometric logic

2022-05-02 · Litan Kumar Das, Kumar Sankar Ray, Prakash Chandra Mali

The paper aims to develop a framework for coalgebraic fuzzy geometric logic by adding modalities to the language of fuzzy geometric logic. Using the methods of coalgebra, the modal operators are introduced in the languag…

SweetDreamer: Aligning Geometric Priors in 2D Diffusion for Consistent Text-to-3D

2023-10-04 · Weiyu Li, Rui Chen, Xuelin Chen, Ping Tan

It is inherently ambiguous to lift 2D results from pre-trained diffusion models to a 3D world for text-to-3D generation. 2D diffusion models solely learn view-agnostic priors and thus lack 3D knowledge during the lifting…

3D GenerationText to 3D

Leveraging Previous-Traversal Point Cloud Map Priors for Camera-Based 3D Object Detection and Tracking

2026-04-28 · Markus Käppeler, Özgün Çiçek, Yakov Miron, Abhinav Valada arxiv

Camera-based 3D object detection and tracking are central to autonomous driving, yet precise 3D object localization remains fundamentally constrained by depth ambiguity when no expensive, depth-rich online LiDAR is avail…

Object Localization3D Object DetectionAutonomous Driving