paper-with-me

Papers

Learning Continuous Depth Representation via Geometric Spatial Aggregator

2022-12-07 · Xiaohang Wang, Xuanhong Chen, Bingbing Ni, Zhengyan Tong, Hang Wang

Depth map super-resolution (DSR) has been a fundamental task for 3D computer vision. While arbitrary scale DSR is a more realistic setting in this scenario, previous approaches predominantly suffer from the issue of inefficient real-numbered scale upsampling. To explicitly address this issue, we propose a novel continuous depth representation for DSR. The heart of this representation is our proposed Geometric Spatial Aggregator (GSA), which exploits a distance field modulated by arbitrarily upsampled target gridding, through which the geometric information is explicitly introduced into feature aggregation and target generation. Furthermore, bricking with GSA, we present a transformer-style backbone named GeoDSR, which possesses a principled way to construct the functional mapping between local coordinates and the high-resolution output results, empowering our model with the advantage of arbitrary shape transformation ready to help diverse zooming demand. Extensive experimental results on standard depth map benchmarks, e.g., NYU v2, have demonstrated that the proposed framework achieves significant restoration gain in arbitrary scale depth map super-resolution compared with the prior art. Our codes are available at https://github.com/nana01219/GeoDSR.

📄 PDF Abstract BibTeX arXiv:2212.03499

Code (1)

nana01219/geodsr 공식 구현 pytorch

Tasks

Depth Map Super-ResolutionSuper-Resolution

Similar Papers 제목 키워드 기반

IVGT: Implicit Visual Geometry Transformer for Neural Scene Representation

2026-05-15 · Yuqi Wu, Tianyu Hu, Wenzhao Zheng, Yuanhui Huang 외 arxiv

Reconstructing coherent 3D geometry and appearance from unposed multi-view images is a fundamental yet challenging problem in computer vision. Most existing visual geometry foundation models predict explicit geometry by …

Camera Pose EstimationNovel View Synthesis

HyGE-Occ: Hybrid View-Transformation with 3D Gaussian and Edge Priors for 3D Panoptic Occupancy Prediction

2025-12-22 · Jong Wook Kim, Wonseok Roh, Ha Dam Baek, Pilhyeon Lee 외 arxiv

3D Panoptic Occupancy Prediction aims to reconstruct a dense volumetric scene map by predicting the semantic class and instance identity of every occupied region in 3D space. Achieving such fine-grained 3D understanding …

SplatSSC: Decoupled Depth-Guided Gaussian Splatting for Semantic Scene Completion

2025-08-04 · Rui Qian, Haozhi Cao, Tianchen Deng, Shenghai Yuan 외 arxiv

Monocular 3D Semantic Scene Completion (SSC) is a challenging yet promising task that aims to infer dense geometric and semantic descriptions of a scene from a single image. While recent object-centric paradigms signific…

3D Semantic Scene Completion

SAND: Spatially Adaptive Network Depth for Fast Sampling of Neural Implicit Surfaces

2026-04-15 · Chuanxiang Yang, Junhui Hou, Yuan Liu, Siyu Ren 외 arxiv

Implicit neural representations are powerful for geometric modeling, but their practical use is often limited by the high computational cost of network evaluations. We observe that implicit representations require progre…

DiST-4D: Disentangled Spatiotemporal Diffusion with Metric Depth for 4D Driving Scene Generation

2025-03-19 · Jiazhe Guo, Yikang Ding, Xiwu Chen, Shuo Chen 외

Current generative models struggle to synthesize dynamic 4D driving scenes that simultaneously support temporal extrapolation and spatial novel view synthesis (NVS) without per-scene optimization. A key challenge lies in…

Novel View SynthesisScene Generation