paper-with-me

홈 › Papers

DepthFocus: Controllable Depth Estimation for See-Through Scenes

2025-11-21 · Junhong Min, Jimin Kim, Minwook Kim, Cheol-Hui Min, Youngpil Jeon, Minyong Choi arxiv

Depth in the real world is rarely singular. Transmissive materials create layered ambiguities that confound conventional perception systems. Existing models remain passive; conventional approaches typically estimate static depth maps anchored to the nearest surface, and even recent multi-head extensions suffer from a representational bottleneck due to fixed feature representations. This stands in contrast to human vision, which actively shifts focus to perceive a desired depth. We introduce \textbf{DepthFocus}, a steerable Vision Transformer that redefines stereo depth estimation as condition-aware control. Instead of extracting fixed features, our model dynamically modulates its computation based on a physical reference depth, integrating dual conditional mechanisms to selectively perceive geometry aligned with the desired focus. Leveraging a newly curated large-scale synthetic dataset, \textbf{DepthFocus} achieves state-of-the-art results across all evaluated benchmarks, including both standard single-layer and complex multi-layered scenarios. While maintaining high precision in opaque regions, our approach effectively resolves depth ambiguities in transparent and reflective scenes by selectively reconstructing geometry at a target distance. This capability enables robust, intent-driven perception that significantly outperforms existing multi-layer methods, marking a substantial step toward active 3D perception. \noindent \textbf{Project page}: \href{https://junhong-3dv.github.io/depthfocus-project/}{\textbf{this https URL}}.

📄 PDF Abstract BibTeX arXiv:2511.16993

Code (0)

등록된 구현이 없습니다.

Tasks

Stereo Depth Estimation

Similar Papers 제목 키워드 기반

ScaleDepth: Decomposing Metric Depth Estimation into Scale Prediction and Relative Depth Estimation

2024-07-11 · Ruijie Zhu, Chuxin Wang, Ziyang Song, Li Liu 외

Estimating depth from a single image is a challenging visual task. Compared to relative depth estimation, metric depth estimation attracts more attention due to its practical physical significance and critical applicatio…

Depth EstimationMonocular Depth Estimation

SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations

2026-04-09 · Yunnan Wang, Kecheng Zheng, Jianyuan Wang, Minghao Chen 외 arxiv

The convergence of 3D geometric perception and video synthesis has created an unprecedented demand for large-scale video data that is rich in both semantic and spatio-temporal information. While existing datasets have ad…

Monocular Depth EstimationVideo GenerationPoint Tracking

DnD: Dense Depth Estimation in Crowded Dynamic Indoor Scenes

2021-08-12 · ICCV 2021 10 · Dongki Jung, Jaehoon Choi, Yonghan Lee, Deokhwa Kim 외

We present a novel approach for estimating depth from a monocular camera as it moves through complex and crowded indoor environments, e.g., a department store or a metro station. Our approach predicts absolute scale dept…

3D ReconstructionDepth Estimation

InseRF: Text-Driven Generative Object Insertion in Neural 3D Scenes

2024-01-10 · Mohamad Shahbazi, Liesbeth Claessens, Michael Niemeyer, Edo Collins 외

We introduce InseRF, a novel method for generative object insertion in the NeRF reconstructions of 3D scenes. Based on a user-provided textual description and a 2D bounding box in a reference viewpoint, InseRF generates …

3D scene EditingDepth EstimationMonocular Depth EstimationNeRF+2

RealMonoDepth: Self-Supervised Monocular Depth Estimation for General Scenes

2020-04-14 · Mertalp Ocal, Armin Mustafa

We present a generalised self-supervised learning approach for monocular estimation of the real depth across scenes with diverse depth ranges from 1--100s of meters. Existing supervised methods for monocular depth estima…

Depth EstimationMonocular Depth EstimationSelf-Supervised Learning