paper-with-me

Papers

GeoDiff: Geometry-Guided Diffusion for Metric Depth Estimation

2025-10-21 · Tuan Pham, Thanh-Tung Le, Xiaohui Xie, Stephan Mandt arxiv

We introduce a novel framework for metric depth estimation that enhances pretrained diffusion-based monocular depth estimation (DB-MDE) models with stereo vision guidance. While existing DB-MDE methods excel at predicting relative depth, estimating absolute metric depth remains challenging due to scale ambiguities in single-image scenarios. To address this, we reframe depth estimation as an inverse problem, leveraging pretrained latent diffusion models (LDMs) conditioned on RGB images, combined with stereo-based geometric constraints, to learn scale and shift for accurate depth recovery. Our training-free solution seamlessly integrates into existing DB-MDE frameworks and generalizes across indoor, outdoor, and complex environments. Extensive experiments demonstrate that our approach matches or surpasses state-of-the-art methods, particularly in challenging scenarios involving translucent and specular surfaces, all without requiring retraining.

📄 PDF Abstract BibTeX arXiv:2510.18291

Code (0)

등록된 구현이 없습니다.

Tasks

Monocular Depth Estimation

Similar Papers 제목 키워드 기반

Enhancing Underwater Light Field Images via Global Geometry-aware Diffusion Process

2026-01-29 · Yuji Lin, Qian Zhao, Zongsheng Yue, Junhui Hou 외 arxiv

This work studies the challenging problem of acquiring high-quality underwater images via 4-D light field (LF) imaging. To this end, we propose GeoDiff-LF, a novel diffusion-based framework built upon SD-Turbo to enhance…

GeoDiffMM: Geometry-Guided Conditional Diffusion for Motion Magnification

2025-12-09 · Xuedeng Liu, Jiabao Guo, Zheng Zhang, Fei Wang 외 arxiv

Video Motion Magnification (VMM) amplifies subtle macroscopic motions to a perceptible level. Recently, existing mainstream Eulerian approaches address amplification-induced noise via decoupling representation learning s…

Representation Learning

GeoDiffusion: A Training-Free Framework for Accurate 3D Geometric Conditioning in Image Generation

2025-10-25 · Phillip Mueller, Talip Uenlue, Sebastian Schmidt, Marcel Kollovieh 외 arxiv

Precise geometric control in image generation is essential for engineering \& product design and creative industries to control 3D object features accurately in image space. Traditional 3D editing approaches are time-con…

Image GenerationStyle TransferImage Editing

GeoDiff-SAR: A Geometric Prior Guided Diffusion Model for SAR Image Generation

2026-01-07 · Fan Zhang, Xuanting Wu, Fei Ma, Qiang Yin 외 arxiv

Synthetic Aperture Radar (SAR) imaging results are highly sensitive to observation geometries and the geometric parameters of targets. However, existing generative methods primarily operate within the image domain, negle…

Image GenerationPoint Clouds

GeoDiff3D: Self-Supervised 3D Scene Generation with Geometry-Constrained 2D Diffusion Guidance

2026-01-27 · Haozhi Zhu, Miaomiao Zhao, Dingyao Liu, Runze Tian 외 arxiv

3D scene generation is a core technology for gaming, film/VFX, and VR/AR. Growing demand for rapid iteration, high-fidelity detail, and accessible content creation has further increased interest in this area. Existing me…

3D ReconstructionScene Generation3D Generation