paper-with-me

홈 › Papers

PatchRefiner V2: Fast and Lightweight Real-Domain High-Resolution Metric Depth Estimation

2025-01-02 · Zhenyu Li, Wenqing Cui, Shariq Farooq Bhat, Peter Wonka

While current high-resolution depth estimation methods achieve strong results, they often suffer from computational inefficiencies due to reliance on heavyweight models and multiple inference steps, increasing inference time. To address this, we introduce PatchRefiner V2 (PRV2), which replaces heavy refiner models with lightweight encoders. This reduces model size and inference time but introduces noisy features. To overcome this, we propose a Coarse-to-Fine (C2F) module with a Guided Denoising Unit for refining and denoising the refiner features and a Noisy Pretraining strategy to pretrain the refiner branch to fully exploit the potential of the lightweight refiner branch. Additionally, we introduce a Scale-and-Shift Invariant Gradient Matching (SSIGM) loss to enhance synthetic-to-real domain transfer. PRV2 outperforms state-of-the-art depth estimation methods on UnrealStereo4K in both accuracy and speed, using fewer parameters and faster inference. It also shows improved depth boundary delineation on real-world datasets like CityScape, ScanNet++, and KITTI, demonstrating its versatility across domains.

📄 PDF Abstract BibTeX arXiv:2501.01121

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingDepth Estimation

Similar Papers 제목 키워드 기반

PatchRefiner: Leveraging Synthetic Data for Real-Domain High-Resolution Monocular Metric Depth Estimation

2024-06-10 · Zhenyu Li, Shariq Farooq Bhat, Peter Wonka

This paper introduces PatchRefiner, an advanced framework for metric single image depth estimation aimed at high-resolution real-domain inputs. While depth estimation is crucial for applications such as autonomous drivin…

3D ReconstructionAutonomous DrivingDepth Estimation

FCL-GAN: A Lightweight and Real-Time Baseline for Unsupervised Blind Image Deblurring

2022-04-16 · Suiyi Zhao, Zhao Zhang, Richang Hong, Mingliang Xu 외

Blind image deblurring (BID) remains a challenging and significant task. Benefiting from the strong fitting ability of deep learning, paired data-driven supervised BID method has obtained great progress. However, paired …

Blind Image DeblurringDeblurringImage Deblurring

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild

2026-05-21 · Mao Zheng, Zheng Li, Tao Chen, Bo Lv 외 arxiv

Hy-MT2 is a family of fast-thinking multilingual translation models designed for complex real-world scenarios. It includes three model sizes: 1.8B, 7B, and 30B-A3B (MoE), all of which support translation among 33 languag…

FedLiTeCAN : A Federated Lightweight Transformer for Fast and Robust CAN Bus Intrusion Detection

2025-12-30 · Devika S, Pratik Narang, Tejasvi Alladi arxiv

This work implements a lightweight Transformer model for IDS in the domain of Connected and Autonomous Vehicles

Autonomous VehiclesIntrusion Detection

FastSHADE: Fast Self-augmented Hierarchical Asymmetric Denoising for Efficient inference on mobile devices

2026-04-11 · Nikolay Falaleev arxiv

Real-time image denoising is essential for modern mobile photography but remains challenging due to the strict latency and power constraints of edge devices. This paper presents FastSHADE (Fast Self-augmented Hierarchica…

Image Denoising