paper-with-me

홈 › Papers

MobiFuse: A High-Precision On-device Depth Perception System with Multi-Data Fusion

2024-12-18 · Jinrui Zhang, Deyu Zhang, Tingting Long, Wenxin Chen, Ju Ren, Yunxin Liu, Yudong Zhao, Yaoxue Zhang, Youngki Lee

We present MobiFuse, a high-precision depth perception system on mobile devices that combines dual RGB and Time-of-Flight (ToF) cameras. To achieve this, we leverage physical principles from various environmental factors to propose the Depth Error Indication (DEI) modality, characterizing the depth error of ToF and stereo-matching. Furthermore, we employ a progressive fusion strategy, merging geometric features from ToF and stereo depth maps with depth error features from the DEI modality to create precise depth maps. Additionally, we create a new ToF-Stereo depth dataset, RealToF, to train and validate our model. Our experiments demonstrate that MobiFuse excels over baselines by significantly reducing depth measurement errors by up to 77.7%. It also showcases strong generalization across diverse datasets and proves effectiveness in two downstream tasks: 3D reconstruction and 3D segmentation. The demo video of MobiFuse in real-life scenarios is available at the de-identified YouTube link(https://youtu.be/jy-Sp7T1LVs).

📄 PDF Abstract BibTeX arXiv:2412.13848

Code (0)

등록된 구현이 없습니다.

Tasks

3D ReconstructionStereo Matching

Similar Papers 제목 키워드 기반

Monocular Depth Perception Enhancement Based on Joint Shading/Contrast Model and Motion Parallax (JSM)

2026-05-17 · Seungchul Ryu, Hyunjin Yoo, Tara Akhavan arxiv

Stereoscopic 3D displays adopt a binocular depth cue to provide depth perception. However, users should be equipped with expensive special devices to appreciate depth perception based on the binocular depth cues. Also, v…

Towards Robust Driving Perception: A Flexible Scale-Driven Family for Self-Supervised Monocular Depth Estimation

2026-07-01 · Zhaowen Zhu, Li Zhang, Yujie Chen, Tian Zhang 외 arxiv

Self-Supervised Monocular Depth Estimation (MDE) has garnered attention in recent years due to its independence from ground truth. However, most existing models are limited to a single scale and exhibit considerable perf…

Monocular Depth EstimationZero-shot Generalization

QVGGT: Post-Training Quantized Visual Geometry Grounded Transformer

2026-05-29 · Zhizhen Pan, Hesong Wang, Huan Wang arxiv

Estimating 3D attributes directly from images has advanced rapidly with the Visual Geometry Grounded Transformer (VGGT), which predicts camera parameters, depth maps, and point clouds in a single forward pass. However, i…

3D ReconstructionPoint Clouds

Real-time Full-stack Traffic Scene Perception for Autonomous Driving with Roadside Cameras

2022-06-20 · Zhengxia Zou, Rusheng Zhang, Shengyin Shen, Gaurav Pandey 외

We propose a novel and pragmatic framework for traffic scene perception with roadside cameras. The proposed framework covers a full-stack of roadside perception pipeline for infrastructure-assisted autonomous driving, in…

Autonomous DrivingEdge-computingObjectobject-detection+3

Real-time single image depth perception in the wild with handheld devices

2020-06-10 · Filippo Aleotti, Giulio Zaccaroni, Luca Bartolomei, Matteo Poggi 외

Depth perception is paramount to tackle real-world problems, ranging from autonomous driving to consumer applications. For the latter, depth estimation from a single image represents the most versatile solution, since a …

Autonomous DrivingDepth Estimation