Papers Depth Prediction
“Depth Prediction” 태그가 달린 논문 422편 · 필터 해제
MonoMVSNet: Monocular Priors Guided Multi-View Stereo Network
Learning-based Multi-View Stereo (MVS) methods aim to predict depth maps for a sequence of calibrated images to recover dense point clouds. However, existing MVS methods often struggle with challenging regions, such as t…
Depth EstimationDepth PredictionMonocular Depth EstimationBeyond Appearance: Geometric Cues for Robust Video Instance Segmentation
Video Instance Segmentation (VIS) fundamentally struggles with pervasive challenges including object occlusions, motion blur, and appearance variations during temporal association. To overcome these limitations, this wor…
Depth EstimationDepth PredictionInstance SegmentationMonocular Depth Estimation+4RoboScape: Physics-informed Embodied World Model
World models have become indispensable tools for embodied intelligence, serving as powerful simulators capable of generating realistic robotic videos while addressing critical data scarcity challenges. However, current e…
3D geometryDepth EstimationDepth Predictionmodel+1RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation
Dense metric depth estimation using millimeter-wave radar typically requires dense LiDAR supervision, generated via multi-frame projection and interpolation, to guide the learning of accurate depth from sparse radar meas…
Depth EstimationDepth PredictionDiFuse-Net: RGB and Dual-Pixel Depth Estimation using Window Bi-directional Parallax Attention and Cross-modal Transfer Learning
Depth estimation is crucial for intelligent systems, enabling applications from autonomous navigation to augmented reality. While traditional stereo and active depth sensors have limitations in cost, power, and robustnes…
Autonomous NavigationDepth EstimationDepth PredictionDisparity Estimation+2Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
Real-world applications like video gaming and virtual reality often demand the ability to model 3D scenes that users can explore along custom camera trajectories. While significant progress has been made in generating 3D…
3D ReconstructionCamera Pose EstimationDepth EstimationDepth Prediction+3Depth Anything with Any Prior
This work presents Prior Depth Anything, a framework that combines incomplete but precise metric information in depth measurement with relative but complete geometric structures in depth prediction, generating accurate, …
Depth CompletionDepth EstimationDepth PredictionMonocular Depth Estimation+3Edge-Enabled VIO with Long-Tracked Features for High-Accuracy Low-Altitude IoT Navigation
This paper presents a visual-inertial odometry (VIO) method using long-tracked features. Long-tracked features can constrain more visual frames, reducing localization drift. However, they may also lead to accumulated mat…
Depth EstimationDepth PredictionState EstimationMonoCoP: Chain-of-Prediction for Monocular 3D Object Detection
Accurately predicting 3D attributes is crucial for monocular 3D object detection (Mono3D), with depth estimation posing the greatest challenge due to the inherent ambiguity in mapping 2D images to 3D space. While existin…
3D Object DetectionAttributeDepth EstimationDepth Prediction+4In-situ and Non-contact Etch Depth Prediction in Plasma Etching via Machine Learning (ANN & BNN) and Digital Image Colorimetry
Precise monitoring of etch depth and the thickness of insulating materials, such as Silicon dioxide and silicon nitride, is critical to ensuring device performance and yield in semiconductor manufacturing. While conventi…
Depth EstimationDepth PredictionDERD-Net: Learning Depth from Event-based Ray Densities
Event cameras offer a promising avenue for multi-view stereo depth estimation and Simultaneous Localization And Mapping (SLAM) due to their ability to detect blur-free 3D edges at high-speed and over broad illumination c…
Depth EstimationDepth PredictionSimultaneous Localization and MappingStereo Depth EstimationEndo3R: Unified Online Reconstruction from Dynamic Monocular Endoscopic Video
Reconstructing 3D scenes from monocular surgical videos can enhance surgeon's perception and therefore plays a vital role in various computer-assisted surgery tasks. However, achieving scale-consistent reconstruction rem…
Camera Pose EstimationDepth EstimationDepth PredictionDynamic Reconstruction+1Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective Surfaces
Self-supervised monocular depth estimation (SSMDE) has gained attention in the field of deep learning as it estimates depth without requiring ground truth depth maps. This approach typically uses a photometric consistenc…
Depth EstimationDepth PredictionIntrinsic Image DecompositionKnowledge Distillation+1Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video
This paper presents a unified approach to understanding dynamic scenes from casual videos. Large pretrained vision foundation models, such as vision-language, video depth prediction, motion tracking, and segmentation mod…
Camera Pose EstimationDepth EstimationDepth PredictionDynamic Reconstruction+1Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better
Temporal consistency is critical in video prediction to ensure that outputs are coherent and free of artifacts. Traditional methods, such as temporal attention and 3D convolution, may struggle with significant object mot…
ColorizationDepth EstimationDepth PredictionPoint Tracking+2GAA-TSO: Geometry-Aware Assisted Depth Completion for Transparent and Specular Objects
Transparent and specular objects are frequently encountered in daily life, factories, and laboratories. However, due to the unique optical properties, the depth information on these objects is usually incomplete and inac…
Depth CompletionDepth EstimationDepth PredictionRobotic GraspingPow3R: Empowering Unconstrained 3D Reconstruction with Camera and Scene Priors
We present Pow3r, a novel large 3D vision regression model that is highly versatile in the input modalities it accepts. Unlike previous feed-forward models that lack any mechanism to exploit known camera or scene priors …
3D ReconstructionDepth CompletionDepth EstimationDepth Prediction+2Dynamic Point Maps: A Versatile Representation for Dynamic 3D Reconstruction
DUSt3R has recently shown that one can reduce many tasks in multi-view geometry, including estimating camera intrinsics and extrinsics, reconstructing the scene in 3D, and establishing image correspondences, to the predi…
3D Object Tracking3D ReconstructionDepth EstimationDepth Prediction+5Multi-view Reconstruction via SfM-guided Monocular Depth Estimation
In this paper, we present a new method for multi-view geometric reconstruction. In recent years, large vision models have rapidly developed, performing excellently across various tasks and demonstrating remarkable genera…
Depth EstimationDepth PredictionMonocular Depth EstimationGarmentCrafter: Progressive Novel View Synthesis for Single-View 3D Garment Reconstruction and Editing
We introduce GarmentCrafter, a new approach that enables non-professional users to create and modify 3D garments from a single-view image. While recent advances in image generation have facilitated 2D garment design, cre…
3D ReconstructionDepth EstimationDepth PredictionGarment Reconstruction+3