EdgeStereo: A Context Integrated Residual Pyramid Network for Stereo Matching
Recent convolutional neural networks, especially end-to-end disparity estimation models, achieve remarkable performance on stereo matching task. However, existed methods, even with the complicated cascade structure, may fail in the regions of non-textures, boundaries and tiny details. Focus on these problems, we propose a multi-task network EdgeStereo that is composed of a backbone disparity network and an edge sub-network. Given a binocular image pair, our model enables end-to-end prediction of both disparity map and edge map. Basically, we design a context pyramid to encode multi-scale context information in disparity branch, followed by a compact residual pyramid for cascaded refinement. To further preserve subtle details, our EdgeStereo model integrates edge cues by feature embedding and edge-aware smoothness loss regularization. Comparative results demonstrates that stereo matching and edge detection can help each other in the unified model. Furthermore, our method achieves state-of-art performance on both KITTI Stereo and Scene Flow benchmarks, which proves the effectiveness of our design.
Code (0)
등록된 구현이 없습니다.
Tasks
Disparity EstimationEdge DetectionStereo MatchingStereo Matching HandSimilar Papers 제목 키워드 기반
EdgeStereo: An Effective Multi-Task Learning Network for Stereo Matching and Edge Detection
Recently, leveraging on the development of end-to-end convolutional neural networks (CNNs), deep stereo matching networks have achieved remarkable performance far exceeding traditional approaches. However, state-of-the-a…
Disparity EstimationEdge DetectionMulti-Task LearningStereo Matching+1Pyramid Stereo Matching Network
Recent work has shown that depth estimation from a stereo pair of images can be formulated as a supervised learning task to be resolved with convolutional neural networks (CNNs). However, current architectures rely on pa…
Depth EstimationOmnnidirectional Stereo Depth EstimationStereo Depth EstimationStereo-LiDAR Fusion+2Multi-scale Cross-form Pyramid Network for Stereo Matching
Stereo matching plays an indispensable part in autonomous driving, robotics and 3D scene reconstruction. We propose a novel deep learning architecture, which called CFP-Net, a Cross-Form Pyramid stereo matching network f…
3D Feature Matching3D Scene ReconstructionAutonomous DrivingForm+2Cost Volume Pyramid Based Depth Inference for Multi-View Stereo
We propose a cost volume-based neural network for depth inference from multi-view images. We demonstrate that building a cost volume pyramid in a coarse-to-fine manner instead of constructing a cost volume at a fixed res…
3D ReconstructionPoint CloudsA Flexible Recurrent Residual Pyramid Network for Video Frame Interpolation
Video frame interpolation (VFI) aims at synthesizing new video frames in-between existing frames to generate smoother high frame rate videos. Current methods usually use the fixed pre-trained networks to generate interpo…
Optical Flow EstimationVideo Frame Interpolation