LED2-Net: Monocular 360deg Layout Estimation via Differentiable Depth Rendering
Although significant progress has been made in room layout estimation, most methods aim to reduce the loss in the 2D pixel coordinate rather than exploiting the room structure in the 3D space. Towards reconstructing the room layout in 3D, we formulate the task of 360 layout estimation as a problem of predicting depth on the horizon line of a panorama. Specifically, we propose the Differentiable Depth Rendering procedure to make the conversion from layout to depth prediction differentiable, thus making our proposed model end-to-end trainable while leveraging the 3D geometric information, without the need of providing the ground truth depth. Our method achieves state-of-the-art performance on numerous 360 layout benchmark datasets. Moreover, our formulation enables a pre-training step on the depth dataset, which further improves the generalizability of our layout estimation model.
Code (0)
등록된 구현이 없습니다.
Tasks
Depth EstimationDepth PredictionRoom Layout EstimationSimilar Papers 제목 키워드 기반
LED2-Net: Monocular 360 Layout Estimation via Differentiable Depth Rendering
Although significant progress has been made in room layout estimation, most methods aim to reduce the loss in the 2D pixel coordinate rather than exploiting the room structure in the 3D space. Towards reconstructing the …
3D Room Layouts From A Single RGB PanoramaDepth EstimationDepth PredictionRoom Layout EstimationMonocular Differentiable Rendering for Self-Supervised 3D Object Detection
3D object detection from monocular images is an ill-posed problem due to the projective entanglement of depth and scale. To overcome this ambiguity, we present a novel self-supervised method for textured 3D shape reconst…
3D Object Detection3D Object Detection From Monocular Images3D Shape ReconstructionDepth Estimation+5Aperture Supervision for Monocular Depth Estimation
We present a novel method to train machine learning algorithms to estimate scene depths from a single image, by using the information provided by a camera's aperture as supervision. Prior works use a depth sensor's outpu…
Depth EstimationMonocular Depth EstimationSAFT: Shape and Appearance of Fabrics from Template via Differentiable Physical Simulations from Monocular Video
The reconstruction of three-dimensional dynamic scenes is a well-established yet challenging task within the domain of computer vision. In this paper, we propose a novel approach that combines the domains of 3D geometry …
Physical Simulations3D ReconstructionRefinement of Monocular Depth Maps via Multi-View Differentiable Rendering
The accurate reconstruction of per-pixel depth for an image is vital for many tasks in computer graphics, computer vision, and robotics. In this paper, we present a novel approach to generate view consistent and detailed…
Depth EstimationMonocular Depth Estimation