paper-with-me

Papers

Manhattan Scene Understanding via XSlit Imaging

2013-06-01 · CVPR 2013 6 · Jinwei Ye, Yu Ji, Jingyi Yu

A Manhattan World (MW) [3] is composed of planar surfaces and parallel lines aligned with three mutually orthogonal principal axes. Traditional MW understanding algorithms rely on geometry priors such as the vanishing points and reference (ground) planes for grouping coplanar structures. In this paper, we present a novel single-image MW reconstruction algorithm from the perspective of nonpinhole cameras. We show that by acquiring the MW using an XSlit camera, we can instantly resolve coplanarity ambiguities. Specifically, we prove that parallel 3D lines map to 2D curves in an XSlit image and they converge at an XSlit Vanishing Point (XVP). In addition, if the lines are coplanar, their curved images will intersect at a second common pixel that we call Coplanar Common Point (CCP). CCP is a unique image feature in XSlit cameras that does not exist in pinholes. We present a comprehensive theory to analyze XVPs and CCPs in a MW scene and study how to recover 3D geometry in a complex MW scene from XVPs and CCPs. Finally, we build a prototype XSlit camera by using two layers of cylindrical lenses. Experimental results on both synthetic and real data show that our new XSlitcamera-based solution provides an effective and reliable solution for MW understanding.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

3D geometryScene Understanding

Similar Papers 제목 키워드 기반

Resolving Scale Ambiguity Via XSlit Aspect Ratio Analysis

2015-06-14 · ICCV 2015 12 · Wei Yang, Haiting Lin, Sing Bing Kang, Jingyi Yu

In perspective cameras, images of a frontal-parallel 3D object preserve its aspect ratio invariant to its depth. Such an invariance is useful in photography but is unique to perspective projection. In this paper, we show…

3D Reconstruction

Text-Driven 3D Indoor Scene Synthesis in Non-Manhattan Environments

2026-07-02 · Xianhui Meng, Zirui Song, Yuchen Zhang, Li Zhang 외 arxiv

Large Language Models (LLMs) have demonstrated remarkable capabilities in 3D indoor synthesis for Manhattan environments. However, existing methods often fail to capture plausible object layout patterns in non-Manhattan …

Indoor Scene Synthesis

Single-Shot Cuboids: Geodesics-based End-to-end Manhattan Aligned Layout Estimation from Spherical Panoramas

2021-02-07 · Nikolaos Zioulis, Federico Alvarez, Dimitrios Zarpalas, Petros Daras

It has been shown that global scene understanding tasks like layout estimation can benefit from wider field of views, and specifically spherical panoramas. While much progress has been made recently, all previous approac…

Keypoint EstimationScene Understanding

Rotational Crossed-Slit Light Field

2016-06-01 · CVPR 2016 6 · Nianyi Li, Haiting Lin, Bilin Sun, Mingyuan Zhou 외

Light fields (LFs) are image-based representation that records the radiance along all rays along every direction through every point in space. Traditionally LFs are acquired by using a 2D grid of evenly spaced pinhole…

Stereo MatchingStereo Matching Hand

ManhattanSLAM: Robust Planar Tracking and Mapping Leveraging Mixture of Manhattan Frames

2021-03-28 · Raza Yunus, Yanyan Li, Federico Tombari

In this paper, a robust RGB-D SLAM system is proposed to utilize the structural information in indoor scenes, allowing for accurate tracking and efficient dense mapping on a CPU. Prior works have used the Manhattan World…

Camera Pose EstimationCPUPose EstimationSuperpixels