paper-with-me

홈 › Papers

NeurOCS: Neural NOCS Supervision for Monocular 3D Object Localization

2023-05-28 · CVPR 2023 1 · Zhixiang Min, Bingbing Zhuang, Samuel Schulter, Buyu Liu, Enrique Dunn, Manmohan Chandraker

Monocular 3D object localization in driving scenes is a crucial task, but challenging due to its ill-posed nature. Estimating 3D coordinates for each pixel on the object surface holds great potential as it provides dense 2D-3D geometric constraints for the underlying PnP problem. However, high-quality ground truth supervision is not available in driving scenes due to sparsity and various artifacts of Lidar data, as well as the practical infeasibility of collecting per-instance CAD models. In this work, we present NeurOCS, a framework that uses instance masks and 3D boxes as input to learn 3D object shapes by means of differentiable rendering, which further serves as supervision for learning dense object coordinates. Our approach rests on insights in learning a category-level shape prior directly from real driving scenes, while properly handling single-view ambiguities. Furthermore, we study and make critical design choices to learn object coordinates more effectively from an object-centric view. Altogether, our framework leads to new state-of-the-art in monocular 3D localization that ranks 1st on the KITTI-Object benchmark among published monocular methods.

📄 PDF Abstract BibTeX arXiv:2305.17763

Code (0)

등록된 구현이 없습니다.

Tasks

Monocular 3D Object LocalizationObjectObject Localization

Methods 이 논문이 사용한 방법론

PnP PnP, or Poll and Pool, is sampling module extension for DETR-type architectures that adaptively allocates its computation…

Similar Papers 제목 키워드 기반

OmniNOCS: A unified NOCS dataset and model for 3D lifting of 2D objects

2024-07-11 · Akshay Krishnan, Abhijit Kundu, Kevis-Kokitsi Maninis, James Hays 외

We propose OmniNOCS, a large-scale monocular dataset with 3D Normalized Object Coordinate Space (NOCS) maps, object masks, and 3D bounding box annotations for indoor and outdoor scenes. OmniNOCS has 20 times more object …

ObjectPrediction

DL2Fence: Integrating Deep Learning and Frame Fusion for Enhanced Detection and Localization of Refined Denial-of-Service in Large-Scale NoCs

2024-03-20 · Haoyu Wang, Basel Halak, Jianjie Ren, Ahmad Atamli

This study introduces a refined Flooding Injection Rate-adjustable Denial-of-Service (DoS) model for Network-on-Chips (NoCs) and more importantly presents DL2Fence, a novel framework utilizing Deep Learning (DL) and Fram…

Object Level Depth Reconstruction for Category Level 6D Object Pose Estimation From Monocular RGB Image

2022-04-04 · Zhaoxin Fan, Zhenbo Song, Jian Xu, Zhicheng Wang 외

Recently, RGBD-based category-level 6D object pose estimation has achieved promising improvement in performance, however, the requirement of depth information prohibits broader applications. In order to relieve this prob…

6D Pose Estimation using RGBObjectPose Estimation

MonoGRNet: A Geometric Reasoning Network for Monocular 3D Object Localization

2018-11-26 · Zengyi Qin, Jinglu Wang, Yan Lu

Detecting and localizing objects in the real 3D space, which plays a crucial role in scene understanding, is particularly challenging given only a single RGB image due to the geometric information loss during imagery pro…

2D Object Detection3D Object DetectionDepth EstimationMonocular 3D Object Detection+5

StereoPose: Category-Level 6D Transparent Object Pose Estimation from Stereo Images via Back-View NOCS

2022-11-03 · Kai Chen, Stephen James, Congying Sui, Yun-hui Liu 외

Most existing methods for category-level pose estimation rely on object point clouds. However, when considering transparent objects, depth cameras are usually not able to capture meaningful data, resulting in point cloud…

ObjectPose EstimationTransparent objects