Point Anywhere: Directed Object Estimation from Omnidirectional Images
One of the intuitive instruction methods in robot navigation is a pointing gesture. In this study, we propose a method using an omnidirectional camera to eliminate the user/object position constraint and the left/right constraint of the pointing arm. Although the accuracy of skeleton and object detection is low due to the high distortion of equirectangular images, the proposed method enables highly accurate estimation by repeatedly extracting regions of interest from the equirectangular image and projecting them onto perspective images. Furthermore, we found that training the likelihood of the target object in machine learning further improves the estimation accuracy.
Code (1)
Tasks
Objectobject-detectionObject DetectionPositionRobot NavigationSimilar Papers 제목 키워드 기반
A Flexible Framework for Virtual Omnidirectional Vision to Improve Operator Situation Awareness
During teleoperation of a mobile robot, providing good operator situation awareness is a major concern as a single mistake can lead to mission failure. Camera streams are widely used for teleoperation but offer limited f…
Scene UnderstandingHelvipad: A Real-World Dataset for Omnidirectional Stereo Depth Estimation
Despite progress in stereo depth estimation, omnidirectional imaging remains underexplored, mainly due to the lack of appropriate data. We introduce Helvipad, a real-world dataset for omnidirectional stereo depth estimat…
Depth CompletionDepth EstimationOmnnidirectional Stereo Depth EstimationStereo Depth EstimationDistortion-Tolerant Monocular Depth Estimation On Omnidirectional Images Using Dual-cubemap
Estimating the depth of omnidirectional images is more challenging than that of normal field-of-view (NFoV) images because the varying distortion can significantly twist an object's shape. The existing methods suffer fro…
Depth EstimationMonocular Depth EstimationGaze Target Estimation Anywhere with Concepts
Estimating human gaze targets from images in-the-wild is an important and formidable task. Existing approaches primarily employ brittle, multi-stage pipelines that require explicit inputs, like head bounding boxes and hu…
Gaze Target EstimationGaze EstimationApplications of Deep Learning for Top-View Omnidirectional Imaging: A Survey
A large field-of-view fisheye camera allows for capturing a large area with minimal numbers of cameras when they are mounted on a high position facing downwards. This top-view omnidirectional setup greatly reduces the wo…
Activity RecognitionDeep LearningMiscellaneousobject-detection+3