paper-with-me

홈 › Papers

Explaining the Ambiguity of Object Detection and 6D Pose From Visual Data

2018-12-01 · ICCV 2019 10 · Fabian Manhardt, Diego Martin Arroyo, Christian Rupprecht, Benjamin Busam, Tolga Birdal, Nassir Navab, Federico Tombari

3D object detection and pose estimation from a single image are two inherently ambiguous problems. Oftentimes, objects appear similar from different viewpoints due to shape symmetries, occlusion and repetitive textures. This ambiguity in both detection and pose estimation means that an object instance can be perfectly described by several different poses and even classes. In this work we propose to explicitly deal with this uncertainty. For each object instance we predict multiple pose and class outcomes to estimate the specific pose distribution generated by symmetries and repetitive textures. The distribution collapses to a single outcome when the visual appearance uniquely identifies just one valid pose. We show the benefits of our approach which provides not only a better explanation for pose ambiguity, but also a higher accuracy in terms of pose estimation.

📄 PDF Abstract BibTeX arXiv:1812.00287

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object DetectionObjectobject-detectionObject DetectionPose Estimationvalid

Similar Papers 제목 키워드 기반

Explaining YOLO: Leveraging Grad-CAM to Explain Object Detections

2022-11-22 · Armin Kirchknopf, Djordje Slijepcevic, Ilkay Wunderlich, Michael Breiter 외

We investigate the problem of explainability for visual object detectors. Specifically, we demonstrate on the example of the YOLO object detector how to integrate Grad-CAM into the model architecture and analyze the resu…

Object

DoRO: Disambiguation of referred object for embodied agents

2022-07-28 · Pradip Pramanick, Chayan Sarkar, Sayan Paul, Ruddra dev Roychoudhury 외

Robotic task instructions often involve a referred object that the robot must locate (ground) within the environment. While task intent understanding is an essential part of natural language understanding, less effort is…

Natural Language UnderstandingObject

AF-XRAY: Visual Explanation and Resolution of Ambiguity in Legal Argumentation Frameworks

2025-07-14 · Yilin Xia, Heng Zheng, Shawn Bowers, Bertram Ludäscher arxiv

Argumentation frameworks (AFs) provide formal approaches for legal reasoning, but identifying sources of ambiguity and explaining argument acceptance remains challenging for non-experts. We present AF-XRAY, an open-sourc…

Legal Reasoning

The mutual exclusivity bias of bilingual visually grounded speech models

2025-06-04 · Dan Oneata, Leanne Nortje, Yevgen Matusevych, Herman Kamper

Mutual exclusivity (ME) is a strategy where a novel word is associated with a novel object rather than a familiar one, facilitating language learning in children. Recent work has found an ME bias in a visually grounded s…

Deep Variation-structured Reinforcement Learning for Visual Relationship and Attribute Detection

2017-03-08 · CVPR 2017 7 · Xiaodan Liang, Lisa Lee, Eric P. Xing

Despite progress in visual perception tasks such as image classification and detection, computers still struggle to understand the interdependency of objects in the scene as a whole, e.g., relations between objects or th…

Attributeimage-classificationImage ClassificationObject+5