paper-with-me

Papers

HS-Pose: Hybrid Scope Feature Extraction for Category-level Object Pose Estimation

2023-03-28 · CVPR 2023 1 · Linfang Zheng, Chen Wang, Yinghan Sun, Esha Dasgupta, Hua Chen, Ales Leonardis, Wei zhang, Hyung Jin Chang

In this paper, we focus on the problem of category-level object pose estimation, which is challenging due to the large intra-category shape variation. 3D graph convolution (3D-GC) based methods have been widely used to extract local geometric features, but they have limitations for complex shaped objects and are sensitive to noise. Moreover, the scale and translation invariant properties of 3D-GC restrict the perception of an object's size and translation information. In this paper, we propose a simple network structure, the HS-layer, which extends 3D-GC to extract hybrid scope latent features from point cloud data for category-level object pose estimation tasks. The proposed HS-layer: 1) is able to perceive local-global geometric structure and global information, 2) is robust to noise, and 3) can encode size and translation information. Our experiments show that the simple replacement of the 3D-GC layer with the proposed HS-layer on the baseline method (GPV-Pose) achieves a significant improvement, with the performance increased by 14.5% on 5d2cm metric and 10.3% on IoU75. Our method outperforms the state-of-the-art methods by a large margin (8.3% on 5d2cm, 6.9% on IoU75) on the REAL275 dataset and runs in real-time (50 FPS).

📄 PDF Abstract BibTeX arXiv:2303.15743

Code (1)

lynne-zheng-linfang/hs-pose 공식 구현 pytorch

Tasks

Pose EstimationTranslation

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

SCOPE: Semantic Conditioning for Sim2Real Category-Level Object Pose Estimation in Robotics

2025-09-29 · Peter Hönig, Stefan Thalhammer, Jean-Baptiste Weibel, Matthias Hirschmanner 외 arxiv

Object manipulation requires accurate object pose estimation. In open environments, robots encounter unknown objects, which requires semantic understanding in order to generalize both to known categories and beyond. To r…

Pose Estimation

SIM-OFE: Structure Information Mining and Object-aware Feature Enhancement for Fine-Grained Visual Categorization

2024-09-18 · journal 2024 9 · Hongbo Sun, Xiangteng He, Jinglin Xu, Yuxin Peng

Fine-grained visual categorization (FGVC) aims to distinguish visual objects from multiple subcategories of the coarse-grained category. Subtle inter-class differences among various subcategories make the FGVC task more …

Fine-Grained Image ClassificationFine-Grained Visual CategorizationObject

Efficient Scopeformer: Towards Scalable and Rich Feature Extraction for Intracranial Hemorrhage Detection

2023-02-01 · Yassine Barhoumi, Nidhal C. Bouaynaya, Ghulam Rasool

The quality and richness of feature maps extracted by convolution neural networks (CNNs) and vision Transformers (ViTs) directly relate to the robust model performance. In medical computer vision, these information-rich …

Computed Tomography (CT)Style Transfer

A Conservative OCR-Enabled Workflow for R214 Sodium Screening of South African Packaged Foods

2026-09-14 · Mayimunah Nagayi, Alice Scaria Khan, Tamryn Frank, Rina Swart 외 arxiv

Using food package images to monitor sodium and salt content against South Africa's R214 sodium limits is challenging when screening decisions require product identity, nutrition facts panel evidence, reporting basis, an…

SCOPE:Planning for Hybrid Querying over Clinical Trial Data

2026-04-28 · Suparno Roy Chowdhury, Manan Roy Choudhury, Tejas Anvekar, Muhammad Ali Khan 외 arxiv

We study clinical trial table reasoning, where answers are not directly stored in visible cells but must be reasoned from semantic understanding through normalization, classification, extraction, or lightweight domain re…

Answer Generation