paper-with-me

홈 › Papers

A Learning-based Variable Size Part Extraction Architecture for 6D Object Pose Recovery in Depth

2017-01-09 · Caner Sahin, Rigas Kouskouridas, Tae-Kyun Kim

State-of-the-art techniques for 6D object pose recovery depend on occlusion-free point clouds to accurately register objects in 3D space. To deal with this shortcoming, we introduce a novel architecture called Iterative Hough Forest with Histogram of Control Points that is capable of estimating the 6D pose of occluded and cluttered objects given a candidate 2D bounding box. Our Iterative Hough Forest (IHF) is learnt using parts extracted only from the positive samples. These parts are represented with Histogram of Control Points (HoCP), a "scale-variant" implicit volumetric description, which we derive from recently introduced Implicit B-Splines (IBS). The rich discriminative information provided by the scale-variant HoCP features is leveraged during inference. An automatic variable size part extraction framework iteratively refines the object's initial pose that is roughly aligned due to the extraction of coarsest parts, the ones occupying the largest area in image pixels. The iterative refinement is accomplished based on finer (smaller) parts that are represented with more discriminative control point descriptors by using our Iterative Hough Forest. Experiments conducted on a publicly available dataset report that our approach show better registration performance than the state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:1701.02166

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Efficient Joint Learning for Clinical Named Entity Recognition and Relation Extraction Using Fourier Networks: A Use Case in Adverse Drug Events

2023-02-08 · Anthony Yazdani, Dimitrios Proios, Hossein Rouhizadeh, Douglas Teodoro

Current approaches for clinical information extraction are inefficient in terms of computational costs and memory consumption, hindering their application to process large-scale electronic health records (EHRs). We propo…

GPUnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Efficient Motion Modelling with Variable-sized blocks from Hierarchical Cuboidal Partitioning

2022-08-28 · Priyabrata Karmakar, Manzur Murshed, Manoranjan Paul, David Taubman

Motion modelling with block-based architecture has been widely used in video coding where a frame is divided into fixed-sized blocks that are motion compensated independently. This often leads to coding inefficiency as f…

4k

Learngene Search Across Multiple Datasets for Building Variable-Sized Models

2026-05-06 · Boyu Shi, Junbo Zhou, Chang Liu, Xu Yang 외 arxiv

Deep learning methods are widely used under diverse resource constraints, resulting in models of varying sizes, such as the Vision Transformer (ViT) series. Deploying these models typically requires costly pretraining an…

Justlookup: One Millisecond Deep Feature Extraction for Point Clouds By Lookup Tables

2019-08-14 · Hongxin Lin, Zelin Xiao, Yang Tan, Hongyang Chao 외

Deep models are capable of fitting complex high dimensional functions while usually yielding large computation load. There is no way to speed up the inference process by classical lookup tables due to the high-dimensiona…

CPU

Disentangling semantics in language through VAEs and a certain architectural choice

2020-12-24 · Ghazi Felhi, Joseph Le Roux, Djamé Seddah

We present an unsupervised method to obtain disentangled representations of sentences that single out semantic content. Using modified Transformers as building blocks, we train a Variational Autoencoder to translate the …

Open Information ExtractionSentence