paper-with-me

Papers

Recognizing Objects From Any View With Object and Viewer-Centered Representations

2020-06-01 · CVPR 2020 6 · Sainan Liu, Vincent Nguyen, Isaac Rehg, Zhuowen Tu

In this paper, we tackle an important task in computer vision: any view object recognition. In both training and testing, for each object instance, we are only given its 2D image viewed from an unknown angle. We propose a computational framework by designing object and viewer-centered neural networks (OVCNet) to recognize an object instance viewed from an arbitrary unknown angle. OVCNet consists of three branches that respectively implement object-centered, 3D viewer-centered, and in-plane viewer-centered recognition. We evaluate our proposed OVCNet using two metrics with unseen views from both seen and novel object instances. Experimental results demonstrate the advantages of OVCNet over classic 2D-image-based CNN classifiers, 3D-object (inferred from 2D image) classifiers, and competing multi-view based approaches. It gives rise to a viable and practical computing framework that combines both viewpoint-dependent and viewpoint-independent features for object recognition from any view.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectObject Recognition

Similar Papers 제목 키워드 기반

Pixels, voxels, and views: A study of shape representations for single view 3D object shape prediction

2018-04-17 · CVPR 2018 6 · Daeyun Shin, Charless C. Fowlkes, Derek Hoiem

The goal of this paper is to compare surface-based and volumetric 3D object shape representations, as well as viewer-centered and object-centered reference frames for single-view 3D shape prediction. We propose a new alg…

Object

AUTO3D: Novel view synthesis through unsupervisely learned variational viewpoint and global 3D representation

2020-07-13 · ECCV 2020 8 · Xiaofeng Liu, Tong Che, Yiqun Lu, Chao Yang 외

This paper targets on learning-based novel view synthesis from a single or limited 2D images without the pose supervision. In the viewer-centered coordinates, we construct an end-to-end trainable conditional variational …

3D ReconstructionDecoderNovel View Synthesis

3D Scene Reconstruction with Multi-layer Depth and Epipolar Transformers

2019-02-18 · ICCV 2019 10 · Daeyun Shin, Zhile Ren, Erik B. Sudderth, Charless C. Fowlkes

We tackle the problem of automatically reconstructing a complete 3D model of a scene from a single RGB image. This challenging task requires inferring the shape of both visible and occluded surfaces. Our approach utilize…

3D Scene Reconstruction

Gaze Prediction in Dynamic 360° Immersive Videos

2018-06-01 · CVPR 2018 6 · Yanyu Xu, Yanbing Dong, Junru Wu, Zhengzhong Sun 외

This paper explores gaze prediction in dynamic $360^circ$ immersive videos, emph{i.e.}, based on the history scan path and VR contents, we predict where a viewer will look at an upcoming time. To tackle this problem, we …

Gaze PredictionPrediction

Position on LLM-Assisted Peer Review: Addressing Reviewer Gap through Mentoring and Feedback

2026-01-14 · JungMin Yun, JuneHyoung Kwon, MiHyeon Kim, YoungBin Kim arxiv

The rapid expansion of AI research has intensified the Reviewer Gap, threatening the peer-review sustainability and perpetuating a cycle of low-quality evaluations. This position paper critiques existing LLM approaches t…