paper-with-me

홈 › Papers

WarpNet: Weakly Supervised Matching for Single-view Reconstruction

2016-04-19 · CVPR 2016 6 · Angjoo Kanazawa, David W. Jacobs, Manmohan Chandraker

We present an approach to matching images of objects in fine-grained datasets without using part annotations, with an application to the challenging problem of weakly supervised single-view reconstruction. This is in contrast to prior works that require part annotations, since matching objects across class and pose variations is challenging with appearance features alone. We overcome this challenge through a novel deep learning architecture, WarpNet, that aligns an object in one image with a different object in another. We exploit the structure of the fine-grained dataset to create artificial data for training this network in an unsupervised-discriminative learning approach. The output of the network acts as a spatial prior that allows generalization at test time to match real images across variations in appearance, viewpoint and articulation. On the CUB-200-2011 dataset of bird categories, we improve the AP over an appearance-only network by 13.6%. We further demonstrate that our WarpNet matches, together with the structure of fine-grained datasets, allow single-view reconstructions with quality comparable to using annotated point correspondences.

📄 PDF Abstract BibTeX arXiv:1604.05592

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

WSCF-MVCC: Weakly-supervised Calibration-free Multi-view Crowd Counting

2025-12-02 · Bin Li, Daijie Chen, Qi Zhang arxiv

Multi-view crowd counting can effectively mitigate occlusion issues that commonly arise in single-image crowd counting. Existing deep-learning multi-view crowd counting methods project different camera view images onto a…

Crowd Counting

Weakly Supervised Video Individual CountingWeakly Supervised Video Individual Counting

2023-12-10 · Xinyan Liu, Guorong Li, Yuankai Qi, Ziheng Yan 외

Video Individual Counting (VIC) aims to predict the number of unique individuals in a single video. % Existing methods learn representations based on trajectory labels for individuals, which are annotation-expensive. % T…

Contrastive LearningVideo Individual Counting

Weakly Supervised Video Individual Counting

2024-01-01 · CVPR 2024 1 · Xinyan Liu, Guorong Li, Yuankai Qi, Ziheng Yan 외

Video Individual Counting (VIC) aims to predict the number of unique individuals in a single video. Existing methods learn representations based on trajectory labels for individuals which are annotation-expensive. To…

Contrastive LearningVideo Individual Counting

DewarpNet: Single-Image Document Unwarping With Stacked 3D and 2D Regression Networks

2019-10-01 · ICCV 2019 10 · Sagnik Das, Ke Ma, Zhixin Shu, Dimitris Samaras 외

Capturing document images with hand-held devices in unstructured environments is a common practice nowadays. However, "casual" photos of documents are usually unsuitable for automatic information extraction, mainly due t…

3D geometryLocal DistortionMS-SSIMOptical Character Recognition (OCR)+2

Zero-shot Inexact CAD Model Alignment from a Single Image

2025-07-04 · Pattaramanee Arsomngern, Sasikarn Khwanmuang, Matthias Nießner, Supasorn Suwajanakorn arxiv

One practical approach to infer 3D scene structure from a single image is to retrieve a closely matching 3D model from a database and align it with the object in the image. Existing methods rely on supervised training wi…