paper-with-me

홈 › Papers

Discriminate-and-Rectify Encoders: Learning from Image Transformation Sets

2017-03-14 · Andrea Tacchetti, Stephen Voinea, Georgios Evangelopoulos

The complexity of a learning task is increased by transformations in the input space that preserve class identity. Visual object recognition for example is affected by changes in viewpoint, scale, illumination or planar transformations. While drastically altering the visual appearance, these changes are orthogonal to recognition and should not be reflected in the representation or feature encoding used for learning. We introduce a framework for weakly supervised learning of image embeddings that are robust to transformations and selective to the class distribution, using sets of transforming examples (orbit sets), deep parametrizations and a novel orbit-based loss. The proposed loss combines a discriminative, contrastive part for orbits with a reconstruction error that learns to rectify orbit transformations. The learned embeddings are evaluated in distance metric-based tasks, such as one-shot classification under geometric transformations, as well as face verification and retrieval under more realistic visual variability. Our results suggest that orbit sets, suitably computed or observed, can be used for efficient, weakly-supervised learning of semantically relevant image embeddings.

📄 PDF Abstract BibTeX arXiv:1703.04775

Code (0)

등록된 구현이 없습니다.

Tasks

Face VerificationObject RecognitionRetrievalWeakly-supervised Learning

Similar Papers 제목 키워드 기반

Rectifying homographies for stereo vision: analytical solution for minimal distortion

2022-02-28 · Pasquale Lafiosca, Marta Ceccaroni

Stereo rectification is the determination of two image transformations (or homographies) that map corresponding points on the two images, projections of the same point in the 3D space, onto the same horizontal line in th…

On a Generalization of the Average Distance Classifier

2020-01-08 · Sarbojit Roy, Soham Sarkar, Subhajit Dutta

In high dimension, low sample size (HDLSS)settings, the simple average distance classifier based on the Euclidean distance performs poorly if differences between the locations get masked by the scale differences. To rect…

A DNN Framework For Text Image Rectification From Planar Transformations

2016-11-14 · Chengzhe Yan, Jie Hu, Chang-Shui Zhang

In this paper, a novel neural network architecture is proposed attempting to rectify text images with mild assumptions. A new dataset of text images is collected to verify our model and open to public. We explored the ca…

Steering Self-Supervised Feature Learning Beyond Local Pixel Statistics

2020-04-05 · CVPR 2020 6 · Simon Jenni, Hailin Jin, Paolo Favaro

We introduce a novel principle for self-supervised feature learning based on the discrimination of specific transformations of an image. We argue that the generalization capability of learned features depends on what ima…

On the Transformation of Latent Space in Autoencoders

2019-01-24 · Jaehoon Cha, Kyeong Soo Kim, Sanghyuk Lee

Noting the importance of the latent variables in inference and learning, we propose a novel framework for autoencoders based on the homeomorphic transformation of latent variables, which could reduce the distance between…

Denoising