Object Localization based on Structural SVM using Privileged Information
We propose a structured prediction algorithm for object localization based on Support Vector Machines (SVMs) using privileged information. Privileged information provides useful high-level knowledge for image understanding and facilitates learning a reliable model even with a small number of training examples. In our setting, we assume that such information is available only at training time since it may be difficult to obtain from visual data accurately without human supervision. Our goal is to improve performance by incorporating privileged information into ordinary learning framework and adjusting model parameters for better generalization. We tackle object localization problem based on a novel structural SVM using privileged information, where an alternating loss-augmented inference procedure is employed to handle the term in the objective function corresponding to privileged information. We apply the proposed algorithm to the Caltech-UCSD Birds 200-2011 dataset, and obtain encouraging results suggesting further investigation into the benefit of privileged information in structured prediction.
Code (0)
등록된 구현이 없습니다.
Tasks
ObjectObject LocalizationStructured PredictionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Realistic PointGoal Navigation via Auxiliary Losses and Information Bottleneck
We propose a novel architecture and training paradigm for training realistic PointGoal Navigation -- navigating to a target coordinate in an unseen environment under actuation and sensor noise without access to ground-tr…
PointGoal NavigationLearning to Transfer Privileged Information
We introduce a learning framework called learning using privileged information (LUPI) to the computer vision field. We focus on the prototypical computer vision problem of teaching computers to recognize objects in image…
General ClassificationLearning with Privileged Information for Multi-Label Classification
In this paper, we propose a novel approach for learning multi-label classifiers with the help of privileged information. Specifically, we use similarity constraints to capture the relationship between available informati…
Action Unit DetectionClassificationFacial Action Unit DetectionGeneral Classification+4Multi Teacher Privileged Knowledge Distillation for Multimodal Expression Recognition
Human emotion is a complex phenomenon conveyed and perceived through facial expressions, vocal tones, body language, and physiological signals. Multimodal emotion recognition systems can perform well because they can lea…
Emotion RecognitionKnowledge DistillationMultimodal Emotion RecognitionDistilling Privileged Multimodal Information for Expression Recognition using Optimal Transport
Deep learning models for multimodal expression recognition have reached remarkable performance in controlled laboratory environments because of their ability to learn complementary and redundant semantic information. How…
DiversityKnowledge DistillationOrdinal Classification