paper-with-me

Papers

Multi-view metric learning for multi-instance image classification

2016-10-21 · Dewei Li, Yingjie Tian

It is critical and meaningful to make image classification since it can help human in image retrieval and recognition, object detection, etc. In this paper, three-sides efforts are made to accomplish the task. First, visual features with bag-of-words representation, not single vector, are extracted to characterize the image. To improve the performance, the idea of multi-view learning is implemented and three kinds of features are provided, each one corresponds to a single view. The information from three views is complementary to each other, which can be unified together. Then a new distance function is designed for bags by computing the weighted sum of the distances between instances. The technique of metric learning is explored to construct a data-dependent distance metric to measure the relationships between instances, meanwhile between bags and images, more accurately. Last, a novel approach, called MVML, is proposed, which optimizes the joint probability that every image is similar with its nearest image. MVML learns multiple distance metrics, each one models a single view, to unifies the information from multiple views. The method can be solved by alternate optimization iteratively. Gradient ascent and positive semi-definite projection are utilized in the iterations. Distance comparisons verified that the new bag distance function is prior to previous functions. In model evaluation, numerical experiments show that MVML with multiple views performs better than single view condition, which demonstrates that our model can assemble the complementary information efficiently and measure the distance between images more precisely. Experiments on influence of parameters and instance number validate the consistency of the method.

📄 PDF Abstract BibTeX arXiv:1610.06671

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classificationimage-classificationImage ClassificationImage RetrievalMetric LearningMULTI-VIEW LEARNINGobject-detectionObject DetectionRetrieval

Similar Papers 제목 키워드 기반

Simultaneous multi-view instance detection with learned geometric soft-constraints

2019-07-25 · ICCV 2019 10 · Ahmed Samy Nassar, Sebastien Lefevre, Jan D. Wegner

We propose to jointly learn multi-view geometry and warping between views of the same object instances for robust cross-view object detection. What makes multi-view object instance detection difficult are strong changes …

multi-view detectionObjectobject-detectionObject Detection

Aerial Lifting: Neural Urban Semantic and Building Instance Lifting from Aerial Imagery

2024-03-18 · CVPR 2024 1 · Yuqi Zhang, GuanYing Chen, Jiaxing Chen, Shuguang Cui

We present a neural radiance field method for urban-scale semantic and building-level instance segmentation from aerial images by lifting noisy 2D labels to 3D. This is a challenging problem due to two primary reasons. F…

Instance SegmentationNeRFNovel View SynthesisSegmentation+1

Recognizing Objects From Any View With Object and Viewer-Centered Representations

2020-06-01 · CVPR 2020 6 · Sainan Liu, Vincent Nguyen, Isaac Rehg, Zhuowen Tu

In this paper, we tackle an important task in computer vision: any view object recognition. In both training and testing, for each object instance, we are only given its 2D image viewed from an unknown angle. We propose …

ObjectObject Recognition

SegVGGT: Joint 3D Reconstruction and Instance Segmentation from Multi-View Images

2026-03-20 · Jinyuan Qu, Hongyang Li, Lei Zhang arxiv

3D instance segmentation methods typically rely on high-quality point clouds or posed RGB-D scans, requiring complex multi-stage processing pipelines, and are highly sensitive to reconstruction noise. While recent feed-f…

Multi-View 3D Reconstruction3D Instance SegmentationPoint Clouds

SiM3D: Single-instance Multiview Multimodal and Multisetup 3D Anomaly Detection Benchmark

2025-06-26 · Alex Costanzino, Pierluigi Zama Ramirez, Luigi Lella, Matteo Ragaglia 외

We propose SiM3D, the first benchmark considering the integration of multiview and multimodal information for comprehensive 3D anomaly detection and segmentation (ADS), where the task is to produce a voxel-based Anomaly …

3D Anomaly Detection3D Anomaly Detection and SegmentationAnomaly Detection