paper-with-me

Papers

Discriminative Cross-View Binary Representation Learning

2018-04-04 · Liu Liu, Hairong Qi

Learning compact representation is vital and challenging for large scale multimedia data. Cross-view/cross-modal hashing for effective binary representation learning has received significant attention with exponentially growing availability of multimedia content. Most existing cross-view hashing algorithms emphasize the similarities in individual views, which are then connected via cross-view similarities. In this work, we focus on the exploitation of the discriminative information from different views, and propose an end-to-end method to learn semantic-preserving and discriminative binary representation, dubbed Discriminative Cross-View Hashing (DCVH), in light of learning multitasking binary representation for various tasks including cross-view retrieval, image-to-image retrieval, and image annotation/tagging. The proposed DCVH has the following key components. First, it uses convolutional neural network (CNN) based nonlinear hashing functions and multilabel classification for both images and texts simultaneously. Such hashing functions achieve effective continuous relaxation during training without explicit quantization loss by using Direct Binary Embedding (DBE) layers. Second, we propose an effective view alignment via Hamming distance minimization, which is efficiently accomplished by bit-wise XOR operation. Extensive experiments on two image-text benchmark datasets demonstrate that DCVH outperforms state-of-the-art cross-view hashing algorithms as well as single-view image hashing algorithms. In addition, DCVH can provide competitive performance for image annotation/tagging.

📄 PDF Abstract BibTeX arXiv:1804.01233

Code (0)

등록된 구현이 없습니다.

Tasks

Image RetrievalQuantizationRepresentation LearningRetrieval

Similar Papers 제목 키워드 기반

Image Hashing via Cross-View Code Alignment in the Age of Foundation Models

2025-10-31 · Ilyass Moummad, Kawtar Zaher, Hervé Goëau, Alexis Joly arxiv

Efficient large-scale retrieval requires representations that are both compact and discriminative. Foundation models provide powerful visual and multimodal embeddings, but nearest neighbor search in these high-dimensiona…

DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition

2025-08-05 · Xinzhu Li, Juepeng Zheng, Yikun Chen, Xudong Mao 외 arxiv

Robust gait recognition requires highly discriminative representations, which are closely tied to input modalities. While binary silhouettes and skeletons have dominated recent literature, these 2D representations fall s…

Gait Recognition

Discriminative Anchor Learning for Efficient Multi-view Clustering

2024-09-25 · Yalan Qin, Nan Pu, Hanzhou Wu, Nicu Sebe

Multi-view clustering aims to study the complementary information across views and discover the underlying structure. For solving the relatively high computational cost for the existing approaches, works based on anchor …

Clusteringgraph construction

Deep Sketch-Shape Hashing With Segmented 3D Stochastic Viewing

2019-06-01 · CVPR 2019 6 · Jiaxin Chen, Jie Qin, Li Liu, Fan Zhu 외

Sketch-based 3D shape retrieval has been extensively studied in recent works, most of which focus on improving the retrieval accuracy, whilst neglecting the efficiency. In this paper, we propose a novel framework for eff…

3D Shape Classification3D Shape Representation3D Shape RetrievalRetrieval

Discriminative Feature Learning with Foreground Attention for Person Re-Identification

2018-07-04 · Sanping Zhou, Jinjun Wang, Deyu Meng, Yudong Liang 외

The performance of person re-identification (Re-ID) has been seriously effected by the large cross-view appearance variations caused by mutual occlusions and background clutters. Hence learning a feature representation t…

DecoderMulti-Task LearningPerson Re-IdentificationTriplet