PD-Loss: Proxy-Decidability for Efficient Metric Learning
Deep Metric Learning (DML) aims to learn embedding functions that map semantically similar inputs to proximate points in a metric space while separating dissimilar ones. Existing methods, such as pairwise losses, are hindered by complex sampling requirements and slow convergence. In contrast, proxy-based losses, despite their improved scalability, often fail to optimize global distribution properties. The Decidability-based Loss (D-Loss) addresses this by targeting the decidability index (d') to enhance distribution separability, but its reliance on large mini-batches imposes significant computational constraints. We introduce Proxy-Decidability Loss (PD-Loss), a novel objective that integrates learnable proxies with the statistical framework of d' to optimize embedding spaces efficiently. By estimating genuine and impostor distributions through proxies, PD-Loss combines the computational efficiency of proxy-based methods with the principled separability of D-Loss, offering a scalable approach to distribution-aware DML. Experiments across various tasks, including fine-grained classification and face verification, demonstrate that PD-Loss achieves performance comparable to that of state-of-the-art methods while introducing a new perspective on embedding optimization, with potential for broader applications.
Code (0)
등록된 구현이 없습니다.
Tasks
Computational EfficiencyFace VerificationMetric LearningSimilar Papers 제목 키워드 기반
A Decidability-Based Loss Function
Nowadays, deep learning is the standard approach for a wide range of problems, including biometrics, such as face recognition and speech recognition, etc. Biometric problems often use deep learning models to extract feat…
Face Recognitionspeech-recognitionSpeech RecognitionTripletAsymmetric Proxy Loss for Multi-View Acoustic Word Embeddings
Acoustic word embeddings (AWEs) are discriminative representations of speech segments, and learned embedding space reflects the phonetic similarity between words. With multi-view learning, where text labels are considere…
Metric LearningMULTI-VIEW LEARNINGTripletWord EmbeddingsMasked Proxy Loss For Text-Independent Speaker Verification
Open-set speaker recognition can be regarded as a metric learning problem, which is to maximize inter-class variance and minimize intra-class variance. Supervised metric learning can be categorized into entity-based lear…
Metric LearningSpeaker RecognitionSpeaker VerificationText-Independent Speaker Recognition+2Hierarchical Proxy-based Loss for Deep Metric Learning
Proxy-based metric learning losses are superior to pair-based losses due to their fast convergence and low training complexity. However, existing proxy-based losses focus on learning class-discriminative features while o…
Image RetrievalMetric LearningRetrievalTowards Improved Proxy-based Deep Metric Learning via Data-Augmented Domain Adaptation
Deep Metric Learning (DML) plays an important role in modern computer vision research, where we learn a distance metric for a set of image representations. Recent DML techniques utilize the proxy to interact with the cor…
Domain AdaptationMetric LearningRetrieval