paper-with-me

홈 › Papers

Identification of Indian Languages using Ghost-VLAD pooling

2020-02-05 · Krishna D N, Ankita Patil, M. S. P Raj, Sai Prasad H S, Prabhu Aashish Garapati

In this work, we propose a new pooling strategy for language identification by considering Indian languages. The idea is to obtain utterance level features for any variable length audio for robust language recognition. We use the GhostVLAD approach to generate an utterance level feature vector for any variable length input audio by aggregating the local frame level features across time. The generated feature vector is shown to have very good language discriminative features and helps in getting state of the art results for language identification task. We conduct our experiments on 635Hrs of audio data for 7 Indian languages. Our method outperforms the previous state of the art x-vector [11] method by an absolute improvement of 1.88% in F1-score and achieves 98.43% F1-score on the held-out test data. We compare our system with various pooling approaches and show that GhostVLAD is the best pooling approach for this task. We also provide visualization of the utterance level embeddings generated using Ghost-VLAD pooling and show that this method creates embeddings which has very good language discriminative features.

📄 PDF Abstract BibTeX arXiv:2002.01664

Code (0)

등록된 구현이 없습니다.

Tasks

Language Identification

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

GhostVLAD for set-based face recognition

2018-10-23 · Yujie Zhong, Relja Arandjelović, Andrew Zisserman

The objective of this paper is to learn a compact representation of image sets for template-based face recognition. We make the following contributions: first, we propose a network architecture which aggregates and embed…

Face RecognitionFace Verification

Ghost-dil-NetVLAD: A Lightweight Neural Network for Visual Place Recognition

2021-12-22 · Qingyuan Gong, Yu Liu, Liqiang Zhang, Renhe Liu

Visual place recognition (VPR) is a challenging task with the unbalance between enormous computational cost and high recognition performance. Thanks to the practical feature extraction ability of the lightweight convolut…

Visual Place Recognition

End-to-end Language Identification using NetFV and NetVLAD

2018-09-09 · Jinkun Chen, Weicheng Cai, Danwei Cai, Zexin Cai 외

In this paper, we apply the NetFV and NetVLAD layers for the end-to-end language identification task. NetFV and NetVLAD layers are the differentiable implementations of the standard Fisher Vector and Vector of Locally Ag…

Language Identification

Writer Identification and Writer Retrieval Based on NetVLAD with Re-ranking

2020-12-11 · Shervin Rasoulzadeh, Bagher BabaAli

This paper addresses writer identification and writer retrieval which is considered as a challenging problem in the document analysis and recognition field. In this work, a novel pipeline is proposed for the problem at h…

Re-RankingRetrievalTripletWriter Retrieval

ILID: Native Script Language Identification for Indian Languages

2025-07-16 · Yash Ingle, Pruthwik Mishra arxiv

The language identification task is a crucial fundamental step in NLP. Often it serves as a pre-processing step for widely used NLP applications such as multilingual machine translation, information retrieval, question a…

Language IdentificationInformation RetrievalMachine TranslationText Summarization