GTM-UVigo Systems for the Query-by-Example Search on Speech Task at MediaEval 2015
In this paper, we present the systems developed by GTMUVigo team for the query by example search on speech task (QUESST) at MediaEval 2015. The systems consist in a fusion of 11 dynamic time warping based systems that use phoneme posteriorgrams for speech representation; the primary system introduces a technique to select the most relevant phonetic units on each phoneme decoder, leading to an improvement of the search results.
Code (1)
Tasks
DecoderDynamic Time WarpingKeyword SpottingSimilar Papers 제목 키워드 기반
LSE\_UVIGO: A Multi-source Database for Spanish Sign Language Recognition
This paper presents LSE{\_}UVIGO, a multi-source database designed to foster research on Sign Language Recognition. It is being recorded and compiled for Spanish Sign Language (LSE acronym in Spanish) and contains also s…
Sign Language RecognitionELiRF at MediaEval 2015: Query by Example Search on Speech Task (QUESST)
n this paper, we present the systems that the Natural Language Engineering and Pattern Recognition group (ELiRF) has submitted to the MediaEval 2015 Query by Example Search on Speech Task. All of them are based on a Subs…
Dynamic Time WarpingKeyword SpottingELiRF at MediaEval 2014: Query by Example Search on Speech Task (QUESST)
In this paper, we present the systems that the Natural Language Engineering and Pattern Recognition group (ELiRF) has submitted to the MediaEval 2014 Query by Example Search on Speech task. All of them are based on a Sub…
Dynamic Time WarpingKeyword SpottingSemantic query-by-example speech search using visual grounding
A number of recent studies have started to investigate how speech systems can be trained on untranscribed speech by leveraging accompanying images at training time. Examples of tasks include keyword prediction and within…
RetrievalSemantic RetrievalVisual GroundingCUNY Systems for the Query-by-Example Search on Speech Task at MediaEval 2015
This paper describes two query-by-example systems developed by Speech Lab, Queens College (CUNY). Our systems aimed to respond with quick search results from the selected reference files. Three phonetic recognizers (Czec…
Keyword Spotting