Low-Resource Contextual Topic Identification on Speech
In topic identification (topic ID) on real-world unstructured audio, an audio instance of variable topic shifts is first broken into sequential segments, and each segment is independently classified. We first present a general purpose method for topic ID on spoken segments in low-resource languages, using a cascade of universal acoustic modeling, translation lexicons to English, and English-language topic classification. Next, instead of classifying each segment independently, we demonstrate that exploring the contextual dependencies across sequential segments can provide large improvements. In particular, we propose an attention-based contextual model which is able to leverage the contexts in a selective manner. We test both our contextual and non-contextual models on four LORELEI languages, and on all but one our attention-based contextual model significantly outperforms the context-independent models.
Code (0)
등록된 구현이 없습니다.
Tasks
General ClassificationTopic ClassificationTranslationSimilar Papers 제목 키워드 기반
Topic Identification for Speech without ASR
Modern topic identification (topic ID) systems for speech use automatic speech recognition (ASR) to produce speech transcripts, and perform supervised classification on such ASR outputs. However, under resource-limited c…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)General ClassificationMulti-Label Classification+3Topic Identification For Spontaneous Speech: Enriching Audio Features With Embedded Linguistic Information
Traditional topic identification solutions from audio rely on an automatic speech recognition system (ASR) to produce transcripts used as input to a text-based model. These approaches work well in high-resource scenarios…
Automatic Speech Recognitionspeech-recognitionSpeech RecognitionCross-lingual topic prediction for speech using translations
Given a large amount of unannotated speech in a low-resource language, can we classify the speech utterances by topic? We consider this question in the setting where a small amount of speech in the low-resource language …
HumanitarianPredictionSpeech-to-TextSpeech-to-Text Translation+1Assessing the impact of contextual information in hate speech detection
In recent years, hate speech has gained great relevance in social networks and other virtual media because of its intensity and its relationship with violent acts against members of protected groups. Due to the great amo…
Hate Speech DetectionAutomatic Speech Recognition and Topic Identification for Almost-Zero-Resource Languages
Automatic speech recognition (ASR) systems often need to be developed for extremely low-resource languages to serve end-uses such as audio content categorization and search. While universal phone recognition is natural t…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Humanitarianspeech-recognition+1