paper-with-me

Papers

The Slovene BNSI Broadcast News database and reference speech corpus GOS: Towards the uniform guidelines for future work

2014-05-01 · LREC 2014 5 · Andrej {\v{Z}}gank, Ana Zwitter Vitez, Darinka Verdonik

The aim of the paper is to search for common guidelines for the future development of speech databases for less resourced languages in order to make them the most useful for both main fields of their use, linguistic research and speech technologies. We compare two standards for creating speech databases, one followed when developing the Slovene speech database for automatic speech recognition ― BNSI Broadcast News, the other followed when developing the Slovene reference speech corpus GOS, and outline possible common guidelines for future work. We also present an add-on for the GOS corpus, which enables its usage for automatic speech recognition.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

TUKE-BNews-SK: Slovak Broadcast News Corpus Construction and Evaluation

2014-05-01 · LREC 2014 5 · Mat{\'u}{\v{s}} Pleva, Jozef Juh{\'a}r

This article presents an overview of the existing acoustical corpuses suitable for broadcast news automatic transcription task in the Slovak language. The TUKE-BNews-SK database created in our department was built to sup…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

New bilingual speech databases for audio diarization

2014-05-01 · LREC 2014 5 · David Tavarez, Eva Navas, Daniel Erro, Ibon Saratxaga 외

This paper describes the process of collecting and recording two new bilingual speech databases in Spanish and Basque. They are designed primarily for speaker diarization in two different application domains: broadcast n…

speaker-diarizationSpeaker DiarizationSpeaker Recognition

Multi-task Learning for Cross-Lingual Sentiment Analysis

2022-12-14 · Gaurish Thakkar, Nives Mikelic Preradovic, Marko Tadic

This paper presents a cross-lingual sentiment analysis of news articles using zero-shot and few-shot learning. The study aims to classify the Croatian news articles with positive, negative, and neutral sentiments using t…

ArticlesFew-Shot LearningMulti-Task LearningSentiment Analysis+1

A Computational Analysis of the Dehumanisation of Migrants from Syria and Ukraine in Slovene News Media

2024-04-10 · Jaya Caporusso, Damar Hoogland, Mojca Brglez, Boshko Koloski 외

Dehumanisation involves the perception and or treatment of a social group's members as less than human. This phenomenon is rarely addressed with computational linguistic techniques. We adapt a recently proposed approach …

Gigafida 2.0: The Reference Corpus of Written Standard Slovene

2020-05-01 · LREC 2020 5 · Simon Krek, {\v{S}}pela Arhar Holdt, Toma{\v{z}} Erjavec, Jaka {\v{C}}ibej 외

We describe a new version of the Gigafida reference corpus of Slovene. In addition to updating the corpus with new material and annotating it with better tools, the focus of the upgrade was also on its transformation fro…