Automatic word stress annotation of Russian unrestricted text
Code (0)
등록된 구현이 없습니다.
Tasks
TransliterationSimilar Papers 제목 키워드 기반
Balalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech
We introduce Balalaika, an open-source, data-centric pipeline for processing audio and producing prosody-aware annotations. It combines semantic VAD for context-preserving segmentation, multi-ASR ensembling with ROVER co…
Speech DenoisingRuDSI: graph-based word sense induction dataset for Russian
We present RuDSI, a new benchmark for word sense induction (WSI) in Russian. The dataset was created using manual annotation and semi-automatic clustering of Word Usage Graphs (WUGs). Unlike prior WSI datasets for Russia…
ClusteringGraph ClusteringWord Sense InductionAutomated Word Stress Detection in Russian
In this study we address the problem of automated word stress detection in Russian using character level models and no part-speech-taggers. We use a simple bidirectional RNN with LSTM nodes and achieve the accuracy of 90…
TheRuSLan: Database of Russian Sign Language
In this paper, a new Russian sign language multimedia database TheRuSLan is presented. The database includes lexical units (single words and phrases) from Russian sign language within one subject area, namely, {``}food p…
Gesture RecognitionSign Language RecognitionSense-Annotated Corpus for Russian
We present a sense-annotated corpus for Russian. The resource was obtained my manually annotating texts from the OpenCorpora corpus, an open corpus for the Russian language, by senses of Russian wordnet RuWordNet. The an…
Word Sense Disambiguation