Towards a Speech Recognizer for Komi, an Endangered and Low-Resource Uralic Language
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Extracting a Semantic Database with Syntactic Relations for Finnish to Boost Resources for Endangered Uralic Languages
This paper introduces the second version of SemFi, a semantic database for Finnish with syntactic relations. The previous version of SemFi has been used in poem generation, and thus it has application area in NLG applica…
TranslationEvaluating OpenAI GPT Models for Translation of Endangered Uralic Languages: A Comparison of Reasoning and Non-Reasoning Architectures
The evaluation of Large Language Models (LLMs) for translation tasks has primarily focused on high-resource languages, leaving a significant gap in understanding their performance on low-resource and endangered languages…
Using Graph-Based Methods to Augment Online Dictionaries of Endangered Languages
Many endangered Uralic languages have multilingual machine readable dictionaries saved in an XML format. However, the dictionaries cover translations very inconsistently between language pairs, for instance, the Livonian…
Sentiment Analysis Using Aligned Word Embeddings for Uralic Languages
In this paper, we present an approach for translating word embeddings from a majority language into 4 minority languages: Erzya, Moksha, Udmurt and Komi-Zyrian. Furthermore, we align these word embeddings and present a n…
Sentiment AnalysisWord EmbeddingsEvaluating Transferability of BERT Models on Uralic Languages
Transformer-based language models such as BERT have outperformed previous models on a large number of English benchmarks, but their evaluation is often limited to English or a small number of well-resourced languages. In…
Hyperparameter OptimizationNERPOS