paper-with-me

Papers

A Processing Platform Relating Data and Tools for Romanian Language

2020-05-01 · LREC 2020 5 · Vasile P{\u{a}}i{\textcommabelow{s}}, Radu Ion, Dan Tufi{\textcommabelow{s}}

This paper presents RELATE (http://relate.racai.ro), a high-performance natural language platform designed for Romanian language. It is meant both for demonstration of available services, from text-span annotations to syntactic dependency trees as well as playing or automatically synthesizing Romanian words, and for the development of new annotated corpora. It also incorporates the search engines for the large COROLA reference corpus of contemporary Romanian and the Romanian wordnet. It integrates multiple text and speech processing modules and exposes their functionality through a web interface designed for the linguist researcher. It makes use of a scheduler-runner architecture, allowing processing to be distributed across multiple computing nodes. A series of input/output converters allows large corpora to be loaded, processed and exported according to user preferences.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RELATE: A Modern Processing Platform for Romanian Language

2024-10-29 · Vasile Păiş, Radu Ion, Andrei-Marius Avram, Maria Mitrofan 외

This paper presents the design and evolution of the RELATE platform. It provides a high-performance environment for natural language processing activities, specially constructed for Romanian language. Initially developed…

Challenges in Creating a Representative Corpus of Romanian Micro-Blogging Text

2022-06-01 · CMLC (LREC) 2022 6 · Vasile Pais, Maria Mitrofan, Verginica Barbu Mititelu, Elena Irimia 외

Following the successful creation of a national representative corpus of contemporary Romanian language, we turned our attention to the social media text, as present in micro-blogging platforms. In this paper, we present…

Tools and resources for Romanian text-to-speech and speech-to-text applications

2018-02-15 · Tiberiu Boros, Stefan Daniel Dumitrescu, Vasile Pais

In this paper we introduce a set of resources and tools aimed at providing support for natural language processing, text-to-speech synthesis and speech recognition for Romanian. While the tools are general purpose and ca…

speech-recognitionSpeech RecognitionSpeech SynthesisSpeech-to-Text+3

Clustering Word Embeddings with Self-Organizing Maps. Application on LaRoSeDa - A Large Romanian Sentiment Data Set

2021-04-01 · EACL 2021 2 · Anca Tache, Gaman Mihaela, Radu Tudor Ionescu

Romanian is one of the understudied languages in computational linguistics, with few resources available for the development of natural language processing tools. In this paper, we introduce LaRoSeDa, a Large Romanian Se…

ClusteringSentiment AnalysisSentiment ClassificationText Categorization+1

Clustering Word Embeddings with Self-Organizing Maps. Application on LaRoSeDa -- A Large Romanian Sentiment Data Set

2021-01-11 · Anca Maria Tache, Mihaela Gaman, Radu Tudor Ionescu

Romanian is one of the understudied languages in computational linguistics, with few resources available for the development of natural language processing tools. In this paper, we introduce LaRoSeDa, a Large Romanian Se…

ClusteringSentiment AnalysisSentiment ClassificationText Categorization+1