ORTOLANG an infrastructure for sharing of written and speech language resources (ORTOLANG : une infrastructure de mutualisation de ressources linguistiques \'ecrites et orales) [in French]
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Mediapi-RGB: Enabling Technological Breakthroughs in French Sign Language (LSF) Research through an Extensive Video-Text Corpus
We introduce Mediapi-RGB, a new dataset of French Sign Language (LSF) along with the first LSF-to-French machine translation model. With 86 hours of video, it the largest LSF corpora with translation. The corpus consists…
Machine TranslationSign Language TranslationTranslationUWSpeech: Speech to Speech Translation for Unwritten Languages
Existing speech to speech translation systems heavily rely on the text of target language: they usually translate source language either to target text and then synthesize target speech from text, or directly to target s…
speech-recognitionSpeech RecognitionSpeech-to-Speech TranslationTranslationTagging a Norwegian Dialect Corpus
This paper describes an evaluation of five data-driven part-of-speech (PoS) taggers for spoken Norwegian. The taggers all rely on different machine learning mechanisms: decision trees, hidden Markov models (HMMs), condit…
POSCrowdsourcing Dialect Characterization through Twitter
We perform a large-scale analysis of language diatopic variation using geotagged microblogging datasets. By collecting all Twitter messages written in Spanish over more than two years, we build a corpus from which a care…
On the Role of Style in Parsing Speech with Neural Models
The differences in written text and conversational speech are substantial; previous parsers trained on treebanked text have given very poor results on spontaneous speech. For spoken language, the mismatch in style also e…