paper-with-me

Papers

A Web Tool for Building Parallel Corpora of Spoken and Sign Languages

2016-05-01 · LREC 2016 5 · Alex Becker, Fabio Kepler, C, Sara eias

In this paper we describe our work in building an online tool for manually annotating texts in any spoken language with SignWriting in any sign language. The existence of such tool will allow the creation of parallel corpora between spoken and sign languages that can be used to bootstrap the creation of efficient tools for the Deaf community. As an example, a parallel corpus between English and American Sign Language could be used for training Machine Learning models for automatic translation between the two languages. Clearly, this kind of tool must be designed in a way that it eases the task of human annotators, not only by being easy to use, but also by giving smart suggestions as the annotation progresses, in order to save time and effort. By building a collaborative, online, easy to use annotation tool for building parallel corpora between spoken and sign languages we aim at helping the development of proper resources for sign languages that can then be used in state-of-the-art models currently used in tools for spoken languages. There are several issues and difficulties in creating this kind of resource, and our presented tool already deals with some of them, like adequate text representation of a sign and many to many alignments between words and signs.

📄 PDF Abstract BibTeX

Code (1)

https://bitbucket.org/unipampa/signcorpus 공식 구현

Tasks

Translation

Similar Papers 제목 키워드 기반

Parallel Corpus for Japanese Spoken-to-Written Style Conversion

2020-05-01 · LREC 2020 5 · Mana Ihori, Akihiko Takashima, Ryo Masumura

With the increase of automatic speech recognition (ASR) applications, spoken-to-written style conversion that transforms spoken-style text into written-style text is becoming an important technology to increase the reada…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Punctuation Restorationspeech-recognition+1

Machine Translation for Nko: Tools, Corpora and Baseline Results

2023-10-24 · Moussa Koulako Bala Doumbouya, Baba Mamadi Diané, Solo Farabado Cissé, Djibrila Diané 외

Currently, there is no usable machine translation system for Nko, a language spoken by tens of millions of people across multiple West African countries, which holds significant cultural and educational value. To address…

Machine TranslationTranslation

Multi-source morphosyntactic tagging for spoken Rusyn

2017-04-01 · WS 2017 4 · Yves Scherrer, Achim Rabus

This paper deals with the development of morphosyntactic taggers for spoken varieties of the Slavic minority language Rusyn. As neither annotated corpora nor parallel corpora are electronically available for Rusyn, we pr…

Morphological TaggingPart-Of-Speech Tagging

Resource Evaluation for Usable Speech Interfaces: Utilizing Human-Human Dialogue

2012-05-01 · LREC 2012 5 · Pepi Stavropoulou, Dimitris Spiliotopoulos, Georgios Kouroupetroglou

Human-human spoken dialogues are considered an important tool for effective speech interface design and are often used for stochastic model training in speech based applications. However, the less restricted nature of hu…

Dialogue ManagementLanguage ModellingSentencespeech-recognition+3

Building Subject-aligned Comparable Corpora and Mining it for Truly Parallel Sentence Pairs

2015-09-29 · Krzysztof Wołk, Krzysztof Marasek

Parallel sentences are a relatively scarce but extremely useful resource for many applications including cross-lingual retrieval and statistical machine translation. This research explores our methodology for mining such…

ArticlesMachine TranslationRetrievalSentence+1