paper-with-me

홈 › Papers

RoCoISLR: A Romanian Corpus for Isolated Sign Language Recognition

2025-11-16 · Cătălin-Alexandru Rîpanu, Andrei-Theodor Hotnog, Giulia-Stefania Imbrea, Dumitru-Clementin Cercel arxiv

Automatic sign language recognition plays a crucial role in bridging the communication gap between deaf communities and hearing individuals; however, most available datasets focus on American Sign Language. For Romanian Isolated Sign Language Recognition (RoISLR), no large-scale, standardized dataset exists, which limits research progress. In this work, we introduce a new corpus for RoISLR, named RoCoISLR, comprising over 9,000 video samples that span nearly 6,000 standardized glosses from multiple sources. We establish benchmark results by evaluating seven state-of-the-art video recognition models-I3D, SlowFast, Swin Transformer, TimeSformer, Uniformer, VideoMAE, and PoseConv3D-under consistent experimental setups, and compare their performance with that of the widely used WLASL2000 corpus. According to the results, transformer-based architectures outperform convolutional baselines; Swin Transformer achieved a Top-1 accuracy of 34.1%. Our benchmarks highlight the challenges associated with long-tail class distributions in low-resource sign languages, and RoCoISLR provides the initial foundation for systematic RoISLR research.

📄 PDF Abstract BibTeX arXiv:2511.12767

Code (0)

등록된 구현이 없습니다.

Tasks

Sign Language Recognition

Similar Papers 제목 키워드 기반

Adapting the TTL Romanian POS Tagger to the Biomedical Domain

2017-09-01 · RANLP 2017 9 · Maria Mitrofan, Radu Ion

This paper presents the adaptation of the Hidden Markov Models-based TTL part-of-speech tagger to the biomedical domain. TTL is a text processing platform that performs sentence splitting, tokenization, POS tagging, chun…

ChunkingDomain AdaptationLemmatizationnamed-entity-recognition+8

RSC: A Romanian Read Speech Corpus for Automatic Speech Recognition

2020-05-01 · LREC 2020 5 · Alex Georgescu, ru-Lucian, Horia Cucu, Andi Buzo 외

Although many efforts have been made in the last decade to enhance the speech and language resources for Romanian, this language is still considered under-resourced. While for many other languages there are large speech …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Tools for Building a Corpus to Study the Historical and Geographical Variation of the Romanian Language

2017-09-01 · RANLP 2017 9 · Victoria Bobicev, C{\u{a}}t{\u{a}}lina M{\u{a}}r{\u{a}}nduc, Cenel Augusto Perez

Contemporary standard language corpora are ideal for NLP. There are few morphologically and syntactically annotated corpora for Romanian, and those existing or in progress only deal with the Contemporary Romanian standar…

Romanian TimeBank: An Annotated Parallel Corpus for Temporal Information

2012-05-01 · LREC 2012 5 · Corina For{\u{a}}scu, Dan Tufi{\c{s}}

The paper describes the main steps for the construction, annotation and validation of the Romanian version of the TimeBank corpus. Starting from the English TimeBank corpus ― the reference annotated corpus in the tempo…

Information RetrievalMachine TranslationQuestion AnsweringTAG

Introducing RONEC -- the Romanian Named Entity Corpus

2019-09-03 · Stefan Daniel Dumitrescu, Andrei-Marius Avram

We present RONEC - the Named Entity Corpus for the Romanian language. The corpus contains over 26000 entities in ~5000 annotated sentences, belonging to 16 distinct classes. The sentences have been extracted from a copy-…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)