paper-with-me

Papers

findsylls: A Language-Agnostic Toolkit for Syllable-Level Speech Tokenization and Embedding

2026-03-27 · Héctor Javier Vázquez Martínez arxiv

Syllable-level units offer compact and linguistically meaningful representations for spoken language modeling and unsupervised word discovery, but research on syllabification remains fragmented across disparate implementations, datasets, and evaluation protocols. We introduce findsylls, a modular, language-agnostic toolkit that unifies classical syllable detectors and end-to-end syllabifiers under a common interface for syllable segmentation, embedding extraction, and multi-granular evaluation. The toolkit implements and standardizes widely used methods (e.g., Sylber, VG-HuBERT) and allows their components to be recombined, enabling controlled comparisons of representations, algorithms, and token rates. We demonstrate findsylls on English and Spanish corpora and on new hand-annotated data from Kono, an underdocumented Central Mande language, illustrating how a single framework can support reproducible syllable-level experiments across both high-resource and under-resourced settings.

📄 PDF Abstract BibTeX arXiv:2603.26292

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Cross-Lingual Multi-Granularity Framework for Interpretable Parkinson's Disease Diagnosis from Speech

2025-10-04 · Ilias Tougui, Mehdi Zakroum, Mounir Ghogho arxiv

Parkinson's Disease (PD) affects over 10 million people worldwide, with speech impairments in up to 89% of patients. Current speech-based detection systems analyze entire utterances, potentially overlooking the diagnosti…

Syllable-level lyrics generation from melody exploiting character-level language model

2023-10-02 · Zhe Zhang, Karol Lasocki, Yi Yu, Atsuhiro Takasu

The generation of lyrics tightly connected to accompanying melodies involves establishing a mapping between musical notes and syllables of lyrics. This process requires a deep understanding of music constraints and seman…

Language ModelingLanguage ModellingSentence

Syllable based DNN-HMM Cantonese Speech to Text System

2024-02-13 · LREC 2016 5 · Timothy Wong, Claire Li, Sam Lam, Billy Chiu 외

This paper reports our work on building up a Cantonese Speech-to-Text (STT) system with a syllable based acoustic model. This is a part of an effort in building a STT system to aid dyslexic students who have cognitive de…

speech-recognitionSpeech RecognitionSpeech-to-Text

Syllabification by Phone Categorization

2018-07-15 · Jacob Krantz, Maxwell Dulin, Paul De Palma, Mark VanDam

Syllables play an important role in speech synthesis, speech recognition, and spoken document retrieval. A novel, low cost, and language agnostic approach to dividing words into their corresponding syllables is presented…

Retrievalspeech-recognitionSpeech RecognitionSpeech Synthesis

CoPaSul Manual - Contour-based parametric and superpositional intonation stylization

2016-12-14 · Uwe D. Reichel

The purposes of the CoPaSul toolkit are (1) automatic prosodic annotation and (2) prosodic feature extraction from syllable to utterance level. CoPaSul stands for contour-based, parametric, superpositional intonation sty…

Clustering