paper-with-me

Papers

Multilingual MFA: Forced Alignment on Low-Resource Related Languages

2025-04-09 · Alessio Tosolini, Claire Bowern

We compare the outcomes of multilingual and crosslingual training for related and unrelated Australian languages with similar phonological inventories. We use the Montreal Forced Aligner to train acoustic models from scratch and adapt a large English model, evaluating results against seen data, unseen data (seen language), and unseen data and language. Results indicate benefits of adapting the English baseline model for previously unseen languages.

📄 PDF Abstract BibTeX arXiv:2504.07315

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Grapheme-Based Cross-Language Forced Alignment: Results with Uralic Languages

2021-05-01 · NoDaLiDa 2021 5 · Juho Leinonen, Sami Virpioja, Mikko Kurimo

Forced alignment is an effective process to speed up linguistic research. However, most forced aligners are language-dependent, and under-resourced languages rarely have enough resources to train an acoustic model for an…

Phoneme- and Word-Level Metrics Using Self-Supervised Speech Representations for Forced Alignment Evaluation

2026-08-28 · V. S. D. S. Mahesh Akavarapu, Michael Daniel, Gerhard Jäger arxiv

Forced alignment evaluation typically requires manually annotated timestamps, limiting large-scale and multilingual analysis. We introduce two corpus-level metrics based on self-supervised (SSL) speech representations fo…

Multilingual Word-Level Forced Alignment with Self-Supervised Representations and Learned Dynamic Programming

2026-06-09 · Roy Weber, Meidan Zehavi, Rotem Rousso, Joseph Keshet arxiv

We present a method for accurate multilingual word-level forced alignment, consisting of an alignment encoder and a learned alignment decoder. The encoder integrates two representations: one from the Massively Multilingu…

Semiautomatic Speech Alignment for Under-Resourced Languages

2022-06-01 · EURALI (LREC) 2022 6 · Juho Leinonen, Niko Partanen, Sami Virpioja, Mikko Kurimo

Cross-language forced alignment is a solution for linguists who create speech corpora for very low-resource languages. However, cross-language is an additional challenge making a complex task, forced alignment, even more…

The taste of IPA: Towards open-vocabulary keyword spotting and forced alignment in any language

2023-11-14 · Jian Zhu, Changbing Yang, Farhan Samir, Jahurul Islam

In this project, we demonstrate that phoneme-based models for speech processing can achieve strong crosslinguistic generalizability to unseen languages. We curated the IPAPACK, a massively multilingual speech corpora wit…

Keyword Spotting