paper-with-me

Papers

Fluent Alignment with Disfluent Judges: Post-training for Lower-resource Languages

2025-12-09 · David Samuel, Lilja Øvrelid, Erik Velldal, Andrey Kutuzov arxiv

We propose a post-training method for lower-resource languages that preserves the fluency of language models even when aligned by disfluent reward models. Preference optimization is now a well-researched topic, but previous work has mostly addressed models for English and Chinese. Lower-resource languages lack both datasets written by native speakers and instruction-tuned language models capable of generating fluent synthetic data. To address this, we focus on developing a fluent preference-aligned language model without any instruction-tuning data in the target language. Our approach uses an on-policy training method, which we compare with two common alternatives: supervised finetuning on machine-translated data and multilingual finetuning. We conduct a case study on Norwegian Bokmål and evaluate fluency through native-speaker assessments. The results show that the on-policy aspect is crucial and outperforms the alternatives without relying on any hard-to-obtain data.

📄 PDF Abstract BibTeX arXiv:2512.08777

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Generating Fluent Translations from Disfluent Text Without Access to Fluent References: IIT Bombay@IWSLT2020

2020-07-01 · WS 2020 7 · Nikhil Saini, Jyotsana Khatri, Preethi Jyothi, Pushpak Bhattacharyya

Machine translation systems perform reasonably well when the input is well-formed speech or text. Conversational speech is spontaneous and inherently consists of many disfluencies. Producing fluent translations of disflu…

DenoisingMachine TranslationTranslation

Planning and Generating Natural and Diverse Disfluent Texts as Augmentation for Disfluency Detection

2020-11-01 · EMNLP 2020 11 · Jingfeng Yang, Diyi Yang, Zhaoran Ma

Existing approaches to disfluency detection heavily depend on human-annotated data. Numbers of data augmentation methods have been proposed to alleviate the dependence on labeled data. However, current augmentation appro…

Data Augmentation

Augmenting Automatic Speech Recognition Models with Disfluency Detection

2024-09-16 · Robin Amann, Zhaolin Li, Barbara Bruno, Jan Niehues

Speech disfluency commonly occurs in conversational and spontaneous speech. However, standard Automatic Speech Recognition (ASR) models struggle to accurately recognize these disfluencies because they are typically train…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Zero-shot Disfluency Detection for Indian Languages

2022-10-01 · COLING 2022 10 · Rohit Kundu, Preethi Jyothi, Pushpak Bhattacharyya

Disfluencies that appear in the transcriptions from automatic speech recognition systems tend to impair the performance of downstream NLP tasks. Disfluency correction models can help alleviate this problem. However, the …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Inclusive ASR for Disfluent Speech: Cascaded Large-Scale Self-Supervised Learning with Targeted Fine-Tuning and Data Augmentation

2024-06-14 · Dena Mujtaba, Nihar R. Mahapatra, Megan Arney, J. Scott Yaruss 외

Automatic speech recognition (ASR) systems often falter while processing stuttering-related disfluencies -- such as involuntary blocks and word repetitions -- yielding inaccurate transcripts. A critical barrier to progre…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationSelf-Supervised Learning+2