paper-with-me

홈 › Papers

From Babble to Words: Pre-Training Language Models on Continuous Streams of Phonemes

2024-10-30 · Zébulon Goriely, Richard Diehl Martinez, Andrew Caines, Lisa Beinborn, Paula Buttery

Language models are typically trained on large corpora of text in their default orthographic form. However, this is not the only option; representing data as streams of phonemes can offer unique advantages, from deeper insights into phonological language acquisition to improved performance on sound-based tasks. The challenge lies in evaluating the impact of phoneme-based training, as most benchmarks are also orthographic. To address this, we develop a pipeline to convert text datasets into a continuous stream of phonemes. We apply this pipeline to the 100-million-word pre-training dataset from the BabyLM challenge, as well as to standard language and grammatical benchmarks, enabling us to pre-train and evaluate a model using phonemic input representations. Our results show that while phoneme-based training slightly reduces performance on traditional language understanding tasks, it offers valuable analytical and practical benefits.

📄 PDF Abstract BibTeX arXiv:2410.22906

Code (1)

codebyzeb/Corpus-Phonemizer 공식 구현

Tasks

Language Acquisition

Similar Papers 제목 키워드 기반

Nonnegative HMM for Babble Noise Derived from Speech HMM: Application to Speech Enhancement

2017-09-16 · Nasser Mohammadiha, Arne Leijon

Deriving a good model for multitalker babble noise can facilitate different speech processing algorithms, e.g. noise reduction, to reduce the so-called cocktail party difficulty. In the available systems, the fact that t…

Speech Enhancement

Stream State-tying for Sign Language Recognition

2024-04-21 · Jiyong Ma, Wen Gao, Chunli Wang

In this paper, a novel approach to sign language recognition based on state tying in each of data streams is presented. In this framework, it is assumed that hand gesture signal is represented in terms of six synchronous…

Gesture RecognitionHand Gesture RecognitionHand-Gesture RecognitionPosition+2

Ambient Search: A Document Retrieval System for Speech Streams

2016-12-01 · COLING 2016 12 · Benjamin Milde, Jonas Wacker, Stefan Radomski, Max M{\"u}hlh{\"a}user 외

We present Ambient Search, an open source system for displaying and retrieving relevant documents in real time for speech input. The system works ambiently, that is, it unobstructively listens to speech streams in the ba…

ArticlesInformation RetrievalKeyphrase GenerationRetrieval+1

StockBabble: A Conversational Financial Agent to support Stock Market Investors

2021-06-15 · Suraj Sharma, Joseph Brennan, Jason R. C. Nurse

We introduce StockBabble, a conversational agent designed to support understanding and engagement with the stock market. StockBabble's value and novelty is in its ability to empower retail investors -- many of which may …

A Data-Driven Investigation of Noise-Adaptive Utterance Generation with Linguistic Modification

2022-10-19 · Anupama Chingacham, Vera Demberg, Dietrich Klakow

In noisy environments, speech can be hard to understand for humans. Spoken dialog systems can help to enhance the intelligibility of their output, either by modifying the speech synthesis (e.g., imitate Lombard speech) o…

Speech SynthesisText Generation