paper-with-me

홈 › Papers

Is word-to-phone mapping better than phone-phone mapping for handling English words?

2013-08-01 · ACL 2013 8 · Naresh Kumar Elluru, An Vadapalli, aswarup, Raghavendra Elluru, Hema Murthy, Kishore Prahallad
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Synthesis

Similar Papers 제목 키워드 기반

Phoneme-Based Contextualization for Cross-Lingual Speech Recognition in End-to-End Models

2019-06-21 · Ke Hu, Antoine Bruguier, Tara N. Sainath, Rohit Prabhavalkar 외

Contextual automatic speech recognition, i.e., biasing recognition towards a given context (e.g. user's playlists, or contacts), is challenging in end-to-end (E2E) models. Such models maintain a limited number of candida…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Finding phonemes: improving machine lip-reading

2017-10-03 · Helen L. Bear, Richard W. Harvey, Yuxuan Lan

In machine lip-reading there is continued debate and research around the correct classes to be used for recognition. In this paper we use a structured approach for devising speaker-dependent viseme classes, which enables…

Lip ReadingPhoneme Recognition

Phoneme-to-viseme mappings: the good, the bad, and the ugly

2018-05-08 · Helen L. Bear, Richard Harvey

Visemes are the visual equivalent of phonemes. Although not precisely defined, a working definition of a viseme is "a set of phonemes which have identical appearance on the lips". Therefore a phoneme falls into one visem…

LEARNING PHONEME-LEVEL DISCRETE SPEECH REPRESENTATION WITH WORD-LEVEL SUPERVISION

2021-09-29 · Liming Wang, Siyuan Feng, Mark A. Hasegawa-Johnson, Chang D. Yoo

Phonemes are defined by their relationship to words: changing a phoneme changes the word. Learning a phoneme inventory with little supervision has been a long-standing challenge with important applications to under-reso…

Representation LearningSelf-Supervised Learning

PWESuite: Phonetic Word Embeddings and Tasks They Facilitate

2023-04-05 · Vilém Zouhar, Kalvin Chang, Chenxuan Cui, Nathaniel Carlson 외

Mapping words into a fixed-dimensional vector space is the backbone of modern NLP. While most word embedding methods successfully encode semantic information, they overlook phonetic information that is crucial for many t…

RetrievalWord Embeddings