paper-with-me

홈 › Papers

Sounds Wilde. Phonetically Extended Embeddings for Author-Stylized Poetry Generation

2018-10-01 · WS 2018 10 · Aleksey Tikhonov, Ivan Yamshchikov

This paper addresses author-stylized text generation. Using a version of a language model with extended phonetic and semantic embeddings for poetry generation we show that phonetics has comparable contribution to the overall model performance as the information on the target author. Phonetic information is shown to be important for English and Russian language. Humans tend to attribute machine generated texts to the target author.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeLanguage ModelingLanguage ModellingText GenerationWord Embeddings

Similar Papers 제목 키워드 기반

Beyond Orthography: Automatic Recovery of Short Vowels and Dialectal Sounds in Arabic

2024-08-05 · Yassine El Kheir, Hamdy Mubarak, Ahmed Ali, Shammur Absar Chowdhury

This paper presents a novel Dialectal Sound and Vowelization Recovery framework, designed to recognize borrowed and dialectal sounds within phonologically diverse and dialect-rich languages, that extends beyond its stand…

Designing the Latvian Speech Recognition Corpus

2014-05-01 · LREC 2014 5 · M{\=a}rcis Pinnis, Ilze Auzi{\c{n}}a, K{\=a}rlis Goba

In this paper the authors present the first Latvian speech corpus designed specifically for speech recognition purposes. The paper outlines the decisions made in the corpus designing process through analysis of related w…

speech-recognitionSpeech RecognitionSpeech Synthesis

An Information-Theoretic Analysis of Self-supervised Discrete Representations of Speech

2023-06-04 · Badr M. Abdullah, Mohammed Maqsood Shaik, Bernd Möbius, Dietrich Klakow

Self-supervised representation learning for speech often involves a quantization step that transforms the acoustic input into discrete units. However, it remains unclear how to characterize the relationship between these…

QuantizationRepresentation Learning

BiPhone: Modeling Inter Language Phonetic Influences in Text

2023-07-06 · Abhirut Gupta, Ananya B. Sai, Richard Sproat, Yuri Vasilevski 외

A large number of people are forced to use the Web in a language they have low literacy in due to technology asymmetries. Written text in the second language (L2) from such users often contains a large number of errors t…

Learning acoustic word embeddings with phonetically associated triplet network

2018-11-07 · Hyungjun Lim, Younggwan Kim, Youngmoon Jung, Myunghun Jung 외

Previous researches on acoustic word embeddings used in query-by-example spoken term detection have shown remarkable performance improvements when using a triplet network. However, the triplet network is trained using on…

TripletWord Embeddings