paper-with-me

Papers

Analyzing analytical methods: The case of phonology in neural models of spoken language

2020-04-15 · ACL 2020 6 · Grzegorz Chrupała, Bertrand Higy, Afra Alishahi

Given the fast development of analysis techniques for NLP and speech processing systems, few systematic studies have been conducted to compare the strengths and weaknesses of each method. As a step in this direction we study the case of representations of phonology in neural network models of spoken language. We use two commonly applied analytical techniques, diagnostic classifiers and representational similarity analysis, to quantify to what extent neural activation patterns encode phonemes and phoneme sequences. We manipulate two factors that can affect the outcome of analysis. First, we investigate the role of learning by comparing neural activations extracted from trained versus randomly-initialized models. Second, we examine the temporal scope of the activations by probing both local activations corresponding to a few milliseconds of the speech signal, and global activations pooled over the whole utterance. We conclude that reporting analysis results with randomly initialized models is crucial, and that global-scope methods tend to yield more consistent results and we recommend their use as a complement to local-scope diagnostic methods.

📄 PDF Abstract BibTeX arXiv:2004.07070

Code (1)

gchrupala/analyzing-analytical-methods 공식 구현 pytorch

Tasks

Diagnostic

Similar Papers 제목 키워드 기반

Improve Bilingual TTS Using Dynamic Language and Phonology Embedding

2022-12-07 · Fengyu Yang, Jian Luan, Yujun Wang

In most cases, bilingual TTS needs to handle three types of input scripts: first language only, second language only, and second language embedded in the first language. In the latter two situations, the pronunciation an…

Encoding of lexical tone in self-supervised models of spoken language

2024-03-25 · Gaofei Shen, Michaela Watkins, Afra Alishahi, Arianna Bisazza 외

Interpretability research has shown that self-supervised Spoken Language Models (SLMs) encode a wide variety of features in human speech from the acoustic, phonetic, phonological, syntactic and semantic levels, to speake…

A phonetic model of non-native spoken word processing

2021-01-27 · EACL 2021 2 · Yevgen Matusevych, Herman Kamper, Thomas Schatz, Naomi H. Feldman 외

Non-native speakers show difficulties with spoken word processing. Many studies attribute these difficulties to imprecise phonological encoding of words in the lexical memory. We test an alternative hypothesis: that some…

Attribute

How Generative Spoken Language Modeling Encodes Noisy Speech: Investigation from Phonetics to Syntactics

2023-06-01 · Joonyong Park, Shinnosuke Takamichi, Tomohiko Nakamura, Kentaro Seki 외

We examine the speech modeling potential of generative spoken language modeling (GSLM), which involves using learned symbols derived from data rather than phonemes for speech analysis and synthesis. Since GSLM facilitate…

Language ModelingLanguage ModellingResynthesis

Mapping Phonology to Semantics: A Computational Model of Cross-Lingual Spoken-Word Recognition

2022-10-01 · VarDial (COLING) 2022 10 · Iuliia Zaitova, Badr Abdullah, Dietrich Klakow

Closely related languages are often mutually intelligible to various degrees. Therefore, speakers of closely related languages are usually capable of (partially) comprehending each other’s speech without explicitly learn…