paper-with-me

홈 › Papers

Language-agnostic, automated assessment of listeners' speech recall using large language models

2025-03-02 · Björn Herrmann

Speech-comprehension difficulties are common among older people. Standard speech tests do not fully capture such difficulties because the tests poorly resemble the context-rich, story-like nature of ongoing conversation and are typically available only in a country's dominant/official language (e.g., English), leading to inaccurate scores for native speakers of other languages. Assessments for naturalistic, story speech in multiple languages require accurate, time-efficient scoring. The current research leverages modern large language models (LLMs) in native English speakers and native speakers of 10 other languages to automate the generation of high-quality, spoken stories and scoring of speech recall in different languages. Participants listened to and freely recalled short stories (in quiet/clear and in babble noise) in their native language. LLM text-embeddings and LLM prompt engineering with semantic similarity analyses to score speech recall revealed sensitivity to known effects of temporal order, primacy/recency, and background noise, and high similarity of recall scores across languages. The work overcomes limitations associated with simple speech materials and testing of closed native-speaker groups because recall data of varying length and details can be mapped across languages with high accuracy. The full automation of speech generation and recall scoring provides an important step towards comprehension assessments of naturalistic speech with clinical applicability.

📄 PDF Abstract BibTeX arXiv:2503.01045

Code (0)

등록된 구현이 없습니다.

Tasks

Prompt EngineeringSemantic SimilaritySemantic Textual Similarity

Similar Papers 제목 키워드 기반

Advancing Hearing Assessment: An ASR-Based Frequency-Specific Speech Test for Diagnosing Presbycusis

2025-05-28 · Stefan Bleeck

Traditional audiometry often fails to fully characterize the functional impact of hearing loss on speech understanding, particularly supra-threshold deficits and frequency-specific perception challenges in conditions lik…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diagnosticspeech-recognition+1

Perceptual Implications of Automatic Anonymization in Pathological Speech

2025-05-01 · Soroosh Tayebi Arasteh, Saba Afza, Tri-Thien Nguyen, Lukas Buess 외

Automatic anonymization techniques are essential for ethical sharing of pathological speech data, yet their perceptual consequences remain understudied. This study presents the first comprehensive human-centered analysis…

Diagnostic

Covertly improving intelligibility with data-driven adaptations of speech timing

2026-03-31 · Paige Tuttösí, Angelica Lim, H. Henny Yeung, Yue Wang 외 arxiv

Human talkers often address listeners with language-comprehension challenges, such as hard-of-hearing or non-native adults, by globally slowing down their speech. However, it remains unclear whether this strategy actuall…

CCATMos: Convolutional Context-aware Transformer Network for Non-intrusive Speech Quality Assessment

2022-11-04 · Yuchen Liu, Li-Chia Yang, Alex Pawlicki, Marko Stamenovic

Speech quality assessment has been a critical component in many voice communication related applications such as telephony and online conferencing. Traditional intrusive speech quality assessment requires the clean refer…

HASA-net: A non-intrusive hearing-aid speech assessment network

2021-11-10 · Hsin-Tien Chiang, Yi-Chiao Wu, Cheng Yu, Tomoki Toda 외

Without the need of a clean reference, non-intrusive speech assessment methods have caught great attention for objective evaluations. Recently, deep neural network (DNN) models have been applied to build non-intrusive sp…