paper-with-me

Papers

Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility

2025-05-22 · Sheng-Fu Wang, Laurent Prevot, Jou-an Chi, Ri-Sheng Huang, Shu-Kai Hsieh

The achievements of Large Language Models in Natural Language Processing, especially for high-resource languages, call for a better understanding of their characteristics from a cognitive perspective. Researchers have attempted to evaluate artificial models by testing their ability to predict behavioral (e.g., eye-tracking fixations) and physiological (e.g., brain responses) variables during language processing (e.g., reading/listening). In this paper, we propose using spontaneous speech corpora to derive production variables (speech reductions, prosodic prominences) and applying them in a similar fashion. More precisely, we extract. We then test models trained with a standard procedure on different pretraining datasets (written, spoken, and mixed genres) for their ability to predict these two variables. Our results show that, after some fine-tuning, the models can predict these production variables well above baselines. We also observe that spoken genre training data provides more accurate predictions than written genres. These results contribute to the broader effort of using high-quality speech corpora as benchmarks for LLMs.

📄 PDF Abstract BibTeX arXiv:2505.16277

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CogBench: A Large Language Model Benchmark for Multilingual Speech-Based Cognitive Impairment Assessment

2025-08-05 · Rui Feng, Zhiyao Luo, Wei Wang, Yuting Song 외 arxiv

Automatic assessment of cognitive impairment from spontaneous speech offers a promising, non-invasive avenue for early cognitive screening. However, current approaches often lack generalizability when deployed across dif…

Evaluating ASR robustness to spontaneous speech errors: A study of WhisperX using a Speech Error Database

2025-08-18 · John Alderete, Macarious Kin Fung Hui, Aanchan Mohan arxiv

The Simon Fraser University Speech Error Database (SFUSED) is a public data collection developed for linguistic and psycholinguistic research. Here we demonstrate how its design and annotations can be used to test and ev…

Speech Recognition

Early Dementia Detection Using Multiple Spontaneous Speech Prompts: The PROCESS Challenge

2024-12-05 · Fuxiang Tao, Bahman Mirheidari, Madhurananda Pahar, Sophie Young 외

Dementia is associated with various cognitive impairments and typically manifests only after significant progression, making intervention at this stage often ineffective. To address this issue, the Prediction and Recogni…

Evaluating Sampling-based Filler Insertion with Spontaneous TTS

2022-06-01 · LREC 2022 6 · Siyang Wang, Joakim Gustafson, Éva Székely

Inserting fillers (such as “um”, “like”) to clean speech text has a rich history of study. One major application is to make dialogue systems sound more spontaneous. The ambiguity of filler occurrence and inter-speaker di…

Multi-modal fusion with gating using audio, lexical and disfluency features for Alzheimer's Dementia recognition from spontaneous speech

2021-06-17 · Morteza Rohanian, Julian Hough, Matthew Purver

This paper is a submission to the Alzheimer's Dementia Recognition through Spontaneous Speech (ADReSS) challenge, which aims to develop methods that can assist in the automated prediction of severity of Alzheimer's Disea…

Prediction