paper-with-me

Papers

Study on the Correlation between Objective Evaluations and Subjective Speech Quality and Intelligibility

2023-07-10 · Hsin-Tien Chiang, Kuo-Hsuan Hung, Szu-Wei Fu, Heng-Cheng Kuo, Ming-Hsueh Tsai, Yu Tsao

Subjective tests are the gold standard for evaluating speech quality and intelligibility; however, they are time-consuming and expensive. Thus, objective measures that align with human perceptions are crucial. This study evaluates the correlation between commonly used objective measures and subjective speech quality and intelligibility using a Chinese speech dataset. Moreover, new objective measures are proposed that combine current objective measures using deep learning techniques to predict subjective quality and intelligibility. The proposed deep learning model reduces the amount of training data without significantly affecting prediction performance. We analyzed the deep learning model to understand how objective measures reflect subjective quality and intelligibility. We also explored the impact of including subjective speech quality ratings on speech intelligibility prediction. Our findings offer valuable insights into the relationship between objective measures and human perceptions.

📄 PDF Abstract BibTeX arXiv:2307.04517

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

How Stylistic Similarity Shapes Preferences in Dialogue Dataset with User and Third Party Evaluations

2025-07-15 · Ikumi Numaya, Shoji Moriya, Shiki Sato, Reina Akama 외 arxiv

Recent advancements in dialogue generation have broadened the scope of human-bot interactions, enabling not only contextually appropriate responses but also the analysis of human affect and sensitivity. While prior work …

Dialogue Generation

SI-FID: Only One Objective Indicator for Evaluating Stitched Images

2024-04-22 · Xinrui Zhang, Shengwei Guo, Guobing Sun

Image quality evaluation accurately is vital in developing image stitching algorithms as it directly reflects the algorithms progress. However, commonly used objective indicators always produce inconsistent and even conf…

Contrastive LearningData AugmentationImage Stitching

A Text-to-Speech Pipeline, Evaluation Methodology, and Initial Fine-Tuning Results for Child Speech Synthesis

2022-03-22 · Rishabh Jain, Mariam Yiwere, Dan Bigioi, Peter Corcoran 외

Speech synthesis has come a long way as current text-to-speech (TTS) models can now generate natural human-sounding speech. However, most of the TTS research focuses on using adult speech data and there has been very lim…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+4

Explainable AI-Powered Framework for Video-Based Skill Assessment in Cataract Surgery

2026-08-18 · Mohammad Javad Ahmadi, Hamid D. Taghirad arxiv

Persistent shortages in the surgical workforce and inherent limitations of traditional training methods highlight the necessity of automated, data-driven approaches in surgical education. This study addresses these chall…

A Study on the Relationship Between Depth Map Quality and the Overall 3D Video Quality OF Experience

2018-03-14

The emergence of multiview displays has made the need for synthesizing virtual views more pronounced, since it is not practical to capture all of the possible views when filming multiview content. View synthesis is perfo…