paper-with-me

홈 › Papers

Increasing the Accessibility of Time-Aligned Speech Corpora with Spokes Mix

2018-05-01 · LREC 2018 5 · Piotr P{\k{e}}zik
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Recognition

Similar Papers 제목 키워드 기반

Speaking at the Right Level: Literacy-Controlled Counterspeech Generation with RAG-RL

2025-09-01 · Xiaoying Song, Anirban Saha Anik, Dibakar Barua, Pengcheng Luo 외 arxiv

Health misinformation spreading online poses a significant threat to public health. Researchers have explored methods for automatically generating counterspeech to health misinformation as a mitigation strategy. Existing…

Reinforcement Learning

Who Gets Left Behind? Auditing Disability Inclusivity in Large Language Models

2025-08-31 · Deepika Dash, Yeshil Bangera, Mithil Bangera, Gouthami Vadithya 외 arxiv

Large Language Models (LLMs) are increasingly used for accessibility guidance, yet many disability groups remain underserved by their advice. To address this gap, we present taxonomy aligned benchmark1 of human validated…

TVD: A Reproducible and Multiply Aligned TV Series Dataset

2014-05-01 · LREC 2014 5 · Anindya Roy, Camille Guinaudeau, Herv{\'e} Bredin, Claude Barras

We introduce a new dataset built around two TV series from different genres, The Big Bang Theory, a situation comedy and Game of Thrones, a fantasy drama. The dataset has multiple tracks extracted from diverse sources, i…

Dynamic Time WarpingInformation RetrievalRetrievalSentiment Analysis+1

Exploring Generative Error Correction for Dysarthric Speech Recognition

2025-05-26 · Moreno La Quatra, Alkis Koudounas, Valerio Mario Salerno, Sabato Marco Siniscalchi

Despite the remarkable progress in end-to-end Automatic Speech Recognition (ASR) engines, accurately transcribing dysarthric speech remains a major challenge. In this work, we proposed a two-stage framework for the Speec…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Facial Expression-Enhanced TTS: Combining Face Representation and Emotion Intensity for Adaptive Speech

2024-09-24 · Yunji Chu, Yunseob Shim, Unsang Park

We propose FEIM-TTS, an innovative zero-shot text-to-speech (TTS) model that synthesizes emotionally expressive speech, aligned with facial images and modulated by emotion intensity. Leveraging deep learning, FEIM-TTS tr…

Emotional Speech SynthesisSpeech Synthesistext-to-speechText to Speech