paper-with-me

홈 › Papers

Disfluency Detection with a Semi-Markov Model and Prosodic Features

2015-05-01 · HLT 2015 5 · Dan Klein, Greg Durrett, James Ferguson
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Giving Attention to the Unexpected: Using Prosody Innovations in Disfluency Detection

2019-04-08 · NAACL 2019 6 · Vicky Zayats, Mari Ostendorf

Disfluencies in spontaneous speech are known to be associated with prosodic disruptions. However, most algorithms for disfluency detection use only word transcripts. Integrating prosodic cues has proved difficult because…

Integrating Disfluency-based and Prosodic Features with Acoustics in Automatic Fluency Evaluation of Spontaneous Speech

2020-05-01 · LREC 2020 5 · Huaijin Deng, Youchao Lin, Takehito Utsuro, Akio Kobayashi 외

This paper describes an automatic fluency evaluation of spontaneous speech. In the task of automatic fluency evaluation, we integrate diverse features of acoustics, prosody, and disfluency-based ones. Then, we attempt to…

Identification of primary and collateral tracks in stuttered speech

2020-03-02 · LREC 2020 5 · Rachid Riad, Anne-Catherine Bachoud-Lévi, Frank Rudzicz, Emmanuel Dupoux

Disfluent speech has been previously addressed from two main perspectives: the clinical perspective focusing on diagnostic, and the Natural Language Processing (NLP) perspective aiming at modeling these events and detect…

Diagnostic

Parsing Speech: A Neural Approach to Integrating Lexical and Acoustic-Prosodic Information

2017-04-24 · NAACL 2018 6 · Trang Tran, Shubham Toshniwal, Mohit Bansal, Kevin Gimpel 외

In conversational speech, the acoustic signal provides cues that help listeners disambiguate difficult parses. For automatically parsing spoken utterances, we introduce a model that integrates transcribed text and acoust…

Sentence

A novel multimodal dynamic fusion network for disfluency detection in spoken utterances

2022-11-27 · Sreyan Ghosh, Utkarsh Tyagi, Sonal Kumar, Manan Suri 외

Disfluency, though originating from human spoken utterances, is primarily studied as a uni-modal text-based Natural Language Processing (NLP) task. Based on early-fusion and self-attention-based multimodal interaction be…

multimodal interaction