paper-with-me

홈 › Papers

Chunk Different Kind of Spoken Discourse: Challenges for Machine Learning

2020-05-01 · LREC 2020 5 · Iris Eshkol-Taravella, Mariame Maarouf, Flora Badin, Marie Skrovec, Isabelle Tellier

This paper describes the development of a chunker for spoken data by supervised machine learning using the CRFs, based on a small reference corpus composed of two kinds of discourse: prepared monologue vs. spontaneous talk in interaction. The methodology considers the specific character of the spoken data. The machine learning uses the results of several available taggers, without correcting the results manually. Experiments show that the discourse type (monologue vs. free talk), the speech nature (spontaneous vs. prepared) and the corpus size can influence the results of the machine learning process and must be considered while interpreting the results.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Similar Papers 제목 키워드 기반

Polysemy in Spoken Conversations and Written Texts

2022-06-01 · LREC 2022 6 · Aina Garí Soler, Matthieu Labeau, Chloé Clavel

Our discourses are full of potential lexical ambiguities, due in part to the pervasive use of words having multiple senses. Sometimes, one word may even be used in more than one sense throughout a text. But, to what exte…

Fillers in Spoken Language Understanding: Computational and Psycholinguistic Perspectives

2023-01-25 · Tanvi Dinkar, Chloé Clavel, Ioana Vasilescu

Disfluencies (i.e. interruptions in the regular flow of speech), are ubiquitous to spoken discourse. Fillers ("uh", "um") are disfluencies that occur the most frequently compared to other kinds of disfluencies. Yet, to t…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1

Annotating Discourse Relations in Spoken Language: A Comparison of the PDTB and CCR Frameworks

2016-05-01 · LREC 2016 5 · Ines Rehbein, Merel Scholman, Vera Demberg

In discourse relation annotation, there is currently a variety of different frameworks being used, and most of them have been developed and employed mostly on written data. This raises a number of questions regarding int…

Relation

STITCH: Simultaneous Thinking and Talking with Chunked Reasoning for Spoken Language Models

2025-07-21 · Cheng-Han Chiang, Xiaofei Wang, Linjie Li, Chung-Ching Lin 외 arxiv

Spoken Language Models (SLMs) are designed to take speech inputs and produce spoken responses. However, current SLMs lack the ability to perform an internal, unspoken thinking process before responding. In contrast, huma…

Towards Modelling Coherence in Spoken Discourse

2020-12-31 · Rajaswa Patil, Yaman Kumar Singla, Rajiv Ratn Shah, Mika Hama 외

While there has been significant progress towards modelling coherence in written discourse, the work in modelling spoken discourse coherence has been quite limited. Unlike the coherence in text, coherence in spoken disco…