Semi-automatically Alignment of Predicates between Speech and OntoNotes data
Speech data currently receives a growing attention and is an important source of information. We still lack suitable corpora of transcribed speech annotated with semantic roles that can be used for semantic role labeling (SRL), which is not the case for written data. Semantic role labeling in speech data is a challenging and complex task due to the lack of sentence boundaries and the many transcription errors such as insertion, deletion and misspellings of words. In written data, SRL evaluation is performed at the sentence level, but in speech data sentence boundaries identification is still a bottleneck which makes evaluation more complex. In this work, we semi-automatically align the predicates found in transcribed speech obtained with an automatic speech recognizer (ASR) with the predicates found in the corresponding written documents of the OntoNotes corpus and manually align the semantic roles of these predicates thus obtaining annotated semantic frames in the speech data. This data can serve as gold standard alignments for future research in semantic role labeling of speech data.
Code (0)
등록된 구현이 없습니다.
Tasks
Semantic Role LabelingSentenceSimilar Papers 제목 키워드 기반
Revisiting the Entropy Semiring for Neural Speech Recognition
In streaming settings, speech recognition models have to map sub-sequences of speech to text before the full audio stream becomes available. However, since alignment information between speech and text is rarely availabl…
speech-recognitionSpeech RecognitionSpeech-to-TextSemiautomatic Speech Alignment for Under-Resourced Languages
Cross-language forced alignment is a solution for linguists who create speech corpora for very low-resource languages. However, cross-language is an additional challenge making a complex task, forced alignment, even more…
Uncovering Hidden Semantics of Set Information in Knowledge Bases
Knowledge Bases (KBs) contain a wealth of structured information about entities and predicates. This paper focuses on set-valued predicates, i.e., the relationship between an entity and a set of entities. In KBs, this in…
PositionQuestion AnsweringTransAlign: Fully Automatic and Effective Entity Alignment for Knowledge Graphs
The task of entity alignment between knowledge graphs (KGs) aims to identify every pair of entities from two different KGs that represent the same entity. Many machine learning-based methods have been proposed for this t…
Entity AlignmentEntity EmbeddingsKnowledge GraphsAutoAlign: Fully Automatic and Effective Knowledge Graph Alignment enabled by Large Language Models
The task of entity alignment between knowledge graphs (KGs) aims to identify every pair of entities from two different KGs that represent the same entity. Many machine learning-based methods have been proposed for this t…
Entity AlignmentEntity EmbeddingsKnowledge Graphs