LIA-RAG: a system based on graphs and divergence of probabilities applied to Speech-To-Text Summarization
This paper aims to introduces a new algorithm for automatic speech-to-text summarization based on statistical divergences of probabilities and graphs. The input is a text from speech conversations with noise, and the output a compact text summary. Our results, on the pilot task CCCS Multiling 2015 French corpus are very encouraging
Code (0)
등록된 구현이 없습니다.
Tasks
RAGSpeech-to-TextText SummarizationSimilar Papers 제목 키워드 기반
Landing Probabilities of Random Walks for Seed-Set Expansion in Hypergraphs
We describe the first known mean-field study of landing probabilities for random walks on hypergraphs. In particular, we examine clique-expansion and tensor methods and evaluate their mean-field characteristics over a cl…
Probability-turbulence divergence: A tunable allotaxonometric instrument for comparing heavy-tailed categorical distributions
Real-world complex systems often comprise many distinct types of elements as well as many more types of networked interactions between elements. When the relative abundances of types can be measured well, we often observ…
Balance Divergence for Knowledge Distillation
Knowledge distillation has been widely adopted in computer vision task processing, since it can effectively enhance the performance of lightweight student networks by leveraging the knowledge transferred from cumbersome …
image-classificationImage ClassificationKnowledge DistillationSemantic SegmentationRegularizing End-to-End Speech Translation with Triangular Decomposition Agreement
End-to-end speech-to-text translation (E2E-ST) is becoming increasingly popular due to the potential of its less error propagation, lower latency, and fewer parameters. Given the triplet training corpus $\langle speech, …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+5Voice Conversion Using Sequence-to-Sequence Learning of Context Posterior Probabilities
Voice conversion (VC) using sequence-to-sequence learning of context posterior probabilities is proposed. Conventional VC using shared context posterior probabilities predicts target speech parameters from the context po…
speech-recognitionSpeech RecognitionSpeech SynthesisVoice Conversion