paper-with-me

홈 › Papers

LIA-RAG: a system based on graphs and divergence of probabilities applied to Speech-To-Text Summarization

2016-01-26 · Elvys Linhares Pontes, Juan-Manuel Torres-Moreno, Andréa Carneiro Linhares

This paper aims to introduces a new algorithm for automatic speech-to-text summarization based on statistical divergences of probabilities and graphs. The input is a text from speech conversations with noise, and the output a compact text summary. Our results, on the pilot task CCCS Multiling 2015 French corpus are very encouraging

📄 PDF Abstract BibTeX arXiv:1601.07124

Code (0)

등록된 구현이 없습니다.

Tasks

RAGSpeech-to-TextText Summarization

Similar Papers 제목 키워드 기반

Landing Probabilities of Random Walks for Seed-Set Expansion in Hypergraphs

2019-10-20 · Eli Chien, Pan Li, Olgica Milenkovic

We describe the first known mean-field study of landing probabilities for random walks on hypergraphs. In particular, we examine clique-expansion and tensor methods and evaluate their mean-field characteristics over a cl…

Probability-turbulence divergence: A tunable allotaxonometric instrument for comparing heavy-tailed categorical distributions

2020-08-30 · P. S. Dodds, J. R. Minot, M. V. Arnold, T. Alshaabi 외

Real-world complex systems often comprise many distinct types of elements as well as many more types of networked interactions between elements. When the relative abundances of types can be measured well, we often observ…

Balance Divergence for Knowledge Distillation

2025-01-14 · Yafei Qi, Chen Wang, Zhaoning Zhang, Yaping Liu 외

Knowledge distillation has been widely adopted in computer vision task processing, since it can effectively enhance the performance of lightweight student networks by leveraging the knowledge transferred from cumbersome …

image-classificationImage ClassificationKnowledge DistillationSemantic Segmentation

Regularizing End-to-End Speech Translation with Triangular Decomposition Agreement

2021-12-21 · Yichao Du, Zhirui Zhang, Weizhi Wang, Boxing Chen 외

End-to-end speech-to-text translation (E2E-ST) is becoming increasingly popular due to the potential of its less error propagation, lower latency, and fewer parameters. Given the triplet training corpus $\langle speech, …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+5

Voice Conversion Using Sequence-to-Sequence Learning of Context Posterior Probabilities

2017-04-10 · Hiroyuki Miyoshi, Yuki Saito, Shinnosuke Takamichi, Hiroshi Saruwatari

Voice conversion (VC) using sequence-to-sequence learning of context posterior probabilities is proposed. Conventional VC using shared context posterior probabilities predicts target speech parameters from the context po…

speech-recognitionSpeech RecognitionSpeech SynthesisVoice Conversion