paper-with-me

Papers

Topic Identification for Speech without ASR

2017-03-22 · Chunxi Liu, Jan Trmal, Matthew Wiesner, Craig Harman, Sanjeev Khudanpur

Modern topic identification (topic ID) systems for speech use automatic speech recognition (ASR) to produce speech transcripts, and perform supervised classification on such ASR outputs. However, under resource-limited conditions, the manually transcribed speech required to develop standard ASR systems can be severely limited or unavailable. In this paper, we investigate alternative unsupervised solutions to obtaining tokenizations of speech in terms of a vocabulary of automatically discovered word-like or phoneme-like units, without depending on the supervised training of ASR systems. Moreover, using automatic phoneme-like tokenizations, we demonstrate that a convolutional neural network based framework for learning spoken document representations provides competitive performance compared to a standard bag-of-words representation, as evidenced by comprehensive topic ID evaluations on both single-label and multi-label classification tasks.

📄 PDF Abstract BibTeX arXiv:1703.07476

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)General ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATIONspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Topic Identification and Discovery on Text and Speech

2015-09-01 · EMNLP 2015 9 · Ch May, ler, Francis Ferraro, Alan McCree 외
Dimensionality ReductionSpeech RecognitionTopic Models

Topic Identification For Spontaneous Speech: Enriching Audio Features With Embedded Linguistic Information

2023-07-21 · Dejan Porjazovski, Tamás Grósz, Mikko Kurimo

Traditional topic identification solutions from audio rely on an automatic speech recognition system (ASR) to produce transcripts used as input to a text-based model. These approaches work well in high-resource scenarios…

Automatic Speech Recognitionspeech-recognitionSpeech Recognition

Low-Resource Contextual Topic Identification on Speech

2018-07-17 · Chunxi Liu, Matthew Wiesner, Shinji Watanabe, Craig Harman 외

In topic identification (topic ID) on real-world unstructured audio, an audio instance of variable topic shifts is first broken into sequential segments, and each segment is independently classified. We first present a g…

General ClassificationTopic ClassificationTranslation

Hate Speech Detection in Turkish and Arabic: A Comprehensive Study

2026-06-30 · Somaiyeh Dehghan, Gökçe Uludoğan, Mehmet Umut Şen, Elif Erol 외 arxiv

Online hate speech has been linked to a global rise in violence against minorities, including incidents such as mass shootings, lynchings, and ethnic cleansing. Societies grappling with this issue, particularly when hate…

Hate Speech Detection

Speech Translation with Speech Foundation Models and Large Language Models: What is There and What is Missing?

2024-02-19 · Marco Gaido, Sara Papi, Matteo Negri, Luisa Bentivogli

The field of natural language processing (NLP) has recently witnessed a transformative shift with the emergence of foundation models, particularly Large Language Models (LLMs) that have revolutionized text-based NLP. Thi…

Speech-to-TextSpeech-to-Text Translation