paper-with-me

Papers

Modernizing Open-Set Speech Language Identification

2022-05-20 · Mustafa Eyceoz, Justin Lee, Homayoon Beigi

While most modern speech Language Identification methods are closed-set, we want to see if they can be modified and adapted for the open-set problem. When switching to the open-set problem, the solution gains the ability to reject an audio input when it fails to match any of our known language options. We tackle the open-set task by adapting two modern-day state-of-the-art approaches to closed-set language identification: the first using a CRNN with attention and the second using a TDNN. In addition to enhancing our input feature embeddings using MFCCs, log spectral features, and pitch, we will be attempting two approaches to out-of-set language detection: one using thresholds, and the other essentially performing a verification task. We will compare both the performance of the TDNN and the CRNN, as well as our detection approaches.

📄 PDF Abstract BibTeX arXiv:2205.10397

Code (0)

등록된 구현이 없습니다.

Tasks

Language IdentificationSpeech Language Identification

Similar Papers 제목 키워드 기반

MERLIon CCS Challenge Evaluation Plan

2023-05-31 · Leibny Paola Garcia Perera, Y. H. Victoria Chua, Hexin Liu, Fei Ting Woon 외

This paper introduces the inaugural Multilingual Everyday Recordings- Language Identification on Code-Switched Child-Directed Speech (MERLIon CCS) Challenge, focused on developing robust language identification and langu…

Language IdentificationTask 2

A Machine Translation Approach for Modernizing Historical Documents Using Backtranslation

2018-10-01 · IWSLT (EMNLP) 2018 10 · Miguel Domingo, Francisco Casacuberta

Human language evolves with the passage of time. This makes historical documents to be hard to comprehend by contemporary people and, thus, limits their accessibility to scholars specialized in the time period in which a…

Machine TranslationTranslation

OWSM-CTC: An Open Encoder-Only Speech Foundation Model for Speech Recognition, Translation, and Language Identification

2024-02-20 · Yifan Peng, Yui Sudo, Muhammad Shakeel, Shinji Watanabe

There has been an increasing interest in large speech models that can perform multiple tasks in a single model. Such models usually adopt an encoder-decoder or decoder-only architecture due to their popularity and good p…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderHallucination+5

Robust Open-Set Spoken Language Identification and the CU MultiLang Dataset

2023-08-29 · Mustafa Eyceoz, Justin Lee, Siddharth Pittie, Homayoon Beigi

Most state-of-the-art spoken language identification models are closed-set; in other words, they can only output a language label from the set of classes they were trained on. Open-set spoken language identification syst…

Language IdentificationSpoken language identification

Spoken Language Identification System for English-Mandarin Code-Switching Child-Directed Speech

2023-06-01 · Shashi Kant Gupta, Sushant Hiray, Prashant Kukde

This work focuses on improving the Spoken Language Identification (LangId) system for a challenge that focuses on developing robust language identification systems that are reliable for non-standard, accented (Singaporea…

DecoderLanguage IdentificationSpoken language identification