paper-with-me

홈 › Papers

Wiki-En-ASR-Adapt: Large-scale synthetic dataset for English ASR Customization

2023-09-29 · Alexandra Antonova

We present a first large-scale public synthetic dataset for contextual spellchecking customization of automatic speech recognition (ASR) with focus on diverse rare and out-of-vocabulary (OOV) phrases, such as proper names or terms. The proposed approach allows creating millions of realistic examples of corrupted ASR hypotheses and simulate non-trivial biasing lists for the customization task. Furthermore, we propose injecting two types of ``hard negatives" to the simulated biasing lists in training examples and describe our procedures to automatically mine them. We report experiments with training an open-source customization model on the proposed dataset and show that the injection of hard negative biasing phrases decreases WER and the number of false alarms.

📄 PDF Abstract BibTeX arXiv:2309.17267

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

WEC: Deriving a Large-scale Cross-document Event Coreference dataset from Wikipedia

2021-04-11 · NAACL 2021 4 · Alon Eirew, Arie Cattan, Ido Dagan

Cross-document event coreference resolution is a foundational task for NLP applications involving multi-text processing. However, existing corpora for this task are scarce and relatively small, while annotating only mode…

coreference-resolutionCoreference ResolutionEvent Coreference Resolution

Utilizing citation index and synthetic quality measure to compare Wikipedia languages across various topics

2025-05-22 · Włodzimierz Lewoniewski, Krzysztof Węcel, Witold Abramowicz

This study presents a comparative analysis of 55 Wikipedia language editions employing a citation index alongside a synthetic quality measure. Specifically, we identified the most significant Wikipedia articles within di…

Articles

TAPEX: Table Pre-training via Learning a Neural SQL Executor

2021-07-16 · ICLR 2022 4 · Qian Liu, Bei Chen, Jiaqi Guo, Morteza Ziyadi 외

Recent progress in language model pre-training has achieved a great success via leveraging large-scale unstructured textual data. However, it is still a challenge to apply pre-training on structured tabular data due to t…

Language ModelingLanguage ModellingSemantic ParsingTable-based Fact Verification

Pre-training Cross-lingual Open Domain Question Answering with Large-scale Synthetic Supervision

2024-02-26 · Fan Jiang, Tom Drummond, Trevor Cohn

Cross-lingual open domain question answering (CLQA) is a complex problem, comprising cross-lingual retrieval from a multilingual knowledge base, followed by answer generation in the query language. Both steps are usually…

Answer GenerationCross-Lingual Question AnsweringDecoderMachine Translation+5

Understanding the Limits of Lifelong Knowledge Editing in LLMs

2025-03-07 · Lukas Thede, Karsten Roth, Matthias Bethge, Zeynep Akata 외

Keeping large language models factually up-to-date is crucial for deployment, yet costly retraining remains a challenge. Knowledge editing offers a promising alternative, but methods are only tested on small-scale or syn…

Benchmarkingknowledge editing