paper-with-me

홈 › Papers

Seq-2-Seq based Refinement of ASR Output for Spoken Name Capture

2022-03-29 · Karan Singla, Shahab Jalalvand, Yeon-Jun Kim, Ryan Price, Daniel Pressel, Srinivas Bangalore

Person name capture from human speech is a difficult task in human-machine conversations. In this paper, we propose a novel approach to capture the person names from the caller utterances in response to the prompt "say and spell your first/last name". Inspired from work on spell correction, disfluency removal and text normalization, we propose a lightweight Seq-2-Seq system which generates a name spell from a varying user input. Our proposed method outperforms the strong baseline which is based on LM-driven rule-based approach.

📄 PDF Abstract BibTeX arXiv:2203.15833

Code (0)

등록된 구현이 없습니다.

Tasks

Text Normalization

Similar Papers 제목 키워드 기반

Lost in Transcription: How Speech-to-Text Errors Derail Code Understanding

2026-01-20 · Jayant Havare, Ashish Mittal, Srikanth Tamilselvam, Ganesh Ramakrishnan arxiv

Code understanding is a foundational capability in software engineering tools and developer workflows. However, most existing systems are designed for English-speaking users interacting via keyboards, which limits access…

Speech RecognitionQuestion Answering

Medical Spoken Named Entity Recognition

2024-06-19 · Khai Le-Duc, David Thulke, Hung-Phong Tran, Long Vo-Dang 외

Spoken Named Entity Recognition (NER) aims to extract named entities from speech and categorise them into types like person, location, organization, etc. In this work, we present VietMed-NER - the first spoken NER datase…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1

MTL-SLT: Multi-Task Learning for Spoken Language Tasks

2022-05-01 · NLP4ConvAI (ACL) 2022 5 · Zhiqi Huang, Milind Rao, Anirudh Raju, Zhe Zhang 외

Language understanding in speech-based systems has attracted extensive interest from both academic and industrial communities in recent years with the growing demand for voice-based applications. Prior works focus on ind…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModellingMulti-Task Learning+4

How Does That Sound? Multi-Language SpokenName2Vec Algorithm Using Speech Generation and Deep Learning

2020-05-24 · Aviad Elyashar, Rami Puzis, Michael Fire

Searching for information about a specific person is an online activity frequently performed by many users. In most cases, users are aided by queries containing a name and sending back to the web search engines for findi…

English-Indonesian Neural Machine Translation for Spoken Language Domains

2019-07-01 · ACL 2019 7 · Meisyarah Dwiastuti

In this work, we conduct a study on Neural Machine Translation (NMT) for English-Indonesian (EN-ID) and Indonesian-English (ID-EN). We focus on spoken language domains, namely colloquial and speech languages. We build NM…

Domain AdaptationMachine TranslationNMTTranslation