paper-with-me

홈 › Papers

AVATAR: Robust Voice Search Engine Leveraging Autoregressive Document Retrieval and Contrastive Learning

2023-09-04 · Yi-Cheng Wang, Tzu-Ting Yang, Hsin-Wei Wang, Bi-Cheng Yan, Berlin Chen

Voice, as input, has progressively become popular on mobiles and seems to transcend almost entirely text input. Through voice, the voice search (VS) system can provide a more natural way to meet user's information needs. However, errors from the automatic speech recognition (ASR) system can be catastrophic to the VS system. Building on the recent advanced lightweight autoregressive retrieval model, which has the potential to be deployed on mobiles, leading to a more secure and personal VS assistant. This paper presents a novel study of VS leveraging autoregressive retrieval and tackles the crucial problems facing VS, viz. the performance drop caused by ASR noise, via data augmentations and contrastive learning, showing how explicit and implicit modeling the noise patterns can alleviate the problems. A series of experiments conducted on the Open-Domain Question Answering (ODSQA) confirm our approach's effectiveness and robustness in relation to some strong baseline systems.

📄 PDF Abstract BibTeX arXiv:2309.01395

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Contrastive LearningOpen-Domain Question AnsweringQuestion AnsweringRetrievalspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Pre-Avatar: An Automatic Presentation Generation Framework Leveraging Talking Avatar

2022-10-13 · Aolan Sun, xulong Zhang, Tiandong Ling, Jianzong Wang 외

Since the beginning of the COVID-19 pandemic, remote conferencing and school-teaching have become important tools. The previous applications aim to save the commuting cost with real-time interactions. However, our applic…

text-to-speechText to Speech

AvatarPointillist: AutoRegressive 4D Gaussian Avatarization

2026-04-06 · Hongyu Liu, Xuan Wang, Zijian Wu, Yating Wang 외 arxiv

We introduce AvatarPointillist, a novel framework for generating dynamic 4D Gaussian avatars from a single portrait image. At the core of our method is a decoder-only Transformer that autoregressively generates a point c…

ALIVE: An Avatar-Lecture Interactive Video Engine with Content-Aware Retrieval for Real-Time Interaction

2025-12-24 · Md Zabirul Islam, Md Motaleb Hossen Manik, Ge Wang arxiv

Traditional lecture videos offer flexibility but lack mechanisms for real-time clarification, forcing learners to search externally when confusion arises. Recent advances in large language models and neural avatars provi…

Semantic Similarity

AutoAvatar: Autoregressive Neural Fields for Dynamic Avatar Modeling

2022-03-25 · Ziqian Bai, Timur Bagautdinov, Javier Romero, Michael Zollhöfer 외

Neural fields such as implicit surfaces have recently enabled avatar modeling from raw scans without explicit temporal correspondences. In this work, we exploit autoregressive modeling to further extend this notion to ca…

3D Human Dynamics3D Human ReconstructionHuman Dynamics

StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars

2025-12-26 · Zhiyao Sun, Ziqiao Peng, Yifeng Ma, Yi Chen 외 arxiv

Real-time, streaming interactive avatars represent a critical yet challenging goal in digital human research. Although diffusion-based human avatar generation methods achieve remarkable success, their non-causal architec…