paper-with-me

홈 › Papers

VoCopilot: Voice-Activated Tracking of Everyday Interactions

2023-12-15 · Sheen An Goh, Manoj Gulati, Ambuj Varshney

Voice plays an important role in our lives by facilitating communication, conveying emotions, and indicating health. Therefore, tracking vocal interactions can provide valuable insight into many aspects of our lives. This paper presents our ongoing efforts to design a new vocal tracking system we call VoCopilot. VoCopilot is an end-to-end system centered around an energy-efficient acoustic hardware and firmware combined with advanced machine learning models. As a result, VoCopilot is able to continuously track conversations, record them, transcribe them, and then extract useful insights from them. By utilizing large language models, VoCopilot ensures the user can extract useful insights from recorded interactions without having to learn complex machine learning techniques. In order to protect the privacy of end users, VoCopilot uses a novel wake-up mechanism that only records conversations of end users. Additionally, all the rest of pipeline can be run on a commodity computer (Mac Mini M2). In this work, we show the effectiveness of VoCopilot in real-world environment for two use cases.

📄 PDF Abstract BibTeX arXiv:2312.10265

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

StepWrite: Adaptive Planning for Speech-Driven Text Generation

2025-08-06 · Hamza El Alaoui, Atieh Taheri, Yi-Hao Peng, Jeffrey P. Bigham arxiv

People frequently use speech-to-text systems to compose short texts with voice. However, current voice-based interfaces struggle to support composing more detailed, contextually complex texts, especially in scenarios whe…

Text GenerationFact Checking

Smartwatch-derived Acoustic Markers for Deficits in Cognitively Relevant Everyday Functioning

2023-09-11 · Yasunori Yamada, Kaoru Shinkawa, Masatomo Kobayashi, Miyuki Nemoto 외

Detection of subtle deficits in everyday functioning due to cognitive impairment is important for early detection of neurodegenerative diseases, particularly Alzheimer's disease. However, current standards for assessment…

Multi-task Voice Activated Framework using Self-supervised Learning

2021-10-03 · Shehzeen Hussain, Van Nguyen, Shuhua Zhang, Erik Visser

Self-supervised learning methods such as wav2vec 2.0 have shown promising results in learning speech representations from unlabelled and untranscribed speech data that are useful for speech recognition. Since these repre…

Emotion ClassificationKeyword SpottingMulti-Task LearningSelf-Supervised Learning+3

SurfaceXR: Fusing Smartwatch IMUs and Egocentric Hand Pose for Seamless Surface Interactions

2026-03-19 · Vasco Xu, Brian Chen, Eric J. Gonzalez, Andrea Colaço 외 arxiv

Mid-air gestures in Extended Reality (XR) often cause fatigue and imprecision. Surface-based interactions offer improved accuracy and comfort, but current egocentric vision methods struggle due to hand tracking challenge…

Gesture Recognition

Coimagining the Future of Voice Assistants with Cultural Sensitivity

2024-03-26 · Katie Seaborn, Yuto Sawa, Mizuki Watanabe

Voice assistants (VAs) are becoming a feature of our everyday life. Yet, the user experience (UX) is often limited, leading to underuse, disengagement, and abandonment. Co-designing interactions for VAs with potential en…

Sensitivity