paper-with-me

홈 › Papers

People are poorly equipped to detect AI-powered voice clones

2024-10-03 · Sarah Barrington, Emily A. Cooper, Hany Farid

As generative artificial intelligence (AI) continues its ballistic trajectory, everything from text to audio, image, and video generation continues to improve at mimicking human-generated content. Through a series of perceptual studies, we report on the realism of AI-generated voices in terms of identity matching and naturalness. We find human participants cannot consistently identify recordings of AI-generated voices. Specifically, participants perceived the identity of an AI-voice to be the same as its real counterpart approximately 80% of the time, and correctly identified a voice as AI generated only about 60% of the time.

📄 PDF Abstract BibTeX arXiv:2410.03791

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

Design and Implementation of an OCR-Powered Pipeline for Table Extraction from Invoices

2025-07-09 · Parshva Dhilankumar Patel

This paper presents the design and development of an OCR-powered pipeline for efficient table extraction from invoices. The system leverages Tesseract OCR for text recognition and custom post-processing logic to detect, …

Boundary DetectionOptical Character Recognition (OCR)Table Extraction

Deaf and Hard of Hearing Access to Intelligent Personal Assistants: Comparison of Voice-Based Options with an LLM-Powered Touch Interface

2026-01-21 · Paige S. DeVries, Michaela Okosi, Ming Li, Nora Dunphy 외 arxiv

We investigate intelligent personal assistants (IPAs) accessibility for deaf and hard of hearing (DHH) people who can use their voice in everyday communication. The inability of IPAs to understand diverse accents includi…

Speech Recognition

Voice Recognition Robot with Real-Time Surveillance and Automation

2023-12-07 · Lochan Basyal

Voice recognition technology enables the execution of real-world operations through a single voice command. This paper introduces a voice recognition system that involves converting input voice signals into corresponding…

FaVoA: Face-Voice Association Favours Ambiguous Speaker Detection

2021-09-01 · Hugo Carneiro, Cornelius Weber, Stefan Wermter

The strong relation between face and voice can aid active speaker detection systems when faces are visible, even in difficult settings, when the face of a speaker is not clear or when there are several people in the same…

Active Speaker Detection

A Semantic Web Framework for Automated Smart Assistants: COVID-19 Case Study

2020-07-01 · Yusuf Sermet, Ibrahim Demir

COVID-19 pandemic elucidated that knowledge systems will be instrumental in cases where accurate information needs to be communicated to a substantial group of people with different backgrounds and technological resource…

Natural Language Understanding