paper-with-me

홈 › Papers

HASP: A High-Performance Adaptive Mobile Security Enhancement Against Malicious Speech Recognition

2018-09-04 · Zirui Xu, Fuxun Yu, ChenChen Liu, Xiang Chen

Nowadays, machine learning based Automatic Speech Recognition (ASR) technique has widely spread in smartphones, home devices, and public facilities. As convenient as this technology can be, a considerable security issue also raises -- the users' speech content might be exposed to malicious ASR monitoring and cause severe privacy leakage. In this work, we propose HASP -- a high-performance security enhancement approach to solve this security issue on mobile devices. Leveraging ASR systems' vulnerability to the adversarial examples, HASP is designed to cast human imperceptible adversarial noises to real-time speech and effectively perturb malicious ASR monitoring by increasing the Word Error Rate (WER). To enhance the practical performance on mobile devices, HASP is also optimized for effective adaptation to the human speech characteristics, environmental noises, and mobile computation scenarios. The experiments show that HASP can achieve optimal real-time security enhancement: it can lead an average WER of 84.55% for perturbing the malicious ASR monitoring, and the data processing speed is 15x to 40x faster compared to the state-of-the-art methods. Moreover, HASP can effectively perturb various ASR systems, demonstrating a strong transferability.

📄 PDF Abstract BibTeX arXiv:1809.01697

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Mobile Securityspeech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Do Dogs have Whiskers? A New Knowledge Base of hasPart Relations

2020-06-12 · Sumithra Bhakthavatsalam, Kyle Richardson, Niket Tandon, Peter Clark

We present a new knowledge-base of hasPart relationships, extracted from a large corpus of generic statements. Complementary to other resources available, it is the first which is all three of: accurate (90% precision), …

SynthASpoof: Developing Face Presentation Attack Detection Based on Privacy-friendly Synthetic Data

2023-03-05 · Meiling Fang, Marco Huber, Naser Damer

Recently, significant progress has been made in face presentation attack detection (PAD), which aims to secure face recognition systems against presentation attacks, owing to the availability of several face PAD datasets…

DiversityDomain GeneralizationFace Presentation Attack DetectionFace Recognition

Multi-modal Learning with Missing Modality via Shared-Specific Feature Modelling

2023-07-26 · CVPR 2023 1 · Hu Wang, Yuanhong Chen, Congbo Ma, Jodie Avery 외

The missing modality issue is critical but non-trivial to be solved by multi-modal models. Current methods aiming to handle the missing modality problem in multi-modal tasks, either deal with missing modalities only duri…

Classificationdomain classificationImage SegmentationMedical Image Segmentation+2

Harnessing LLM Agents with Skill Programs

2026-05-18 · Hongjun Liu, Yifei Ming, Shafiq Joty, Chen Zhao arxiv

Equipping LLM agents with reusable skills derived from past experience has become a popular and successful approach for tackling complex and long-horizon tasks. However, such lessons are often encoded as textual guidance…

Hybrid attention structure preserving network for reconstruction of under-sampled OCT images

2024-06-01 · Zezhao Guo, Zhanfang Zhao

Optical coherence tomography (OCT) is a non-invasive, high-resolution imaging technology that provides cross-sectional images of tissues. Dense acquisition of A-scans along the fast axis is required to obtain high digita…

Super-Resolution