paper-with-me

Papers

Cloud-Based Face and Speech Recognition for Access Control Applications

2020-04-23 · Nathalie Tkauc, Thao Tran, Kevin Hernandez-Diaz, Fernando Alonso-Fernandez

This paper describes the implementation of a system to recognize employees and visitors wanting to gain access to a physical office through face images and speech-to-text recognition. The system helps employees to unlock the entrance door via face recognition without the need of tag-keys or cards. To prevent spoofing attacks and increase security, a randomly generated code is sent to the employee, who then has to type it into the screen. On the other hand, visitors and delivery persons are provided with a speech-to-text service where they utter the name of the employee that they want to meet, and the system then sends a notification to the right employee automatically. The hardware of the system is constituted by two Raspberry Pi, a 7-inch LCD-touch display, a camera, and a sound card with a microphone and speaker. To carry out face recognition and speech-to-text conversion, the cloud-based platforms Amazon Web Services and the Google Speech-to-Text API service are used respectively. The two-step face authentication mechanism for employees provides an increased level of security and protection against spoofing attacks without the need of carrying key-tags or access cards, while disturbances by visitors or couriers are minimized by notifying their arrival to the right employee, without disturbing other co-workers by means of ring-bells.

📄 PDF Abstract BibTeX arXiv:2004.11168

Code (0)

등록된 구현이 없습니다.

Tasks

Face Recognitionspeech-recognitionSpeech RecognitionSpeech-to-TextTAG

Similar Papers 제목 키워드 기반

Unmanned Aerial Vehicle Control Through Domain-based Automatic Speech Recognition

2020-09-09 · Ruben Contreras, Angel Ayala, Francisco Cruz

Currently, unmanned aerial vehicles, such as drones, are becoming a part of our lives and reaching out to many areas of society, including the industrialized world. A common alternative to control the movements and actio…

Action RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+1

Adaptive Edge-Cloud Inference for Speech-to-Action Systems Using ASR and Large Language Models

2025-12-14 · Mohammad Jalili Torkamani, Israt Zarin arxiv

Voice-based interaction has emerged as a natural and intuitive modality for controlling IoT devices. However, speech-driven edge devices face a fundamental trade-off between cloud-based solutions, which offer stronger la…

Speech Recognition

Emotion Recognition in Speech using Cross-Modal Transfer in the Wild

2018-08-16 · Samuel Albanie, Arsha Nagrani, Andrea Vedaldi, Andrew Zisserman

Obtaining large, human labelled speech datasets to train models for emotion recognition is a notoriously challenging task, hindered by annotation cost and label ambiguity. In this work, we consider the task of learning e…

Emotion RecognitionFacial Emotion RecognitionFacial Expression Recognition (FER)Speech Emotion Recognition

Cloud-Enabled IoT System for Real-Time Environmental Monitoring and Remote Device Control Using Firebase

2026-01-24 · Abdul Hasib, A. S. M. Ahsanul Sarkar Akib arxiv

The proliferation of Internet of Things (IoT) devices has created unprecedented opportunities for remote monitoring and control applications across various domains. Traditional monitoring systems often suffer from limita…

Encrypted Speech Recognition using Deep Polynomial Networks

2019-05-11 · Shi-Xiong Zhang, Yifan Gong, Dong Yu

The cloud-based speech recognition/API provides developers or enterprises an easy way to create speech-enabled features in their applications. However, sending audios about personal or company internal information to the…

speech-recognitionSpeech Recognition