Cloud-Based Face and Speech Recognition for Access Control Applications
This paper describes the implementation of a system to recognize employees and visitors wanting to gain access to a physical office through face images and speech-to-text recognition. The system helps employees to unlock the entrance door via face recognition without the need of tag-keys or cards. To prevent spoofing attacks and increase security, a randomly generated code is sent to the employee, who then has to type it into the screen. On the other hand, visitors and delivery persons are provided with a speech-to-text service where they utter the name of the employee that they want to meet, and the system then sends a notification to the right employee automatically. The hardware of the system is constituted by two Raspberry Pi, a 7-inch LCD-touch display, a camera, and a sound card with a microphone and speaker. To carry out face recognition and speech-to-text conversion, the cloud-based platforms Amazon Web Services and the Google Speech-to-Text API service are used respectively. The two-step face authentication mechanism for employees provides an increased level of security and protection against spoofing attacks without the need of carrying key-tags or access cards, while disturbances by visitors or couriers are minimized by notifying their arrival to the right employee, without disturbing other co-workers by means of ring-bells.
Code (0)
등록된 구현이 없습니다.
Tasks
Face Recognitionspeech-recognitionSpeech RecognitionSpeech-to-TextTAGSimilar Papers 제목 키워드 기반
Unmanned Aerial Vehicle Control Through Domain-based Automatic Speech Recognition
Currently, unmanned aerial vehicles, such as drones, are becoming a part of our lives and reaching out to many areas of society, including the industrialized world. A common alternative to control the movements and actio…
Action RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+1Adaptive Edge-Cloud Inference for Speech-to-Action Systems Using ASR and Large Language Models
Voice-based interaction has emerged as a natural and intuitive modality for controlling IoT devices. However, speech-driven edge devices face a fundamental trade-off between cloud-based solutions, which offer stronger la…
Speech RecognitionEmotion Recognition in Speech using Cross-Modal Transfer in the Wild
Obtaining large, human labelled speech datasets to train models for emotion recognition is a notoriously challenging task, hindered by annotation cost and label ambiguity. In this work, we consider the task of learning e…
Emotion RecognitionFacial Emotion RecognitionFacial Expression Recognition (FER)Speech Emotion RecognitionCloud-Enabled IoT System for Real-Time Environmental Monitoring and Remote Device Control Using Firebase
The proliferation of Internet of Things (IoT) devices has created unprecedented opportunities for remote monitoring and control applications across various domains. Traditional monitoring systems often suffer from limita…
Encrypted Speech Recognition using Deep Polynomial Networks
The cloud-based speech recognition/API provides developers or enterprises an easy way to create speech-enabled features in their applications. However, sending audios about personal or company internal information to the…
speech-recognitionSpeech Recognition