HMM-based phoneme speech recognition system for the control and command of industrial robots
In recent years, the integration of human-robot interaction with speech recognition has gained a lot of pace in the manufacturing industries. Conventional methods to control the robots include semi-autonomous, fully autonomous, and wired methods. Operating through a teaching pendant or a joystick is easy to implement but is not effective when the robot is deployed to perform complex repetitive tasks. Speech and touch are natural ways of communicating for humans and speech recognition, being the best option, is a heavily researched technology. In this study, we aim at developing a stable and robust speech recognition system to allow humans to communicate with machines (robotic arms) in a seamless manner. This paper investigates the potential of the linear predictive coding technique to develop a stable and robust HMM-based phoneme speech recognition system for applications in robotics. Our system is divided into three segments: a microphone array, a voice module, and a robotic arm with three degrees of freedom (DOF). To validate our approach, we performed experiments with simple and complex sentences for various robotic activities such as manipulating a cube and picking and placing tasks. Moreover, we also analyzed the test results to rectify problems including accuracy and recognition scores.
Code (0)
등록된 구현이 없습니다.
Tasks
Industrial RobotsRobust Speech Recognitionspeech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
Unmanned Aerial Vehicle Control Through Domain-based Automatic Speech Recognition
Currently, unmanned aerial vehicles, such as drones, are becoming a part of our lives and reaching out to many areas of society, including the industrialized world. A common alternative to control the movements and actio…
Action RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+1CONATION: English Command Input/Output System for Computers
In this information technology age, a convenient and user friendly interface is required to operate the computer system on very fast rate. In the human being, speech being a natural mode of communication has potential to…
speech-recognitionSpeech RecognitionInner speech recognition through electroencephalographic signals
This work focuses on inner speech recognition starting from EEG signals. Inner speech recognition is defined as the internalized process in which the person thinks in pure meanings, generally associated with an auditory …
EEGElectroencephalogram (EEG)speech-recognitionSpeech RecognitionSpanish and English Phoneme Recognition by Training on Simulated Classroom Audio Recordings of Collaborative Learning Environments
Audio recordings of collaborative learning environments contain a constant presence of cross-talk and background noise. Dynamic speech recognition between Spanish and English is required in these environments. To elimina…
Data AugmentationPhoneme Recognitionspeech-recognitionSpeech Recognition+1SPEAK YOUR MIND! Towards Imagined Speech Recognition With Hierarchical Deep Learning
Speech-related Brain Computer Interface (BCI) technologies provide effective vocal communication strategies for controlling devices through speech commands interpreted from brain signals. In order to infer imagined speec…
Brain Computer InterfaceGeneral Classificationspeech-recognitionSpeech Recognition+1