首页 /研究 /Speaker Recognition for Robotic Control via an IoT Device
LEARNING

Speaker Recognition for Robotic Control via an IoT Device

Zhanibek Kozhirbayev, Berat A. Erol, Altynbek Sharipbay, Mo Jamshidi

发表年份
2018
引用次数
25

摘要

Speaker Recognition is considered as one of the primary tasks in speech processing. Nowadays, the speaker identification method has been extensively appealing for its broad application in many fields, such as smart environments, securing the cyber-physical systems, speech communications, and robotic controls. Researchers are targeting to perform an effective method that makes it possible to obtain the recognition ability that is close to the hearing of human. In order to get high accuracy, challenges of large-scale applications of speaker identification are overcome through applying techniques not only traditional models based on the GMM, but also deep learning methods. Aiming at effectively dealing with this challenge, in this paper, we present a novel model to increase the recognition accuracy of the short utterance speaker recognition system. We developed a technique to train a Neural Network (NN) on the extracted Mel-Frequency Cepstral Coefficient (MFCC) features from audio samples. Therefore, the recognition system gains the significant accuracy. The model was trained using open-source high-level neural networks API Keras.

关键词

Computer scienceMel-frequency cepstrumSpeaker recognitionSpeech recognitionArtificial neural networkArtificial intelligenceSpeaker identificationIdentification (biology)Feature extractionUtterance

相关论文

查看 LEARNING 分类全部论文