Speech emotion recognition based on convolution neural network combined with random forest
Zheng Li, Qiao Li, Hua Ban, Shuhua Liu
- 发表年份
- 2018
- 引用次数
- 63
摘要
The key to speech emotion recognition is extraction of speech emotion features. In this paper, a new network model (CNN-RF) based on convolution neural network combined with random forest is proposed. Firstly, the convolution neural network is used as the feature extractor to extract the speech emotion feature from the normalized spectrogram, used random forest classification algorithm to classify the speech emotion features. The result of experiment shows that CNN-RF model is superior to the traditional CNN model. Secondly, Improved the Record Sound command box of Nao and applied the CNN-RF model to Nao robot. Finally, Nao robot can "try to figure out" a human's psychology through speech emotion recognition and also know about people's happiness, anger, sadness and joy, achieving a more intelligent human-computer interaction.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002