首页    期刊浏览 2025年02月17日 星期一
登录注册

文章基本信息

  • 标题:Bimodal Emotion Recognition Model for Minnan Songs
  • 本地全文:下载
  • 作者:Zhenglong Xiang ; Xialei Dong ; Yuanxiang Li
  • 期刊名称:Information
  • 电子版ISSN:2078-2489
  • 出版年度:2020
  • 卷号:11
  • 期号:3
  • 页码:145-158
  • DOI:10.3390/info11030145
  • 出版社:MDPI Publishing
  • 摘要:Most of the existing research papers study the emotion recognition of Minnan songs from the perspectives of music analysis theory and music appreciation. However, these investigations do not explore any possibility of carrying out an automatic emotion recognition of Minnan songs. In this paper, we propose a model that consists of four main modules to classify the emotion of Minnan songs by using the bimodal data—song lyrics and audio. In the proposed model, an attention-based Long Short-Term Memory (LSTM) neural network is applied to extract lyrical features, and a Convolutional Neural Network (CNN) is used to extract the audio features from the spectrum. Then, two kinds of extracted features are concatenated by multimodal compact bilinear pooling, and finally, the concatenated features are input to the classifying module to determine the song emotion. We designed three experiment groups to investigate the classifying performance of combinations of the four main parts, the comparisons of proposed model with the current approaches and the influence of a few key parameters on the performance of emotion recognition. The results show that the proposed model exhibits better performance over all other experimental groups. The accuracy, precision and recall of the proposed model exceed 0.80 in a combination of appropriate parameters.
  • 关键词:bimodal emotion recognition; Minnan songs; attention-based LSTM; convolutional neural network; Mel spectrum bimodal emotion recognition ; Minnan songs ; attention-based LSTM ; convolutional neural network ; Mel spectrum
国家哲学社会科学文献中心版权所有