Speech-Music Classification Model Based on Improved Neural Network and Beat Spectrum

被引:0
|
作者
Huang, Chun [1 ]
Wei, HeFu [2 ]
机构
[1] Gen Educ & Int Coll, Chongqing Coll Elect Engn, Chongqing 400031, Peoples R China
[2] Arts Coll Sichuan Univ, Chengdu 401331, Peoples R China
关键词
Vocal music; classification model; beat spectrum; feature parameter extraction; cosine similarity; convolutional neural network; FREQUENCY;
D O I
10.14569/IJACSA.2023.0140706
中图分类号
TP301 [理论、方法];
学科分类号
081202 ;
摘要
A speech-music classification method according to a developed neural system and beat spectrum is proposed to achieve accurate classification of speech-music through preemphasis, endpoint detection, framing, windowing and other steps to preprocess and collect vocal music signals. After fast Fourier transforms and triangle filter processing, the Mel frequency cepstrum coefficient (MFCC) is obtained, and a discrete cosine transform is performed to obtain the signal MFCC characteristic parameters. After calculating the similarity of feature parameters through cosine similarity, the signal similarity matrix is obtained, based on which the vocal music beat spectrum is obtained. The residual structure is optimized by adding Swish and max-out activation functions, respectively, between convolutional neural network layers to build residual convolution layers and deepen the number of convolution layers. The connected time series classification (CTC) is used as the objective loss function. It is applied to the softmax layer to build a deep optimization residual convolutional neural network for speech-music classification model. The pitch spectrum of vocal music is used as the input information of the model to realize the vocal music classification. The experiment proves that the classification accuracy of the design model is higher than 99%; when the iteration reaches 1200, the training loss approaches 0; when the signal-to-noise ratio is 180dB, the sensitivity and specificity are 99.98% and 99.96%, respectively; the accuracy of voice music classification is higher than 99%, and the running time is 0.48 seconds. It has been proven that the model has high classification accuracy, low training loss, good sensitivity and special effects, and can effectively achieve the classification of speech-music.
引用
收藏
页码:52 / 64
页数:13
相关论文
共 50 条
  • [41] ECG beat classification by a novel hybrid neural network
    Dokur, Z
    Ölmez, T
    COMPUTER METHODS AND PROGRAMS IN BIOMEDICINE, 2001, 66 (2-3) : 167 - 181
  • [42] Application of InP neural network to ECG beat classification
    Ölmez, T
    Dokur, Z
    NEURAL COMPUTING & APPLICATIONS, 2003, 11 (3-4): : 144 - 155
  • [43] Neural network modeling of speech and music signals
    Robel, A
    ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 9: PROCEEDINGS OF THE 1996 CONFERENCE, 1997, 9 : 779 - 785
  • [44] A hybrid neural network model based on optimized margin softmax loss function for music classification
    Jingxian Li
    Lixin Han
    Xin Wang
    Yang Wang
    Jianhua Xia
    Yi Yang
    Bing Hu
    Shu Li
    Hong Yan
    Multimedia Tools and Applications, 2024, 83 : 43871 - 43906
  • [45] A hybrid neural network model based on optimized margin softmax loss function for music classification
    Li, Jingxian
    Han, Lixin
    Wang, Xin
    Wang, Yang
    Xia, Jianhua
    Yang, Yi
    Hu, Bing
    Li, Shu
    Yan, Hong
    MULTIMEDIA TOOLS AND APPLICATIONS, 2023, 83 (15) : 43871 - 43906
  • [46] Network traffic classification method based on improved capsule neural network
    Zhang, Fan
    Wang, Yong
    Miao, Ye
    2018 14TH INTERNATIONAL CONFERENCE ON COMPUTATIONAL INTELLIGENCE AND SECURITY (CIS), 2018, : 174 - 178
  • [47] Bidirectional Recurrent Neural Network and Convolutional Neural Network (BiRCNN) for ECG Beat Classification
    Xie, Pengwei
    Wang, Guijin
    Zhang, Chenshuang
    Chen, Ming
    Yang, Huazhong
    Lv, Tingting
    Sang, Zhenhua
    Zhang, Ping
    2018 40TH ANNUAL INTERNATIONAL CONFERENCE OF THE IEEE ENGINEERING IN MEDICINE AND BIOLOGY SOCIETY (EMBC), 2018, : 2555 - 2558
  • [48] An Efficient Hidden Markov Model with Periodic Recurrent Neural Network Observer for Music Beat Tracking
    Song, Guangxiao
    Wang, Zhijie
    ELECTRONICS, 2022, 11 (24)
  • [49] Music Style Classification Algorithm Based on Music Feature Extraction and Deep Neural Network
    Zhang, Kedong
    WIRELESS COMMUNICATIONS & MOBILE COMPUTING, 2021, 2021
  • [50] Insect Detection and Classification Based on an Improved Convolutional Neural Network
    Xia, Denan
    Chen, Peng
    Wang, Bing
    Zhang, Jun
    Xie, Chengjun
    SENSORS, 2018, 18 (12)