Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Speech-Music Classification Model Based on Improved Neural Network and Beat Spectrum

Author 1: Chun Huang Author 2: Wei HeFu
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 14, No. 7 · Published 2023

DOI: https://doi.org/10.14569/IJACSA.2023.0140706

Abstract

A speech-music classification method according to a developed neural system and beat spectrum is proposed to achieve accurate classification of speech-music through pre-emphasis, endpoint detection, framing, windowing and other steps to preprocess and collect vocal music signals. After fast Fourier transforms and triangle filter processing, the Mel frequency cepstrum coefficient (MFCC) is obtained, and a discrete cosine transform is performed to obtain the signal MFCC characteristic parameters. After calculating the similarity of feature parameters through cosine similarity, the signal similarity matrix is obtained, based on which the vocal music beat spectrum is obtained. The residual structure is optimized by adding Swish and max-out activation functions, respectively, between convolutional neural network layers to build residual convolution layers and deepen the number of convolution layers. The connected time series classification (CTC) is used as the objective loss function. It is applied to the softmax layer to build a deep optimization residual convolutional neural network for speech-music classification model. The pitch spectrum of vocal music is used as the input information of the model to realize the vocal music classification. The experiment proves that the classification accuracy of the design model is higher than 99%; when the iteration reaches 1200, the training loss approaches 0; when the signal-to-noise ratio is 180dB, the sensitivity and specificity are 99.98% and 99.96%, respectively; the accuracy of voice music classification is higher than 99%, and the running time is 0.48 seconds. It has been proven that the model has high classification accuracy, low training loss, good sensitivity and special effects, and can effectively achieve the classification of speech-music.

Keywords

How to Cite this Article

Huang, C., & HeFu, W. (2023). Speech-Music Classification Model Based on Improved Neural Network and Beat Spectrum. International Journal of Advanced Computer Science and Applications, 14(7). https://doi.org/10.14569/IJACSA.2023.0140706

Huang, Chun, and Wei HeFu. "Speech-Music Classification Model Based on Improved Neural Network and Beat Spectrum." International Journal of Advanced Computer Science and Applications, vol. 14, no. 7, 2023, https://doi.org/10.14569/IJACSA.2023.0140706.

@article{Huang2023,
  title     = {Speech-Music Classification Model Based on Improved Neural Network and Beat Spectrum},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {14},
  number    = {7},
  year      = {2023},
  publisher = {The Science and Information Organization},
  author    = {Chun Huang and Wei HeFu},
  doi       = {10.14569/IJACSA.2023.0140706},
  url       = {https://doi.org/10.14569/IJACSA.2023.0140706}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.