Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Towards Robust Combined Deep Architecture for Speech Recognition : Experiments on TIMIT

Author 1: Hinda DRIDI Author 2: Kais OUNI
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 11, No. 4 · Published 2020 · Cited by 9

DOI: https://doi.org/10.14569/IJACSA.2020.0110469

Abstract

Over the last years, many researchers have engaged in improving accuracies on Automatic Speech Recognition (ASR) task by using deep learning. In state-of-the-art speech recognizers, both Long Short-Term Memory (LSTM) and Gated Recurrent Unit (GRU) based Reccurent Neural Network (RNN) have achieved improved performances compared to Convolutional Neural Network (CNN) and Deep Neural Network (DNN). Due to the strong complementarity of CNN, LSTM-RNN and DNN, they may be combined in one architecture called Convolutional Long Short-Term Memory, Deep Neural Network (CLDNN). Similarly we propose to combine CNN, GRU-RNN and DNN in a single deep architecture called Convolutional Gated Recurrent Unit, Deep Neural Network (CGDNN). In this paper, we present our experiments for phoneme recognition task tested on TIMIT data set. A phone error rate of 15.72% has been reached using the proposed CGDNN model. The achieved result confirms the superiority of CGDNN over all their baselines networks used alone and also over the CLDNN architecture.

Keywords

How to Cite this Article

DRIDI, H., & OUNI, K. (2020). Towards Robust Combined Deep Architecture for Speech Recognition : Experiments on TIMIT. International Journal of Advanced Computer Science and Applications, 11(4). https://doi.org/10.14569/IJACSA.2020.0110469

DRIDI, Hinda, and Kais OUNI. "Towards Robust Combined Deep Architecture for Speech Recognition : Experiments on TIMIT." International Journal of Advanced Computer Science and Applications, vol. 11, no. 4, 2020, https://doi.org/10.14569/IJACSA.2020.0110469.

@article{DRIDI2020,
  title     = {Towards Robust Combined Deep Architecture for Speech Recognition : Experiments on TIMIT},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {11},
  number    = {4},
  year      = {2020},
  publisher = {The Science and Information Organization},
  author    = {Hinda DRIDI and Kais OUNI},
  doi       = {10.14569/IJACSA.2020.0110469},
  url       = {https://doi.org/10.14569/IJACSA.2020.0110469}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.