Efficient feature extraction and classification for the development of Pashto speech recognition system

Irfan Ahmed, Muhammad Abeer Irfan, Abid Iqbal, Amaad Khalil, Salman Ilahi Siddiqui

Research output: Contribution to journalArticlepeer-review

Abstract

In this work, a novel framework for the efficient feature extraction and recognition of Pashto speech signals is proposed. The targeted language is one of the low-resource languages and prone to higher Automatic Speech Recognition (ASR) errors due to the availability of its colloquial dialects. We devised a framework which not only employed classical Machine Learning (ML) models for speech recognition tasks, but also achieved a higher level of performance accuracy by using the optimal feature extraction techniques. The designed frameworks for feature extraction are based on two well-know feature extraction techniques: Discrete Wavelet Transform (DWT )coefficients and Mel-Frequency Cepstral Coefficients (MFCC). In our work, we deployed classical ML models i.e., Support Vector Machine (SVM) and K-Nearest Neighbors (k-NN), due to their efficiency in terms of computation complexity, energy efficiency, and higher accuracy as compared to other ML and Deep Learning (DL) model. Hence, our proposed framework exhibited improved performance level when trained on a Pashto isolated words dataset.

Original languageEnglish
Pages (from-to)54081-54096
Number of pages16
JournalMultimedia Tools and Applications
Volume83
Issue number18
DOIs
Publication statusPublished - May 2024
Externally publishedYes

Keywords

  • Automatic speech recognition (ASR)
  • DWT
  • Feature extraction
  • k-NN
  • Machine learning (ML)
  • MFCC
  • SVM

Fingerprint

Dive into the research topics of 'Efficient feature extraction and classification for the development of Pashto speech recognition system'. Together they form a unique fingerprint.

Cite this