Skip to content
Research Article Open access CC BY 4.0

Evaluation of Machine Learning Algorithms using Combined Feature Extraction Techniques for Speaker Identification

Unwana Ubong Iwok, Kingsley Monday Udofia, Akaninyene Bernard Obot, Kufre Michael Udofia, Unyime Anietie Michael, Akpabio Itoro Kingsley

Journal of Engineering Research and Reports · pp. 197–216 · Published 19 Sep 2023

10.9734/jerr/2023/v25i8973

Abstract

The aim of this study is to evaluate and compare machine learning algorithms when various feature extraction techniques are employed together and determine the optimal feature combinations for the models studied. The TIMIT online database was used where 5 male and 5 female non-native English speakers from five American locations were selected. Each speaker had ten 3-second utterances, totaling 500. Mel frequency cepstral coefficients (MFCC), linear predictive cepstral coefficients (LPCC), gammatone frequency cepstral coefficients (GFCC), discrete wavelet transforms (DWT) and pitch features were extracted using MATLAB and concatenated. The concatenated features were used to train and evaluate three classifier models—Random Forest (RF), Linear Discriminant Analysis (LDA), and Logistic Regression (LR)—using Python software. The results obtained showed that as the number of features combinations increased, the models’ performances improved as well. This improved performance was observed when all the cepstral features were part of the combinations. This implies that cepstral features are more robust and improve speaker identification systems. The best average score of accurate predictions of ≈ 76% for the LR model was obtained for the MGL (39 features) features combination and dropped to 70% for the highest number of feature combinations MGDLP (53 features). This indicates that more training data improves system performance, however, too much data does not translate to even better performance because the system will eventually achieve its peak performance. This information is useful for applications where limited data can present a problem.

Speaker identification feature extraction machine learning evaluation metrics

Cited by 4

Speaker Identification From a Tribal Language Using Several Learning Techniques

Subrat Kumar Nayak, Ajit Kumar Nayak, Smitaprava Mishra · 2024 3rd Odisha International Conference on Electrical Power Engineering, Communication and Computing Technology (ODICON) · 2024

Enhancement of Speaker Identification System Based on Voice Active Detection Techniques using Machine Learning

Khine Zin Oo, Lwin Nyein Thu, Zaw Htet Aung · 2024 Conference of Young Researchers in Electrical and Electronic Engineering (ElCon) · 2024

Speaker Identification System Using Artificial Neural Network By Enhancing the Signals

Khine Zin Oo, Lwin Nyein Thu · 2024 3rd International Conference on Artificial Intelligence For Internet of Things (AIIoT) · 2024

Article metrics

Real usage data collected on this platform.

0

Page views

0

PDF downloads

0

Outbound clicks

4

Citations

Views by country

Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".

No views recorded yet.

Traffic sources

Referring site, by host.

No traffic recorded yet.

Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.