Рақамли технологияларнинг назарий ва амалий масалалари Том 9 № 1 (2026) · с. 80-89
Performance evaluation of speaker identification algorithms using speech signal features
Шукуров, К.Э., Хасанов, У.К.
Аннотация
This article analyzes the effectiveness of using different models in speaker recognition processes and selects the best one for the system. In terms of accuracy and speed performance of the system, the classical MFCC + cosine similarity and modern x-vector, ECAPA-TDNN + PLDA architectures are compared. Based on the data set generated from different speakers, the accuracy, f1-score, EER, latency, and GPU load indicators of the models are evaluated. According to the experimental results, the ECAPA-TDNN model outperforms the other models with an accuracy of 95.7%. Since the speaker recognition stage is also important for speaker separation systems, accuracy indicators are of high relevance. The ECAPA-TDNN + PLDA model offers good solutions in terms of using computational resources, working with large data sets, and analyzing their data.
идентификация говорящегоECAPA-TDNNx-векторMFCCлогарифмическое сходствоPLDAкосинусное сходствоглубокое обучениевектор признаковречевая биометрия
Источник метаданных: OAI-PMH архив журнала · Sindex не хранит полный текст, а даёт ссылку на источник.