Robust Automatic Speaker Identification System Using Shuffled MFCC Features
IEEE International Conference on Machine Learning and Applied Network Technologies (ICMLANT), pp. 1–6
Abstract
Speaker Identification using deep learning is an ongoing active topic that has wide applications in voice authenticated and activated systems, forensic investigations, security, and remote transactions. Mel frequency cepstral coefficients (MFCC) feature-based speaker identification systems are proven to per- form very well with clean speech conditions. However, there is rapid degradation in the robustness with adverse conditions and short speech duration. To overcome this limitation, we propose a new speaker identification pipeline system grounded on novel MFCC based features called shuffled MFCC (SHMFCC) along with a new data augmentation approach. Moreover, the system uses a simple and efficient deep neural network model with variable-duration input speech signals and is proved to perform remarkably at different datasets with different environmental and noise conditions and for a high number of speakers.
Authors 3
-
Affiliation as printed
ISEK Research Area, RWTH University, Aachen, Germany
-
Affiliation as printed
ICE Institute, RWTH University, Aachen, Germany
-
Affiliation as printed
ISEK Research Area, RWTH University, Aachen, Germany
Cited by 8 stored of 8
-
Semi-supervised Algorithms in Resource-constrained Edge Devices: An Overview and Experimental Comparison2022 IEEE International Conferences on Internet of Things (iThings) and IEEE Green Computing & Communications (GreenCom) and IEEE Cyber, Physical & Social Computing (CPSCom) and IEEE Smart Data (SmartData) and IEEE Congress on Cybermatics (Cybermatics) conference-paper Computer Science Machine Learning and Algorithms5citations
8 results
No patents citing this paper on Lens.org (checked 2026-10-06).