In Speaker Recognition (SR) system, feature extraction is one of the crucial steps where the particular speaker related information are extracted. The state of the art algorithm for this purpose is Mel Frequency Cepstral Coefficient (MFCC), and its complementary feature, Inverted Mel Frequency Cepstral Coefficient (IMFCC). MFCC is based on mel scale and IMFCC is based on inverted mel (imel) scale. In this paper, another complementary set of features are proposed which is also based on mel-imel scale, and the filtering operation makes these set of features different from MFCC and IMFCC. On the background of this proposed features, the filter banks are placed linearly on the nonlinear scale which makes the features different from the state-of-theart feature extraction techniques. We call these two features as mMFCC, and mIMFCC. mMFCC is based on mel scale, whereas, mIMFCC is based on imel. mMFCC is compared with MFCC and mIMFCC is compared with IMFCC. The result has been verified on two standard databases YOHO, and POLYCOST using Gaussian Mixture Model (GMM) as the speaker modeling paradigm.
In Speaker Recognition (SR) system, feature extraction is one of the crucial steps where the particular speaker related information is extracted. The state of the art algorithm for this purpose is Mel Frequency Cepstral Coefficient (MFCC), and its complementary feature, Inverted Mel Frequency Cepstral Coefficient (IMFCC). MFCC is based on mel scale and IMFCC is based on inverted mel (imel) scale. There are two another set of features we proposed as mMFCC and mIMFCC. In state-of-the-art system, we neglect the DC co-efficient of DCT from the feature set. In this paper, the DC coefficient and its effect on recognition accuracy on MFCC-IMFCC, as well as, mMFCC-mIMFCC has been studied. This has been verified on two standard different types of databases, like, YOHO for clean speech signal and POLYCOST for telephone based speech. The recognition accuracy of the proposed feature is better than their respective baseline feature when the DC coefficient was included, as well as, when it was not included.
scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.