Back to Search Start Over

Articulation based admissible wavelet packet feature based on human cochlear frequency response for TIMIT speech recognition

Authors :
Astik Biswas
P.K. Sahu
Anirban Bhowmick
Mahesh Chandra
Source :
Ain Shams Engineering Journal, Vol 5, Iss 4, Pp 1189-1198 (2014)
Publication Year :
2014
Publisher :
Elsevier, 2014.

Abstract

To deal with non-stationary and quasi-stationary signals, wavelet transform has been used as an effective tool for the time-frequency analysis. In the recent years, wavelet transform has been used extensively for feature extraction in noisy speech recognition. These filters have the benefit of having frequency bands spacing similar to the auditory Equivalent Rectangular Bandwidth (ERB) scale. Central frequencies of ERB are equally distributed with the frequency response of the human cochlea. This paper deals with the speaker-independent Automatic Speech Recognition (ASR) system for continuous speech. This Hidden Markov Model (HMM) based ASR system was developed for English using recordings of four regions taken from TIMIT database. A new set of features were derived using wavelet packet transform’s multi-resolution capabilities and having an advantage of ERB filter based on the human cochlea. New set of wavelet features have shown significant improvements in the noisy environment, especially at low SNR values.

Details

Language :
English
ISSN :
20904479
Volume :
5
Issue :
4
Database :
Directory of Open Access Journals
Journal :
Ain Shams Engineering Journal
Publication Type :
Academic Journal
Accession number :
edsdoj.fcda9a00ade74373add712ebf5e6df6b
Document Type :
article
Full Text :
https://doi.org/10.1016/j.asej.2014.07.006