Please use this identifier to cite or link to this item:
Title: A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system
Authors: Dai, Peng
Soon, Ing Yann
Keywords: DRNTU::Engineering::Electrical and electronic engineering
Issue Date: 2011
Source: Dai, P., & Soon, I. Y. (2012). A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system. Speech Communication, 54(3), 402-413.
Series/Report no.: Speech communication
Abstract: In this paper, a novel hybrid feature extraction algorithm is proposed, which implements forward masking, lateral inhibition, and temporal integration with a simple 2D psychoacoustic filter. The proposed algorithm consists of two key parts, the 2D psychoacoustic filter and cepstral mean variance normalization (CMVN). Mathematical derivation is provided to show the correctness of the 2D psychoacoustic filter based on the characteristic functions of masking effects. The effectiveness of the proposed algorithm is tested on the AURORA2 database. Extensive comparison is made against lateral inhibition (LI), forward masking (FM), CMVN, RASTA filter, the ETSI standard advanced front-end feature extraction algorithm (AFE), and the temporal warped 2D psychoacoustic filter. Experimental results show significant improvements from the proposed algorithm, a relative improvement of nearly 46.78% over the baseline mel-frequency cepstral coefficients (MFCC) system in noisy conditions.
ISSN: 0167-6393
DOI: 10.1016/j.specom.2011.10.004
Rights: © 2011 Elsevier B.V.
Fulltext Permission: none
Fulltext Availability: No Fulltext
Appears in Collections:EEE Journal Articles

Citations 50

Updated on Jan 15, 2023

Web of ScienceTM
Citations 20

Updated on Feb 4, 2023

Page view(s) 20

Updated on Feb 6, 2023

Google ScholarTM




Items in DR-NTU are protected by copyright, with all rights reserved, unless otherwise indicated.