A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system

In this paper, a novel hybrid feature extraction algorithm is proposed, which implements forward masking, lateral inhibition, and temporal integration with a simple 2D psychoacoustic filter. The proposed algorithm consists of two key parts, the 2D psychoacoustic filter and cepstral mean variance nor...

Full description

Saved in:
Bibliographic Details
Main Authors: Dai, Peng, Soon, Ing Yann
Other Authors: School of Electrical and Electronic Engineering
Format: Article
Language:English
Published: 2013
Subjects:
Online Access:https://hdl.handle.net/10356/95803
http://hdl.handle.net/10220/11933
Tags: Add Tag
No Tags, Be the first to tag this record!
Institution: Nanyang Technological University
Language: English
id sg-ntu-dr.10356-95803
record_format dspace
spelling sg-ntu-dr.10356-958032020-03-07T14:02:45Z A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system Dai, Peng Soon, Ing Yann School of Electrical and Electronic Engineering DRNTU::Engineering::Electrical and electronic engineering In this paper, a novel hybrid feature extraction algorithm is proposed, which implements forward masking, lateral inhibition, and temporal integration with a simple 2D psychoacoustic filter. The proposed algorithm consists of two key parts, the 2D psychoacoustic filter and cepstral mean variance normalization (CMVN). Mathematical derivation is provided to show the correctness of the 2D psychoacoustic filter based on the characteristic functions of masking effects. The effectiveness of the proposed algorithm is tested on the AURORA2 database. Extensive comparison is made against lateral inhibition (LI), forward masking (FM), CMVN, RASTA filter, the ETSI standard advanced front-end feature extraction algorithm (AFE), and the temporal warped 2D psychoacoustic filter. Experimental results show significant improvements from the proposed algorithm, a relative improvement of nearly 46.78% over the baseline mel-frequency cepstral coefficients (MFCC) system in noisy conditions. 2013-07-22T03:22:32Z 2019-12-06T19:21:49Z 2013-07-22T03:22:32Z 2019-12-06T19:21:49Z 2011 2011 Journal Article Dai, P., & Soon, I. Y. (2012). A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system. Speech Communication, 54(3), 402-413. 0167-6393 https://hdl.handle.net/10356/95803 http://hdl.handle.net/10220/11933 10.1016/j.specom.2011.10.004 en Speech communication © 2011 Elsevier B.V.
institution Nanyang Technological University
building NTU Library
country Singapore
collection DR-NTU
language English
topic DRNTU::Engineering::Electrical and electronic engineering
spellingShingle DRNTU::Engineering::Electrical and electronic engineering
Dai, Peng
Soon, Ing Yann
A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system
description In this paper, a novel hybrid feature extraction algorithm is proposed, which implements forward masking, lateral inhibition, and temporal integration with a simple 2D psychoacoustic filter. The proposed algorithm consists of two key parts, the 2D psychoacoustic filter and cepstral mean variance normalization (CMVN). Mathematical derivation is provided to show the correctness of the 2D psychoacoustic filter based on the characteristic functions of masking effects. The effectiveness of the proposed algorithm is tested on the AURORA2 database. Extensive comparison is made against lateral inhibition (LI), forward masking (FM), CMVN, RASTA filter, the ETSI standard advanced front-end feature extraction algorithm (AFE), and the temporal warped 2D psychoacoustic filter. Experimental results show significant improvements from the proposed algorithm, a relative improvement of nearly 46.78% over the baseline mel-frequency cepstral coefficients (MFCC) system in noisy conditions.
author2 School of Electrical and Electronic Engineering
author_facet School of Electrical and Electronic Engineering
Dai, Peng
Soon, Ing Yann
format Article
author Dai, Peng
Soon, Ing Yann
author_sort Dai, Peng
title A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system
title_short A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system
title_full A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system
title_fullStr A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system
title_full_unstemmed A temporal frequency warped (TFW) 2D psychoacoustic filter for robust speech recognition system
title_sort temporal frequency warped (tfw) 2d psychoacoustic filter for robust speech recognition system
publishDate 2013
url https://hdl.handle.net/10356/95803
http://hdl.handle.net/10220/11933
_version_ 1681046147291938816