This paper proposes modifications to the Multi-resolution RASTA (MRASTA) feature extraction technique for the automatic speech recognition (ASR). By emulating asymmetries of the temporal receptive field (TRF) profiles of auditory mid-brain neurons, we obtain more than relative improvement in word error rate on OGI-Digits database. Experiments on TIMIT database confirm that proposed modifications are indeed useful.
David Atienza Alonso, Andreas Peter Burg, Pablo Garcia del Valle, Loris Gérard Duch, Shrikanth Ganapathy
Hossein Asadi, Nooshin Sadat Mirzadeh