ISCA Archive SLTU 2018
ISCA Archive SLTU 2018

Development of Assamese Continuous Speech Recognition System

Barsha Deka, S R Nirmala, Samudravijaya K.

This paper describes the development of a continuous speech recognition system for Assamese, an under-resourced language of North-East India. The Speech corpus used in this work consists of 5658 spoken utterances collected from 27 speakers over telephone channel. The baseline speech recognition system was implemented using conventional hidden Markov model in conjunction with Gaussian mixture model, employing Mel-frequency cepstral coefficients as features. ASR systems using subspace Gaussian mixture model and deep neural networks together with hidden Markov model were implemented. The systems were evaluated with 3-fold cross validation method. The average word error rate of the best ASR system is 4.3%.