SPEECH RECOGNITION FRONT-END FOR SEGMENTING AND CLUSTERING CONTINUOUS BANGLA SPEECH

dc.contributor.authorRahman, Md. Mijanur
dc.contributor.authorKhan, Md. Farukuzzaman
dc.contributor.authorMoni, Mohammad Ali
dc.date.accessioned2012-11-10T06:28:36Z
dc.date.accessioned2019-05-28T09:52:26Z
dc.date.available2012-11-10T06:28:36Z
dc.date.available2019-05-28T09:52:26Z
dc.date.issued2010-01-01
dc.description.abstractThis research is concerned with the development of speech recognition front-end for segmenting and clustering continuous Bangla speech sentence to some predefined clusters. From the study of different previous research works it was observed that the front-end is an important part of any speech recognition system. In our work, the original speech sentences were recorded and stored as RIFF (.wav) file format. Then a segmentation approach was used to segment the continuous speech into uniquely identifiable and meaningful units. Among the different techniques, the word/sub-word segmentation is simple and produces very good results. This is why this technique was selected for speech segmentation to obtain improved performance. After segmentation, the segmented words were clustered into different clusters according to the number of syllables and the sizes of the segmented words. The test database contained 758 words/sub-words segmented from 120 sentences. Each sentence was recorded from six different speakers and saved as a different wave file. The developed system achieved the segmentation accuracy rate at about 95%.
dc.identifier.otherhttp://dspace.daffodilvarsity.edu.bd:8080/handle/20.500.11948/513
dc.identifier.urihttp://hdl.handle.net/20.500.11948/513
dc.language.isoen
dc.publisherDaffodil International University
dc.sourceDIU Institutional Repository
dc.subjectFront-end, Phonemic and Word segmentation, Clustering, End Point Detection.
dc.titleSPEECH RECOGNITION FRONT-END FOR SEGMENTING AND CLUSTERING CONTINUOUS BANGLA SPEECH
dc.typeArticle

Files

Original bundle

Now showing 1 - 1 of 1
Thumbnail Image
Name:
Speech recognition front-end.pdf
Size:
301.93 KB
Format:
Adobe Portable Document Format