A study on Korean connected digit recognition and short-term cepstral mean normalization한국어 연속 숫자 음성 인식과 단구간 켑스트럼 평균 정규화에 관한 연구

Cited 0 time in webofscience Cited 0 time in scopus
  • Hit : 661
  • Download : 0
DC FieldValueLanguage
dc.contributor.advisorHahn, Min-Soo-
dc.contributor.advisor한민수-
dc.contributor.authorKim, Sang-Jin-
dc.contributor.author김상진-
dc.date.accessioned2011-12-28T02:54:56Z-
dc.date.available2011-12-28T02:54:56Z-
dc.date.issued2002-
dc.identifier.urihttp://library.kaist.ac.kr/search/detail/view.do?bibCtrlNo=392127&flag=dissertation-
dc.identifier.urihttp://hdl.handle.net/10203/54775-
dc.description학위논문(석사) - 한국정보통신대학원대학교 : 공학부, 2002, [ xi, 97 p. ]-
dc.description.abstractAlthough many researchers have studied about digit recognition, it is still away from commercial applications in Korea. It is well known that Korean digit recognition is more difficult than English digit recognition, even worse in continuous digits. In this paper, I studied about various techniques to improve the recognition, especially one of the environmental compensation preprocessing methods, called the cepstral mean normalization, with some acoustic-phonetic models. I found that the recognition results varied depending on the windows size for the cepstral mean normalization, and not always the long-term cepstral mean normalization produces the best results. This can be interpreted as if we use the short-term cepstral mean normalization technique with a proper window size for Korean digit recognition, we can get the better results than the conventional cepstral mean normalization. The reason could be the variation of the phone length caused by the short-term cepstral mean normalization, and this variation is believed to improve the recognition rate. Monophone, triphone, whole-word, tri-word, and phonological-rule- considered digit models in Korean pronunciation, are tested in various numbers of states and mixtures. Mel-frequency cepstral coefficients (MFCC) and perceptual linear prediction (PLP) cepstral coefficients are extracted as the feature vectors. Long-term and short-term cepstral mean normalization/ subtraction(CMN/CMS) processing, and relative spectral (RASTA) processing is used for the channel noise compensation. Kalman filtering is applied for additive noise reduction. Linear discriminant analysis (LDA) transformation for the digit recognition is also tested in the end.eng
dc.languageeng-
dc.publisher한국정보통신대학원대학교-
dc.subjectShort-Term Cepstral-
dc.subjectConnected Digit Recognition-
dc.subject인식 시스템-
dc.subject연속 숫자 음성 인식-
dc.subjectST-CMN-
dc.titleA study on Korean connected digit recognition and short-term cepstral mean normalization-
dc.title.alternative한국어 연속 숫자 음성 인식과 단구간 켑스트럼 평균 정규화에 관한 연구-
dc.typeThesis(Master)-
dc.identifier.CNRN392127/225023-
dc.description.department한국정보통신대학원대학교 : 공학부, -
dc.identifier.uid020003853-
dc.contributor.localauthorHahn, Min-Soo-
dc.contributor.localauthor한민수-
Appears in Collection
School of Engineering-Theses_Master(공학부 석사논문)
Files in This Item
There are no files associated with this item.

qr_code

  • mendeley

    citeulike


rss_1.0 rss_2.0 atom_1.0