論文 ·日本語 ·未確認

Japanese Speaker-Independent Homonyms Speech Recognition

Jin’ichi Murakami Haseo Hotta

刊行年
2011-01-01
収録
『Procedia - Social and Behavioral Sciences』 27 pp. 306-313
出版
Elsevier BV
言語
英語
openalex
W2045977232
doi
10.1016/j.sbspro.2011.10.612
mag
2045977232
issn
1877-0428
URL
https://www.sciencedirect.com/science/article/pii/S1877042811024396/pdf

要旨

Japanese has homonyms such as “hashi” ((Chop-sticks)) and “hashi” ((Bridge)). Word speech recognition has been studied for a long time, but homonym speech recognition in Japanese has not been studied. In this paper, we studied speaker-independent homonym speech recognition. For homonym speech recognition, pitch extraction has been normally used to estimate a pitch frequency.. However, we did not use pitch extraction in our study. Instead, we used an accent model that was a phoneme label with more length, Mora position, accent type and accent high or low. It means that we used the effect of pitch on formant. The results of the experiments were that 89% accuracy was obtained by using MFCC, full covariance HMM, and the accent model.

主題

この書誌の出所

  • openalex— W2045977232(2026-08-13取得)

引用キー: MurakamiHotta2011JapaneseSpeakerIndependent

書誌 56,535件 語別索引 17,251件 資源 113件 研究者 303名 JSON