論文 ·日本語 ·未確認
Japanese Speaker-Independent Homonyms Speech Recognition
Jin’ichi Murakami ・ Haseo Hotta
- 刊行年
- 2011-01-01
- 収録
- 『Procedia - Social and Behavioral Sciences』 27 pp. 306-313
- 出版
- Elsevier BV
- 言語
- 英語
- openalex
- W2045977232
- doi
- 10.1016/j.sbspro.2011.10.612
- mag
- 2045977232
- issn
- 1877-0428
- URL
- https://www.sciencedirect.com/science/article/pii/S1877042811024396/pdf
要旨
Japanese has homonyms such as “hashi” ((Chop-sticks)) and “hashi” ((Bridge)). Word speech recognition has been studied for a long time, but homonym speech recognition in Japanese has not been studied. In this paper, we studied speaker-independent homonym speech recognition. For homonym speech recognition, pitch extraction has been normally used to estimate a pitch frequency.. However, we did not use pitch extraction in our study. Instead, we used an accent model that was a phoneme label with more length, Mora position, accent type and accent high or low. It means that we used the effect of pitch on formant. The results of the experiments were that 89% accuracy was obtained by using MFCC, full covariance HMM, and the accent model.
主題
この書誌の出所
- openalex— W2045977232(2026-08-13取得)
引用キー: MurakamiHotta2011JapaneseSpeakerIndependent