論文 ·日本語 ·未確認
A Statistical Approach to Automatic Phonetic Transcription of Japanese Orthographic Words.
Wei-Bin Chang ・ Sachiko Morishita
- 刊行年
- 2003-01-01
- 収録
- 『Journal of Natural Language Processing』 10(4) pp. 55-63
- 言語
- 英語
- openalex
- W1984324186
- doi
- 10.5715/jnlp.10.4_55
- mag
- 1984324186
- issn
- 1340-7619
- URL
- https://www.jstage.jst.go.jp/article/jnlp1994/10/4/10_4_55/_pdf
要旨
We address the problem of automatically transcribing Japanese orthographic words into symbols representing their pronunciations. Such a function is necessary for commercial continuous speech recognition systems since there are constant needs to create new recognition lexica for new applications or purposes. Simple look-up schemes are not adequate to deal with Japanese, while methods based on morphological analysis require in-depth linguistic knowledge and development effort. In this paper, we propose a statistical approach which is based on an N-gram language model. It is assumed that the pronunciation of a character only depends on the previous one to two characters and their pronunciations. Given an orthographic word, our method outputs the most likely phonetic transcription. It is shown that our approach provides superior performance to the public-domain conversion tool KAKASI on ten out of twelve test sets.
主題
この書誌の出所
- openalex— W1984324186(2026-08-13取得)
引用キー: ChangMorishita2003StatisticalApproachAutomatic