論文 ·日本語 ·未確認

A Statistical Approach to Automatic Phonetic Transcription of Japanese Orthographic Words.

Wei-Bin Chang Sachiko Morishita

刊行年
2003-01-01
収録
『Journal of Natural Language Processing』 10(4) pp. 55-63
言語
英語
openalex
W1984324186
doi
10.5715/jnlp.10.4_55
mag
1984324186
issn
1340-7619
URL
https://www.jstage.jst.go.jp/article/jnlp1994/10/4/10_4_55/_pdf

要旨

We address the problem of automatically transcribing Japanese orthographic words into symbols representing their pronunciations. Such a function is necessary for commercial continuous speech recognition systems since there are constant needs to create new recognition lexica for new applications or purposes. Simple look-up schemes are not adequate to deal with Japanese, while methods based on morphological analysis require in-depth linguistic knowledge and development effort. In this paper, we propose a statistical approach which is based on an N-gram language model. It is assumed that the pronunciation of a character only depends on the previous one to two characters and their pronunciations. Given an orthographic word, our method outputs the most likely phonetic transcription. It is shown that our approach provides superior performance to the public-domain conversion tool KAKASI on ten out of twelve test sets.

主題

この書誌の出所

  • openalex— W1984324186(2026-08-13取得)

引用キー: ChangMorishita2003StatisticalApproachAutomatic

書誌 56,535件 語別索引 17,251件 資源 113件 研究者 303名 JSON