論文 ·日本語 ·未確認
Japanese Dictation Toolkit. 1997 version.
Tatsuya Kawahara ・ Akinobu Lee ・ Tetsunori Kobayashi ・ Kazuya Takeda ・ Nobuaki Minematsu ・ Katsunobu Itou ・ Akinori Ito ・ Mikio Yamamoto ・ Atsushi Yamada ・ Takehito Utsuro ・ Kiyohiro Shikano
- 刊行年
- 1999-01-01
- 収録
- 『Journal of the Acoustical Society of Japan (E)』 20(3) pp. 233-239
- 言語
- 英語
- OpenAlex
- W2020933619
- DOI
- 10.1250/ast.20.233
- MAG
- 2020933619
- ISSN
- 0388-2861
- URL
- https://www.jstage.jst.go.jp/article/ast1980/20/3/20_3_233/_pdf
要旨
The Japanese Dictation Toolkit has been designed and developed as a baseline platform for Japanese LVCSR (Large Vocabulary Continuous Speech Recognition). The platform consists of a standard recognition engine, Japanese phone models and Japanese statistical language models. We set up a variety of Japanese phone HMMs from a contextindependent monophone to a triphone model of thousands of states. They are trained with ASJ (The Acoustical Society of Japan) databases. A lexicon and word N-gram (2-gram and 3-gram) models are constructed with a corpus of Mainichi newspaper. The recognition engine JULIUS is developed for evaluation of both acoustic and language models. As an integrated system of these modules, we have implemented a baseline 5, 000-word dictation system and evaluated various components. The software repository is available to the public.
主題
この書誌の出所
- openalex— W2020933619(2026-08-14取得)
引用
Tatsuya Kawahara・Akinobu Lee・Tetsunori Kobayashi・Kazuya Takeda・Nobuaki Minematsu・Katsunobu Itou・Akinori Ito・Mikio Yamamoto・Atsushi Yamada・Takehito Utsuro・Kiyohiro Shikano(1999-01-01) Japanese Dictation Toolkit. 1997 version. 『Journal of the Acoustical Society of Japan (E)』 20(3) pp. 233-239