論文 ·日本語 ·未確認

Learning Lexicons from Spoken Utterances Based on Statistical Model Selection

Ryo Taguchi Naoto Iwahashi Kotaro Funakoshi Mikio Nakano Takashi Nose Tsuneo Nitta

刊行年
2010-01-01
収録
『Transactions of the Japanese Society for Artificial Intelligence』 25(4) pp. 549-559
出版
The Japanese Society for Artificial Intelligence
言語
英語
openalex
W1965921560
doi
10.1527/tjsai.25.549
mag
1965921560
issn
1346-0714
URL
https://www.jstage.jst.go.jp/article/tjsai/25/4/25_4_549/_pdf

要旨

This paper proposes a method for the unsupervised learning of lexicons from pairs of a spoken utterance and an object as its meaning under the condition that any priori linguistic knowledge other than acoustic models of Japanese phonemes is not used. The main problems are the word segmentation of spoken utterances and the learning of the phoneme sequences of the words. To obtain a lexicon, a statistical model, which represents the joint probability of an utterance and an object, is learned based on the minimum description length (MDL) principle. The model consists of three parts: a word list in which each word is represented by a phoneme sequence, a word-bigram model, and a word-meaning model. Through alternate learning processes of these parts, acoustically, grammatically, and semantically appropriate units of phoneme sequences that cover all utterances are acquired as words. Experimental results show that our model can acquire phoneme sequences of object words with about 83.6% accuracy.

主題

この書誌の出所

  • openalex— W1965921560(2026-08-13取得)

引用キー: Taguchi2010LearningLexiconsSpoken

書誌 56,535件 語別索引 17,251件 資源 113件 研究者 303名 JSON