本文へ移動

論文 ·日本語 ·未確認

A part of speech estimation method for Japanese unknown words using a statistical model of morphology and context

Masaaki Nagata

刊行年
1999-01-01
言語
英語
OpenAlex
W2125513361
DOI
10.3115/1034678.1034725
MAG
2125513361
URL
https://dl.acm.org/doi/pdf/10.3115/1034678.1034725

要旨

We present a statistical model of Japanese unknown words consisting of a set of length and spelling models classified by the character types that constitute a word. The point is quite simple: different character sets should be treated differently and the changes between character types are very important because Japanese script has both ideograms like Chinese (kanji) and phonograms like English (katakana). Both word segmentation accuracy and part of speech tagging accuracy are improved by the proposed model. The model can achieve 96.6% tagging accuracy if unknown words are correctly segmented.

主題

この書誌の出所

  • openalex— W2125513361(2026-08-14取得)

引用

Masaaki Nagata(1999-01-01) A part of speech estimation method for Japanese unknown words using a statistical model of morphology and context pp. 277-284

Nagata1999PartSpeechEstimation
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON