本文へ移動

論文 ·日本語 ·未確認

A Japanese sentence analyzer

Newton Maruyama Masayuki Morohashi Shigeki UMEDA Eiichro Sumita

刊行年
1988-03-01
収録
『IBM Journal of Research and Development』 32(2) pp. 238-250
出版
IBM
言語
英語
OpenAlex
W2074736143
DOI
10.1147/rd.322.0238
MAG
2074736143
ISSN
0018-8646
URL
https://doi.org/10.1147/rd.322.0238

要旨

This paper presents the design of a broad-coverage Japanese sentence analyzer which can be part of various Japanese processing systems. The sentence analyzer comprises two components: the lexical analyzer and the syntactic analyzer. Lexical analysis, i.e., segmenting a sentence into words, is a formidable problem for a language like Japanese, because it has no explicit delimiters (blanks) between written words. In practical applications, this task is made more difficult by the occurrence of words not listed in a dictionary. We have developed a five-layered knowledge source and used it successfully in the lexical analyzer, resulting in very accurate segmentation, even in cases where there are unknown words. The syntactic analyzer has two modules: One consists of an augmented context-free grammar and the PLNLP parser; the other is the dependency structure constructor, which converts the phrase structures to dependency structures. The dependency structures represent various key linguistic relations in a more direct way. The dependency structures have semantically important information such as tense, aspect, and modality, as well as preference scores reflecting relative ranking of parse acceptability.

主題

この書誌の出所

  • openalex— W2074736143(2026-08-14取得)

引用

Newton Maruyama・Masayuki Morohashi・Shigeki UMEDA・Eiichro Sumita(1988-03-01) A Japanese sentence analyzer 『IBM Journal of Research and Development』 32(2) pp. 238-250 IBM

Maruyama1988JapaneseSentenceAnalyzer
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON