論文 ·日本語 ·未確認
A Japanese sentence analyzer
Newton Maruyama ・ Masayuki Morohashi ・ Shigeki UMEDA ・ Eiichro Sumita
- 刊行年
- 1988-03-01
- 収録
- 『IBM Journal of Research and Development』 32(2) pp. 238-250
- 出版
- IBM
- 言語
- 英語
- OpenAlex
- W2074736143
- DOI
- 10.1147/rd.322.0238
- MAG
- 2074736143
- ISSN
- 0018-8646
- URL
- https://doi.org/10.1147/rd.322.0238
要旨
This paper presents the design of a broad-coverage Japanese sentence analyzer which can be part of various Japanese processing systems. The sentence analyzer comprises two components: the lexical analyzer and the syntactic analyzer. Lexical analysis, i.e., segmenting a sentence into words, is a formidable problem for a language like Japanese, because it has no explicit delimiters (blanks) between written words. In practical applications, this task is made more difficult by the occurrence of words not listed in a dictionary. We have developed a five-layered knowledge source and used it successfully in the lexical analyzer, resulting in very accurate segmentation, even in cases where there are unknown words. The syntactic analyzer has two modules: One consists of an augmented context-free grammar and the PLNLP parser; the other is the dependency structure constructor, which converts the phrase structures to dependency structures. The dependency structures represent various key linguistic relations in a more direct way. The dependency structures have semantically important information such as tense, aspect, and modality, as well as preference scores reflecting relative ranking of parse acceptability.
主題
この書誌の出所
- openalex— W2074736143(2026-08-14取得)
引用
Newton Maruyama・Masayuki Morohashi・Shigeki UMEDA・Eiichro Sumita(1988-03-01) A Japanese sentence analyzer 『IBM Journal of Research and Development』 32(2) pp. 238-250 IBM