本文へ移動

論文 ·日本語 ·未確認

ASPEC: Asian Scientific Paper Excerpt Corpus

Toshiaki Nakazawa Manabu Yaguchi Kiyotaka Uchimoto Masao Utiyama Eiichiro Sumita Sadao Kurohashi Hitoshi Isahara

刊行年
2016-05-01
言語
英語
OpenAlex
W2576482813
DOI
10.63317/3iku22jbpbzm
MAG
2576482813
URL
http://www.lrec-conf.org/proceedings/lrec2016/pdf/621_Paper.pdf

要旨

In this paper, we describe the details of the ASPEC (Asian Scientific Paper Excerpt Corpus), which is the first large-size parallel corpus of scientific paper domain.ASPEC was constructed in the Japanese-Chinese machine translation project conducted between 2006 and 2010 using the Special Coordination Funds for Promoting Science and Technology.It consists of a Japanese-English scientific paper abstract corpus of approximately 3 million parallel sentences (ASPEC-JE) and a Chinese-Japanese scientific paper excerpt corpus of approximately 0.68 million parallel sentences (ASPEC-JC).ASPEC is used as the official dataset for the machine translation evaluation workshop WAT (Workshop on Asian Translation).

主題

この書誌の出所

  • openalex— W2576482813(2026-08-14取得)

引用

Toshiaki Nakazawa・Manabu Yaguchi・Kiyotaka Uchimoto・Masao Utiyama・Eiichiro Sumita・Sadao Kurohashi・Hitoshi Isahara(2016-05-01) ASPEC: Asian Scientific Paper Excerpt Corpus pp. 2204-2208

Nakazawa2016ASPEC
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON