本文へ移動

論文 ·日本語 ·未確認

Cross-Lingual Image Caption Generation

Takashi Miyazaki Nobuyuki Shimizu

刊行年
2016-01-01
言語
英語
OpenAlex
W2509490957
DOI
10.18653/v1/p16-1168
MAG
2509490957
URL
https://www.aclweb.org/anthology/P16-1168.pdf

要旨

Automatically generating a natural language description of an image is a fundamental problem in artificial intelligence. This task involves both computer vision and natural language processing and is called "image caption generation." Research on image caption generation has typically focused on taking in an image and generating a caption in English as existing image caption corpora are mostly in English. The lack of corpora in languages other than English is an issue, especially for morphologically rich languages such as Japanese. There is thus a need for corpora sufficiently large for image captioning in other languages. We have developed a Japanese version of the MS COCO caption dataset and a generative model based on a deep recurrent architecture that takes in an image and uses this Japanese version of the dataset to generate a caption in Japanese. As the Japanese portion of the corpus is small, our model was designed to transfer the knowledge representation obtained from the English portion into the Japanese portion. Experiments showed that the resulting bilingual comparable corpus has better performance than a monolingual corpus, indicating that image understanding using a resource-rich language benefits a resource-poor language.

主題

この書誌の出所

  • openalex— W2509490957(2026-08-14取得)

引用

Takashi Miyazaki・Nobuyuki Shimizu(2016-01-01) Cross-Lingual Image Caption Generation pp. 1780-1790

MiyazakiShimizu2016CrossLingualImage
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON