本文へ移動

論文 ·日本語 ·未確認

Design and Construction of Japanese Multimodal Utterance Corpus with Improved Emotion Balance and Naturalness

Daisuke Horii Akinori Ito Takashi Nose

刊行年
2022-11-07
収録
『2022 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC)』 pp. 245-250
言語
英語
OpenAlex
W4312120891
DOI
10.23919/apsipaasc55919.2022.9980272
URL
https://doi.org/10.23919/apsipaasc55919.2022.9980272

要旨

This paper describes the development of a corpus of multimodal emotional behaviors. So far, many databases of multimodal affective behaviors have been developed. These databases are divided into spontaneous and acted behavior databases. Acted behavior databases can easily collect words with a balanced number of emotions; however, it has been pointed out that acted speech differs from spontaneous speech. In this work, we aim to collect acted multimodal emotional utterances that sound as natural as possible. To this end, we first collected scenes from tweets in which emotional balance was considered. Then, we performed an initial corpus collection, demonstrating that we could collect various emotional utterances. Next, we collected the corpus using a crowdsourcing platform. Then, we evaluated the naturalness of the collected speech by comparing it with the naturalness of the read speech database (JTES) and the spontaneous speech database (SMOC). As a result, the collected corpus was more natural than JTES, which indicates that the recording program effectively collected naturally-sounding emotional behavior corpus.

主題

この書誌の出所

  • openalex— W4312120891(2026-08-14取得)

引用

Daisuke Horii・Akinori Ito・Takashi Nose(2022-11-07) Design and Construction of Japanese Multimodal Utterance Corpus with Improved Emotion Balance and Naturalness 『2022 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC)』 pp. 245-250

Horii2022DesignConstructionJapanese
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON