本文へ移動

論文 ·対照・比較 ·未確認

Japanese-English Code-Switching Speech Data Construction

Sahoko Nakayama Takatomo Kano Quoc Truong Sakriani Sakti Satoshi Nakamura

刊行年
2018-05-01
言語
英語
OpenAlex
W2936842814
DOI
10.1109/icsda.2018.8693044
MAG
2936842814
URL
https://doi.org/10.1109/icsda.2018.8693044

要旨

As the number of Japanese-English bilingual speakers continues to increase, code-switching phenomena also happen more frequently. The units and locations of switches may vary widely from single word switches to whole phrases (beyond the length of the loanword units). Therefore, speech recognition systems must be developed that can handle not only Japanese or English but also Japanese-English code-switching. Consequently, a large-scale code-switching speech database is required for model training. But collecting natural conversation dialogues of Japanese-English data is both time-consuming and expensive. This paper presents the construction of Japanese-English code-switching speech data by utilizing a Japanese and English text-to-speech system from a bilingual speaker. Various switching units are also investigated including units of words and phrases. As a result, we successfully constructed over 280-k speech utterances of Japanese-English code-switching.

主題

この書誌の出所

  • openalex— W2936842814(2026-08-14取得)

引用

Sahoko Nakayama・Takatomo Kano・Quoc Truong・Sakriani Sakti・Satoshi Nakamura(2018-05-01) Japanese-English Code-Switching Speech Data Construction pp. 67-71

Nakayama2018JapaneseEnglishCode
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON