論文 ·対照・比較 ·未確認
Japanese-English Code-Switching Speech Data Construction
Sahoko Nakayama ・ Takatomo Kano ・ Quoc Truong ・ Sakriani Sakti ・ Satoshi Nakamura
- 刊行年
- 2018-05-01
- 言語
- 英語
- OpenAlex
- W2936842814
- DOI
- 10.1109/icsda.2018.8693044
- MAG
- 2936842814
- URL
- https://doi.org/10.1109/icsda.2018.8693044
要旨
As the number of Japanese-English bilingual speakers continues to increase, code-switching phenomena also happen more frequently. The units and locations of switches may vary widely from single word switches to whole phrases (beyond the length of the loanword units). Therefore, speech recognition systems must be developed that can handle not only Japanese or English but also Japanese-English code-switching. Consequently, a large-scale code-switching speech database is required for model training. But collecting natural conversation dialogues of Japanese-English data is both time-consuming and expensive. This paper presents the construction of Japanese-English code-switching speech data by utilizing a Japanese and English text-to-speech system from a bilingual speaker. Various switching units are also investigated including units of words and phrases. As a result, we successfully constructed over 280-k speech utterances of Japanese-English code-switching.
主題
この書誌の出所
- openalex— W2936842814(2026-08-14取得)
引用
Sahoko Nakayama・Takatomo Kano・Quoc Truong・Sakriani Sakti・Satoshi Nakamura(2018-05-01) Japanese-English Code-Switching Speech Data Construction pp. 67-71