本文へ移動

論文 ·日本語 ·未確認

Speech analysis of sung-speech and lyric recognition in monophonic singing

Dairoku Kawai Kazumasa Yamamoto Seiichi Nakagawa

刊行年
2016-03-01
言語
英語
OpenAlex
W2405834540
DOI
10.1109/icassp.2016.7471679
MAG
2405834540
URL
https://doi.org/10.1109/icassp.2016.7471679

要旨

Lyric recognition in singing is challenging because of a number of problems, including a lack of singing databases, superposed musical instruments and different spectral variations. First of all, we investigated the difference of spectral variations among read speech, spontaneous speech and sung speech and we found that sung speech recognition was the most difficult. Next, we consider Japanese lyric recognition in monophonic singing that contains no musical instruments. To express singing well, we use an n-gram language model with a lyrics corpus, singing-adapted acoustic models, and plural pronunciation lexicons for vowel-lengthening. We also compare GMM-HMM and DNN-HMM acoustic models. We obtained a remarkable improvement on lyric recognition in comparison with the baseline system for spontaneous speech recognition.

主題

この書誌の出所

  • openalex— W2405834540(2026-08-14取得)

引用

Dairoku Kawai・Kazumasa Yamamoto・Seiichi Nakagawa(2016-03-01) Speech analysis of sung-speech and lyric recognition in monophonic singing 105 pp. 271-275

Kawai2016SpeechAnalysisSung
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON