論文 ·用例に日本語 ·未確認
Language-specific phonetic realisation of stop voicing contrasts in English and Japanese synthesised speech
James T. Tanner ・ Yasuaki Shinohara ・ Faith Chiu
- 刊行年
- 2025-08-01
- 収録
- 『JASA Express Letters』 5(8)
- 出版
- Acoustical Society of America
- 言語
- 英語
- OpenAlex
- W4413744657
- DOI
- 10.1121/10.0039066
- PubMed
- 40862682
- ISSN
- 2691-1191
- URL
- https://doi.org/10.1121/10.0039066
要旨
Speech synthesis has improved dramatically over recent years, enabled by large datasets and advances in neural network architectures. Little is known, however, about how synthesised speech patterns are realized from a phonetic perspective. By synthesising speech in two languages with differing implementations of stop voicing, we observe that synthesised speech broadly follows expected patterns for each language, though partially diverges for specific segments. Synthesising speakers into the opposing language also results in stops similar to target language distributions. These findings demonstrate the capability of speech synthesis models to encode phonetic information and further motivate questions regarding the phonetics of synthesised speech.
主題
この書誌の出所
- openalex— W4413744657(2026-08-14取得)
引用
James T. Tanner・Yasuaki Shinohara・Faith Chiu(2025-08-01) Language-specific phonetic realisation of stop voicing contrasts in English and Japanese synthesised speech 『JASA Express Letters』 5(8) Acoustical Society of America