本文へ移動

論文 ·日本語 ·未確認

Development of a Low-Latency and Real-Time Automatic Speech Recognition System

Chee Siang Leow Tomoaki Hayakawa Hiromitsu Nishizaki Norihide Kitaoka

刊行年
2020-10-13
言語
英語
OpenAlex
W3115127521
DOI
10.1109/gcce50665.2020.9291818
MAG
3115127521
URL
https://doi.org/10.1109/gcce50665.2020.9291818

要旨

In this study, a real-time automatic speech recognition (ASR) system based on the Kaldi ASR toolkit, with low-latency and customizable models, without any internet connection, was developed. The proposed ASR system includes a voice activity detection (VAD) module and an audio transmitter as a front-end speech processing and a decoder for the received audio signals. The ASR system was evaluated in terms of ASR accuracy and speech processing speed. Consequently, the ASR system achieved high ASR accuracy on the CSJ (Corpus of Spontaneous Japanese) test set with super low-latency.

主題

この書誌の出所

  • openalex— W3115127521(2026-08-14取得)

引用

Chee Siang Leow・Tomoaki Hayakawa・Hiromitsu Nishizaki・Norihide Kitaoka(2020-10-13) Development of a Low-Latency and Real-Time Automatic Speech Recognition System pp. 925-928

Leow2020DevelopmentLowLatency
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON