論文 ·日本語 ·未確認
Sequential Linefeed Insertion into Lecture Transcriptions for Real‐Time Captioning
Tomohiro Ohno ・ Masaki Murata ・ Shigeki Matsubara
- 刊行年
- 2015-01-13
- 収録
- 『Electronics and Communications in Japan』 98(2) pp. 20-31
- 出版
- Wiley
- 言語
- 英語
- OpenAlex
- W1978495056
- DOI
- 10.1002/ecj.11616
- MAG
- 1978495056
- ISSN
- 0424-8368
- URL
- https://doi.org/10.1002/ecj.11616
要旨
SUMMARY To generate readable captions for Japanese spoken monologues such as lectures in real time, it is necessary to sequentially display captions that have proper linefeeds inserted. This paper proposes a technique for sequentially inserting proper linefeeds into a lecture transcript whenever a bunsetsu, which is a linguistic unit shorter than a sentence in Japanese and that roughly corresponds to a basic phrase in English, is identified. Under the assumption that linefeeds are inserted at bunsetsu boundaries, this technique can reduce the delay time of captioning to the utmost possible. This technique statistically judges whether or not a linefeed should be inserted into each bunsetsu boundary by using the information that is available at the time. We conducted experiments on linefeed insertion using a Japanese lecture corpus. The experimental results confirmed that our method, which is a bunsetsu‐based linefeed insertion method, was almost as accurate as the sentence‐based linefeed insertion method. In addition, we conducted comparative evaluations using four baseline methods. The results confirmed that our method could insert linefeeds more accurately than the simple methods that are thought to have the same delay time as our method.
主題
この書誌の出所
- openalex— W1978495056(2026-08-14取得)
引用
Tomohiro Ohno・Masaki Murata・Shigeki Matsubara(2015-01-13) Sequential Linefeed Insertion into Lecture Transcriptions for Real‐Time Captioning 『Electronics and Communications in Japan』 98(2) pp. 20-31 Wiley