本文へ移動

論文 ·日本語 ·未確認

Sequential Linefeed Insertion into Lecture Transcriptions for Real‐Time Captioning

Tomohiro Ohno Masaki Murata Shigeki Matsubara

刊行年
2015-01-13
収録
『Electronics and Communications in Japan』 98(2) pp. 20-31
出版
Wiley
言語
英語
OpenAlex
W1978495056
DOI
10.1002/ecj.11616
MAG
1978495056
ISSN
0424-8368
URL
https://doi.org/10.1002/ecj.11616

要旨

SUMMARY To generate readable captions for Japanese spoken monologues such as lectures in real time, it is necessary to sequentially display captions that have proper linefeeds inserted. This paper proposes a technique for sequentially inserting proper linefeeds into a lecture transcript whenever a bunsetsu, which is a linguistic unit shorter than a sentence in Japanese and that roughly corresponds to a basic phrase in English, is identified. Under the assumption that linefeeds are inserted at bunsetsu boundaries, this technique can reduce the delay time of captioning to the utmost possible. This technique statistically judges whether or not a linefeed should be inserted into each bunsetsu boundary by using the information that is available at the time. We conducted experiments on linefeed insertion using a Japanese lecture corpus. The experimental results confirmed that our method, which is a bunsetsu‐based linefeed insertion method, was almost as accurate as the sentence‐based linefeed insertion method. In addition, we conducted comparative evaluations using four baseline methods. The results confirmed that our method could insert linefeeds more accurately than the simple methods that are thought to have the same delay time as our method.

主題

この書誌の出所

  • openalex— W1978495056(2026-08-14取得)

引用

Tomohiro Ohno・Masaki Murata・Shigeki Matsubara(2015-01-13) Sequential Linefeed Insertion into Lecture Transcriptions for Real‐Time Captioning 『Electronics and Communications in Japan』 98(2) pp. 20-31 Wiley

Ohno2015SequentialLinefeedInsertion
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON