Skip to content

Detail of publication

Citation

Pražák, A. and Psutka Josef V. and Hoidekr, J. and Kanis, J. and Müller, L. and Psutka, J. : Automatic online subtitling of the Czech parliament meetings . Lecture Notes in Artificial Intelligence, Lecture notes in artificial intelligence. 0302-9743 ; 4188, 4188, p. 501-508, Springer, Berlin, 2006.

Abstract

This paper describes a LVCSR system for automatic online subtitling (closed captioning) of TV transmissions of the Czech Parliament meetings. The recognition system is based on Hidden Markov Models, lexical trees and bigram language model. The acoustic model is trained on 40 hours of parliament speech and the language model on more than 10M tokens of parliament speech trancriptions. The first part of the article is focused on text normalization and class-based language model preparation. The second part describes the recognition network and its decoding with respect to real-time operation demands using up to 100k vocabulary. The third part outlines the application framework allowing generation and displaying of subtitles for any audio/video source. Finally, experimental results obtained on parliament speeches with recognition accuracy varying from 80 to 95 % (according to the discussed topic) are reported and discussed.

Detail of publication

Title: Automatic online subtitling of the Czech parliament meetings
Author: Pražák, A. ; Psutka Josef V. ; Hoidekr, J. ; Kanis, J. ; Müller, L. ; Psutka, J.
Language: English
Date of publication: 11 Sep 2006
Year: 2006
Type of publication: Papers in journals
Title of journal or book: Lecture Notes in Artificial Intelligence
Edition: Lecture notes in artificial intelligence. 0302-9743 ; 4188
Series: 4188
Page: 501 - 508
ISBN: 0302-9743
Publisher: Springer
Address: Berlin
Date: 11 Sep 2006 - 15 Sep 2006
/ 2011-06-09 12:57:21 /

Keywords

ASR, online, subtitling, parliament, czech

BibTeX

@ARTICLE{PrazakA_2006_Automaticonline,
 author = {Pra\v{z}\'{a}k, A. and Psutka Josef V. and Hoidekr, J. and Kanis, J. and M\"{u}ller, L. and Psutka, J.},
 title = {Automatic online subtitling of the Czech parliament meetings},
 year = {2006},
 publisher = {Springer},
 journal = {Lecture Notes in Artificial Intelligence},
 address = {Berlin},
 pages = {501-508},
 series = {4188},
 ISBN = {0302-9743},
 url = {http://www.kky.zcu.cz/en/publications/PrazakA_2006_Automaticonline},
}